The Text Dilemma: AI Struggles to Find Training Data for Chatbots

TL;DR Summary
AI developers, including OpenAI, may be facing a shortage of text to train chatbots and large language models. Stuart Russell, an AI expert and professor at UC Berkeley, warned that the strategy of training AI models with vast amounts of text is "starting to hit a brick wall." Concerns have been raised about the data-collection practices of AI developers, including OpenAI, as they scrape public and private sources for text data. A study suggests that high-quality language data could be depleted by 2026. Lawsuits have also been filed against OpenAI, alleging the use of copyrighted materials and personal data in training their models.
- AI may be 'running out of text in the universe' to train chatbots Business Insider
- Generative AI Goes 'MAD' When Trained on AI-Created Data Over Five Times Tom's Hardware
- ChatGPT can write a paper in an hour — but there are downsides Nature.com
- Generative AI tools are quickly 'running out of text' to train themselves on, UC Berkeley professor warns Business Insider India
- View Full Coverage on Google News
Reading Insights
Total Reads
0
Unique Readers
11
Time Saved
3 min
vs 4 min read
Condensed
85%
686 → 103 words
Want the full story? Read the original article
Read on Business Insider