"Apple Unveils MM1: The Future of Multimodal AI"

TL;DR Summary
Apple researchers have quietly revealed MM1, a set of multimodal large language models designed for captioning images, answering visual questions, and natural language inference. The MM1 family supports up to 30 billion parameters and achieves competitive performance after supervised fine-tuning on established multimodal benchmarks. Apple's work on MM1 indicates a significant advance in AI, and the company is rumored to be working on an LLM framework code-named "Ajax" as part of a $1 billion AI R&D push. Apple is expected to share more details on its AI efforts at the upcoming WWDC developer show in June.
- Apple Quietly Reveals MM1, a Multimodal LLM Thurrott.com
- Apple researchers achieve breakthroughs in multimodal AI as company ramps up investments VentureBeat
- New Apple AI training method retains privacy, and could make a future Siri more flexible AppleInsider
- Apple Announces MM1: A Family of Multimodal LLMs Up To 30B Parameters that are SoTA in Pre-Training Metrics and Perform Competitively after Fine-Tuning MarkTechPost
- Apple's AI ups its game as researchers make big ‘breakthrough’ The Times of India
Reading Insights
Total Reads
0
Unique Readers
20
Time Saved
2 min
vs 3 min read
Condensed
79%
448 → 96 words
Want the full story? Read the original article
Read on Thurrott.com