OpenAI's Releases GPT-6 Astra
The new LLM nearly saturates math and abstract reasoning tests, leads agentic science benchmarks, and arrives with a much higher risk profile.
The latest news on AI and machine learning. We cover generative models, predictive analytics, industry leaders, and the ethical issues they raise.
The new LLM nearly saturates math and abstract reasoning tests, leads agentic science benchmarks, and arrives with a much higher risk profile.
Google’s third Flash release in 43 days adds stronger coding, longer agent loops, and a restricted cybersecurity model for trusted defenders.
Meta says its latest agent model sustains longer workflows while using fewer tools and tokens, with more cautious user collaboration.
Alibaba’s 2.4-trillion-parameter MoE adds stronger agent coordination, multimodal handling, and cache rates as low as $0.17 per million tokens.
Gemini’s new processing mode selectively inspects video timelines, cutting costs while improving answers on long and detail-heavy footage.
The Gemini-powered app separates objects and text for targeted edits, then brings collaborative image work into Docs, Slides, and Drive.
Anthropic’s new Fable and Mythos models pair stronger long-horizon performance with cheaper cache reads and tighter controls for frontier research.
The 330-million-parameter foundation model forecasts related series and known future events together, producing the full horizon in one pass.
MIT researchers used over 1,000 hours of smartwatch conversations to test whether LLMs can anticipate a person’s next communicative move.