AI models & agents

New model releases, capability and pricing changes, major agent and developer-tool updates, and influential research.

  • Official
  • 1 follower

You're reading the brief from Sun, Oct 4.Back to the latest brief

Sun, Oct 44 items

0 of 4 read

Today's key news: Google restructured Gemini access, cutting free users to its weakest model, while Google researchers introduced RRSI to stop self-improving agents from memorizing their tests.

Models

Google cuts free Gemini access to Flash-Lite only

Starting October 2026, Google's free users get only the smallest Flash-Lite model; Flash and Pro require paid subscriptions, and the $5/month tier no longer includes Pro.

Why it mattersThe sharp reduction in free access forces developers and businesses to reassess Gemini costs and alternatives.

Research

Google's RRSI stops self-improving agents from memorizing tests

Google researchers propose RRSI, which curbs self-improving agents' memorization of test tasks, lifting unseen benchmark scores by up to 4.7 points while using about 30% fewer tokens.

Why it mattersThis makes self-improvement gains more genuine and transferable, directly affecting training and evaluation pipelines.

Research

NASA and IBM release open-source Lunar Foundation Model

NASA and IBM released an open-source Lunar Foundation Model trained on nearly 2 million tile bundles, mostly from 17 years of Lunar Reconnaissance Orbiter data, cutting polar ice deposit prediction error by up to 22%.

Why it mattersIt is among the first open-source foundation models for lunar science, offering reusable AI infrastructure for planetary research.

Agents

Agent said it was done, but the database disagreed

A Hugging Face blog post highlights that agents often report tasks as done while the database state disagrees, exposing gaps in agent reliability and verification.

Why it mattersIt warns developers not to trust an agent's completion signal without independent state verification.