🤖 Models AI News
New AI model releases and updates. From GPT and Claude to open-source models, we score and categorize the top new AI model news.
- OpenAI rolls out GPT-6 Astra to top-tier ChatGPT plans at half the rate of GPT-5.6 Sol
OpenAI has rolled out GPT-6 Astra to Pro, Enterprise, and Business Premium users, with Plus users expected to follow soon. Message allowances for the standard model are roughly half of what GPT-5.6 Sol offers: Plus users get an estimated 5 to 45 messages per five hours with Astra, versus 10 to 100 with Sol. Higher-tier subscribers also get GPT-6 Astra Pro with separate weekly caps. Free and Go users don't get access. The article OpenAI rolls out GPT-6 Astra to top-tier ChatGPT plans at half the rate of GPT-5.6 Sol appeared first on The Decoder .
Score: 88🤖 ModelsSep 5, 2026https://the-decoder.com/openai-rolls-out-gpt-6-astra-to-top-tier-chatgpt-plans-at-half-the-rate-of-gpt-5-6-sol/ - Google ships Gemini 3.8 Flash with stronger agentic coding at $0.75 per million tokens
Google shipped Gemini 3.8 Flash on September 2, 2026, pricing the model at $0.75 per million input tokens and $3.75 per million output tokens, the same introductory rate the company used for Gemini 3.7 Flash just three weeks earlier. The launch, announced just three weeks after the previous Flash release, marks Google’s third Flash release ... Read more
- Why is there so much worry about OpenAI Astra, and what issues could ‘recurrent depth’ reasoning cause? The experts weigh in
As OpenAI unveils GPT-6 Astra, cybersecurity experts question whether the model's 'recurrent depth' reasoning was properly tested.
- Astra appears to think without showing its work, and the people arguing about it co-wrote the warning
AI safety researchers say OpenAI’s Astra appears to do less of its reasoning in visible text, and OpenAI’s chief scientist has warned against a race into unmonitorability. He co-authored a 2025 position paper asking developers to evaluate and report exactly this, which the EU’s code of practice turns into a filing to the AI Office. […] This story continues at The Next Web
- ICYMI: the 7 biggest tech stories of the week, from the best gadgets of IFA 2026 to GPT-6 Astra and 'the AGI era'
We're looking back over the last week to pick out the most important stories published on TechRadar.
- IIT-Madras’ Bodhan AI launches four Indian language models
Positioned as ‘Digital Public Goods,’ the models are released with open-weights
- World Labs unveils Atlas, a single AI model that generates, reconstructs, and simulates 3D worlds from just a few photos
World Labs, co-founded by AI researcher Fei-Fei Li, has announced Atlas, a world model that generates, reconstructs, and simulates 3D scenes from just a few images. The company claims it beats specialized models by anchoring all inputs in 3D space rather than processing them as flat sequences. Atlas can also generate robot training data entirely in simulation. The article World Labs unveils Atlas, a single AI model that generates, reconstructs, and simulates 3D worlds from just a few photos appeared first on The Decoder .
- Researchers fear safety disaster ahead of OpenAI’s Astra release
A report triggered concerns about a safety ‘race to the bottom.’
Score: 82🤖 ModelsSep 2, 2026https://www.theverge.com/ai-artificial-intelligence/988334/openai-astra-ai-monitoring-safety - Frontier AI at a cost: what Anthropic’s Fable 5.1 means for the US-China model race
Anthropic’s powerful new Claude Fable 5.1 model has widened its lead in performance benchmarks over Chinese rivals, even as budget-friendly open-weight models from China continue to gain commercial traction globally. The American company said on Wednesday that Fable 5.1, along with a restricted-access version called Mythos 5.1, was the world’s most advanced model for software coding and complex knowledge work. The standard version scored 66 on Artificial Analysis’ Intelligence Index, taking the...
- With Gemini 3.8 Flash, Google reminds everyone it's still in the race
AI model scores well, runs fast, and doesn't cost too much (yet)
- Google says its new Gemini 3.8 Flash model ‘works harder’ but might cost more
Gemini 3.8 Flash Cyber is also launching to Google’s new Fairwind Program.
Score: 74🤖 ModelsSep 2, 2026https://www.theverge.com/ai-artificial-intelligence/988742/google-gemini-3-8-flash - Meta Unveils New AI Model in Accelerated Effort to Catch Rivals
Meta Unveils New AI Model in Accelerated Effort to Catch Rivals theinformation.com
Score: 68🤖 ModelsSep 2, 2026https://www.theinformation.com/briefings/meta-unveils-new-ai-model-accelerated-effort-catch-rivals - Introducing Gemini 3.8 Flash and 3.8 Flash Cyber
a hero image reading "Gemini 3.8 Flash and 3.8 Flash Cyber"
Score: 67🤖 ModelsSep 2, 2026https://blog.google/innovation-and-ai/models-and-research/gemini-models/3-8-flash-and-3-8-flash-cyber/ - Claude Fable 5.1 and Mythos 5.1 arrive with better coding skills and cheaper cache pricing
Anthropic launched Claude Fable 5.1 and Mythos 5.1, offering stronger coding and research performance alongside a 75% cut to cache read pricing for developers.
- iFLYTEK Open-Sources Million-Token Context Models for On-Device AI
iFLYTEK's wholly-owned subsidiary launched and open-sourced Spark X2.5-4B and X2.5-1.7B, which it says are the first edge models to natively support up to one million tokens of context.
Score: 67🤖 ModelsSep 2, 2026https://pandaily.com/iflytek-spark-x2.5-million-token-edge-models-sep2026 - OpenAI and Anthropic are launching new, more powerful updates amid panic over ‘rogue’ AI systems
ChatGPT maker has suggested that upcoming update might be too powerful to control
Score: 67🤖 ModelsSep 2, 2026https://www.independent.co.uk/tech/openai-anthropic-claude-fable-astra-rogue-b3043605.html - Nvidia and CrowdStrike Develop New Cybersecurity AI Models
Plus, models from Anthropic, Google and World Labs
- Alibaba upgrades Qwen3.8-Max with a new 0902 snapshot
Alibaba has released Qwen3.8-Max-0902, an upgraded snapshot of its Qwen3.8-Max foundation model. The update was post-trained for coding and Cowork-style tasks and is available through Alibaba’s Qwen services and API channels. Alibaba said the model’s front-end CodeArena score rose by 22 points to 1,691, placing it first on the leaderboard. The model retains a context […]
Score: 63🤖 ModelsSep 2, 2026https://technode.com/2026/09/02/alibaba-upgrades-qwen38-max-with-new-0902-snapshot/ - Google releases Gemini 3.8 Flash, its third Flash model in six weeks
Google's Pro model updates are seemingly paused, but there's yet another Gemini Flash today.
Score: 62🤖 ModelsSep 2, 2026https://arstechnica.com/ai/2026/09/google-releases-gemini-3-8-flash-its-third-flash-model-in-six-weeks/ - Muse Spark 1.3: Meta reaches the frontier
Meta unveils Muse Spark 1.3, a new LLM pushing the limits of language understanding.
- Anthropic Joins AI Price War With Release of Fable 5.1
The vendor reduced cache read pricing for Fable 5.1 and said the model is similar to Mythos 5.1, with different safeguards.
Score: 56🤖 ModelsSep 2, 2026https://aibusiness.com/generative-ai/anthropic-joins-ai-price-war-release-of-fable-5-1 - Zuck's Muse to Spark joy with open weights release 'soon'
While you wait, Meta says it’s taught the model to stop wasting tokens and ask for help a bit more often
Score: 56🤖 ModelsSep 2, 2026https://www.theregister.com/ai-and-ml/2026/09/02/zucks-muse-to-spark-joy-with-open-weights-release-soon/5294093 - Gemini 3.8 Flash could land any day now, and it could put the vibe back into vibe coding
It could finally close the gap with Claude's Fable, GPT-5.6 Sol, and DeepSeek V4.
Score: 54🤖 ModelsSep 2, 2026https://www.androidauthority.com/google-gemini-3-8-flash-vibe-coding-3705946/ - Tencent’s Hy4 model gains in open-source AI rankings after ecosystem-driven training
Tencent Holdings’ use of its vast product ecosystem to train its new Hy4 preview model gives it an edge in developing AI agents and brings its flagship model suite back into the top tier of open-source offerings, according to analysts. The Chinese tech giant’s “differentiated product-plus-model strategy”, where preview models were first deployed across Tencent’s suite of products, enabled it to collect user data before feeding the information back into subsequent rounds of training, Goldman...
- Fable 5.1 on Frontier Coding Tasks: Efficient Successes, Distinct Failure Modes
We evaluated Fable 5.1 on a series of frontier coding tasks from our proprietary Terminal-Bench+ dataset and compared the results against Opus 5. Fable remained competitive across most categories and was materially more efficient on successful runs, while its gap was concentrated in a small set of terminal-heavy and build/dependency tasks. Because category sizes are small and uneven, we treat... The post Fable 5.1 on Frontier Coding Tasks: Efficient Successes, Distinct Failure Modes appeared first on Snorkel AI .
- Claude Mythos only model to complete full cyber kill chain, experts say
Cyber Weapon Index finds AI attacks 'imminent'
- Gemini 3.8 Flash now available on AI Gateway
Gemini 3.8 Flash from Google is now available on AI Gateway. The model is 50% off through December 31st. It has a 1M token context window, accepts text, image, PDF, and video input, returns text, and supports tool calling and web search. Maximum output is 65,536 tokens. Gemini 3.8 Flash improves on prior Flash models at software engineering, agent work, and multi-step reasoning, at the same speed and cost as the previous release. Thinking is on by default. To use Gemini 3.8 Flash, set model to google/gemini-3.8-flash : To use it in a coding agent, see the coding agents guide , then run vercel ai-gateway coding-agents setup to connect agents like Claude Code, OpenCode, Cursor, Pi, and more and select google/gemini-3.8-flash inside the agent. Try Gemini 3.8 Flash in the model playground . AI Gateway reflects provider pricing with no markup and does not charge a platform fee on inference, including on Bring Your Own Key (BYOK) requests. You can view all language models available on AI Gateway. Read more
Score: 41🤖 ModelsSep 2, 2026https://vercel.com/changelog/gemini-3-8-flash-now-available-on-ai-gateway - 😺 Claude Fable 5.1 can do the work. The hard part is managing it.
We stress-tested Anthropic’s new model live: browser games, computer use, 3D workflows, runaway initiative, and the new problem of managing AI judgment.
- Muse Spark 1.3 now available on AI Gateway
Muse Spark 1.3 from Meta is now available on AI Gateway, in both the standard and contributor pricing tiers. This model improves on prior Muse Spark models at agent work and coding, with a 1M token context window and text, image, and PDF input. On coding it takes fewer turns and writes less filler than the previous release. To use Muse Spark 1.3, set model to meta/muse-spark-1.3 : Contributor tier Muse Spark 1.3 Contributor is a pricing tier on the same model rather than a separate one, with the same weights, capabilities, and context window. The difference is that Meta uses the inputs and outputs sent to this tier to train and improve its models, and pricing is lower in exchange. Model Input Output Cached input meta/muse-spark-1.3 $1.25 $4.25 $0.15 meta/muse-spark-1.3-contributor $0.10 $0.20 $0.002 Rates are per million tokens and unchanged from Muse Spark 1.2 on both tiers. To use it in a coding agent, see the coding agents guide , then run vercel ai-gateway coding-agents setup to connect Claude Code, Codex, Cursor, and more, then select meta/muse-spark-1.3 in the agent. Try Muse Spark 1.3 in the model playground . You can view all language models available on AI Gateway. Read more
- The Sequence Learning Loop - Issue 925: Learn About Fable and Mythos 5.1, GLM-5.3-Flash, and Qwen 3.8
Three releases, three different bets. Let’s dive in.
- OpenAI Teases Astra, Says It Is the First Model to Meet ‘Critical Cybersecurity Threshold'
OpenAI published a blog post on Tuesday detailing its new model Astra. The company says that Astra meets the Critical cybersecurity capability threshold under our Preparedness Framework. It is claimed to be capable of spotting previously unknown security flaws and developing ways to exploit them across many well-protected systems without human intervention. Astra will...
- OpenAI Astra arrives soon, and the company is already promoting its critical risks
OpenAI confirmed that its unreleased Astra model has reached a dangerous new milestone, even as it preps the model for public release.
- OpenAI to launch new model with 'stronger safeguards' after hack
ChatGPT maker OpenAI said Tuesday it was preparing to release its newest powerful model, known as Astra, after implementing "stronger safeguards" following a rogue cyberattack involving a different AI model.
- Gemini 3.8 Flash is Google's third budget model in six weeks while frontier models remain MIA
Google's Gemini 3.8 Flash, the third Flash model in six weeks, matches Claude Opus 5 on some agentic coding benchmarks at lower cost. But its "working harder" reasoning burns about 30 percent more output tokens per task, making it pricier in practice than its predecessor despite identical token rates. The article Gemini 3.8 Flash is Google's third budget model in six weeks while frontier models remain MIA appeared first on The Decoder .
- Fei-Fei Li’s World Labs Unveils New World Model
Fei-Fei Li’s World Labs Unveils New World Model theinformation.com
- Fei-Fei Li’s World Labs debuts Atlas, a world model showcase for advanced spatial intelligence
World Labs Inc., the high-profile and well-funded artificial intelligence startup co-founded by the renowned computer vision pioneer Fei-Fei Li, has just dropped Atlas, which promises to be a game-changer in the world of “world models.” In a blog post, World Labs explained that Atlas is a breakthrough multimodal world model that aims to bridge the […] The post Fei-Fei Li’s World Labs debuts Atlas, a world model showcase for advanced spatial intelligence appeared first on SiliconANGLE .
- Meta Superintelligence Labs Releases Muse Voice Transcribe: One Real-Time Model for Streaming ASR, Diarization, and Endpointing
Meta Superintelligence Labs Releases Muse Voice Transcribe: One Real-Time Model for Streaming ASR, Diarization, and Endpointing MarkTechPost
- Google DeepMind Releases Gemini 3.8 Flash and Gemini 3.8 Flash Cyber: One Core Model, Two Access Envelopes
Google DeepMind Releases Gemini 3.8 Flash and Gemini 3.8 Flash Cyber: One Core Model, Two Access Envelopes MarkTechPost
- Introducing Gemini 3.8 Flash and 3.8 Flash Cyber
Introducing Gemini 3.8 Flash and 3.8 Flash Cyber
- Google has released Gemini 3.8 Flash, its fourth Flash model in under four months
Google launches Gemini 3.8 Flash, a fast LLM designed for rapid inference.
- Gemini 3.8 Flash rolling out three weeks after last release
After the last model release three weeks ago, Google today is rolling out Gemini 3.8 Flash. This marks the third Flash update in three months. more…
- 😺Anthropic launched Fable 5.1: and now, the agents cost less
PLUS: Atlas invents camera angles after filming, and hobby shops are booming.
- Anthropic rolls out Fable 5.1 after sandbox escapes
Anthropic releases Fable 5.1 following sandbox escape incidents.
- Anthropic launches Claude Fable 5.1 and Mythos 5.1
Fable 5.1 is now generally available, while Mythos 5.1 is only available through Anthropic’s 'trusted access' programmes. Read more: Anthropic launches Claude Fable 5.1 and Mythos 5.1
- Anthropic unveils Claude Fable 5.1 and Mythos 5.1 for coding and knowledge work
Anthropic has launched Claude Fable 5.1 and Claude Mythos 5.1, state-of-the-art AI models designed for coding tasks. These models achieve better performance with lower computational costs for users. Claude Fable 5.1 is now publicly accessible, while Mythos 5.1 remains available for select trusted programs. Additionally, the introduction of Enterprise Frontier Safeguards enhances data privacy significantly, marking a notable advancement in AI for scientific research and cybersecurity fields.
- Anthropic releases Claude Fable 5.1 and Mythos 5.1
Anthropic has released Claude Fable 5.1 and Mythos 5.1, addressing issues such as performance, data retention, safeguards and price. The updates also target scientific discovery, as the company tested the capabilities of the models across many domains. Internal testing shows that Fable 5.1 outperforms Fable 5 when is comes to coding, knowledge work and problem-solving... … continue reading The post Anthropic releases Claude Fable 5.1 and Mythos 5.1 appeared first on SD Times .
- Google AI Releases Gemini Omni 1.1 Flash: 40-Second Scene Extension, First/Last Frame Control, and 4K Upscaling
Google AI Releases Gemini Omni 1.1 Flash: 40-Second Scene Extension, First/Last Frame Control, and 4K Upscaling MarkTechPost
- China’s Tencent releases new open-source AI model for coding, research tasks
China’s Tencent releases new open-source AI model for coding, research tasks
- Tencent unveils AI model it says outperforms Z.ai, Moonshot
Tencent demonstrated Hy4 Preview for game creation, 3D web design, and accounting compliance.
Score: 61🤖 ModelsAug 29, 2026https://www.techinasia.com/tencent-cloud-expands-ai-agent-suite-indonesia - Gnani AI launches Artha sovereign AI stack with 30-billion-parameter Evon 3.3
Gnani AI launched Artha, a sovereign AI stack for Indian companies and public institutions. This stack features the Evon 3.3 language model and the Plexus agentic platform. Evon 3.3 is an open-weights model trained on Indic languages and domain-specific data. The Artha stack allows organisations to run AI models within their own infrastructure.