Skip to contentSkip to stories

Updated

#Model release

Oct 8

Oct 8Thu
  1. LeiphoneAI score58

    Alibaba's Qwen Roadmap Targets 5T to 10T Parameters Amid Self-Improving Model Work

    AIAt the Apsara Conference, Alibaba's Qwen team outlined a roadmap of Qwen4 followed by Qwen4.5 and Qwen5, aiming for 5T to 10T parameters. The article notes that Qwen3.8 reached 2.4T parameters and that Qwen3.8-Flash activates 6B parameters per inference while cutting training cost to one-ninth. It also describes Qwen3.8-Max running model-driven experiments in chip design and inference optimization, and multimodal updates including a video model slated for November.

  2. QbitAIAI score80

    GPT-6 rolls out to free ChatGPT users with interactive answer interfaces

    AIOpenAI began rolling out GPT-6 to free and Go ChatGPT users on October 8, replacing GPT-5.6 Luna with GPT-6 Luna, while paid users receive GPT-6 Sol. The update adds Intelligent UI, which generates charts, buttons, and interactive tools inside chat answers. OpenAI's safety report shows gains on jailbreak and instruction-hierarchy tests but also regressions in some self-harm, sexual, and emotional-dependence evaluations, including for under-18 users.

  3. Testing CatalogAI score36

    Gemini Agent for Business may add Claude Opus 5 and Sonnet 5.5

    AIGoogle's recently announced Gemini Agent for Gemini Business is reportedly set to offer Gemini Argon 4, Gemini Flash 3.8, Claude Opus 5, and Claude Sonnet 5.5. If accurate, it would mark the first time Claude models appear on Google's platform alongside Google's own models, which the post frames as a way for Google to compete for enterprise customers.

  4. Testing CatalogAI score58

    Daily AI brief covers Mistral, Google, OpenAI, Anthropic, Microsoft and other vendor news for October 8

    AIThis daily brief from TestingCatalog collects recent AI announcements from several vendors, including Mistral Large 4, Claude Haiku 5.5, and GitHub stacked pull requests becoming generally available. The author states the brief was composed with Grok and cherry-picked news with post-editing. Items are mostly short product and pricing notices rather than detailed reporting, and several claims are unverified announcements.

Oct 7

Oct 7Wed
  1. GeekParkAI score46

    ChatGPT Adds Intelligent UI That Generates Interactive Tools, Google Launches Playground

    AIOpenAI said on October 7 that ChatGPT's new Intelligent UI will automatically combine text, charts, buttons and forms into interactive interfaces such as calculators and mini-games, rolling out to Plus, Pro, Business and Enterprise users from October 7 and to Free and Go users from October 8. Google also launched Playground, an experimental platform where users create, modify and play browser games from natural-language descriptions, initially for U.S. users aged 18 and older.

  2. KhazixAI score60

    Claude Max subscribers get monthly API credits usable across Claude models

    AISubscribers to Claude's Max plan can claim monthly API credits: $100 for the $100 tier and $200 for the $200 tier. The credits work for any Claude model and can be used in the user's own apps and other agents. The author argues that bundling monthly API credits alongside a broad model lineup will make it hard for other model companies to compete.

  3. Semafor · TechnologyAI score56

    Reflection AI and Mistral launch open models to challenge China's lead

    AIReflection AI and Mistral each unveiled new open-source models this week, aiming to beat other Western open models, though they trail top Chinese and closed systems on prominent benchmarks. Reflection CEO Misha Laskin says the target is regulated industries and governments that cannot or will not use Chinese models. The outcome depends on whether businesses and agencies accept less advanced models for some tasks in exchange for lower cost and more control.

  4. Testing CatalogAI score47

    Daily AI brief covers Mistral Large 4, Google, OpenAI, and Anthropic updates

    AIMistral released Mistral Large 4 "Le Chonk", a 1T-parameter (49B active) multimodal model, with open weights planned in about three weeks. Google rolled out Nano Banana 2.1 across Gemini, AI Studio, and the Gemini API, and released EmbeddingGemma 2, a 740M-parameter open multimodal embedding model under Apache 2.0. OpenAI launched the Decisions API in beta with gpt-6-luna, returning typed answers 10x faster than the Responses API.

  5. Latent SpaceAI score72

    OpenAI publishes 722 math manuscripts from an unreleased internal model

    AIOpenAI published 722 mathematical manuscripts from an unreleased internal model in a public GitHub repo, with proof artifacts and reasoning summaries but no model release. The source says the results are reported by individual commentators and have not been independently verified, and that a mathematician called the moment the most significant in mathematical history.

Oct 6

Oct 6Tue
  1. Julien ChaumondAI score70

    Mistral Large 4 announced with open weights due end of October

    AIJulien Chaumond reposted Mistral's announcement of Mistral Large 4, a 1T-parameter natively multimodal model with 49B active parameters. Mistral says it is available via API today, with open weights scheduled for release at the end of October, and is working privately with cybersecurity partners.

    Why it matters: The post lays out Mistral Large 4's scale, multimodal design, and availability timeline, which helps readers gauge the open-weights landscape outside China.

  2. Guillaume LampleAI score26

    Mistral's ML4 trained on 3,800 NVIDIA Grace Blackwell GPUs in Europe

    AIMistral says its ML4 model was trained on 3,800 NVIDIA Grace Blackwell GPUs in its European datacenters, including its Bruyères-le-Châtel cluster built with Series B funding. The company is investing further, with Series C and D clusters coming online soon to support longer training, more ambitious post-training, and faster iteration. Mistral expects large and rapid improvements in the weeks and months ahead.

Oct 5

Oct 5Mon
  1. GeekParkAI score38

    OpenAI Launches 28-Day Codex and ChatGPT Work Improvement Plan, Adds Visual Ads in ChatGPT

    AIOpenAI says it will ship one meaningful Codex and Work improvement each day for 28 days starting October 5, or else offer a "reset" without specifying what that reset covers. The company also plans to test visual ads in ChatGPT image generation in the U.S. starting in late October, with ads kept separate from generated images and not affecting answers.

  2. KrASIA · Big TechAI score68

    US and China AI release cycles shorten as AI takes on more R&D work

    AINikkei found the average gap between upgraded high-performance model releases among five US and four Chinese developers fell from 125 days (January 2023 to March 2026) to 44 days (April to September 2026). Anthropic said its Claude AI led 26% of its R&D efforts as of August and was involved in more than 90% of R&D activities, while OpenAI reported AI agents working more hours than human researchers in August.

  3. Alex HeathAI score52

    Reflection's founders discuss building a DeepSeek of the West with Beam

    AIReflection is set to release Beam, its first open-weight AI model, aiming to become a Western counterpart to DeepSeek. The source says Beam is trained from scratch for coding, reasoning, and AI agents, with benchmarks placing it alongside the strongest open models and more efficient token economics. Reflection has raised $4.6 billion from investors including Nvidia, Sequoia, and Lightspeed, and the interview covers its monetization plans for open-weight models.

Oct 4

Oct 4Sun
  1. Tibor BlahoAI score62

    OpenAI and Anthropic weekly roundup covers DevDay, Sonnet 5.5, and FTC probe

    AIOpenAI announced over 20 updates at DevDay 2026, including always-on agents on GPT-6 Astra and GPT-6.1 Sol, which arrived in the API and at a fifth of Astra's price. Anthropic launched Claude Sonnet 5.5, priced the same as Sonnet 5 but over 30 percent faster. Reuters reported an FTC probe into Anthropic, OpenAI and other labs over rogue AI agents.

  2. Tibor BlahoAI score37

    OpenAI and Anthropic weekly: DevDay 2026, Sonnet 5.5, FTC probe

    AIOpenAI announced more than 20 DevDay 2026 updates, including always-on agents on GPT-6 Astra, GPT-6.1 Sol priced at a fifth of Astra's API price, and an Ultrafast tier with up to 8x faster token generation in Codex. Anthropic launched Claude Sonnet 5.5 at $2/$10 per million tokens, 30%+ faster than Sonnet 5 and with thinking always on. Reuters reported the FTC is probing OpenAI, Anthropic and other labs over rogue AI agents.

Oct 3

Oct 3Sat
  1. Latent SpaceAI score52

    Latent Space daily roundup covers GPT-6.1 Sol, Sonnet 5.5, agent harnesses, and eval integrity debates

    AIThis Latent Space AINews roundup compiles a weekend's AI news from Twitter and Reddit rather than a single announcement. It covers OpenAI's GPT-6.1 Sol pricing and Agent Arena placement, Anthropic's Sonnet 5.5 debut, Meta's open-sourced Muse hardware firmware, and several research and benchmark items, many reported with unverified claims.

Oct 1

Oct 1Thu
  1. Don't Worry About the Vase (Zvi Mowshowitz)AI score62

    AI #188: Gemini 4 Argon, GPT-6.1 Sol, and Anthropic's IPO Filing

    AIGoogle says Gemini 4 Argon is rolling out at $2/$10 per million tokens, though the author has not yet been able to access the model to test it. OpenAI pulled GPT-6.1 Astra over alignment failures and released GPT-6.1 Sol, which it prices at the same $2/$10 and says shows substantial alignment improvements over GPT-6 Sol. The post also covers Anthropic's leaked IPO prospectus, which reportedly lists roughly $518 billion in compute commitments, and a court ruling upholding the Department of War's supply chain risk designation of Anthropic.

Sep 30

Sep 30Wed
  1. Google · Innovation & AIAI score46

    Google AI Flu Model Ranks First in CDC FluSight Hospitalization Forecasts

    AIA flu forecasting model built with Google AI ranked first among 39 eligible models in the CDC's FluSight 2025-26 season evaluation for predicting U.S. flu-related hospital admissions. The model was developed using Empirical Research Assistance (ERA), an AI tool that generates optimization algorithms, and ERA's underlying technology is now available to trusted testers.

Sep 29

Sep 29Tue
  1. Tibor BlahoAI score78

    OpenAI's DevDay 2026 brings dots agents, GPT-6.1 Sol, and Ultrafast speed tier

    AIOpenAI announced more than 20 updates at DevDay 2026, including dots always-on agents, GPT-6.1 Sol, Ultrafast token generation, ChatGPT Space, and a $500/month Pro 500 plan. GPT-6.1 Sol is priced at $2 input and $10 output per 1M tokens and is available in the API as gpt-6.1-sol. Ultrafast generates tokens up to 8x faster in Codex and up to 6x faster in the API.

    Why it matters: The post lists dozens of OpenAI DevDay 2026 changes across models, agents, plans, and APIs, useful for scanning what shipped and who gets access.

Sep 27

Sep 27Sun
  1. Tibor BlahoAI score85

    OpenAI releases GPT-6 Sol and Luna as Anthropic launches Claude Opus 5.5

    AIOpenAI released GPT-6 Sol and Luna, priced 50 percent below GPT-5.6 promo API pricing, and rolling out in ChatGPT Work, Codex and the API, not yet in regular Chat. Anthropic released Claude Opus 5.5, described as roughly Claude Fable 5.1 level for 40 percent less than Opus 5 and over 30 percent faster, with Sonnet 5.5 and Haiku 5.5 due in coming weeks.

    Why it matters: The recap puts OpenAI and Anthropic releases side by side, with pricing and capability claims that help compare the two launches.

  2. Tibor BlahoAI score71

    OpenAI and Anthropic ship GPT-6 Sol and Luna and Claude Opus 5.5 in the same week

    AIOpenAI released GPT-6 Sol and Luna at API prices 50% below GPT-5.6 promotional pricing, and Anthropic released Claude Opus 5.5 the same day at 40% less than Opus 5. The roundup also covers Claude Code cloud sessions reaching general availability, the Claude Marketplace launch, OpenAI's new misalignment disclosures after the Hugging Face incident, and DevDay on September 29. The post is a relayed weekly digest, and it includes the author's closing promotion for AIPRM, which is not part of the reported news.

  3. The SequenceAI score55

    Opus 5.5 cuts costs while Meta and US–China talks widen AI's reach

    AIAnthropic released Claude Opus 5.5 at about 40% lower cost than Opus 5, priced at $4/$20 per MTok input/output. Meta said Muse is coming to its AI glasses in the coming months, while Washington and Beijing held their first AI dialogue and discussed an incident-notification channel. The newsletter argues that costs, interfaces, experiments, and diplomacy increasingly determine how much value AI creates.

Sep 25

Sep 25Fri
  1. Anthropic ResearchAI score67

    Claude computes a nine-loop physics amplitude that experts had not reached

    AIAnthropic researchers used Claude Science to compute the nine-loop six-particle amplitude in planar N=4 super Yang-Mills, a toy-model result that physicist Lance Dixon checked. The work reportedly cost roughly one or two thousand dollars, with about $100 of compute for the bootstrap calculation, and a similar result was reached by Song He's group.

    Why it matters: The guest post shows a frontier physics calculation done with modest compute, which helps readers gauge what current AI can handle in research and what it still cannot.