Skip to contentSkip to stories

Updated

#Model release

Showing low-relevance items too. Hide low-relevance items

Oct 1

Oct 1Thu
  1. Microsoft AIOfficialAI score15

    Microsoft AI announces three new models now available today

    AIMicrosoft AI says three models are available today, with more details provided in a linked post. The post does not name the models or give specifications, benchmarks, or pricing.

  2. Cloudflare Blog · AIOfficialAI score58

    Cloudflare releases open-source Clef decision models and an RL fine-tuning service

    AICloudflare released Clef and Clef-flash, two decision models hosted on Workers AI and open-sourced on Hugging Face under Apache 2.0, and launched a reinforcement learning fine-tuning service. In Cloudflare's tests, Clef classified a domain in 2.2s versus 4.7s for gpt-oss-120b, and the models are Jev-API compatible. The company is offering fine-tuning first through a forward-deployed engineering team, with a self-serve platform planned later.

  3. Don't Worry About the Vase (Zvi Mowshowitz)BlogAI score62

    AI #188: Gemini 4 Argon, GPT-6.1 Sol, and Anthropic's IPO Filing

    AIGoogle says Gemini 4 Argon is rolling out at $2/$10 per million tokens, though the author has not yet been able to access the model to test it. OpenAI pulled GPT-6.1 Astra over alignment failures and released GPT-6.1 Sol, which it prices at the same $2/$10 and says shows substantial alignment improvements over GPT-6 Sol. The post also covers Anthropic's leaked IPO prospectus, which reportedly lists roughly $518 billion in compute commitments, and a court ruling upholding the Department of War's supply chain risk designation of Anthropic.

  4. OpenRouter · New modelsBlogAI score36

    Pareto 26.10 Preview: A Multimodal Model for Research, Coding and Agents

    AIPareto 26.10 Preview is a multimodal composite model built for research, coding, and agentic workflows. It is described as delivering frontier-level performance across a broad range of general-purpose tasks, though the source excerpt is a preview and provides no benchmark scores, parameter counts, pricing, or availability details.

  5. AI SupremacyBlogAI score50

    Google announces Gemini 4 Argon, its first frontier model since February

    AIGoogle announced Gemini 4 Argon, a model it says is built to sustain deep reasoning across complex, long-horizon workflows, roughly seven months after its last flagship release in February. The article says cybersecurity testing will be completed after October 1, with no benchmark scores, pricing, or availability details provided.

Sep 30

Sep 30Wed
  1. indigoXAI score81

    Google's Gemini 4 Argon debuts with limited access pending US government approval

    AIGoogle has announced Gemini 4 Argon, initially available only to trusted cyber defenders through its Fairwind Program while US government approval is pending. The author says the model is aimed at long-running software engineering, enterprise knowledge work, and cybersecurity tasks, with a 1 million token output limit. The post also gives promotional pricing of $2 per million input tokens and $10 per million output tokens, rising to $4 and $20 afterward, alongside a benchmark comparison.

    Why it matters: The post places Gemini 4 Argon's benchmark table beside GPT-6 Astra and Claude models, showing where each leads across coding, knowledge work, and cybersecurity tasks.

    Image from @indigox's post
  2. Nathan LambertXAI score26

    Nathan Lambert welcomes Google's Gemini 4 as frontier competition

    AINathan Lambert says he is pleased to see Google surprise people with Gemini 4. He argues that more labs at the frontier benefit consumers through competition and reduce the concentration of power. He says he is eager to see how the model performs in real-world scenarios.

  3. Dongxi NLPXAI score18

    Gemini 4 Argon release highlighted among September frontier model launches

    AIDongxi NLP jokes that frontier AI labs now release new models every few days, calling it their pastime rather than a boast. The post lists a September release calendar including Grok 4.7, GPT-6 Sol and Luna, Claude Opus 5.5, Claude Sonnet 5.5, GPT-6.1 Sol, and Gemini 4 Argon, which it says launched today.

  4. Yi TayXAI score46

    Gemini 4 Argon launches, reportedly outperforming astra and fable on many tasks

    AIGoogle DeepMind introduced Gemini 4 Argon, a new frontier model built for coding, enterprise knowledge work, and cybersecurity defense, rolling out to trusted testers through its Fairwind Program. Yi Tay says Gemini 4 outperforms astra and fable on many tasks, though the post gives no benchmark figures.

  5. Google · Innovation & AIOfficialAI score46

    Google AI Flu Model Ranks First in CDC FluSight Hospitalization Forecasts

    AIA flu forecasting model built with Google AI ranked first among 39 eligible models in the CDC's FluSight 2025-26 season evaluation for predicting U.S. flu-related hospital admissions. The model was developed using Empirical Research Assistance (ERA), an AI tool that generates optimization algorithms, and ERA's underlying technology is now available to trusted testers.

  6. Dongxi NLPXAI score9

    Gemini 4 model named Argon, following periodic table naming

    AIA post notes that Gemini 4's new model is called Argon, with a sequence of Gemini names based on periodic table elements: Sulfur, Chlorine, Argon, Potassium, and Calcium. The author remarks that naming models after periodic table elements is quite simple.

  7. whXAI score67

    Gemini 4 Argon previewed with frontier coding and cyber defense claims

    AIThe post quotes Google's Sundar Pichai introducing Gemini 4 Argon as an early look at the next model. It claims frontier performance in complex workflows, cyber defense, and software engineering, and says Google teams are using it for tasks from coding to quantum computing. The author adds that on FrontierSWE the model is very self-critical and often says "Eureka!", a personality they describe as a large improvement over previous Gemini models.

    Image from @nrehiew_'s post
  8. Josh WoodwardXAI score8

    Gemini 4 Argon announced by Google's Josh Woodward

    AIJosh Woodward, who leads Google's Gemini app, announces "Gemini 4 Argon" in a brief post without giving further details. The post offers no specifications, benchmarks, pricing, or availability dates.

    Image from @joshwoodward's post
  9. koray kavukcuogluXAI score62

    Google's Koray Kavukcuoglu Announces Gemini 4 Argon for Trusted Defenders First

    AIGoogle is sharing Gemini 4 Argon first with trusted defenders in its Fairwind Program, with frontier capabilities in coding, knowledge work, and cyber security defense. The model is also being rolled out to the US government, and the company plans broader availability as testing progresses and safeguards allow.

    Why it matters: The author is a Google DeepMind leader announcing the model directly, so the rollout limits to trusted defenders and government are the key detail to note.

  10. Varun MohanXAI score40

    Google announces Gemini 4 Argon, a new frontier model for software tasks

    AIGoogle announced Gemini 4 Argon, a new frontier model that delivers frontier performance across complex software tasks, according to Varun Mohan. Thousands of Googlers have been using it internally in Antigravity, and it is rolling out first to trusted cyber defenders in the Fairwind Program, with broader availability to follow as soon as possible.

  11. Logan KilpatrickXAI score45

    Google's Logan Kilpatrick celebrates Gemini 4 Argon model launch

    AILogan Kilpatrick, Google's Gemini lead, praised the teams behind Gemini 4 Argon and said he looks forward to wider adoption. The post links to Google's blog announcement for the model, but it gives no benchmark scores, pricing, or capability details.

  12. Google AIOfficialAI score8

    Google points readers to a Gemini 4 Argon announcement post

    AIGoogle AI's post links to a Google blog article about Gemini 4 Argon under its Gemini models and research section. The post itself contains only a "Learn more" call to action and the URL, so no features, benchmarks, or availability details can be confirmed from it.

  13. Google AIOfficialAI score72

    Google announces Gemini 4 Argon, a frontier model with 1M output tokens

    AIGoogle AI announced Gemini 4 Argon, a new frontier model built for deep reasoning across long, complex workflows in software engineering, legal and finance knowledge work, and cybersecurity defense. Google says it is expanding the model's output token limit to 1M tokens. Argon is rolling out first to trusted cyber defenders in the Fairwind Program, with broader availability to follow as soon as possible.

    Why it matters: The benchmark table compares Gemini 4 Argon against GPT-6 Astra and Claude models across knowledge work, coding, and multimodal tasks, showing where it leads and trails.

    Image from @GoogleAI's post
  14. Google DeepMindOfficialAI score88

    Google DeepMind releases Gemini 4 Argon to trusted cyber defenders first

    AIGoogle DeepMind announced Gemini 4 Argon, rolling out first to trusted cyber defenders through its Fairwind Program. Argon will launch at an introductory price of $2 per million input tokens and $10 per million output tokens, with output limits raised to 1M tokens. The post cites a 77.9% score on DeepSWE v1.1 and 91.7% on LVBench, and says broad availability will follow safeguard testing.

    Why it matters: The post pairs Argon's benchmark claims with the phased release, pricing, and safeguard details, helping readers weigh its frontier-level capabilities against its access limits.

  15. Google · Gemini appOfficialAI score91

    Google announces Gemini 4 Argon, rolling out first to trusted cyber defenders

    AIGoogle announced Gemini 4 Argon, a new frontier model rolling out first to trusted cyber defenders through its Fairwind Program. The model's output limit rises to 1M tokens from 64K, and its introductory API price is $2 per million input tokens and $10 per million output tokens. Google says broader availability to developers, enterprises, and consumers will follow after more testing of guardrails.

    Why it matters: The post pairs benchmark claims with a phased access plan, pricing, and safety measures, which helps readers judge how quickly Argon may reach developers.

  16. PerplexityOfficialAI score20

    Perplexity's embedding preview tops ConTEB benchmark average nDCG@10

    AIPerplexity's embedding preview achieves the highest average nDCG@10 among tested models on ConTEB, though not on every task. It also outperforms voyage-context-4 on chunk retrieval while using 8x less storage per vector, at 1 KB (1024 dims, int8) versus 8 KB (2048 dims, float32).

    Image from @perplexity_ai's post
  17. MiniMax (official)OfficialAI score44

    HeyGen Video launches on MiniMax H3 at $0.01 per second

    AIHeyGen has released HeyGen Video, a production-quality video product built on MiniMax H3 and post-trained by HeyGen. Pricing starts at $0.01 per second through October, a 50% discount, aimed at businesses that need video without production-level costs.

  18. FireworksOfficialAI score34

    GLM 5.3 Flash now available for training on Fireworks' Serverless API

    AIFireworks AI has made GLM 5.3 Flash available for training through its Serverless Training API, open to all users. The model supports both vision and text inputs. Fireworks says it performs well on its benchmarks for agentic coding, document analysis, and tool use while remaining cost-efficient to serve.

  19. MiniMax (official)OfficialAI score38

    Creatify's Boreal-H3 ad video model built on MiniMax H3

    AICreatify Labs released Boreal-H3, a video model built on MiniMax H3 and post-trained specifically for advertising. Reported results include 85.3% reference fidelity, brief success rising from 28% to 50%, and identity match improving from 83% to 94%. Visible defects per clip dropped 70%, while generation time and estimated cost fell 20%.

  20. FireworksOfficialAI score30

    Fireworks' Ember-1 matches Kimi K3 on Vals with fewer reasoning tokens

    AIFireworks' Ember-1 performs close to Kimi K3 on Vals' finance and legal benchmarks while using fewer reasoning tokens per turn. Fewer tokens per turn lower cost and speed up agent loops, and Ember-1 is available on Fireworks Serverless.

  21. IdeogramOfficialAI score23

    Ideogram 4.5 Performs Strongly Across General Image Editing Tasks

    AIIdeogram 4.5 was built for precise, targeted editing but also performs very well across general editing tasks. Design Arena ranks it 15th in Image Editing with an Elo of 1250, placing it in the same performance band as MAI-Image-2.6 and Gemini 3 Pro Image Preview. It is especially strong at typography edits, such as modifying text in infographics.

  22. Ant LingOfficialAI score38

    Ling-3.1-flash ports C image library to Rust with 8.015× speedup

    AIAnt Ling reports that its Ling-3.1-flash model completed a roughly 20-hour Rust port of a C image library. After a performance regression caused by busy-waiting workers and a parallelism adjustment, the model recovered and reached an 8.015× speedup. All 30 correctness checks passed.

    Image from @AntLingAGI's post
  23. Ant LingOfficialAI score46

    Ant Ling releases Ling-3.1-flash with 1M-token context, plans open-source

    AIAnt Ling introduced Ling-3.1-flash, a model with about 560B total parameters, about 25B active per token, and up to a 1M-token context window. The company plans to open-source the model soon. It reports 1,673 Elo on GDPVal-AA v2.1, 75.16 on FrontierSWE, and 65.35 on HealthBench Professional across work, coding, and healthcare tasks.

    Image from @AntLingAGI's post