Microsoft AI announces three new models now available today
AIMicrosoft AI says three models are available today, with more details provided in a linked post. The post does not name the models or give specifications, benchmarks, or pricing.
Updated
Updated
Showing low-relevance items too. Hide low-relevance items
AIMicrosoft AI says three models are available today, with more details provided in a linked post. The post does not name the models or give specifications, benchmarks, or pricing.
AIMicrosoft AI announced three new models: MAI-Transcribe-2-Streaming, MAI-Voice-2.1, and MAI-Voice-2.1-Flash. The source says the streaming transcription model is accurate and aims for natural speech with less waiting between conversational turns for building voice agents.

AICloudflare released Clef and Clef-flash, two decision models hosted on Workers AI and open-sourced on Hugging Face under Apache 2.0, and launched a reinforcement learning fine-tuning service. In Cloudflare's tests, Clef classified a domain in 2.2s versus 4.7s for gpt-oss-120b, and the models are Jev-API compatible. The company is offering fine-tuning first through a forward-deployed engineering team, with a self-serve platform planned later.
AIGoogle says Gemini 4 Argon is rolling out at $2/$10 per million tokens, though the author has not yet been able to access the model to test it. OpenAI pulled GPT-6.1 Astra over alignment failures and released GPT-6.1 Sol, which it prices at the same $2/$10 and says shows substantial alignment improvements over GPT-6 Sol. The post also covers Anthropic's leaked IPO prospectus, which reportedly lists roughly $518 billion in compute commitments, and a court ruling upholding the Department of War's supply chain risk designation of Anthropic.
AIPareto 26.10 Preview is a multimodal composite model built for research, coding, and agentic workflows. It is described as delivering frontier-level performance across a broad range of general-purpose tasks, though the source excerpt is a preview and provides no benchmark scores, parameter counts, pricing, or availability details.
AIGoogle announced Gemini 4 Argon, a model it says is built to sustain deep reasoning across complex, long-horizon workflows, roughly seven months after its last flagship release in February. The article says cybersecurity testing will be completed after October 1, with no benchmark scores, pricing, or availability details provided.
AIGoogle has announced Gemini 4 Argon, initially available only to trusted cyber defenders through its Fairwind Program while US government approval is pending. The author says the model is aimed at long-running software engineering, enterprise knowledge work, and cybersecurity tasks, with a 1 million token output limit. The post also gives promotional pricing of $2 per million input tokens and $10 per million output tokens, rising to $4 and $20 afterward, alongside a benchmark comparison.
Why it matters: The post places Gemini 4 Argon's benchmark table beside GPT-6 Astra and Claude models, showing where each leads across coding, knowledge work, and cybersecurity tasks.

AINathan Lambert says he is pleased to see Google surprise people with Gemini 4. He argues that more labs at the frontier benefit consumers through competition and reduce the concentration of power. He says he is eager to see how the model performs in real-world scenarios.
AIDongxi NLP jokes that frontier AI labs now release new models every few days, calling it their pastime rather than a boast. The post lists a September release calendar including Grok 4.7, GPT-6 Sol and Luna, Claude Opus 5.5, Claude Sonnet 5.5, GPT-6.1 Sol, and Gemini 4 Argon, which it says launched today.
AIGoogle DeepMind introduced Gemini 4 Argon, a new frontier model built for coding, enterprise knowledge work, and cybersecurity defense, rolling out to trusted testers through its Fairwind Program. Yi Tay says Gemini 4 outperforms astra and fable on many tasks, though the post gives no benchmark figures.
AIRunway's Head of Robotics Andy Chen outlined the company's approach to robotics and premiered Praxis-1, an open-weight world action model. Early access requests are available via the Praxis-1 page on Runway's website.
AIDemis Hassabis, Google DeepMind's CEO, expressed pride in the team's relentless work on a new release, pointing readers to a blog post for details. The post itself provides no specific model names, figures, or features.
AIAnt Ling's Ling 3.1 Flash is now available on Vercel AI Gateway, free through October 13. The model is built for coding and multi-step analysis.
AIA flu forecasting model built with Google AI ranked first among 39 eligible models in the CDC's FluSight 2025-26 season evaluation for predicting U.S. flu-related hospital admissions. The model was developed using Empirical Research Assistance (ERA), an AI tool that generates optimization algorithms, and ERA's underlying technology is now available to trusted testers.
AIA post notes that Gemini 4's new model is called Argon, with a sequence of Gemini names based on periodic table elements: Sulfur, Chlorine, Argon, Potassium, and Calcium. The author remarks that naming models after periodic table elements is quite simple.
AIThe post quotes Google's Sundar Pichai introducing Gemini 4 Argon as an early look at the next model. It claims frontier performance in complex workflows, cyber defense, and software engineering, and says Google teams are using it for tasks from coding to quantum computing. The author adds that on FrontierSWE the model is very self-critical and often says "Eureka!", a personality they describe as a large improvement over previous Gemini models.

AIJosh Woodward, who leads Google's Gemini app, announces "Gemini 4 Argon" in a brief post without giving further details. The post offers no specifications, benchmarks, pricing, or availability dates.

AIGoogle is sharing Gemini 4 Argon first with trusted defenders in its Fairwind Program, with frontier capabilities in coding, knowledge work, and cyber security defense. The model is also being rolled out to the US government, and the company plans broader availability as testing progresses and safeguards allow.
Why it matters: The author is a Google DeepMind leader announcing the model directly, so the rollout limits to trusted defenders and government are the key detail to note.
AIGoogle announced Gemini 4 Argon, a new frontier model that delivers frontier performance across complex software tasks, according to Varun Mohan. Thousands of Googlers have been using it internally in Antigravity, and it is rolling out first to trusted cyber defenders in the Fairwind Program, with broader availability to follow as soon as possible.
AILogan Kilpatrick, Google's Gemini lead, praised the teams behind Gemini 4 Argon and said he looks forward to wider adoption. The post links to Google's blog announcement for the model, but it gives no benchmark scores, pricing, or capability details.
AIGoogle AI's post links to a Google blog article about Gemini 4 Argon under its Gemini models and research section. The post itself contains only a "Learn more" call to action and the URL, so no features, benchmarks, or availability details can be confirmed from it.
AIGoogle AI announced Gemini 4 Argon, a new frontier model built for deep reasoning across long, complex workflows in software engineering, legal and finance knowledge work, and cybersecurity defense. Google says it is expanding the model's output token limit to 1M tokens. Argon is rolling out first to trusted cyber defenders in the Fairwind Program, with broader availability to follow as soon as possible.
Why it matters: The benchmark table compares Gemini 4 Argon against GPT-6 Astra and Claude models across knowledge work, coding, and multimodal tasks, showing where it leads and trails.

AIGoogle DeepMind's Argon supports a 1M token output limit, enabling deeper reasoning on long, multi-step problems in a single pass. Early testers are providing feedback before a broader rollout to developers, enterprises, and consumers.

AIGoogle is rolling out its Argon model with frontier safeguards to the US government and a set of trusted cyber defenders through its Fairwind Program. Sundar Pichai says broader availability will follow as soon as it can be done safely.
AIGoogle DeepMind announced Gemini 4 Argon, rolling out first to trusted cyber defenders through its Fairwind Program. Argon will launch at an introductory price of $2 per million input tokens and $10 per million output tokens, with output limits raised to 1M tokens. The post cites a 77.9% score on DeepSWE v1.1 and 91.7% on LVBench, and says broad availability will follow safeguard testing.
Why it matters: The post pairs Argon's benchmark claims with the phased release, pricing, and safeguard details, helping readers weigh its frontier-level capabilities against its access limits.
AIGoogle announced Gemini 4 Argon, a new frontier model rolling out first to trusted cyber defenders through its Fairwind Program. The model's output limit rises to 1M tokens from 64K, and its introductory API price is $2 per million input tokens and $10 per million output tokens. Google says broader availability to developers, enterprises, and consumers will follow after more testing of guardrails.
Why it matters: The post pairs benchmark claims with a phased access plan, pricing, and safety measures, which helps readers judge how quickly Argon may reach developers.
AIReplicate now hosts Ideogram 4.5, an image model built for hyper-precise, multi-turn editing. It is aimed at editorial mockups, brand assets, and imagery without artifacts.

AIPerplexity has published the pplx-embed-v2-context-9b-preview model on Hugging Face. The post links to the model page but gives no further details on its capabilities, benchmarks, or pricing.
AIPerplexity's embedding preview achieves the highest average nDCG@10 among tested models on ConTEB, though not on every task. It also outperforms voyage-context-4 on chunk retrieval while using 8x less storage per vector, at 1 KB (1024 dims, int8) versus 8 KB (2048 dims, float32).

AIPerplexity says its pplx-embed-v2-context-9b-preview model leads on Answer and Evidence retrieval at every cutoff. In document retrieval the gap is smaller, with its context v1 4B model slightly ahead at Document@3 and Document@5.

AIPerplexity introduced a new training method for contextual embedding models that encode each document chunk with the whole document in view. Its pplx-embed-v2-context-9b-preview sets a new state of the art on the ConTEB benchmark and on turbopuffer's privately held context-bench.

AIHeyGen has released HeyGen Video, a production-quality video product built on MiniMax H3 and post-trained by HeyGen. Pricing starts at $0.01 per second through October, a 50% discount, aimed at businesses that need video without production-level costs.
AIFireworks AI has made GLM 5.3 Flash available for training through its Serverless Training API, open to all users. The model supports both vision and text inputs. Fireworks says it performs well on its benchmarks for agentic coding, document analysis, and tool use while remaining cost-efficient to serve.
AICreatify Labs released Boreal-H3, a video model built on MiniMax H3 and post-trained specifically for advertising. Reported results include 85.3% reference fidelity, brief success rising from 28% to 50%, and identity match improving from 83% to 94%. Visible defects per clip dropped 70%, while generation time and estimated cost fell 20%.
AIFireworks' Ember-1 performs close to Kimi K3 on Vals' finance and legal benchmarks while using fewer reasoning tokens per turn. Fewer tokens per turn lower cost and speed up agent loops, and Ember-1 is available on Fireworks Serverless.
AIIdeogram 4.5 was built for precise, targeted editing but also performs very well across general editing tasks. Design Arena ranks it 15th in Image Editing with an Elo of 1250, placing it in the same performance band as MAI-Image-2.6 and Gemini 3 Pro Image Preview. It is especially strong at typography edits, such as modifying text in infographics.
AIAnt Ling asked developers what they will build with Ling-3.1-flash and invited them to share demos and feedback in its official Discord server. The post gives no specifications, benchmarks, pricing, or release details.
AIAnt Ling reports that its Ling-3.1-flash model completed a roughly 20-hour Rust port of a C image library. After a performance regression caused by busy-waiting workers and a parallelism adjustment, the model recovered and reached an 8.015× speedup. All 30 correctness checks passed.

AILing-3.1-flash built a Lua-to-x86-64 ELF compiler from scratch in about 17 hours, passing 178 of 182 tests for 97.8%. That is within 1.2 percentage points of the 99.0% shown for Claude Fable 5 and above GLM-5.3-Flash at 95.9%.

AIAnt Ling introduced Ling-3.1-flash, a model with about 560B total parameters, about 25B active per token, and up to a 1M-token context window. The company plans to open-source the model soon. It reports 1,673 Elo on GDPVal-AA v2.1, 75.16 on FrontierSWE, and 65.35 on HealthBench Professional across work, coding, and healthcare tasks.
