Liquid AI's d1 model now available on Vercel AI Gateway
AIVercel says Liquid AI's d1 model is live on AI Gateway under the identifier liquid/d1. The model supports vision inputs for classifying, routing, and scoring decisions.
Updated
Updated
Items with an AI score under 20 are hidden. Show low-relevance items
AIVercel says Liquid AI's d1 model is live on AI Gateway under the identifier liquid/d1. The model supports vision inputs for classifying, routing, and scoring decisions.
AIReports say Google's next model, Gemini 4 Argon, has appeared on several platforms and is already open to some Pro users. The source gives no benchmark scores, pricing, context length, or official release date.
AIA short post from the AI Era account says Gemini can call Opus and presents this as the start of a super-agent era. The source gives no version details, benchmarks, pricing, or availability, so this summary reports only the claim.
AIOpenAI has released an Ultrafast mode for GPT-6.1 Sol, according to AI Era. The mode is available simultaneously through the API, Codex and ChatGPT Work.
AIHidden references to a Gemini 4 Argon model with low, medium, and high reasoning efforts have appeared recently in Antigravity, according to testingcatalog. Business Insider reported that Google employees are testing an internal Gemini 4 checkpoint called "Carbon," which performs at the Opus 5.5 level on coding tasks.

AIOpenRouter has released Microsoft-Decision-1, a model Microsoft says posts the highest accuracy across 36 blind benchmarks of about 150K questions. Microsoft says it runs 4.5x faster than the runner-up and 35x faster than GPT-6 Sol, with decisions flipping on only 1.3% of perturbed inputs. The model is post-trained from Qwen3.5-9B, costs $0.042 per million input tokens, has free output and a 32K context window.
AIHiDream-O1-Video-1.0 ranks #6 in Artificial Analysis's Image to Video with Audio leaderboard, just behind Dreamina Seedance 2.0 720p. HiDream says the model generates 1080p videos of 5 to 20 seconds with synchronized audio, priced at $5.80 per minute ($0.10 per second) on the HiHarness API. It is also available in vivago R1 Studio.
AIGoogle is rolling out new versions of its Gemini 4 AI models to staff ahead of a public release. The latest version is called Carbon, and the post says staff are enthusiastic about it.
AITypeSafe AI raised $870 million at a $7.5 billion valuation, led by Andreessen Horowitz with participation from Sequoia and DCVC. The company says Jev, released September 15, is used by a third of Fortune 500 companies, and it is a transformer model that outputs probabilities rather than text. TypeSafe claims Jev runs faster and uses far fewer tokens than LLMs, positioning it for automation tasks.
AITinker adds GLM-5.3-Flash and DeepSeek-v4.1-Flash, both of which natively accept image inputs and use efficient attention architecture. GLM-5.3-Flash costs 4-5 times less on Tinker than GLM-5.3. Long-context options for Qwen3.5-4B and Qwen3.6-35B-A3B are also live.
AIStepFun's Step 5 Preview, a 600B-total, 27B-active MoE model with 1M context and vision, is free to try on Nous Portal for one week. Nous Research says it scored 33.89 on the Hermes Index, the same score as GPT-6 Luna.
AIAlibaba's Qwen team released Qwen-Image-2.1-Turbo, an accelerated checkpoint of Qwen-Image-2.1 that generates and edits images in 8 denoising steps instead of 40. The model keeps the same 7B architecture and offers a hosted API at CNY 0.1 per image, while its weights are under a Qwen Research License that requires separate permission for commercial self-hosting.
AIBusiness Insider reports that Google is internally testing a new Gemini 4 checkpoint named Carbon. The checkpoint reportedly matches Opus 5.5 in coding.

AIQwen has released its Turbo image model with open weights, giving image app builders an option to run it in-house instead of relying on hosted API pricing. The post does not specify model size, benchmarks, or pricing.
AIMicrosoft has made Microsoft-Decision-1 available on Microsoft Foundry, a model post-trained on Qwen3.5-9B for fast, single-pass decision scoring. Microsoft says it achieved the highest accuracy across a 36-benchmark comparison of nearly 150,000 questions, and runs 4.5 times faster than Quyet-1.0-Large and 35 times faster than GPT-6 Sol. Microsoft plans to rebase it on other models, including MAI and OpenAI models.
AICloudflare has published a new model called clef-omni on Hugging Face, according to a post from Julien Chaumond, who owns the account. The post links to the model page but gives no further details about its size, capabilities, or benchmarks.
AIAMD's Data Center team says it worked with Zyphra to train an advanced reasoning model from scratch on AMD hardware. A linked post says Zyphra trains larger reasoning models more efficiently while supporting longer context windows.

AIMicrosoft releases Microsoft-Decision-1, a model for fast decision-making, according to Satya Nadella. Nadella says it outperforms both LLMs and other decision models on structured decision tasks in latency and quality. Microsoft is testing it internally for incident response, quality control, and scientific discovery.

AINous Research says StepFun's Step 5 Preview is free on Nous Portal for the next week. The model is a 600B total, 27B active MoE with a 1M context window and vision support. It scored 33.89 on the Hermes Index, the same score as GPT-6 Luna.
AICloudflare releases Clef-omni, an open-weight decision model that accepts audio, video, image, and text input in a single API call. Clef-flash's price falls from $0.09 to $0.038 per M input tokens, while its hosted context window drops from 64k to 24k. Cloudflare also reports median latency reductions of 1.7 to 2.0 times for the Clef model on Workers AI.
AICline is offering Solar Mini 4 free, a new 35B mixture-of-experts model from Korean lab Upstage with 3B active parameters. It has a 524K context window and runs at 208 tokens per second. Cline says it scores 24 on the AAII, the highest of any model at 3B active and within one point of Nemotron 3 Ultra, which uses 55B active.
AIMicrosoft introduces Microsoft-Decision-1, a new model for fast decision-making that it says outperforms both LLMs and other decision models on structured decision tasks in latency and quality. The company says it is already testing the model across Microsoft for uses including incident response, quality control, and scientific discovery.
AIaccording to a post from Kilo. The model has 600B parameters, with 27B active per token, and supports a 1M-token context and vision.
AIMistral releases Voxtral Mini 4B Realtime Arabic, a streaming speech-to-text model for Arabic dialects and Modern Standard Arabic under the Apache 2.0 License. The model has about 4.4 billion parameters, is fine-tuned from Voxtral-Mini-4B-Realtime-2602, and reaches an average 8.82% Character Error Rate across seven Arabic benchmarks at a 480 ms transcription delay. It can be run with vLLM or Transformers 5.2.0 or later.
AIMistral AI has released LIDstral-Arabic, a fast classifier that identifies Modern Standard Arabic, Arabic dialects, and non-Arabic languages written in Arabic script across 51 classes. On Moroccan Darija, it scores 88.67% F1, versus 72.65% for LahjatBERT ALDi CL and 71.28% for GlotLID v3, across 84,870 evaluation examples. The model runs on CPU and is available from a private Hugging Face repository under Apache 2.0.
AIElvis Saravia says he has tested StepFun's Step 5 Preview as a coding agent since early access and found that it checks its own work and stops when tasks are done. The post says the model is built for engineering tasks such as bug fixing, multi-file features, and refactoring, plus frontend generation and financial report output.

AIMeta spent heavily on Anthropic's Claude models, with internal use reaching up to 60,000 employees and a projected $10 billion yearly spend, according to The Algorithmic Bridge. The author says Meta then released Muse Spark, which scored 52 on the Artificial Analysis intelligence benchmark, on par with Claude Opus 4.6.
AIAmazon Bedrock Managed Agents, powered by OpenAI, entered public preview, and OpenAI's GPT-6 Astra, GPT-6.1 Sol, and GPT-6.1 Luna became generally available on Amazon Bedrock. AWS also released Strands Decider 2B, a 2B-parameter open source decision model that answers in about 115ms locally, and said the Strands harness uses 28 percent fewer tokens than popular harnesses while matching their accuracy.
AIKilo says StepFun has announced Step 5 Preview, which is free to use in Kilo for one week. The post lists 600B total parameters with 27B active per token, a 1M-token context window with vision, and highlights strong coding and finance performance at lower cost.

AIQwen has published Qwen-Image-2.1-Turbo on Hugging Face, an accelerated checkpoint of Qwen-Image-2.1 for text-to-image generation and image editing with 8 denoising steps. The checkpoint uses the same 7B visual generation architecture, loads directly with QwenImage21Pipeline in Diffusers, and includes its recommended sampling schedule. It defaults to CFG=1 and uses prefix KV caching to reuse text and reference-image context across steps.
AIModelScope announces Qwen-Image-2.1-Turbo, an accelerated checkpoint that keeps the 7B visual architecture and runs image generation and editing in 8 denoising steps. The source says it uses CFG=1 and prefix KV caching to reuse text and reference-image context across steps, supports 2048 resolution with square, portrait, landscape, and widescreen presets, and loads through QwenImage21Pipeline in Diffusers. It is released under the Qwen Research License Agreement.
Why it matters: The source names a concrete speedup path, 8 sampling steps and CFG=1 with prefix KV caching, which matters to anyone weighing image generation latency.

AIGoogle AI announces that SynthID.com is now available globally in English for verifying AI-generated images, video, and audio. The post also lists Nano Banana 2.1, EmbeddingGemma 2, Gemma 4 with BOTANIC-1, a Gemini business agent, and Guided Vision in Gemini Live.
AIAlibaba's Qwen team says Qwen-Image-2.1-Turbo is an accelerated checkpoint of Qwen-Image-2.1 on the same 7B architecture, now with open weights. It generates 2K images from text in 8 denoising steps and supports natural-language edits, with Pro and Turbo APIs also live.
AIMeta has released Muse, its A.I. agent app, after delaying it for months over safety concerns. According to the source, new competition pushed Mark Zuckerberg to proceed with the launch.
AIGoogle Cloud has introduced the Gemini agent, a single cloud-hosted agent that handles Q&A, knowledge work, media creation, and coding from one prompt box and one API. It routes jobs across Gemini and Claude models today, with other private and open models planned. Governance covers per-agent identity, role-based access, audit logging, and hard per-project spend caps, but the source gives no reproducible benchmarks, pricing, or general availability date.
AITencent has open-sourced Youtu-Parsing-Omni, a 5B-parameter omni-modal model that outputs a single structured JSON covering layout, text, tables, formulas, ASR, OCR, and video segments. It scores 96.96 Overall on OmniDocBench, the highest among the compared models, and ships with weights on Hugging Face, a vLLM plugin, and inference examples.
AINVIDIA has released a deployment model for Agile One S SSD pickup, based on GR00T N1.7 checkpoint 58000 and using three cameras: ego, left wrist, and right wrist. The repository republishes ONNX graphs, external tensor files, and two existing TensorRT BF16 engines without retraining or re-export, and the original export reported a numerical warning that full FP32, node, and BF16 parity did not pass all tolerances. The files are not a certified robot deployment or safety qualification.
AINVIDIA published the Agile One S Walk GR00T N2 checkpoint 1680, a walking deployment model with four cameras, on Hugging Face. The repository includes ONNX graphs, TensorRT BF16 plans/engines, and the original checkpoint files, republished without retraining or re-export. The shared Cosmos-Reason1-7B dependency and the Isaac/GR00T runtime must be set up separately, and the files are not a robot safety qualification.
AIJetBrains released Mellum2.1, a 12B mixture-of-experts coding model with 2.5B active parameters under Apache 2.0, emphasizing agentic programming. Under high load, its inference throughput in tokens is nearly twice that of Qwen3.5-9B in JetBrains' comparison, and multi-token prediction (MTP) speeds single-request responses by about 1.6x. The model is available on Hugging Face for local or private-infrastructure deployment, with GGUF and vLLM MTP support announced for later.