Skip to contentSkip to stories

Updated

#Model release

Items with an AI score under 20 are hidden. Show low-relevance items

Oct 9

TodayOct 9Fri
  1. IdeogramOfficialAI score45

    Ideogram 4.5 keeps edited images intact across 30 consecutive edits

    AIArtificial Analysis ran 30 consecutive real estate staging edits through four image editing models, and Ideogram 4.5 kept most of each image unchanged while others drifted. Ideogram 4.5 and FLUX 3 left 95% or more of the image untouched on small edits, while GPT Image 2.5 Sunburst re-rendered most of the image and left only about a fifth unchanged. Nano Banana 2.1 kept its edits local but gradually darkened the rest of the image.

  2. a16z NewsBlogAI score40

    a16z leads investment in TypeSafe AI, maker of Jev System One model

    AIa16z says it is leading an investment in TypeSafe AI, whose Jev model hands decisions to code as typed values and reached 1 trillion tokens generated three days after launch. The company says Jev costs roughly 1/100 to 1/500 of frontier models and runs 100x faster on classification tasks at comparable accuracy. TypeSafe says 25% of the Fortune 500 have integrated Jev.

  3. LeiphoneNewsAI score42

    Ex-ByteDance intern Tian Keyu's secretive world-model lab reportedly raises at $200M valuation

    AITian Keyu, the Peking University PhD student known as the "ByteDance poisoning intern," has a 10-person world-model lab valued at $200 million after $30 million from Fivesource Capital and IDG, according to Leiphone. The lab plans to train a foundation model on about 100 million hours of video using a 200,000-symbol visual vocabulary, with a 2027 release targeted. Tian says the approach could cut the cost of generating one second of video by at least an order of magnitude.

  4. ModelScopeOfficialAI score28

    Corvus-Gov-3B: a 3B Chinese government-domain dialogue model

    AIModelScope released Corvus-Gov-3B, a compact model tuned for Chinese policy Q&A, public-service consultation, and internal government or enterprise assistants. It was fine-tuned on one million Chinese government-domain dialogue samples and built on Llama 3.2 3B Instruct using LoRA SFT via LLaMA Factory. The model is released under Apache 2.0.

    Image from @ModelScope2022's post
  5. MarkTechPostNewsAI score44

    Underdog Releases Saluki 27B, a 2-Bit Qwen3.8-27B That Beats the Original at Tool Calling

    AIUnderdog has released Saluki 27B under Apache 2.0, a 2-bit GGUF of Qwen3.8-27B that fits in 7.89 GB, versus 54 GB for the full BF16 model. On Underdog Bench, Saluki scores 88 against 84 for the full model, and it raises parallel tool-call accuracy to 42 from 35. It runs on stock llama.cpp, but math and reasoning drop sharply, with AIME 2025 at 79.2 versus 96.7.

  6. vLLMOfficialAI score42

    vLLM Semantic Router team releases Decision 2.0 multi-question classification models

    AIThe vLLM Semantic Router team has released Decision 2.0, which answers multiple questions about one input in a single forward pass and outputs per-option probabilities. The post presents this as useful for routing and classification. A quoted post from Xunzhuo Liu says Decision 2.0 includes six open decision models ranging from 0.6B to 27B parameters, each topping same-size open models on the Jev Decision Index 0.3.

  7. Nace AIXAI score22

    NDI 1.0 document processing model launches for coding agents at 90% lower cost

    AINACE introduces NDI 1.0, a document processing model for coding agents that it says is 90% cheaper and ranks first on the Parse Index. The company says it offers native MCP, SDK, and CLI integration for Claude Code, Codex, Hermes, OpenClaw, and PI, and supports 50 languages. NACE also states the model was trained on over 15M financial files and offers $25 in free API credits to developers.

    Video from @NaceAI's post

Oct 8

Oct 8Thu
  1. Alexander DoriaXAI score46

    LightOnOCR-3 claims state-of-the-art OCR performance under 1B parameters

    AILightOn has released LightOnOCR-3, a family of OCR models in 0.8B and 4B versions that it says lead benchmarks including OlmOCR-Bench and ParseBench, with the 0.8B model positioned as the sub-1B option. The models recognize text, handwriting, images, charts and document structure in one pass, process documents up to twice as fast as LightOnOCR-2, and are released under the Apache 2.0 license.

    Image from @Dorialexander's post
  2. GeneralistOfficialAI score28

    Generalist releases GEN-1.5, a foundation model for physical-world robotics

    AIGeneralist has announced GEN-1.5, its latest foundation model for the physical world. The post provides only a link to the company's blog for further details, so no specifications, benchmarks, or availability information can be confirmed from this source.

  3. Leandro von WerraXAI score70

    Carbon-A open model and database predict 566 million gene candidates across 22,617 species

    AICarbon-A is an open model that predicts gene locations directly from DNA, and it has been used to annotate genomes from over 22,000 species. The release includes a database of 566 million gene candidates, about 16 times the gene annotations in the RefSeq dataset. Wet-lab RNA experiments supported 239 candidates missing from RefSeq across cats, Syrian hamsters, chickens, and Arabidopsis.

    Why it matters: The source ties an open gene-annotation model to specific wet-lab checks and gene counts, helping readers judge how far its predictions extend beyond well-studied genomes.

  4. elvisXAI score32

    Drama 3 voice model offers fine-grained tone and emotion control

    AIFish Audio's Drama 3 voice model lets users direct tone, emotion, and pacing in plain language, and can shift emotion mid-sentence. The poster, who found it remarkably effective in testing, says the control over delivery is unlike anything previously seen. A preview is available through the API as drama-3-preview.

  5. Understanding AI (Timothy B. Lee)BlogAI score67

    TypeSafe AI's Jev returns probabilities over fixed answers instead of text

    AITypeSafe AI released Jev, a model that answers yes/no, multiple-choice, or rating questions by outputting the estimated probability of each option. The author notes this design lets the model be served faster and more cheaply than LLMs and fits ordinary if-statement logic, and says he used it to flag spam comments on his blog in place of Gemini 3 Flash.

Oct 7

Oct 7Wed

Oct 6

Oct 6Tue
  1. GeneralistOfficialAI score28

    Generalist's GEN-1.5 robot repeatedly places and removes an O-ring

    AIGeneralist's GEN-1.5 robot demonstrated putting an O-ring onto a hydraulic rod gland, then repeatedly removing and replacing it. The post is a short demonstration video that points readers to a blog post in the comments for more about GEN-1.5.

    Video from @GeneralistAI's post
  2. Latent SpaceBlogAI score60

    Reflection launches Beam, a 501B-parameter open-weight coding model

    AIReflection announced Beam, a text-only 501B-total, 23B-active MoE model for coding, agentic, and scientific work, trained from scratch with full weights under Apache 2.0 promised this month. Self-reported results include 80.9 on SWE-bench Verified and 3–4x the inference efficiency of GLM 5.2, while the roundup notes that GLM 5.3, Kimi K3, Qwen 3.8 Max, and DeepSeek V4.1 Flash are generally ahead.

Oct 5

Oct 5Mon
  1. PikaOfficialAI score22

    Pika lets users try Ideogram 4.5 on its platform

    AIPika announces that users can try Ideogram 4.5 through its create platform, linking to an Ideogram app page in its image tools section. The post provides no details on features, pricing, or capabilities.

  2. ReflectionOfficialAI score23

    Reflection AI's Beam model pretrained in four weeks on 24T tokens

    AIReflection AI says its Beam model was pretrained in 4 weeks on 24T high-quality tokens, giving it innate coding capabilities. The company credits MoE stability improvements and large-scale data curation and deduplication for a base model it claims outperforms open-source base models of the same class. It presents this strong reasoning foundation as what makes sustained reinforcement learning gains possible.

    Image from @reflection_ai's post
  3. ReflectionOfficialAI score42

    Reflection AI previews Beam, a 500B open model under Apache 2.0

    AIReflection AI says its Beam model, with a 500B form factor, combines strong agentic performance and efficient reasoning for enterprises, governments, and developers. Beam is in final red-teaming and will be released this month under an Apache 2.0 license, with quantized FP8 and NVFP4 versions for efficient deployment. Early access sign-ups are open on the company's platform.

Oct 3

Oct 3Sat
  1. IndexTeam (Bilibili) · new models on Hugging FaceOfficialAI score22

    Index-Nailong-9B-FP4 NVFP4 quantized translation model released on Hugging Face

    AIIndexTeam released Index-Nailong-9B-FP4, an official NVFP4 (W4A4) quantization of the Index-Nailong-9B multilingual translation model, which covers 150 languages. In a validation on an NVIDIA A100 against the BF16 checkpoint, perplexity rose 3.10% (2.4339 to 2.5094), and zh-en and en-zh outputs were semantically equivalent. Full FP4 compute acceleration requires an NVIDIA Blackwell GPU, while older GPUs get memory savings only; the FP8 build is recommended for Hopper and Ampere.

  2. IndexTeam (Bilibili) · new models on Hugging FaceOfficialAI score23

    Index-Homura-9B-FP4 released with NVFP4 quantization for translation model

    AIIndexTeam released Index-Homura-9B-FP4, an official NVFP4 (W4A4) quantization of the Index-Homura-9B translation model from the Index-Translate family. On a fixed corpus, perplexity rose from 2.5386 in BF16 to 2.6245, a 3.38% increase, and zh->en generations matched the original. Full FP4 compute acceleration requires an NVIDIA Blackwell GPU, while older GPUs get only weight-only memory savings and the FP8 build is recommended for them.

  3. Aidan GomezXAI score46

    AlephAlpha releases Kolibri, a German-English model with a technical report

    AIAlephAlpha has released Kolibri, a German-English model with 78B total parameters, 3.46B active, and context up to 1M tokens. The weights are available under Apache 2.0 for running on users' own hardware. Cohere's Aidan Gomez congratulated the team on the model and its detailed technical report.

Oct 2

Oct 2Fri

Oct 1

Oct 1Thu
  1. Vercel DevelopersOfficialAI score24

    Laya decision model free on Vercel AI Gateway through October 31

    AIVercel says the Laya decision model is available free on AI Gateway through October 31, in partnership with Boundless. The post suggests using it to route agent work, triage support requests, and check guardrails.

  2. ReplicateOfficialAI score46

    Replicate powers Tavus's Griffin, a video Turing test-passing model

    AIReplicate says it is powering Griffin from Tavus on its platform. Tavus describes Griffin as the first model to pass the video Turing test, with 48% of live conversation participants believing it was a real human. The model ranks first on NVIDIA's full-duplex AI video benchmark.

  3. OpenRouter · New modelsBlogAI score36

    Apodex 1.1 Mini Released as Free Reasoning Model for Research Tasks

    AIApodex has released Apodex 1.1 Mini, a free reasoning-first model designed for complex, long-horizon research and forecasting tasks. According to the source, it works directly with files, data, code, and tools to produce verifiable results.

Sep 30

Sep 30Wed
  1. MiniMax (official)OfficialAI score44

    HeyGen Video launches on MiniMax H3 at $0.01 per second

    AIHeyGen has released HeyGen Video, a production-quality video product built on MiniMax H3 and post-trained by HeyGen. Pricing starts at $0.01 per second through October, a 50% discount, aimed at businesses that need video without production-level costs.

  2. MiniMax (official)OfficialAI score38

    Creatify's Boreal-H3 ad video model built on MiniMax H3

    AICreatify Labs released Boreal-H3, a video model built on MiniMax H3 and post-trained specifically for advertising. Reported results include 85.3% reference fidelity, brief success rising from 28% to 50%, and identity match improving from 83% to 94%. Visible defects per clip dropped 70%, while generation time and estimated cost fell 20%.

  3. IdeogramOfficialAI score23

    Ideogram 4.5 Performs Strongly Across General Image Editing Tasks

    AIIdeogram 4.5 was built for precise, targeted editing but also performs very well across general editing tasks. Design Arena ranks it 15th in Image Editing with an Elo of 1250, placing it in the same performance band as MAI-Image-2.6 and Gemini 3 Pro Image Preview. It is especially strong at typography edits, such as modifying text in infographics.

  4. IdeogramOfficialAI score46

    Ideogram 4.5 launches with four quality modes at native 2K resolution

    AIIdeogram 4.5 comes in four quality modes, ranging from 0.8¢ to 22¢ per image, all at native 2K resolution. Available now on our launch partners: @superscale_ai @Picsart @cfabricacom @luminal_ai @arena @runware @florafaunaai @krea_ai @trymoda @LeonardoAi @runwayml @pika_labs @fal @ComfyUI @magnific @GammaApp @LumaLabsAI @DesignArena

    Image from @ideogram_ai's post
  5. IdeogramOfficialAI score38

    Ideogram 4.5 launches as a precise image edit model

    AIIdeogram released Ideogram 4.5, which it calls the most precise edit model, claiming it avoids the artifacts, pixel shifts, and color changes that leading models add with each edit. The company says this eliminates artifact buildup and makes multi-turn editing possible. It is live in Ideogram, via the API, and with launch partners, with open weights promised soon.

    Video from @ideogram_ai's post
  6. ModelScopeOfficialAI score46

    IndexTeam releases Index-Translate multilingual translation model family

    AIIndexTeam has released Index-Translate, a multilingual family covering text, speech, dubbing, and long-document translation across 150 languages. Its 9B model scores 0.8789 on FLORES, 0.8209 on instTrans, and 0.7387 on MEME, and the 2B and 9B models are released under Apache 2.0.

    Video from @ModelScope2022's post
  7. ModelScopeOfficialAI score62

    InSpatio-World 1.5 turns images and videos into real-time explorable 4D worlds

    AIInSpatio-World 1.5 from InSpatio_AI turns a single image, four images, a panorama, or a video into a navigable scene with wide viewpoint changes. The 1.3B model scores 68.72 on WorldScore-Dynamic, ranking first among evaluated real-time and interactive methods, with speeds up to 24 FPS. The post says the code is released under Apache 2.0 and that dependencies keep their own licenses.

    Video from @ModelScope2022's post
  8. ollamaOfficialAI score23

    Ollama adds Nimble, Tev1 4B, and Tev1 0.8B models

    AIOllama now lists three models, Nimble, Tev1 4b, and Tev1 0.8b, for use with /v1/systemone. The post provides a library link for each model, with the Tev1 variants listed at 4b and 0.8b.

  9. ollamaOfficialAI score30

    Ollama adds local decision models like Nimble via new API

    AIOllama now supports decision models such as Nimble locally, usable for tasks like ticket triaging, model routing, and content moderation. The model can be installed with ollama pull nimble and accessed through the new local /v1/systemone API, as shown in a real-time Ollama racer demo.

    Video from @ollama's post

Sep 29

Sep 29Tue
  1. ModelScopeOfficialAI score54

    IQuest-Q1 released as 320B MoE model for long-horizon coding agents

    AIModelScope announced IQuest-Q1, a 320B MoE model with 15B active parameters and a 512K context window for agentic coding. The post reports scores of 84.5 on CyberGym, 83.2 on Terminal-Bench 2.1, 64.6 on DeepSWE v1.1, and 63.0 on NL2Repo, and says weights are released under the IQuest-Q1 License.

    Image from @ModelScope2022's post