Skip to contentSkip to stories

Updated

#Model release

Oct 8

Oct 8Thu
  1. OpenAI · YouTubeAI score70

    OpenAI adds Intelligent UI to GPT-6 in ChatGPT Chat tab

    AIOpenAI introduced Intelligent UI for GPT-6 in ChatGPT, which lets the chatbot answer with fully interactive interfaces and quickly build tools for a task. The feature rolled out globally to Plus, Pro, Business, and Enterprise tiers in the Chat tab and expands to Free and Go tiers, with Enterprise access depending on workplace admin settings.

    This story has a top pick“OpenAI rolls out GPT-6 and Intelligent UI to all ChatGPT users”

  2. Alexander DoriaAI score46

    LightOnOCR-3 claims state-of-the-art OCR performance under 1B parameters

    AILightOn has released LightOnOCR-3, a family of OCR models in 0.8B and 4B versions that it says lead benchmarks including OlmOCR-Bench and ParseBench, with the 0.8B model positioned as the sub-1B option. The models recognize text, handwriting, images, charts and document structure in one pass, process documents up to twice as fast as LightOnOCR-2, and are released under the Apache 2.0 license.

  3. OpenAI · YouTubeAI score72

    OpenAI rolls out GPT-6 with Intelligent UI in ChatGPT Chat

    AIOpenAI says GPT-6 in ChatGPT adds Intelligent UI, which lets ChatGPT answer with interactive interfaces and build quick tools for a task. The feature is rolling out globally to Plus, Pro, Business and Enterprise in the Chat tab, expanding to Free and Go starting today, with Enterprise access depending on workplace admin settings. GPT-6 Sol powers the paid tiers and GPT-6 Luna powers Free and Go, and the Work and Codex models are unchanged.

    This story has a top pick“OpenAI rolls out GPT-6 and Intelligent UI to all ChatGPT users”

  4. OpenAI · YouTubeAI score67

    OpenAI launches GPT-6 Intelligent UI for interactive ChatGPT answers

    AIOpenAI's GPT-6 in ChatGPT adds Intelligent UI, which lets ChatGPT answer with interactive interfaces and build quick tools for a task. The feature is rolling out to Plus, Pro, Business, and Enterprise first, with Free and Go tiers following, and Enterprise access depends on workplace admin settings. The update covers only the Chat experience, and the models powering Work and Codex are not changing.

    This story has a top pick“OpenAI rolls out GPT-6 and Intelligent UI to all ChatGPT users”

  5. MarkTechPostAI score58

    JetBrains releases Mellum2.1, a 12B MoE open model for coding agents

    AIJetBrains has released Mellum2.1, a 12B mixture-of-experts thinking model with 2.5B active parameters, under Apache 2.0 on Hugging Face. Post-training reinforcement learning in real software repositories raised SWE-bench Verified from 2.0 to 47.0, according to JetBrains' self-reported results. Qwen3.5-9B still leads on SWE-bench Pro, GPQA Diamond and AIME, and GGUF builds start at 7.0 GB for local use.

  6. Leandro von WerraAI score70

    Carbon-A open model and database predict 566 million gene candidates across 22,617 species

    AICarbon-A is an open model that predicts gene locations directly from DNA, and it has been used to annotate genomes from over 22,000 species. The release includes a database of 566 million gene candidates, about 16 times the gene annotations in the RefSeq dataset. Wet-lab RNA experiments supported 239 candidates missing from RefSeq across cats, Syrian hamsters, chickens, and Arabidopsis.

    Why it matters: The source ties an open gene-annotation model to specific wet-lab checks and gene counts, helping readers judge how far its predictions extend beyond well-studied genomes.

  7. Thomas WolfAI score62

    Carbon-A open model finds 566 million candidate genes across 22,617 species

    AIThe team released Carbon-A, an open model that finds genes directly in DNA, along with a database of 566.34 million candidate genes across 22,617 species. The model reads genomes without needing a close relative, and wet-lab validation in cats, chickens, and arabidopsis is cited, with 239 genes found missing from reference annotations of common species. The authors say the model marks gene locations but does not design DNA or predict gene function.

  8. LeiphoneAI score62

    Claude Haiku 5.5 gains on computer use but still trails Sonnet 5.5 in terminal coding

    AIAnthropic released Claude Haiku 5.5, raising its OSWorld 2.1 score from 15.7% to 72.4% and supporting a 1 million token context window. The article notes Haiku 5.5 still scores 39.2% on Terminal-Bench 4.0 against Sonnet 5.5's 70.6%, and that prompts above 100,000 tokens are priced higher, so migration costs need to be measured on real workloads.

  9. OpenRouterAI score62

    StepFun's Step 5 Preview model is now available on OpenRouter at $1.00 per million input tokens

    AIOpenRouter announces that StepFun's Step 5 Preview is live on its platform, priced at $1.00 per million input tokens and $2.70 per million output tokens. Cache hits are 50% off at launch, bringing them to $0.05 per million. A week of free access is rolling out across partners including opencode, Cline, Nous Research, and Kilo Code.

  10. Understanding AI (Timothy B. Lee)AI score67

    TypeSafe AI's Jev returns probabilities over fixed answers instead of text

    AITypeSafe AI released Jev, a model that answers yes/no, multiple-choice, or rating questions by outputting the estimated probability of each option. The author notes this design lets the model be served faster and more cheaply than LLMs and fits ordinary if-statement logic, and says he used it to flag spam comments on his blog in place of Gemini 3 Flash.

  11. OpenRouter · New modelsAI score54

    StepFun releases Step 5 Preview, a 600B-parameter agentic model

    AIStepFun has released Step 5 Preview, its flagship model for agentic work, built on a sparse Mixture-of-Experts architecture with 27B active and 600B total parameters. The source says it performs strongly in software engineering and professional tasks, but the feed supplied only an excerpt, so benchmark details are not available here.

  12. JetBrains AI BlogAI score62

    JetBrains releases Mellum2.1, an open coding model trained with reinforcement learning

    AIJetBrains released Mellum2.1, a 12B mixture-of-experts model with 2.5B active parameters under the Apache 2.0 license, built for coding agents. Post-training shifted to reinforcement learning across thousands of environments and millions of sandboxed runs, and the model is available on Hugging Face. The source reports gains over Mellum2 on LiveCodeBench, AIME, GPQA Diamond, BFCL v4, IFEval, and SWE-bench Verified, and says it serves almost twice the tokens of Qwen3.5-9B under heavy load.

    Why it matters: The post shows how reinforcement learning in real sandboxed environments changed a compact open model's repository work, with benchmark gains against Mellum2 and two peers.

  13. The DecoderAI score72

    Claude Haiku 5.5 cuts prices but uses more tokens than GPT-6 Luna

    AIAnthropic released Claude Haiku 5.5, its fastest and most affordable small model, at prices up to 90 percent lower for most prompts under 100,000 tokens. Artificial Analysis ranks it first among small-class models on its Intelligence Index with a score of 43, but it consumes about three times the output tokens per task that GPT-6 Luna needs at maximum effort.

  14. GuizangAI score22

    Guizang criticizes Anthropic over Haiku 5.5 pricing against Chinese models

    AI怎么这么多精神 Anthropic 公司人 我发这个信息说了句降价,这个定价专门用来狙击国产模型,说了句恶心,一堆人来骂 好像这模型一便宜就忘了 Anthropic 之前干过啥了 Why are there so many Anthropic people (defenders) here? I posted a message saying just one thing—a price cut—and said this pricing is specifically meant to snipe domestic Chinese models, and that it's disgusting. A bunch of people came to attack me. Seems like once the model gets cheap, people forget what Anthropic did before.