Skip to contentSkip to stories

Updated

All AI news

Oct 8

Oct 8Thu
  1. Thomas WolfAI score62

    Carbon-A open model finds 566 million candidate genes across 22,617 species

    AIThe team released Carbon-A, an open model that finds genes directly in DNA, along with a database of 566.34 million candidate genes across 22,617 species. The model reads genomes without needing a close relative, and wet-lab validation in cats, chickens, and arabidopsis is cited, with 239 genes found missing from reference annotations of common species. The authors say the model marks gene locations but does not design DNA or predict gene function.

  2. LeiphoneAI score62

    Claude Haiku 5.5 gains on computer use but still trails Sonnet 5.5 in terminal coding

    AIAnthropic released Claude Haiku 5.5, raising its OSWorld 2.1 score from 15.7% to 72.4% and supporting a 1 million token context window. The article notes Haiku 5.5 still scores 39.2% on Terminal-Bench 4.0 against Sonnet 5.5's 70.6%, and that prompts above 100,000 tokens are priced higher, so migration costs need to be measured on real workloads.

  3. Understanding AI (Timothy B. Lee)AI score67

    TypeSafe AI's Jev returns probabilities over fixed answers instead of text

    AITypeSafe AI released Jev, a model that answers yes/no, multiple-choice, or rating questions by outputting the estimated probability of each option. The author notes this design lets the model be served faster and more cheaply than LLMs and fits ordinary if-statement logic, and says he used it to flag spam comments on his blog in place of Gemini 3 Flash.

  4. OpenRouter · New modelsAI score54

    StepFun releases Step 5 Preview, a 600B-parameter agentic model

    AIStepFun has released Step 5 Preview, its flagship model for agentic work, built on a sparse Mixture-of-Experts architecture with 27B active and 600B total parameters. The source says it performs strongly in software engineering and professional tasks, but the feed supplied only an excerpt, so benchmark details are not available here.

  5. JetBrains AI BlogAI score62

    JetBrains releases Mellum2.1, an open coding model trained with reinforcement learning

    AIJetBrains released Mellum2.1, a 12B mixture-of-experts model with 2.5B active parameters under the Apache 2.0 license, built for coding agents. Post-training shifted to reinforcement learning across thousands of environments and millions of sandboxed runs, and the model is available on Hugging Face. The source reports gains over Mellum2 on LiveCodeBench, AIME, GPQA Diamond, BFCL v4, IFEval, and SWE-bench Verified, and says it serves almost twice the tokens of Qwen3.5-9B under heavy load.

    Why it matters: The post shows how reinforcement learning in real sandboxed environments changed a compact open model's repository work, with benchmark gains against Mellum2 and two peers.

  6. The DecoderAI score72

    Claude Haiku 5.5 cuts prices but uses more tokens than GPT-6 Luna

    AIAnthropic released Claude Haiku 5.5, its fastest and most affordable small model, at prices up to 90 percent lower for most prompts under 100,000 tokens. Artificial Analysis ranks it first among small-class models on its Intelligence Index with a score of 43, but it consumes about three times the output tokens per task that GPT-6 Luna needs at maximum effort.

  7. Ant LingAI score22

    Ant Ling's Ling-3.1-flash now live on AI/ML API

    AIAnt Ling announced a day-zero collaboration with AI/ML API, making Ling-3.1-flash available there for agentic and cowork scenarios. AI/ML API describes it as a 560B-parameter MoE model with about 25B active per token and up to 1M context, built for agents, coding, and long documents. The model is free to try on AI/ML API until October 13.

  8. Artificial Analysis ArticlesAI score62

    GPT-6 Sol Daybreak Blue leads the Artificial Analysis Cyber Index

    AIArtificial Analysis is adding trusted-access models to its Cyber Index, starting with GPT-6 Sol (Daybreak Blue, max), which is available only through OpenAI's Daybreak program. The model hits no safety blocks across the Index and scores 32 points higher overall than the publicly available GPT-6 Sol (max), with its largest gains on CyberGym-E2E.

    Why it matters: The source shows how safety refusals shape cyber benchmark scores, with the trusted-access model's gains concentrated on CyberGym-E2E, useful for comparing guarded and unguarded models.

Oct 7

Oct 7Wed
  1. IThome · AIAI score72

    Anthropic releases Claude Haiku 5.5, cutting run costs about 75% from Haiku 4.5

    AIAnthropic released Claude Haiku 5.5, which it calls the fastest, cheapest, and most capable Haiku model so far. On average it costs about 75% less to run than Haiku 4.5, with input at $0.10 and output at $0.50 per million tokens for requests up to 100,000 tokens. Anthropic also cut Sonnet 5.5's cache read price from $0.20 to $0.10 per million tokens, which it says lowers run costs by about 20% on many agent tasks.

  2. Simon WillisonAI score62

    Anthropic releases Claude Haiku 5.5, priced like GPT-6 Luna up to 100,000 tokens

    AIAnthropic has released Claude Haiku 5.5, priced at $0.10 input and $0.50 output per million tokens up to 100,000 tokens, matching GPT-6 Luna. Beyond 100,000 tokens the price rises to $0.50 and $2.50, and the author found the new tokenizer uses about 1.25x as many tokens as Haiku 4.5 on the same long prompt. The model cannot disable reasoning and defaults to medium effort.

  3. MarkTechPostAI score67

    Anthropic releases Claude Haiku 5.5, a small model with 1M context

    AIAnthropic has released Claude Haiku 5.5, its cheapest and fastest small model, priced at $0.10 per million input tokens and $0.50 per million output tokens for prompts up to 100K tokens. It keeps a 1M token context window, up to 128K output tokens, and is generally available on the Claude API, Amazon Bedrock, Google Cloud, Microsoft Foundry and Claude Platform on AWS. Anthropic reports 72.4% on OSWorld 2.1 (offline subset) versus 15.7% for Haiku 4.5, and the article notes that non-default temperature, top_p or top_k values return a 400 error.

  4. TiboAI score78

    OpenAI rolls out GPT-6 to all ChatGPT users with an Intelligent UI

    AIOpenAI is releasing a new version of GPT-6 to all ChatGPT users, extending the model beyond text. The post says model and infrastructure improvements were combined to scale it to 1.2 billion users, and it pairs the release with Intelligent UI, which delivers fast, interactive, and visual answers.

    Why it matters: The post names the rollout scope and points to model and infrastructure work behind serving the update, which shows how a large consumer launch is being scaled.

  5. AWS Machine Learning BlogAI score56

    Claude Haiku 5.5 becomes available on Amazon Bedrock and Claude Platform on AWS

    AIAnthropic's Claude Haiku 5.5 is now available on Amazon Bedrock and Claude Platform on AWS. According to Anthropic, it is the fastest and most efficient model in the Claude 5.5 family and costs around 75 percent less than Claude Haiku 4.5 for most tasks. The post also covers pairing it with Claude Opus 5.5 as a subagent layer and provides Boto3, Converse, and Anthropic SDK examples for calling the model.

  6. MarkTechPostAI score60

    Liquid AI releases open-weight d1-3B and d1-omni-600M decision models

    AILiquid AI released two open-weight multimodal decision models, d1-3B and d1-omni-600M, which return probability answers in one forward pass with zero output tokens. d1-3B scores 48.57 on Decision Index v0.2.1 and answers one question in 8 ms on an RTX 4090, while the models are licensed free for commercial use below $10 million in annual revenue.

  7. ThariqAI score67

    Claude Haiku 5.5 returns as a cheaper, faster small model

    AIAnthropic has released Claude Haiku 5.5, which it describes as the cheapest, fastest, and most capable small model it has released. On average it costs around 75% less to run than Claude Haiku 4.5, and the author says it is 10x cheaper than Haiku 4.5 under 100k tokens. It can be tried with computer use, workflows, and the API.

    This story has a top pick“Anthropic releases Claude Haiku 5.5, scoring 43 on the Intelligence Index”

  8. Testing CatalogAI score62

    Anthropic releases Claude Haiku 5.5, its fastest and cheapest model

    AIAnthropic has released Claude Haiku 5.5, which the author describes as its fastest and cheapest model to date. The source says it costs about 75% less to run than Claude Haiku 4.5 and is the first Haiku model with an adjustable effort setting. The attached benchmark table reports Haiku 5.5 scores on tasks including computer use (OSWorld 2.1 offline subset, 72.4%) and Terminal-Bench 4.0 (39.2%), compared with Haiku 4.5 and other models.

  9. OpenAIAI score82

    GPT-6 with Intelligent UI rolls out to ChatGPT Chat tab across tiers

    AIOpenAI is rolling out GPT-6 with Intelligent UI globally to Plus, Pro, Business, and Enterprise users today, with Free and Go users following starting tomorrow. Plus, Pro, Business, and Enterprise tiers are powered by GPT-6 Sol, while Free and Go tiers use GPT-6 Luna, and both are tuned for everyday conversation. The update applies only to the Chat tab, and the models powering Work and Codex are not changing.

    This story has a top pick“OpenAI rolls out GPT-6 and Intelligent UI to all ChatGPT users”