Skip to contentSkip to stories

Updated

#Model release

Oct 8

Oct 8Thu
  1. GuizangAI score22

    Guizang criticizes Anthropic over Haiku 5.5 pricing against Chinese models

    AI怎么这么多精神 Anthropic 公司人 我发这个信息说了句降价,这个定价专门用来狙击国产模型,说了句恶心,一堆人来骂 好像这模型一便宜就忘了 Anthropic 之前干过啥了 Why are there so many Anthropic people (defenders) here? I posted a message saying just one thing—a price cut—and said this pricing is specifically meant to snipe domestic Chinese models, and that it's disgusting. A bunch of people came to attack me. Seems like once the model gets cheap, people forget what Anthropic did before.

  2. meng shaoAI score39

    Claude Haiku 5.5 tops GPT-6 Luna on benchmarks, with 2x faster token output

    AIAnthropic's Claude Haiku 5.5, released alongside Claude Opus 5.5 and Claude Sonnet 5.5, is reported to lead GPT-6 Luna across benchmarks, with OpenRouter measuring roughly twice the token output speed. Anthropic says Haiku 5.5 is its cheapest, fastest, and most capable small model, costing about 75% less to run than Claude Haiku 4.5 on average. The post also notes some CodeX users are reportedly migrating to Claude Code.

Oct 7

Oct 7Wed

Oct 6

Oct 6Tue
  1. Clément DelangueAI score62

    Mistral Large 4 announced with API access today and open weights due end of October

    AIMistral announced Mistral Large 4, a natively multimodal model with 1T parameters and 49B active parameters. It is available via API today, with open weights planned for the end of October. Clément Delangue, Hugging Face's CEO, reacted by noting that the model cannot be the best open-weight model until its weights are actually released.

  2. Yuchen JinAI score72

    Mistral Large 4 launches as a 1T-parameter multimodal model with open weights due end of October

    AIMistral announced Mistral Large 4, a natively multimodal model with 1T parameters and 49B active, available via API today. Mistral claims it is the best open weights model from the US or Europe on aggregated benchmarks, with open weights set for release at the end of October. The author quotes this claim and comments that it appears to beat GLM-5.3.

  3. Sophia YangAI score45

    Mistral Large 4 tops benchmarks across cybersecurity, legal, and agentic tasks

    AIMistral Large 4 is a 1T-parameter natively multimodal model with 49B active parameters, which the Mistral account says leads open-weights models from the US or Europe on aggregated benchmarks. The post claims it beats closed frontier models on visual grounding and posts strong results across cybersecurity, legal, and agentic behavior. It is available via API now, with open weights due at the end of October.

Oct 5

Oct 5Mon
  1. Thomas WolfAI score14

    Thomas Wolf hopes Claude Opus 4.6 stays available for a long time

    AIThomas Wolf, who runs Hugging Face, said he hopes Claude Opus 4.6 remains available for a long time. The post is a brief expression of preference, supported by a quoted post in which David Holz reported that in a self-run "have fun" test across LLMs, Opus 4.6 repeatedly won by imagining brief worlds of contradictions inside falling water droplets, while he felt newer models seemed to have less fun.

Oct 2

Oct 2Fri

Oct 1

Oct 1Thu

Sep 30

Sep 30Wed
  1. whAI score67

    Gemini 4 Argon previewed with frontier coding and cyber defense claims

    AIThe post quotes Google's Sundar Pichai introducing Gemini 4 Argon as an early look at the next model. It claims frontier performance in complex workflows, cyber defense, and software engineering, and says Google teams are using it for tasks from coding to quantum computing. The author adds that on FrontierSWE the model is very self-critical and often says "Eureka!", a personality they describe as a large improvement over previous Gemini models.

Sep 29

Sep 29Tue

Sep 28

Sep 28Mon
  1. Google Cloud · AI & Machine LearningAI score40

    Why startups should pair open models like Gemma 4 with frontier APIs

    AIGoogle Cloud argues startups should combine open-weight models with frontier APIs rather than routing every request to one frontier model. It cites Gemma 4, which spans five sizes including a 31B dense model and a 26B A4B Mixture-of-Experts model, released under Apache 2.0. The article's examples report a 44% latency drop for Cue, from 876 ms to 488 ms, and a $0 server cost for BetterSpeak's on-device Gemma 4 E2B.

Sep 27

Sep 27Sun

Sep 26

Sep 26Sat

Sep 23

Sep 23Wed
  1. howie.seriousAI score22

    Opus 5.5 praised for language, visual taste, code quality, and token efficiency

    AIThe X user howie.serious says Claude Opus 5.5 delivers high language quality, good visual taste, strong code quality, and notably low token usage. He also compares Anthropic's reported roughly $2 trillion IPO valuation with OpenAI's roughly $1.2 trillion fundraising valuation, arguing OpenAI is worth about 0.6 Anthropics and the gap may widen.

Sep 22

Sep 22Tue
  1. Boris ChernyAI score62

    Claude Opus 5.5 ports HAProxy to Rust faster and cheaper than Fable 5.1

    AIAnthropic introduced Claude Opus 5.5 as the first model in its Claude 5.5 family, saying it performs at the level of Claude Fable 5.1 for most tasks at 40% lower run cost than Opus 5. Boris Cherny reports that Opus 5.5 and Fable 5.1 each ported HAProxy from C to Rust and both passed nearly all of its tests, with Opus 5.5 finishing in 9.5 hours versus 12 hours and at 51% less cost.