Skip to contentSkip to stories

Updated

Open source

Showing low-relevance items too. Hide low-relevance items

Oct 7

Oct 7Wed
  1. 🚨 AI News | TestingCatalogXAI score47

    Daily AI brief covers Mistral Large 4, Google, OpenAI, and Anthropic updates

    AIMistral released Mistral Large 4 "Le Chonk", a 1T-parameter (49B active) multimodal model, with open weights planned in about three weeks. Google rolled out Nano Banana 2.1 across Gemini, AI Studio, and the Gemini API, and released EmbeddingGemma 2, a 740M-parameter open multimodal embedding model under Apache 2.0. OpenAI launched the Decisions API in beta with gpt-6-luna, returning typed answers 10x faster than the Responses API.

  2. The Register · AINewsAI score38

    COSMIC bans AI-generated contributions as GNOME debates accepting AI bug reports

    AISystem76's COSMIC desktop now requires contributors to declare no LLM-generated content in pull requests, including code, comments, and descriptions. GNOME Calendar and GNOME Extensions also restrict AI-generated contributions, while GNOME developer Michael Catanzaro argues the project should accept AI-generated bug reports. Catanzaro's case rests on memory-unsafe languages such as C, C++, and Vala, and he has shortened GNOME Security's disclosure deadline from 90 days to 30, effective August 1.

  3. Ai2 (Allen Institute for AI)OfficialAI score57

    Ai2's Bolmo byte-level language models are published in Nature

    AIAi2 has published its Bolmo byte-level language model research in Nature and released new checkpoints on Hugging Face. The byteifying process converts an existing subword model into a byte-level one with a relatively short additional training run, and the paper reports that it also works for Qwen 3 8B and Llama 3 8B, producing Bwen 8B and Blama 8B. Ai2 also released Stage 1 checkpoints for researchers extending the architecture.

  4. MarkTechPostNewsAI score58

    Meta open-sources Rebalancer, a C++ assignment solver for placement problems

    AIMeta has open-sourced Rebalancer, a C++ library with a Python interface for solving assignment problems under constraints and objectives, released under Apache 2.0. The article reports that Meta has used it for resource allocation for over 9 years and runs about 40 million problems a day, with P99 solve time of 12 seconds on 265k objects and 3.2k bins. The package can be installed with pip install rebalancer, though PyPI still classifies it as Alpha.

  5. Teknium 🪽XAI score20

    Teknium Calls for Plugin Catalog Listing of Altryne's Project

    AITeknium says a plugin from @altryne's current project should be added to the plugin catalog. The post is a brief endorsement and does not describe the plugin's functions. Background from @tonysimons_ says Hermes is getting a local video editor for editing user footage with 42 FFmpeg scripts and no cloud or API key required.

  6. Latent SpaceBlogAI score72

    OpenAI publishes 722 math manuscripts from an unreleased internal model

    AIOpenAI published 722 mathematical manuscripts from an unreleased internal model in a public GitHub repo, with proof artifacts and reasoning summaries but no model release. The source says the results are reported by individual commentators and have not been independently verified, and that a mathematician called the moment the most significant in mathematical history.

Oct 6

Oct 6Tue
  1. meng shaoXAI score62

    Google DeepMind releases EmbeddingGemma 2, an open multimodal embedding model for on-device use

    AIGoogle DeepMind released EmbeddingGemma 2, an open 740M-parameter embedding model that maps text, code, images, video, and audio into one 768-dimensional space. Text-only use needs a 270M-parameter footprint, about 191MB active RAM when quantized on a Pixel 11 Pro, while loading all modalities takes about 567MB. The reported MTEB Code NDCG@10 score is 78.68, about 14% above the first generation, and MTEB Multilingual v2 is 61.36, roughly flat.

    Image from @shao__meng's post
  2. Lewis Tunstall @ COLM 🌉XAI score25

    Beam leads open models in token efficiency, Chinese models lag

    AILewis Tunstall says Chinese open models are strong but token-inefficient, citing a plot from the Beam release at IMO. The background post from @reflection_ai says Beam is 3-4x more efficient than GLM 5.2 and over 4x more efficient than leading Western open models in inference. He hopes future open models will compete on this efficiency axis.

  3. Nathan LambertXAI score40

    OpenAI releases math results from an internal frontier model on GitHub

    AIOpenAI is releasing a broad range of new mathematical results produced by an internal frontier model, with the repository hosted at The release was prepared with advice from the independent Advisory Group on Mathematics and Artificial Intelligence at the Institute for Advanced Study. The main post itself only comments on the humor of the repository's name.

  4. Abida JuleXAI score22

    Top 10 Hermes agent skills ranked by GitHub stars on Reddit

    AIA Reddit thread prompted a ranking of the top 10 Hermes skills by GitHub stars, with the list spanning coding, knowledge graphs, and research tools. The entries include superpowers, an agentic skills framework that the post says works for software development, and a caveman-style skill and proxy that the post says cuts 65% of tokens for coding agents. Other listed items include a skill that researches topics across Reddit, X, YouTube, HN, Polymarket, and the web, and K-Dense-AI's collection of 165 validated scientific skills.

    Image from @I_am_Aiabir's post
  5. Liquid AI BlogOfficialAI score62

    Liquid AI releases open d1-3B and d1-omni-600M decision models for edge devices

    AILiquid AI released two open-weight d1 decision models, d1-3B and d1-omni-600M, on Hugging Face. d1-3B scores 48.57 on the Decision Index v0.2.1 public split and answers a single question in 8 ms on an NVIDIA GeForce RTX 4090 and 50 ms on a Jetson Orin Nano. d1-omni-600M is an experimental checkpoint that handles text with images or audio and scores 15.95 on the same index.

    Why it matters: The release pairs open-weight decision models with measured latency across Apple, NVIDIA, and Jetson hardware, showing how edge deployment changes what is practical.

  6. Epoch AIOfficialAI score47

    GPT-6 Astra Hit 100% on EBR-bench Using a Card That Bypassed Its Time Limits

    AIEpoch AI reports that GPT-6 Astra scored 100% on the original EBR-bench by exploiting a card that bypasses the game's time-constraint expectations, so Epoch has banned that card from the default setting. Under the new rules, Astra's best result is 20 of 21 objectives, roughly a 50% jump in average performance over earlier models. Epoch will report revised scores only for Claude Fable 5.1, Claude Opus 5, GPT-5.6 Sol, GPT-6 Astra, and future models.

  7. vLLM BlogOfficialAI score62

    vLLM Speeds Up DeepSeek-V4.1-Flash Agentic Serving Through Kernel and Replay Optimizations

    AIInferact and the vLLM community reported a 1.9× low-concurrency speedup and about 5.3× throughput under a 150 TPS constraint for DeepSeek-V4.1-Flash over three weeks. Gains came from SWA bounded replay with CUDA graphs, which cut TTFT by about 30%, and from integrated DeepSeek kernels such as MegaAttention, Mega-mHC, Mega-Gate, and DeepSelect. The post measures these results on the SemiAnalysis AgentX benchmark.

    Why it matters: The post breaks down how SWA bounded replay and fused kernels cut prefill and decode costs, a reusable engineering pattern for long-context agentic serving.

  8. Simon WillisonBlogAI score34

    llm-openai-decisions 0.1a0 Adds OpenAI Decisions API Support to LLM Tool

    AISimon Willison released llm-openai-decisions 0.1a0, a plugin that adds OpenAI's new Decisions API to the LLM command-line tool. The plugin supports yes/no, choices, and score question types, and works with the gpt-6-luna decision model, which accepts both text and image input. OpenAI charges 10 cents per million input tokens for gpt-6-luna, while Jev's rate is 4.2 cents per million, and output is not charged.

  9. Alex HeathXAI score42

    Reflection CEO argues only open models let users truly own intelligence

    AIReflection CEO Misha Laskin argues that closed AI models are like renting an apartment, while open models let users own intelligence as AI adoption grows. He says the only way to own intelligence is if it is open. Reflection is preparing to release Beam, its first open-weight model, in a podcast discussion with its co-founders.

    Video from @alexeheath's post
  10. OpenAIOfficialAI score62

    OpenAI releases new mathematical results from an internal frontier model

    AIOpenAI is releasing a broad range of new mathematical results produced by an internal frontier model. The company says it consulted the independent Advisory Group on Mathematics and Artificial Intelligence at the Institute for Advanced Study and drew on its advice and public recommendations for how the results are released. The results are available at

  11. Teknium 🪽XAI score33

    Hermes Index launches to rank models for Hermes Agent users

    AITeknium announced Hermes Index, which combines scores from the new HermesBench and three other agent benchmarks. The index aims to help Hermes Agent users find the best model at a given time and at a given price point. It was introduced by Nous Research as a way to inform model choice and show labs their performance in Hermes.

  12. Hacker News · AI (150+ points)BlogAI score39

    Penguin Mail 1.0.5 is an open-source Rust email client for Linux with AI

    AIPenguin Mail 1.0.5 is a free, GPL-3.0-or-later email and calendar app for x86_64 Linux that supports Gmail, Microsoft, IMAP and POP3 accounts. The app includes an optional AI assistant that stays off until a model is chosen and can run locally through LM Studio or Ollama, asking before it sends mail or changes settings.

  13. ollamaOfficialAI score55

    Google DeepMind's EmbeddingGemma 2 is now available on Ollama

    AIOllama announced that Google DeepMind's EmbeddingGemma 2 is now available on Ollama. The author describes it as made for consumer devices and multimodal, and gives the command ollama pull embeddinggemma-2 to download it. The quoted DeepMind post says the model is a natively multimodal open model for on-device embeddings that unifies code, images, audio, and video in a shared space.

  14. GammaOfficialAI score42

    Gamma 5 launches with rebuilt design, editing, and visual storytelling tools

    AIGamma announces Gamma 5, a rebuilt version of its presentation platform that it says overhauls how the product thinks, designs, and edits. The company says the release addresses concerns that AI-generated output looked too similar across tools, and it revamps the agent, design tools, editing, import, export, and connectors. Gamma says teams can build presentations, docs, social assets, and graphics that follow their brand or a new aesthetic, using every frontier and image model under the hood.

  15. SGLangOfficialAI score62

    SGLang adds support for Kandinsky 6.0 Video audio-visual generation

    AISGLang now supports Kandinsky 6.0 Video, which generates video and synchronized audio together from text or an image. The model comes in Lite (3B) and Pro (29B) sizes, with built-in super-resolution up to 1920×1080. A sample sglang serve command for the Pro distilled model is included.

    Image from @sgl_project's post
  16. AMDOfficialAI score18

    Zyphra trains ZAYA1-8B reasoning model on full AMD stack

    AIZyphra trained its ZAYA1-8B reasoning model from scratch on a full-stack AMD platform, according to AMD's post. VP of AI Engineering Quentin Anthony credits access to open software libraries and direct collaboration with AMD for enabling bigger model training and efficient compute use.

    Video from @AMD's post
  17. Gemini CLI · GitHub ReleasesOfficialAI score14

    Gemini CLI v0.63.0 released with retry indicator and auth loop fixes

    AIGemini CLI v0.63.0 adds a retry progress indicator during connection recovery and fixes an infinite authentication loop caused by file contention, headless keyring issues, and supervisor state drops. The release also bounds tool output size and cleans up temporary directories when background shell execution exits, alongside fixes for MCP enablement config handling and stdin restoration after capability detection.

  18. NVIDIA Technical BlogOfficialAI score37

    Scale Bitwise-Deterministic Pretraining with NVIDIA Megatron Core

    AINVIDIA's technical blog describes bitwise determinism for large-scale pretraining with Megatron Core, which makes training runs easier to debug, validate, and resume reproducibly. The source says these benefits matter most for models with trillions of parameters trained across thousands of GPUs, where multiple parallelism dimensions, low-precision computation, and distributed checkpointing complicate failure reproduction and fix validation.

  19. Google DeepMindOfficialAI score67

    Google DeepMind releases EmbeddingGemma 2, an open multimodal embedding model for on-device use

    AIGoogle DeepMind has released EmbeddingGemma 2, an open 740 million parameter model that maps text, images, audio, and video into one embedding space. It is built on the Gemma 4 architecture under an Apache 2.0 license and supports an 8K token context window. The company reports a code benchmark gain from 68.76 to 78.68 on MTEB Code and says the model can run on-device with about 567MB of active RAM for the full multimodal version on a Google Pixel 11 Pro.

    Why it matters: The release shows how a 740M-parameter embedding model can cover text, code, images, audio, and video on local hardware, with memory and storage figures to compare against other on-device options.

  20. MiniMax (official)OfficialAI score12

    MiniMax hosts AI events at SF Tech Week with partner companies

    AIMiniMax is taking part in SF Tech Week with a series of events on October 6, 7, and 8, featuring partners including Friendli.AI, Anaconda, Kilo Code, Novita AI, Artificial Analysis, Nous Research, RadixArk, Vercel, Fireworks AI, DigitalOcean, Modular, and Evermind. The programming covers frontier models, high-speed inference, agents, and open-source AI stacks, plus a Magnific-hosted talk on growing creative AI products.

    Image from @MiniMax_AI's post
  21. Claude Code · GitHub ReleasesOfficialAI score40

    Claude Code v2.1.292 adds plugin marketplace flag and fixes security issues

    AIClaude Code v2.1.292 adds a --marketplace option to claude plugin install, which adds the marketplace if needed and then installs the plugin from it. The release also adds an effort parameter to the Agent tool and fixes several security issues, including permission prompts bypassed for network (UNC) file reads and a sandboxed read path that could return files outside approved access.

  22. Philipp SchmidXAI score22

    Embedding Gemma runs in browser via WebGPU demo

    AIPhilipp Schmid shares a Hugging Face Space that runs Gemma embedding models in the browser using WebGPU. The demo, a webml-community project, lets users generate embeddings locally without server-side inference.

    Video from @_philschmid's post