Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 9

TodayOct 9Fri
  1. SGLangOfficialAI score52

    SGLang adds Rubin optimizations that speed up Kimi K3 inference

    AISGLang says it worked with NVIDIA to optimize attention, MoE, and speculative verification kernels for Kimi K3 inference on early-access Rubin hardware. It reports up to 20% faster FP8 MLA at batch 1 with 128K context, 20% faster KDA verification with bitwise-identical output, and a 5.9% end-to-end speedup from MoE tail fusion that removes 276 kernel launches per decode step. The post also says SGLang powers Miles' end-to-end RL training on Rubin, including agentic RL with 64 concurrent sandboxes on the Vera CPU.

  2. Artificial AnalysisOfficialAI score32

    HiDream-O1-Video-1.0 ranks #6 on Artificial Analysis image-to-video leaderboard

    AIHiDream-O1-Video-1.0 ranks #6 in Artificial Analysis's Image to Video with Audio leaderboard, just behind Dreamina Seedance 2.0 720p. HiDream says the model generates 1080p videos of 5 to 20 seconds with synchronized audio, priced at $5.80 per minute ($0.10 per second) on the HiHarness API. It is also available in vivago R1 Studio.

    GIF from @ArtificialAnlys's post
  3. Artificial AnalysisOfficialAI score18

    HiDream-O1-Video-1.0 is listed on the AA-Video leaderboards

    AIArtificial Analysis invites users to check HiDream-O1-Video-1.0 on its AA-Video-I2V v1.0 image-to-video leaderboard. Users can also vote for the model in the Video Arena.

  4. Andrew CurranXAI score55

    Prime Agent swarm rewrites itself in Rust, reaching input 13x faster

    AIPrime Intellect says Prime Agent used a swarm of over 2,000 agents to rewrite itself end to end in Rust over two weeks. The rewrite ran across 10,000+ sandboxes and over 200 billion GLM-5.3 tokens, and the company says usable input now arrives about 13 times faster with 83% less startup memory.

    Image from @AndrewCurran_'s post
  5. Prime Intellect BlogOfficialAI score65

    Prime Agent is rewritten in Rust by a swarm of agents

    AIPrime Intellect says it rewrote its Prime Agent coding tool in Rust, using a swarm of more than 2,000 agents over two weeks. The company reports cold start to typing about 13 times faster than the TypeScript version, and memory use over 80% lower after startup. Prime Agent remains open source and adds native Windows support in beta and Homebrew installation.

    Why it matters: The post shows how a multi-agent swarm rewrote a coding agent with parity checks, giving a concrete case of agent-driven software engineering with measured results.

  6. Bertholomus AIXAI score22

    DeepSeek TP=2 and TP=4 kernel kits released on GitHub

    AIA GitHub post announces that kernel kits for DeepSeek's TP=2 and TP=4 recipes are now live. The post links to a repository named deepseek-v4.1-tensorfold-tp2-2xgb10 and gives no further details on performance or features.

  7. Rohan PaulXAI score46

    Pine launches cloud computer for AI agents, reports 1/20 token cost

    AIPine has launched a cloud computer built for AI agents, which developers create through an SDK and give jobs in plain language. Running GPT-5.6 Luna, Pine reports about 1/20 the model-token cost of GPT-5.6 Sol with Codex on SaaS-Bench v1.1, scoring 78.3%, the highest in the published comparison. Pine also reports 1/26 the token cost of Opus 5 with Claude Code and 2 to 5 times faster speed in selected preliminary internal tests.

    Image from @rohanpaul_ai's post
  8. Hacker News · AI (150+ points)BlogAI score38

    Show HN: big-arrow-on-the-screen lets AI agents draw arrows and text on macOS

    AIbig-arrow-on-the-screen (bigarrow) is a MIT-licensed macOS command-line tool and skill for Claude Code and Codex that draws arrows, boxes and text over any window. Clicks pass through, keyboard focus stays put, and each arrow removes itself after a set duration or when its agent process ends. The tool only points; it never clicks, types or captures the screen, and it requires no macOS permission to draw.

  9. Prime IntellectOfficialAI score46

    Prime Agent swarm of 2,000+ agents rewrites itself in Rust

    AIPrime Intellect says its Prime Agent orchestrated over 2,000 agents over two weeks to rewrite the agent in Rust. The run used more than 10,000 sandboxes, over 200B GLM-5.3 tokens, and 16,000 agent-to-agent messages. The company says the rewritten agent reaches usable input about 13 times faster and uses 83% less startup memory.

    Video from @PrimeIntellect's post
  10. Prime IntellectOfficialAI score36

    Prime Agent improves its runtime through self-directed benchmark hillclimbing

    AIPrime Intellect says Prime Agent, after reaching feature parity, ran its own runtime benchmark suite and tested candidate changes against the current build. Changes that passed parity checks and independent review became the baseline for the next experiment. The loop moved work off the startup and render paths and released memory after large sessions loaded.

    Video from @PrimeIntellect's post
  11. Prime IntellectOfficialAI score29

    Prime Agent Rust leads six agent harnesses in startup speed and footprint

    AIPrime Intellect says its Prime Agent Rust had the lowest time to usable input, startup memory, and installed size among six agent harnesses it benchmarked. The rewrite splits the codebase into nine crates with enforced dependency boundaries, and changes to shared protocol types are checked across the client, daemon, and session workers.

    Image from @PrimeIntellect's post
  12. Prime IntellectOfficialAI score42

    Prime Intellect plans reusable agent state machines in Prime Agent

    AIPrime Intellect says it is turning the workflow behind a recent rewrite into reusable state machines in Prime Agent, letting users run their own agent teams through implementation, review, and verification. The company also says it is accelerating work on capabilities and evals and connecting Prime Agent with cloud agent swarms and hosted training for autonomous research. This release adds native Windows support in beta and Homebrew installation.

  13. dexXAI score43

    Dex Horthy posts a one-word teaser, "he cook"

    AIDex Horthy (@dexhorthy) posted only the words "he cook" on X, with no further detail. The post is a short reaction and does not describe a product, release, or result on its own. Background from the quoted post by @0xblacklight describes a serverless background agent that created a GitHub pull request from an issue.

  14. OpenRouterOfficialAI score8

    Luna and Astra drive a surge in new Codex usage

    AIOpenRouter says the models Luna and Astra are driving a lot of new Codex usage. The post gives no figures, dates, or pricing for that usage.

    Image from @OpenRouter's post
  15. Prime IntellectOfficialAI score44

    Prime Intellect extends RL training to multi-agent swarms

    AIPrime Intellect says swarms have costs, since messages consume tokens, lose information, and agents must coordinate to avoid duplicated work. The company is extending its RL training infrastructure from individual agents to multi-agent systems, letting developers express arbitrary agent interactions and train them.

  16. GoodfireOfficialAI score36

    Goodfire launches activation monitors that detect undesired model behaviors

    AIGoodfire says its activation monitors use signals from inside a model to detect undesired behaviors, catching more cases, running faster and costing less than text-based monitors. Baseten customers can use them to monitor for prompt injection, actions outside policy, sensitive data exposure and cyber misuse.

  17. GoodfireOfficialAI score25

    Goodfire and Baseten partner on configurable model concern monitoring

    AIGoodfire says teams can configure how their applications respond when a concern is flagged, including logging, additional review, refusal, and re-routing. The company directs model servers and trainers to a partnership post with Baseten for building monitors into their stack.

  18. RadixArkOfficialAI score22

    RadixArk praises Proximal for training coding agents with Miles

    AIRadixArk says Proximal is using Miles to train coding agents and calls it a flexible, scalable foundation for teams running their own training workloads. Proximal says its training framework is built on Miles, with runs on Modal's on-demand GPU clusters and serverless GPUs for inference. Its sandboxing infrastructure runs on Kubernetes and can handle millions of concurrent rollouts.

  19. dexXAI score38

    HumanLayer releases teleport and orchestrate commands with a minimalist UI

    AIHumanLayer announces a new release with /hl:teleport, which moves a local session to any remote host the user owns or launches without losing context. The release also adds /hl:orchestrate, which lets HumanLayer drive its own tasks, including splitting work, forking workflows, and moving artifacts, and it ships a minimalist UI with rounded corners and less visual noise.

    Image from @dexhorthy's post
  20. ClineOfficialAI score39

    Cline offers free access to Upstage's Solar Mini 4 model

    AICline is offering Solar Mini 4 free, a new 35B mixture-of-experts model from Korean lab Upstage with 3B active parameters. It has a 524K context window and runs at 208 tokens per second. Cline says it scores 24 on the AAII, the highest of any model at 3B active and within one point of Nemotron 3 Ultra, which uses 55B active.

  21. Mason HallXAI score22

    Agentzon launches a storefront built for AI agents

    AIAgentzon announces a storefront built for AI agents, drawing on two years of work on agentic commerce products and tools. The post frames the product as a response to what agents want, without listing features, prices, or availability. Background from Paul Graham says Amazon's ban on agents creates an opening for startups to compete with Amazon.

  22. David PawlanXAI score25

    Agentzon launches a shopping site for AI assistants to buy household items

    AIAgentzon launches at agentzon.co as a checkout site for AI assistants, aggregating about 500 household essentials so far. Its founder says most AI assistant checkouts are poor because Amazon blocks assistants from shopping, and users can ask an agent to buy specific brands such as Colgate toothpaste. The site supports browser-use or API checkout, WebMCP, UCP and OpenAPI, and requires no accounts or API keys.

  23. IdeogramOfficialAI score45

    Ideogram 4.5 keeps edited images intact across 30 consecutive edits

    AIArtificial Analysis ran 30 consecutive real estate staging edits through four image editing models, and Ideogram 4.5 kept most of each image unchanged while others drifted. Ideogram 4.5 and FLUX 3 left 95% or more of the image untouched on small edits, while GPT Image 2.5 Sunburst re-rendered most of the image and left only about a fifth unchanged. Nano Banana 2.1 kept its edits local but gradually darkened the rest of the image.

  24. Artificial AnalysisOfficialAI score34

    HeyGen Voice tops Artificial Analysis Controlled Voice TTS Arena leaderboard

    AIHeyGen Voice ranks first on the Artificial Analysis Controlled Voice TTS Arena leaderboard with an Elo of 1,201 across 1,468 appearances, ahead of Qwen-Audio-3.1-TTS-Plus at 1,182 and ElevenLabs' Eleven v4 Turbo at 1,166. On pronunciation robustness it scores 83.1%, ranking #10 of 29 models, and it is priced at $30 per 1M characters and processes 40 characters per second.

    GIF from @ArtificialAnlys's post
  25. Artificial AnalysisOfficialAI score4

    HeyGen Voice sample narrates Arctic tern migration fact

    AIArtificial Analysis posted a sample prompt on HeyGen Voice, a text-to-speech output about the Arctic tern. The sample text states that Arctic terns have the longest migration of any animal, flying roughly 44,000 miles annually between the Arctic and Antarctic circles.

    Video from @ArtificialAnlys's post