Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Jul 2

Jul 2Thu
  1. Cognition Blog (Devin, Windsurf)AI score38

    Cognition launches Devin Security Vulnerability Remediation Program for enterprise backlogs

    AICognition launched the Devin Security Vulnerability Remediation Program, in which its forward-deployed engineers embed with customer teams to deploy Devin to find, validate, and fix vulnerabilities. The program first works through existing scanner backlogs from tools such as Snyk, SonarQube, and Semgrep, shipping validated fixes as pull requests, then adds Devin Security Swarm for continuous discovery of logic flaws. Most engagements run about six weeks, and eligibility is limited to enterprise Devin Cloud customers meeting the program's requirements.

Jul 1

Jul 1Wed
  1. Cognition Blog (Devin, Windsurf)AI score57

    Cognition launches Devin Security Swarm to find, verify, and patch vulnerabilities

    AICognition has launched Devin Security Swarm, which uses parallel agents to find vulnerabilities across a codebase, confirms exploitability in isolated sandboxes, and opens remediation PRs. In an evaluation on 50 real-world GitHub Security Advisory vulnerabilities, Devin reached 72% recall at about $90.23 per run, compared with 68% for Claude Security at $131.87 per run. The product is available starting today, with scan profiles and incremental scans that process only changed code after the first full baseline.

Jun 30

Jun 30Tue

Jun 29

Jun 29Mon
  1. Cognition Blog (Devin, Windsurf)AI score62

    Cognition's Devin Fusion routes coding work between two models to cut cost

    AICognition has released a preview of Devin Fusion, a multi-model harness that runs a frontier main agent alongside a cheaper sidekick agent. On FrontierCode 1.1 Extended, the company reports scores near frontier models at up to 60% lower cost per task, and 41% lower cost when paired with Fable 5, which access was suspended from June 12, 2026.

    Why it matters: The post explains a sidekick architecture with cached persistent contexts, which contrasts with advisor-style tools and shows how cost cuts depend on the main model's delegation behavior.

Jun 28

Jun 28Sun
  1. PaddlePaddleAI score46

    PaddlePaddle announces Unlimited-OCR now runs in vLLM

    AIUnlimited-OCR, Baidu's long-context OCR model, now runs in vLLM, with a recipe provided for developers to try it. The background post says it parses entire books in one pass using Reference Sliding Window Attention (R-SWA), which keeps the KV cache fixed during decoding, and claims 35% faster throughput than DeepSeek-OCR at 6K output tokens.

Jun 27

Jun 27Sat
  1. PaddlePaddleAI score36

    PaddleFormers 1.2 adds DeepSeek-V4 training with 128K+ context support

    AIPaddleFormers 1.2 is released with support for training DeepSeek-V4 and 128K+ long-context training. The update adds Context Parallel, Packing, Document Mask Attention, and the Muon optimizer, plus ultra-fused mHC, CSA, and HCA operators, DeepEP/HybridEP communication, and lossless FP8 training with AutoSubbatch memory balancing. The project is presented as fully open-source and is available on GitHub.

Jun 25

Jun 25Thu

Jun 23

Jun 23Tue

Jun 20

Jun 20Sat

Jun 19

Jun 19Fri

Jun 17

Jun 17Wed
  1. Jim FanAI score64

    ENPIRE lets Codex agents run autonomous research on a robot fleet

    AINVIDIA GEAR's ENPIRE gives eight Codex agents a fleet of robots, GPUs, and a token budget to solve physical tasks with minimal human oversight. The author reports tasks such as tying zip-ties, organizing fine pins, and installing GPUs, and a faster time-to-solution with eight parallel robots than with fewer. Safety uses a kinematic limit that resets a robot leaving its envelope, a torque-limited gripper, and a frozen reward function classifier. The team says everything will be open-sourced.

    Video from @DrJimFan's post

Jun 16

Jun 16Tue
  1. Jim FanAI score62

    Jim Fan's ENPIRE lets Codex agents run autonomous research on robot fleets

    AIJim Fan introduces ENPIRE, which gives eight Codex agents a fleet of robots, GPUs, and a token budget to solve physical tasks autonomously. The post reports that the system can tie zip-ties, organize fine pins, and install GPUs, and that eight robots exploring in parallel improve faster than fewer. The team plans to open-source everything.

    Video from @DrJimFan's post
  2. Xiaomi MiMoAI score38

    Xiaomi launches MiMo Claw, an agent integrated with Kingsoft Office

    AIXiaomi has launched MiMo Claw, an agent built on its flagship MiMo model and integrated with Kingsoft Office for Word, Excel, PowerPoint, and PDF workflows. The company says it consumes 40–60% fewer tokens than comparable solutions, and daily usage has been expanded from 1 hour to 4 hours, with free access and no deployment required. A limited-time subscription is priced at ¥14.9 per month.

Jun 12

Jun 12Fri
  1. Georgi GerganovAI score34

    Gerganov flags locate-anything.cpp, a ggml runtime for NVIDIA's LocateAnything-3B

    AIGeorgi Gerganov highlighted locate-anything.cpp, a native C++/ggml inference implementation of NVIDIA's LocateAnything-3B for open-vocabulary object detection and visual grounding. The project, from the LocalAI team, runs without Python on CPU and GPU, with quantized GGUF weights published on Hugging Face.

Jun 11

Jun 11Thu
  1. OpenRouter BlogAI score74

    OpenRouter Fusion panels beat individual models on the DRACO deep research benchmark

    AIOpenRouter introduced Fusion, a tool that sends a prompt to a panel of models and has a judge model fuse their results into one answer. On 100 DRACO deep research tasks, a Fable 5 and GPT-5.5 panel scored 69.0%, above Fable 5 alone at 65.3%, and a budget panel of Gemini 3 Flash, Kimi K2.6, and DeepSeek V4 Pro reached 64.7% at about half the cost of Fable 5.

    Why it matters: The source gives benchmark scores, panel compositions, and contamination controls, letting readers judge how much of the gain comes from model diversity versus self-synthesis.

Jun 10

Jun 10Wed
  1. Factory NewsAI score58

    Factory launches automated STRIDE-based security review for pull requests in Droid

    AIFactory is rolling out automated security review in Droid, running a STRIDE-based check on every non-draft PR alongside standard code review. Findings include severity, a CWE reference, an explanation, and a suggested fix, posted as inline diff comments. The feature is available today on all plans, and a deeper multi-agent /security-review deep audit is available for full-repository scans.

  2. Zed BlogAI score48

    Zed Unveils DeltaDB, Version Control Built Around Agent Conversations Instead of Commits

    AIZed is building DeltaDB, a version control system that records every operation as a fine-grained delta, linking agent conversations to the code they produce. The company says a beta will arrive in a few weeks, and it invites users to join a waitlist. The system is designed so teammates can collaborate on work in progress without waiting for commits, pull requests, or pushes.

  3. Xiaomi MiMoAI score67

    Xiaomi releases open-source MiMo Code V0.1 terminal coding assistant

    AIXiaomi MiMo has released MiMo Code V0.1, an open-source AI coding assistant for the terminal under the MIT license. It ships with MiMo V2.5, a multimodal model offered free for a limited time with a million-token context window. The tool automatically loads existing Claude Code skills, MCP servers and commands, and reuses API configuration, and it supports providers including Anthropic, OpenAI, DeepSeek, Kimi and GLM.

    Why it matters: The post specifies MiMo Code's Claude Code compatibility and MIT license, which bear directly on whether existing coding-agent setups can migrate without rework.

    Image from @XiaomiMiMo's post
  4. Xiaomi MiMoAI score82

    MiMo Code open-sources a terminal coding agent for long-horizon tasks

    AIXiaomi's MiMo team released MiMo Code, an MIT-licensed terminal coding agent built on OpenCode for long-horizon programming tasks. The design centers on three areas: Max Mode parallel sampling that generates five candidates per turn, Goal-based completion verification, and a memory system that checkpoints session state and rebuilds context. The article reports offline benchmark results and a double-blind A/B test with 1,213 pairs in which MiMo Code's win rate exceeded 65% beyond 200 execution steps.

    Why it matters: The article explains how MiMo Code handles long-horizon coding through computation, checkpointed memory, and cross-session evolution, useful for judging design tradeoffs in coding agents.

Jun 9

Jun 9Tue

Jun 8

Jun 8Mon
  1. Xiaomi MiMoAI score62

    Xiaomi MiMo open-sources a 1T model running over 1,000 tps on 8 GPUs

    AIXiaomi MiMo and the TileRT team say a 1T model exceeds 1,000 tps on a single standard 8-GPU node using general-purpose GPUs. The speedup comes from FP4 quantization and DFlash, a block-masked parallel speculative decoding method that accepts more tokens per verification, with TileRT tailoring its compiler and kernels to these techniques. Open weights for the FP4 + DFlash checkpoint are available on Hugging Face.

  2. Xiaomi MiMoAI score62

    Xiaomi MiMo-V2.5-Pro UltraSpeed claims 1,000+ tokens/s on a 1T model

    AIXiaomi MiMo and TileRT released MiMo-V2.5-Pro-UltraSpeed, which the post says reaches output speeds above 1,000 tokens/s on a 1 trillion parameter MoE model. The post says this runs on a single standard 8-GPGPU node rather than wafer-scale or pure on-chip SRAM hardware. UltraSpeed access is application-based from Jun 8 to Jun 23 (PDT), and the UltraSpeed API costs 3x the standard price.

    Image from @XiaomiMiMo's post

Jun 5

Jun 5Fri

Jun 4

Jun 4Thu

Jun 3

Jun 3Wed
  1. Google LabsAI score31

    Google Labs launches experimental Dreambeans app with personalized daily stories

    AIGoogle Labs has launched Dreambeans, an experimental mobile app that uses Personal Intelligence to connect to users' Google apps and deliver daily collections of personalized stories. The app surfaces relevant topics to help users explore what they care about most. It is available starting today to eligible US-based Google AI Ultra users aged 18 and older, with an open waitlist on the Google Labs website.

    Video from @GoogleLabs's post

Only the first 50 pages are available. Search or browse topics for older items.