Skip to contentSkip to stories

Updated

All AI news

Items with an AI score under 20 are hidden. Show low-relevance items

Sep 16

Sep 16Wed
  1. Greg BrockmanAI score62

    Databricks rolls out Astra to all engineers, reports 60% higher coding spend

    AIDatabricks rolled out Astra to every engineer, about 3,500 people, after a pilot with around 200 users. Engineers given Astra increased coding spend by roughly 60% compared to baseline. The company reports Astra outperforms Opus 5 and Sol 5.6 on highly complex system design tasks, but sees no clear gain on medium or low complexity coding. Astra gets a separate sub-budget in Unity Gateway to encourage selective use.

  2. Matei ZahariaAI score58

    Databricks reports engineering measurements from rolling out Astra to 3,500 engineers

    AIDatabricks rolled out Astra to all of its engineers and reported internal measurements. Engineers given Astra increased overall coding spend by around 60% compared with baseline, and Astra outperformed prior top models on highly complex system design tasks, while the gain on medium or low complexity tasks was unclear.

Sep 15

Sep 15Tue
  1. Lewis Tunstall @ COLM 🌉AI score30

    Periodic Labs advances toward cracking condensed matter physics superconductor problem

    AIPeriodic Labs, the team behind high-throughput materials labs in Menlo Park, reports progress on one of condensed matter physics' hardest problems. Its open-source model Neon, trained with mid-training and RL on 1,300 H200s plus months of lab data, surpasses GPT-6 Astra on the company's analysis benchmark. The work targets materials science challenges including superconductors, magnets, and semiconductors.

  2. Sundar PichaiAI score42

    Google outlines AI for science, weather, languages, and economic research

    AIGoogle says it is focusing AI efforts on health, disaster and weather resilience, learning, and economic opportunity. Recent examples include AlphaGenome Atlas, which maps all 9B possible single-letter genetic changes across the human genome and is openly available to researchers, and WeatherNext 3, described as its most accurate and capable global weather AI model to date. The post also cites AI & Economy ATLAS, an open-access look at global AI usage, and says its translation services now cover nearly 300 languages spoken by 7B people.

    Image from @sundarpichai's post
  3. Cognition Blog (Devin, Windsurf)AI score60

    Cognition and AWS sign multi-year deal to deploy Devin for enterprise modernization

    AICognition and AWS have entered a multi-year Strategic Collaboration Agreement to help enterprises deploy the Devin autonomous engineer in production. Devin can be purchased through AWS Marketplace, and the companies are exploring deeper engineering integrations within customers' AWS environments. Mercedes-Benz reportedly used Devin to analyze more than 200,000 lines of COBOL, reducing an estimated eight-month modernization project to eight days.

    Why it matters: The collaboration shows how an autonomous coding agent is being packaged for enterprise legacy modernization inside existing AWS environments, with concrete customer migration figures.

  4. RadixArkAI score42

    Periodic Labs builds Neon on SGLang and Miles for 2.5x faster inference

    AIPeriodic Labs chose SGLang and Miles to build Neon, an open-source model it says surpasses GPT-6 Astra on its analysis benchmark after mid-training and RL on 1,300 H200s. RadixArk says Periodic extended both frameworks for scientific RL at trillion-parameter scale, delivering more efficient training, lower memory use, and 2.5x faster inference. The work has been contributed back to both projects.

  5. Baseten BlogAI score40

    LangChain uses Baseten Loops to train custom models for LangSmith Engine

    AILangChain is using Baseten Loops, a managed fine-tuning service, to train custom models for LangSmith Engine, its in-platform agent that debugs and improves AI agents. The article says LangChain fine-tunes large open-weight models on agent traces and trains smaller open-weight models such as Qwen for tasks like failure-mode categorization. Baseten Loops supports supervised fine-tuning, reinforcement learning, and long-context workloads, and lets checkpoints be evaluated and deployed directly to inference.

  6. Air Street PressAI score39

    Air Street Capital leads $40 million Series A in Jack & Jill, an AI career agent platform

    AIAir Street Capital led Jack & Jill's $40 million Series A, with Madrona joining and Creandum and Entrepreneurs First investing again, following a $20 million seed round less than a year earlier. Jack & Jill uses AI agents named Jack, which helps candidates plan career moves and search job postings, and Jill, which helps companies recruit from opted-in candidates. The company says it has arranged 25,000 interviews and plans 5,000 more each month.

Sep 14

Sep 14Mon
  1. Factory NewsAI score40

    Factory raises $200M at $5B valuation to scale self-improving enterprise software development

    AIFactory has raised $200M at a $5B valuation from investors including Blackstone, Khosla Ventures, and Sequoia Capital, bringing its total funding to over $400 million. The company says it will use the capital to accelerate research, product, and global go-to-market efforts. Factory says hundreds of thousands of developers use its platform, with customers including Nvidia, Blackstone, and T-Mobile.

  2. Waymo BlogAI score40

    Waymo Partners With Allianz Partners on Insurance and Claims Foundation for European Expansion

    AIWaymo is partnering with Allianz Partners to build insurance, claims, and safety research infrastructure for its planned autonomous ride-hailing expansion in Europe, starting with London and Munich. The collaboration will provide tailored fleet insurance, liability protection, and digital claims handling, plus joint crash analysis and safety modeling research. Waymo cites a 16x reduction in serious injury crashes compared to human drivers in cities where it operates.

  3. InferactAI score42

    Inferact and Google Cloud partner to make TPUs first-class in vLLM

    AIInferact and Google Cloud announce a partnership to make Google TPUs a first-class platform in the vLLM open-source project. The collaboration targets production serving features, optimized kernels, a native PyTorch path via TorchTPU, and day-0 support for frontier model releases. A community program will offer shared TPU capacity and review and design help from vLLM core maintainers, with all outputs released as open source.

    Image from @inferact's post

Sep 13

Sep 13Sun
  1. Waymo BlogAI score44

    Waymo, GO and Nihon Kotsu plan 2027 driverless ride-hailing in Tokyo

    AIWaymo, Nihon Kotsu and GO have agreed to prepare a commercial, fully autonomous ride-hailing service in Tokyo, targeting first public rides in 2027. The service will be available through the GO and Waymo apps, starting with an initial fleet and expanding to around 100 vehicles across key Tokyo neighborhoods. The launch depends on regulatory approval from national and local authorities and completion of ongoing validation.

Sep 12

Sep 12Sat
  1. Sam BowmanAI score46

    Sam Bowman calls for Anthropic-style third-party AI safety access elsewhere

    AISam Bowman says ongoing accountability could open valuable safety possibilities and he would like to see similar arrangements elsewhere. The context is Dario Amodei's announcement that Anthropic will give third-party evaluators permanent, employee-level access to its systems to verify safety measures, report incidents, and assess model alignment during training.

    Image from @sleepinyourhat's post

Sep 11

Sep 11Fri

Sep 10

Sep 10Thu
  1. The Register · AIAI score36

    Oracle Says AI Will Strengthen, Not Replace, Its Application Software Business

    AIOracle co-CEO Mike Sicilia argued on the Q1 FY 2027 earnings call that AI accelerates rather than replaces packaged applications, and promised an "agentic AI accelerator" in October to compress SaaS deployments from years to months. Oracle's SaaS revenue grew 10 percent, while software license revenue fell 3 percent to $5.5 billion and cloud revenue rose 60 percent to $11.6 billion. The company also said GPUs coming up for renewal were renewed or resold at a 20 percent premium and that it turned on 850 MW of new datacenter capacity in the quarter.

  2. Understanding AI (Timothy B. Lee)AI score78

    OpenAI's AI-driven Navier-Stokes result draws anger from mathematicians

    AIOpenAI announced that a swarm of 10,000 agents produced a solution to the Navier-Stokes Millennium Problem, a result that angered mathematicians. NYU mathematician Tristan Buckmaster and Anthropic-employed collaborator Levent Alpöge had been working on related problems and released three draft papers of about 245 pages. Buckmaster said OpenAI's offer to merge efforts required acknowledging an OpenAI model and excluded Alpöge as co-author.

    Why it matters: The piece separates the mathematical result from the collaboration dispute, showing how AI labs' compute spending is straining academic norms around credit and openness.

  3. Cognition Blog (Devin, Windsurf)AI score22

    Cognition Welcomes Dioxus Team to Advance Open-Source Cross-Platform App Framework

    AICognition has welcomed Jonathan Kelley and the Dioxus team, whose framework Cognition used extensively to build and improve Devin's performance. Cognition plans to continue supporting Dioxus, Blitz, Taffy, and Subsecond while increasing investment in Dioxus-Native and Blitz. The Dioxus team will also work on Devin's virtual machine, computer use skills, and testing capabilities.

  4. Mistral AIAI score36

    Cloudera and Mistral Partner to Deliver Sovereign AI on Enterprise Data

    AIMistral AI and Cloudera announced a partnership that integrates Mistral's models with Cloudera's hybrid data platform for enterprise AI. Customers can run inference across private and public cloud, on-prem, and fully air-gapped environments, and train custom models on proprietary data while retaining ownership. Cloudera cited 30 exabytes of customer-managed data on its platform.

  5. DeepSeekAI score37

    DeepSeek to support V4.1-Flash open-source inference and large-scale deployments

    AIDeepSeek says it will work with the open-source community on inference support for DeepSeek-V4.1-Flash and explore more deployment options. The company is inviting organizations planning large-scale deployments with 2,000 GPUs and a storage cluster to get in touch. The model and a technical report are published on Hugging Face.

Sep 9

Sep 9Wed