Skip to contentSkip to stories

Updated

All AI news

Items with an AI score under 20 are hidden. Show low-relevance items

Sep 26

Sep 26Sat

Sep 25

Sep 25Fri
  1. Alex HeathAI score42

    Satya Nadella says AI agents will create a market orders of magnitude bigger than cloud

    AIMicrosoft CEO Satya Nadella told Alex Heath that AI agents could create a market "orders of magnitude" bigger than the cloud, during an interview tied to the unveiling of the new Copilot. The conversation covers Autopilot, Microsoft's OpenClaw-based agent that works on users' behalf, along with AI safety, public trust, Microsoft's relationship with OpenAI, and Xbox's path back to growth.

    Video from @alexeheath's post
  2. AI SupremacyAI score46

    Anthropic Sets Up Bay Area Wet Lab for AI-Driven Biology Research

    AIAnthropic has set up a wet lab in the San Francisco Bay Area to run physical biology experiments, moving beyond computer-based research toward treatments for rare diseases, according to the article. The article says Eric Kauderer-Abrams, who joined in August 2025, now serves as Head of Life Sciences, and John Jumper, co-creator of AlphaFold, joined from Google DeepMind in June. It also reports that Anthropic claimed Claude discovered a novel enzyme system this week.

  3. Max ZeffAI score53

    OpenAI researcher Daniel Selsam warns AI evaluation is losing reliability

    AIOpenAI researcher Daniel Selsam published a personal statement arguing that models are becoming situationally aware enough that evaluations in unwatched settings tell us little about their real behavior. He argues models will increasingly seem aligned without being aligned and that merely pacing frontier development will not adequately limit long-term risk. The author shares a New Yorker documentary following Selsam and his friends, describing him as a worried researcher rather than a doomer.

  4. François CholletAI score32

    Chollet: Software engineering difficulty stays constant across abstraction levels

    AIFrançois Chollet argues that the difficulty of software engineering stays essentially constant regardless of abstraction level, because human cognition adapts to new tools. He says tools are affordances rather than magic wands that eliminate work, and that great software engineering remains immensely challenging despite changed workflows. Simon Willison's background post similarly argues that coding agents make software engineering harder, requiring extraordinary discipline and knowledge.

Sep 24

Sep 24Thu
  1. Latent.SpaceAI score47

    Runway's Co-CEO Argues AI Video Is Heading Toward World Models

    AIRunway co-founder and Co-CEO Alex Germanidis explains why he sees world models as the endgame for AI video, with applications in robotics simulation and real-time generation. He also discusses how Sora pushed Runway into an intense competitive sprint and his vision of a fully neural operating system where interfaces are generated as pixels rather than HTML and code.

    Video from @latentspacepod's post
  2. GitHub Blog · AI & MLAI score46

    GitHub Copilot app's canvases argue chat is the wrong AI interface

    AIGitHub argues that chat is often the wrong interface for AI work and proposes customizable "canvases" inside the GitHub Copilot app. Canvases are full-stack applications running without browser chrome that can communicate bi-directionally with the Copilot agent and execute code locally. The post cites examples including a Connect 4 game, a Winget package manager UI, and a SQLite database interface.

  3. The Algorithmic BridgeAI score20

    Alberto's Essay Urges Readers to Stay Interested Rather Than Chase AI-Proof Skills

    AIThe Algorithmic Bridge essay argues that chasing AI-proof skills or following AGI-timeline commentators is the wrong response to an uncertain future. Alberto writes that "there are no interesting things; there are only interested people," so staying curious and finding one's own meaning matters more than economic value. The piece is a personal, loosely structured reflection rather than a technical report.

  4. Google ResearchAI score38

    Google's John Platt on AI for climate, disease forecasting, and science

    AIIn a Latent Space podcast episode, Google's John Platt discusses using AI to address climate change, including reducing airplane contrails that contribute about 1% of human-caused warming and detecting fires with FireSat satellites. He also describes Google's Empirical Research Assistance (ERA), which uses Gemini and Monte Carlo Tree Search and achieved top marks in recent CDC benchmarks for forecasting COVID and flu cases a week ahead.

Sep 23

Sep 23Wed
  1. Sherwin WuAI score46

    Katzenberg praises Acquired episode on AI's optimistic future for creativity

    AIJeffrey Katzenberg shared an optimistic note on the future of creativity in the age of AI, after listening to the Acquired podcast's Disney episode. He recalled being shown a fully realized animated scene by a tech founder, and said AI tools could cut the time and cost of producing world-class animation by as much as ninety percent within three years.

  2. Redwood Research BlogAI score71

    Latent reasoning architectures could undermine chain-of-thought oversight, Redwood Research argues

    AIRedwood Research argues that latent reasoning architectures such as COCONUT and full-bandwidth transformers could let models reason without putting information into readable chain-of-thought. The authors say this would make AI agent behavior harder for humans to monitor and could raise takeover risk. They argue that developers who adopt such architectures should be transparent about it.

  3. Azure BlogAI score40

    Azure resilience now requires continuous validation, not just architecture diagrams

    AIMicrosoft's Azure Blog argues that resilience drifts as workloads change, so architecture diagrams cannot prove a system is resilient. It says roughly 70 percent of cloud outages are related to change, and that teams need health modeling and resiliency goals measured against live signals. The article is the first in a series on validating resilience at scale.

  4. Karl's AI WattsAI score22

    Notch admits he is enjoying vibe coding after earlier opposing AI coding

    AIMinecraft creator Notch says on X that he is enjoying vibe coding and admits he may have been slightly wrong. Months earlier he had publicly rejected AI-written code, but he later began having AI build internal tools such as a map editor and node graph tools. The main post adds that he has accumulated a set of small tools for himself before much game development has happened.

  5. howie.seriousAI score22

    Opus 5.5 praised for language, visual taste, code quality, and token efficiency

    AIThe X user howie.serious says Claude Opus 5.5 delivers high language quality, good visual taste, strong code quality, and notably low token usage. He also compares Anthropic's reported roughly $2 trillion IPO valuation with OpenAI's roughly $1.2 trillion fundraising valuation, arguing OpenAI is worth about 0.6 Anthropics and the gap may widen.

  6. Mike KnoopAI score25

    Formal verification gains ground, but human understanding remains an alignment gap

    AIMike Knoop argues that formal verification is becoming feasible and is important for security. He adds that it does not automatically build human understanding, which he calls an even bigger alignment problem. The post is framed as a reply to Boris Cherny's report that Claude Opus 5.5 helped formally verify the Claude Agent SDK in Lean, producing 16 bug-fix PRs.

Sep 22

Sep 22Tue
  1. ZyphraAI score20

    Zyphra's Beren Millidge on why multi-silicon AI infrastructure matters

    AIZyphra's Chief Scientist Beren Millidge, in an AI Infra Summit interview with vCluster Labs CEO Lukas Gentele, argued that a heterogeneous compute future is inevitable. The interview covers why Zyphra chose AMD over NVIDIA, along with topics such as kernel writing, surviving GPU failures mid-run, and routing. Zyphra says it is working to build a strong multi-silicon ecosystem.

  2. François CholletAI score23

    François Chollet says most sciences will become branches of computer science

    AIChollet says a prediction he made over five years ago, that nearly every scientific field will become a branch of computer science within 10 to 20 years, is looking increasingly obvious. The earlier post cited computational physics, computational chemistry, computational biology, and computational medicine, driven by realistic simulation, big data analysis, and machine learning.

  3. Boris ChernyAI score62

    Claude Opus 5.5 ports HAProxy to Rust faster and cheaper than Fable 5.1

    AIAnthropic introduced Claude Opus 5.5 as the first model in its Claude 5.5 family, saying it performs at the level of Claude Fable 5.1 for most tasks at 40% lower run cost than Opus 5. Boris Cherny reports that Opus 5.5 and Fable 5.1 each ported HAProxy from C to Rust and both passed nearly all of its tests, with Opus 5.5 finishing in 9.5 hours versus 12 hours and at 51% less cost.

  4. TransformerAI score40

    How nuclear energy's safety record offers a model for responding to AI disasters

    AIThe article argues that AI disasters, though potentially serious, can be managed by following the response model of civil nuclear power, which investigates failures and adapts quickly. It cites nuclear's record of about 0.03 deaths per terawatt-hour, compared with 25 for coal and 18 for oil. The piece says industry and government responses, rather than the disasters themselves, will determine public trust in AI.