Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 1

Oct 1Thu
  1. Guillermo RauchXAI score38

    Guillermo Rauch says verification engineering is the future of software

    AIGuillermo Rauch argues that the future is verification engineering, spanning proofs, end-to-end tests, benchmarks, and linters. He expects some of these tests to be deterministic and others agentic, and he says the approach looks great. The quoted post introduces e2e, an open-source agentic testing framework that mixes deterministic and agentic APIs and runs locally or in CI.

  2. Vaibhav (VB) SrivastavXAI score12

    OpenAI shares a dots demo with improved WiFi

    AIOpenAI's Vaibhav Srivastav posted a short "we're so back" message, linking a demo of "dots" that the OpenAIDevs account described as now running with better WiFi. The post gives no technical details about what dots is, its capabilities, or its performance figures.

  3. ZyphraOfficialAI score20

    Zyphra's Results Explain How NoPE Models Encode Position

    AIZyphra says its results clarify how state-of-the-art NoPE models encode position and which inductive biases support generalization. It adds that global NoPE could enable models to extrapolate to contexts longer than those seen in training, potentially indefinitely.

  4. Dwarkesh PatelXAI score16

    Dwarkesh Patel's podcast revisits Cortés and Pizarro's conquests of Aztec and Inca empires

    AIDwarkesh Patel released a new episode with military historian Si Sheppard on the Spanish conquests of the Aztec and Inca empires. The post highlights Cortés's conquest of the roughly 6 million-strong Aztec empire within about two and a half years, and Pizarro's subsequent conquest of the Inca Empire of some 10 million people, with the Conquistadors' forces made up largely of native allies.

    Video from @dwarkesh_sp's post
  5. Philipp SchmidXAI score25

    Unlimited human usage paired with capped agent usage

    AIThe post notes a new pricing pattern where human use is unlimited while agent usage is limited. It remarks that this is the first time the author has seen such a split, without naming the product or provider.

    Image from @_philschmid's post
  6. Dwarkesh PodcastBlogAI score54

    Si Sheppard on how a few hundred Spanish soldiers toppled the Aztec and Inca empires

    AIDwarkesh Patel interviews military historian Si Sheppard about how a few hundred Spanish conquistadors defeated the Aztec and Inca empires in the 1500s. The episode covers Cortés's conquest of the Aztecs, Pizarro's conquest of the Inca, and the role of horses, steel, diplomacy, and disease. It is a history episode, with the AI takeover comparison raised only as a framing.

  7. TransformerBlogAI score38

    Democrats struggle to agree on a unified AI regulation platform

    AIDemocrats are pushing to make AI regulation a central campaign issue, but the party lacks a unified set of proposals. Lawmakers range from those focused on existential risk, such as Sanders and Casar's bill to ban superintelligent AI until a regulator exists, to those prioritizing workforce, environmental, and corporate-power concerns. Public AI adoption is high, yet attitudes toward it are largely hostile.

  8. The SequenceBlogAI score38

    The Sequence explores creating a futures market for AI compute capacity

    AIThe Sequence argues that AI compute could become a commodity that requires a futures market, because unused GPU-hours cannot be stored and suppliers and buyers face forward-price risk. The piece says realizing this requires defining what is traded, measuring its quality, and building contracts around its risks, drawing on commodity market history.

  9. O'Reilly RadarBlogAI score38

    Conversational AI Interfaces May Matter More Than Full Autonomy for Software

    AIRobert Englander argues that natural language interfaces built on top of deterministic software may prove more valuable than fully autonomous AI agents. He contends that large language models excel at interpreting human intent, while systems of record must still provide the reliability, consistency, and accountability that probabilistic models lack.

  10. One Useful Thing (Ethan Mollick)BlogAI score62

    Ethan Mollick Says Agent Coordination Is Easier Than Expected

    AIEthan Mollick says he was wrong to think coordinating AI agents would require careful human-designed management structures. He points to personal agents like dots and Muse, and to a swarm of thousands of OpenAI agents that solved a Navier-Stokes problem in 88 hours with thin coordination. He argues many management problems stem from human limits, which agents lack, so people should mainly guide direction while agents handle organizing.

  11. StratecheryBlogAI score16

    Stratechery Interview With Jason Del Rey on Amazon, Meta, and Walmart Retail Rivalry

    AIStratechery has published an interview with journalist Jason Del Rey comparing Amazon and Meta, extending the long-running retail rivalry between Amazon and Walmart. The provided text contains only the interview introduction and Stratechery subscription material, so no further details on Muse or specific findings are available.

  12. indigoXAI score8

    Indigo outlines three strategies for bosses: automate work, build networks, invest

    AIThe author, @indigox, shares three linked strategies for business owners: automating company work with an Agent Loop, building personal presence and influence, and investing earnings in next-generation tech companies, mainly US stocks. The post presents these as a complete future plan, with no supporting details or figures.

Sep 30

Sep 30Wed
  1. Hamel HusainXAI score4

    Hamel Husain thanks Lance Martin for updating Claude eval skill

    AIHamel Husain posted a short thank-you emoji reply to Lance Martin's update on an evaluation skill for Claude. Martin says the latest skill now instructs Claude to build a viewer for eval examples, but it does not walk users through the data first, which he agrees could help them prioritize which evals to write.

  2. Allie K. MillerXAI score33

    Jony Ive's OpenAI hardware rumors evolve from voice device to donut

    AISpeculation about Jony Ive's secret OpenAI hardware has shifted from a voice-based desk device in May 2024 to a wearable pin, then a pen, and now a 3D donut with a swivel piece. The post points to the Dots logo as a donut, suggesting the design has moved toward that shape. The author notes that the rumors remain unconfirmed.

  3. hardmaruXAI score26

    David Ha argues AI's future lies in orchestrating multiple models, not bigger ones

    AIIn a Nikkei Asia op-ed, Sakana AI co-founder and CEO David Ha argues that as frontier scaling slows, the next phase of AI will center on mastering orchestration rather than building ever-larger models. He contends that value shifts from individual model weights to intelligence that coordinates multiple models for different situations, and that sovereign AI rests on supply-chain resilience rather than national isolation.

  4. Sakana AIOfficialAI score33

    Sakana AI's David Ha argues the future of AI lies in orchestrators

    AISakana AI co-founder and CEO David Ha published a Nikkei Asia op-ed titled "The future of AI belongs to the orchestrators." He argues that ever-larger models face limits, as open models close the gap within months and frontier inference costs can exceed the hourly wage of the people they assist. He also contends that sovereignty means supply-chain strength, not national isolation.

  5. Guillermo RauchXAI score22

    Vercel AI Gateway rejects dubious token promos, prioritizing trustworthy providers and data privacy

    AIVercel says its AI Gateway turns down "free token" promotions from companies making dubious Zero Data Retention claims, prioritizing the best providers over provider count. The company argues it has the largest trustworthy view of global AI token flows, citing real usage from 400k+ paying customers, thousands of enterprises, and zero markup. It says it invests as heavily in legal, compliance, privacy, and back-office operations as in engineering to serve trillions of tokens daily.

  6. Hamel HusainXAI score38

    Hamel Husain Reviews Claude's New Auto Eval Plugin for Evaluations

    AIHamel Husain has published a longer review of a new Claude Auto Eval plugin after many users asked about it. He invites readers to share their experiences using the plugin and how it went for them. The plugin is part of Claude's ability to help build evaluations and hillclimb on them, as described by @ClaudeDevs.

    Image from @HamelHusain's post
  7. Orange AIXAI score8

    Dacongming shares how to overcome nihilism, drawing on his own experience

    AIThe post says that Dacongming, who two years ago described a sense of emptiness, has since become one of the most fulfilled people the author knows. It presents his experience as guidance for a growing number of people who feel nihilistic, starting with facing that feeling directly. The post is a promotion for an episode, with no concrete AI details.

  8. Nathan LambertXAI score26

    Nathan Lambert welcomes Google's Gemini 4 as frontier competition

    AINathan Lambert says he is pleased to see Google surprise people with Gemini 4. He argues that more labs at the frontier benefit consumers through competition and reduce the concentration of power. He says he is eager to see how the model performs in real-world scenarios.

  9. Dongxi NLPXAI score18

    Gemini 4 Argon release highlighted among September frontier model launches

    AIDongxi NLP jokes that frontier AI labs now release new models every few days, calling it their pastime rather than a boast. The post lists a September release calendar including Grok 4.7, GPT-6 Sol and Luna, Claude Opus 5.5, Claude Sonnet 5.5, GPT-6.1 Sol, and Gemini 4 Argon, which it says launched today.

  10. NewcomerBlogAI score38

    Machine Earning Summit Debates Personal AI Agents and Agentic Commerce in San Francisco

    AIPersonal agents dominated the Machine Earning AI Summit in San Francisco, where founders and investors debated how AI agents will reshape finance and commerce. Speakers predicted that people will spend 40% of their digital time using assistants within a year, rising to 90% within five years, according to Town CEO Jean-Denis Greze. Panelists also stressed that consumers remain uncomfortable letting agents make purchases directly, with guardrails such as spend limits still being built.