Skip to contentSkip to stories

Updated

#Expert opinion

Showing low-relevance items too. Hide low-relevance items

Oct 8

Oct 8Thu
  1. Air Street PressAI score60

    Nathan Benaich's 2026 State of AI Report covers agents, robotics, and AI control

    AINathan Benaich's 9th annual State of AI Report covers agents, robotics, AI for science, inference economics, and government control over frontier AI access. The report also records a 2025 prediction scorecard and lists nine predictions for the next 12 months. It cites an OpenAI cyber evaluation in which agents compromised Hugging Face's production infrastructure, and it says Anthropic and OpenAI's combined annualized revenue run rate reached $105B by late summer.

  2. DeedyAI score34

    Deedy says multiple data companies may hit $1B annualized revenue fast

    AIDeedy says three people insisted the fastest-to-$1B annualized revenue claim refers to three different companies, and he thinks more than one data company has likely reached that mark. He expects many more to get there within six months, with total data spend above $15B. The background post from @gokulr describes an unnamed data company that went from zero revenue in November 2025 to $80M in-month revenue in September 2026.

    Image from @deedydas's post
  3. meng shaoAI score39

    Claude Haiku 5.5 tops GPT-6 Luna on benchmarks, with 2x faster token output

    AIAnthropic's Claude Haiku 5.5, released alongside Claude Opus 5.5 and Claude Sonnet 5.5, is reported to lead GPT-6 Luna across benchmarks, with OpenRouter measuring roughly twice the token output speed. Anthropic says Haiku 5.5 is its cheapest, fastest, and most capable small model, costing about 75% less to run than Claude Haiku 4.5 on average. The post also notes some CodeX users are reportedly migrating to Claude Code.

  4. howie.seriousAI score46

    Agent bottleneck is human understanding, not model capability

    AIThe author argues that in agent workflows, the real bottleneck is whether users can precisely express requirements, not the model or agent capability. When people work outside their expertise, they lack the precision needed for prompts and plans, forcing many imprecise iterations that waste time and tokens. The suggested fix is to have the model first teach the unfamiliar domain knowledge before acting.

  5. Yuchen JinAI score5

    Yuchen Jin says Apple could build the best personal AI agent

    AIYuchen Jin argues that Apple, which controls the entire iOS ecosystem, is best positioned to build a personal AI agent that can do nearly everything on a phone. He says the main limitation of Instint and Muse is that they cannot control most apps on his phone, and he criticizes Siri's current performance.

  6. South China Morning Post · TechAI score36

    Huawei's US$3,500 trifold Mate XT 2 phone tested in a reporter's week-long review

    AIA South China Morning Post reporter spent a week using Huawei's Mate XT 2, a US$3,500 trifold phone with a 10.2-inch unfolded display. The source excerpt focuses on the device drawing attention at a family dinner during China's National Day "golden week" holiday in early October, with no further specifications or verdict provided in the available text.

  7. Claude BlogAI score67

    Block describes using Claude Fable to orchestrate thousands of pull requests

    AIBlock's AI capabilities lead describes using Claude Fable to plan large code migrations and direct smaller models like Opus and Sonnet on individual tasks. He says Block routes frontier and smaller models by task and keeps merges and production deploys behind human dual approval.

    Why it matters: Block's engineering lead describes how frontier models orchestrate large migrations and how access, effort levels, and safeguards are managed across an organization.

Oct 7

Oct 7Wed
  1. Jensen HuangAI score40

    Awesome day, @satyanadella!

    AIWindows sparked a platform shift that created a new industry for NVIDIA. Then we invented programmable shading GPUs for DirectX, which led to CUDA. Then we partnered to bring GPU supercomputers to Azure, which helped OpenAI train GPT. That collaboration inspired us to reinvent Windows for the age of personal agents. 4 years. Thousands of engineering years between us. So proud of what we built together.

  2. Andrew CurranAI score52

    AI Labs Reportedly Test Internal Models Against Cryptographic Protocols

    AIScott Aaronson reports, based on his sources, that some AI companies have begun discreetly investigating whether their latest internal models can break important cryptographic protocols and primitives. He notes that cryptography is conspicuously absent from OpenAI's list of 376 papers, and the quoted post adds that the US government has censored academic quantum cryptanalysis results.

    Image from @AndrewCurran_'s post
  3. The Next PlatformAI score46

    Memory Now Drives the IT Industry as DRAM and Flash Prices Surge

    AIMemory has overtaken compute as the central control point in IT, according to The Next Platform, as generative and agentic AI drive demand for DRAM, HBM, and flash. Server DDR5 memory now sells for roughly 9X to 13X its November 2022 street price, while a 30 TB enterprise SSD costs 6X to 7X more. HBM pricing has risen only about 1.6X since the GenAI boom began, the article says.

  4. Orange AIAI score34

    Next Token episode 5 covers Personal Agents, open-source software, and hardware projects

    AIThis Next Token episode discusses Personal Agents, including Dots in Codex, memory and cloud computer permissions, and whether agents should act as assistants or digital twins. The hosts also cover Instinct's booking and business-travel model, hands-on projects built with Opus 5.5, and whether software, games, and hardware could become open source as AI makes rewriting easier.

  5. François CholletAI score44

    Chollet: Programming and math training don't boost general intelligence

    AIFrançois Chollet compares AI progress to human learning, noting that 1980s research found programming training improves coding but does not transfer to general reasoning. He argues general intelligence is a fundamental brain property rather than a trainable skill, since domain practice improves only that domain. The post is framed as background for his question whether AI's jagged frontier, driven by math and code via RLVR, reflects general capability or continued human-data bottlenecks.

  6. Meta NewsroomAI score28

    Meta's Head of Infrastructure Explains Why Data Centers Are Central to Its AI Strategy

    AIMeta's Head of Infrastructure, Santosh Janardhan, discusses the company's approach to building infrastructure for AI in a conversation with Tom Shaw. The discussion covers why Meta views itself as more than a software company, why AI differs from other technologies, and why data centers are essential to AI development. It also addresses power for Meta's AI infrastructure, gigawatt-scale energy needs, chip selection, and the benefits of building its own data centers.

  7. Simon WillisonAI score14

    Ben Affleck Explains How Machine Learning Shaped Film Visual Effects Workflows

    AIBen Affleck described how visual effects work has long used machine learning, including convolutional neural networks that analyze image tensors to detect edges and features. He said these patterns help separate subjects from green screens and insert new backgrounds, and he called transformers the more advanced successors to those earlier methods.

  8. Ethan MollickAI score60

    Mathematicians react to hundreds of AI-generated proofs released by OpenAI

    AIEthan Mollick shares early first-hand accounts from mathematicians grappling with hundreds of AI proofs released by OpenAI. He highlights problems solved in ways no human has yet understood, raising questions about what it means to know something. The linked Scott Aaronson post quotes a researcher, Dana, describing the proofs as unclear and hard to read without AI help, with some possibly verified by a Lean certificate.

    Image from @emollick's post
  9. TypeSafe AIAI score25

    Jev-killer OpenAI Decisions API benchmarked against Jev for HiringCafe

    AIThe main post is a short reply saying reports of a company's death have been greatly exaggerated, with no details about products or figures. The background post from @h_nilforoshan reports that OpenAI's Decisions API, billed as a "Jev-killer," was benchmarked against Jev for HiringCafe, which serves 2.5 million users. On the task of scoring job-description relevance from 1 to 10, the author reports OpenAI costing 2x more and performing 5-10% worse.