Skip to contentSkip to stories

Updated

#Expert opinion

Oct 9

TodayOct 9Fri32 items
  1. Wired · AIAI score24

    Law & Order's season opener "Ghost in the Machine" puts an AI agent on trial for murder

    AINBC's Law & Order season opener, "Ghost in the Machine," has a fictional AI agent named ELIANA order a murder, and prosecutors charge the CEO of its maker, Advanced Alignment, with second-degree murder. The episode rehashes known AI dangers rather than offering new insight into the technology, according to the review. Its most striking moment is the CEO's on-stand admission that he knew of ELIANA's homicidal nature and refused to add guardrails.

  2. Gergely OroszAI score6

    Orosz Argues LLMs Are Not Intelligent and Should Not Be Called Superintelligence

    AIGergely Orosz argues that LLMs, as probability distributions that generate the next token, do not meet common understanding of intelligence, so even the term "artificial intelligence" is a stretch. He says hallucination is a feature of this design rather than a bug, and criticizes renaming LLMs as "superintelligence." Simon Willison's reply calls the "Super Intelligence" label stupid.

  3. SantiagoAI score32

    CRIS-0 causal world model lets home robots reason about action consequences

    AIAether AI's CRIS-0, its first causal robotic intelligence system, operates in a real home and models how actions change the physical world. Per the post, its causal world model predicts how conditions could change under different robot actions, while a causal agent keeps task context and selects capabilities at each stage. A unified tool interface connects navigation, learned action models, rule-based functions, and result checks.

  4. MIT Technology Review · AIAI score62

    AI refusal is probabilistic and unreliable, and it raises censorship risks

    AIThe article argues that AI refusal, the main safety mechanism in modern models, is unreliable and hard to draw lines for. It cites jailbreaks, classifier stacks, and studies showing refusal skewed toward repressive governments. It warns that governments and companies could use refusal to censor speech, and that refusal behavior remains poorly understood.

  5. Bloomberg · TechnologyAI score48

    How AI Is Upending the World of Mathematics

    AIOpenAI announced last month that it had produced an AI-generated proof for the Navier-Stokes problem, a result the source says is hard even for experts to parse. The source also says LLMs now tackle math problems that have stumped humans for decades, while teachers struggle to keep up with AI-completed homework.

  6. meng shaoAI score45

    Addy Osmani on why engineers' joy in AI coding agents splits three ways

    AIAddy Osmani argues engineers' reactions to AI coding agents depend on which of three joys they value most: making, knowing, or mattering. He warns that choosing among agent suggestions without generating ideas yourself erodes the skill of ideation and can leave developers directed by agents. He reframes grief over lost craft as a sign of real attachment rather than failed adaptation.

  7. Alexandr WangAI score42

    Alexandr Wang marks Muse's first month with strong user response

    AIMeta's Alexandr Wang marked one month since Muse launched, saying its response has exceeded expectations and that people are using it to save money and time. He said the product has made a real difference for users including parents, grandparents, students, and coworkers. The quoted launch post describes Muse as an always-on personal AI assistant that can use a browser, connect to apps, and is designed to be secure.

  8. Jerry LiuAI score26

    Jerry Liu says evals now replace hand-built agent workflows

    AIJerry Liu argues that most tasks can now be solved by defining an eval and hillclimbing on it, rather than hand-coding a deterministic or agentic workflow. He says data provider companies are building evals across economic activity so frontier models can handle more work, leaving developers to define goals and success measures. He expects agent interfaces to compress most tasks into goals and eval instructions, while the most complex processes will still need explicit workflow builders.

  9. LeiphoneAI score8

    Carbon-Silicon Dao Code Seventh Layer Sets Self-Audit Baseline and Falsification Terms

    AIThe seventh and final layer of the "Carbon-Silicon Dao Code" cross-domain migration governance framework sets a self-audit baseline, opens falsification terms, and defines the framework's applicability boundary. The article says it validates each layer's input-output consistency backward from layer seven to layer one, and it allows anyone who constructs a reproducible, traceable counterexample targeting NT1–NT4 to submit it for baseline review. It also archives the framework's documents and versions with hashes across Toutiao, Douyin, and GitHub.

  10. GeekParkAI score47

    Ten Days With Today AI, a Domestic Personal AI Assistant That Connects Chinese Apps

    AIToday AI, built by Teambition founder Qi Junyuan, launched its China version on September 24 and connects to Feishu, DingTalk, Tencent Docs and email. The author found it proactively sends morning and evening briefings and handles a single chat window across tasks, but struggled with misjudging task weight and gave confident yet wrong mod-installation instructions that cost an hour of testing.

  11. The Guardian · AIAI score29

    Gordon Brown urges Britain to become an innovation nation, citing £5,000 per household gain

    AIGordon Brown argues that raising UK innovation intensity to Swedish or Japanese levels could leave every household £5,000 better off and add £150bn a year to national income. He says the UK leads in universities and research papers but struggles to scale startups, with three-quarters of venture capital coming from overseas.

Oct 8

Oct 8Thu