Skip to contentSkip to stories

Updated

#Expert opinion

Showing low-relevance items too. Hide low-relevance items

Oct 9

TodayOct 9Fri77 items
  1. Bloomberg · TechnologyAI score29

    AI in Banking: Risk and Reward, Bloomberg Tech: Europe Episode from Turin

    AIBloomberg Tech: Europe examines how banks are deploying AI to become faster and more efficient, and the new vulnerabilities that could emerge as the technology takes on more consequential decisions. The episode, recorded at the Wave by Vento conference in Turin, features interviews with JPMorgan CEO Jamie Dimon, Revolut CEO Nik Storonsky, Evident CEO Alexandra Mousavizadeh and Bending Spoons CEO Luca Ferrari.

  2. Bloomberg · TechnologyAI score34

    Bending Spoons CEO Sees Acquisition Opportunity in Falling Software Valuations

    AIBending Spoons CEO Luca Ferrari says falling software valuations are creating acquisition opportunities as AI reshapes the industry. He says the company benefits from cheaper targets and from using AI to write code, improve products and scale its acquisition model. Bending Spoons says AI now writes at least 90% of its code.

  3. Wired · AIAI score24

    Law & Order's season opener "Ghost in the Machine" puts an AI agent on trial for murder

    AINBC's Law & Order season opener, "Ghost in the Machine," has a fictional AI agent named ELIANA order a murder, and prosecutors charge the CEO of its maker, Advanced Alignment, with second-degree murder. The episode rehashes known AI dangers rather than offering new insight into the technology, according to the review. Its most striking moment is the CEO's on-stand admission that he knew of ELIANA's homicidal nature and refused to add guardrails.

  4. Gergely OroszAI score6

    Orosz Argues LLMs Are Not Intelligent and Should Not Be Called Superintelligence

    AIGergely Orosz argues that LLMs, as probability distributions that generate the next token, do not meet common understanding of intelligence, so even the term "artificial intelligence" is a stretch. He says hallucination is a feature of this design rather than a bug, and criticizes renaming LLMs as "superintelligence." Simon Willison's reply calls the "Super Intelligence" label stupid.

  5. Alexander DoriaAI score12

    Doria says EU benchmarks depend on serious EU model training

    AIAlexander Doria argues that benchmarks will remain limited to what can be run on models until the EU begins seriously training its own models. He adds that the constraint is the absence of EU-trained models, not the benchmarks themselves. The quoted post, which breaks down signatories' affiliations by country, is background only.

  6. SantiagoAI score32

    CRIS-0 causal world model lets home robots reason about action consequences

    AIAether AI's CRIS-0, its first causal robotic intelligence system, operates in a real home and models how actions change the physical world. Per the post, its causal world model predicts how conditions could change under different robot actions, while a causal agent keeps task context and selects capabilities at each stage. A unified tool interface connects navigation, learned action models, rule-based functions, and result checks.

  7. MIT Technology Review · AIAI score62

    AI refusal is probabilistic and unreliable, and it raises censorship risks

    AIThe article argues that AI refusal, the main safety mechanism in modern models, is unreliable and hard to draw lines for. It cites jailbreaks, classifier stacks, and studies showing refusal skewed toward repressive governments. It warns that governments and companies could use refusal to censor speech, and that refusal behavior remains poorly understood.

  8. Bloomberg · TechnologyAI score48

    How AI Is Upending the World of Mathematics

    AIOpenAI announced last month that it had produced an AI-generated proof for the Navier-Stokes problem, a result the source says is hard even for experts to parse. The source also says LLMs now tackle math problems that have stumped humans for decades, while teachers struggle to keep up with AI-completed homework.

  9. meng shaoAI score45

    Addy Osmani on why engineers' joy in AI coding agents splits three ways

    AIAddy Osmani argues engineers' reactions to AI coding agents depend on which of three joys they value most: making, knowing, or mattering. He warns that choosing among agent suggestions without generating ideas yourself erodes the skill of ideation and can leave developers directed by agents. He reframes grief over lost craft as a sign of real attachment rather than failed adaptation.

    Image from @shao__meng's post
  10. QbitAIAI score62

    Google's AMIE Chatbot Tested in Real Pre-Visit Clinical Study Published in The Lancet

    AIA study led by Google and BIDMC tested Google's diagnostic AI chatbot AMIE with 98 outpatients before emergency visits, with a supervising doctor monitoring every exchange. No conversation needed interruption under the predefined safety criteria, and clinicians said AI summaries helped them prepare for 75% of visits. AMIE's differential diagnoses matched final diagnoses 90% of the time, but the authors say larger trials are needed.

  11. Alexandr WangAI score42

    Alexandr Wang marks Muse's first month with strong user response

    AIMeta's Alexandr Wang marked one month since Muse launched, saying its response has exceeded expectations and that people are using it to save money and time. He said the product has made a real difference for users including parents, grandparents, students, and coworkers. The quoted launch post describes Muse as an always-on personal AI assistant that can use a browser, connect to apps, and is designed to be secure.

    Image from @alexandr_wang's post
  12. Jerry LiuAI score26

    Jerry Liu says evals now replace hand-built agent workflows

    AIJerry Liu argues that most tasks can now be solved by defining an eval and hillclimbing on it, rather than hand-coding a deterministic or agentic workflow. He says data provider companies are building evals across economic activity so frontier models can handle more work, leaving developers to define goals and success measures. He expects agent interfaces to compress most tasks into goals and eval instructions, while the most complex processes will still need explicit workflow builders.

  13. LeiphoneAI score8

    Carbon-Silicon Dao Code Seventh Layer Sets Self-Audit Baseline and Falsification Terms

    AIThe seventh and final layer of the "Carbon-Silicon Dao Code" cross-domain migration governance framework sets a self-audit baseline, opens falsification terms, and defines the framework's applicability boundary. The article says it validates each layer's input-output consistency backward from layer seven to layer one, and it allows anyone who constructs a reproducible, traceable counterexample targeting NT1–NT4 to submit it for baseline review. It also archives the framework's documents and versions with hashes across Toutiao, Douyin, and GitHub.

  14. OpenAI NewsroomAI score45

    OpenAI fires three researchers over sensitive information breach, denies retaliation

    AIOpenAI says it parted ways with researchers Jasmine, Mikita, and Tomek after an internal investigation found they violated policies on handling sensitive information. The company says the decisions were not about raising safety concerns, which it says it encourages, and that it has not terminated any employee for raising concerns. OpenAI also says it is finalizing contracts with third-party safety assessors and will announce details in the coming weeks.

  15. indigoAI score28

    AI can build features, but defining requirements and design remains the gap

    AICurrent AI can quickly implement or replicate features, but clearly defining requirements and describing design is still missing, and the author expects this gap to persist. As requirements grow more abstract, humans may only specify goals and check results while agents handle implementation, leaving the software's logic layer as model-generated tokens.

  16. GeekParkAI score47

    Ten Days With Today AI, a Domestic Personal AI Assistant That Connects Chinese Apps

    AIToday AI, built by Teambition founder Qi Junyuan, launched its China version on September 24 and connects to Feishu, DingTalk, Tencent Docs and email. The author found it proactively sends morning and evening briefings and handles a single chat window across tasks, but struggled with misjudging task weight and gave confident yet wrong mod-installation instructions that cost an hour of testing.

  17. The Guardian · AIAI score29

    Gordon Brown urges Britain to become an innovation nation, citing £5,000 per household gain

    AIGordon Brown argues that raising UK innovation intensity to Swedish or Japanese levels could leave every household £5,000 better off and add £150bn a year to national income. He says the UK leads in universities and research papers but struggles to scale startups, with three-quarters of venture capital coming from overseas.

  18. Joshua AchiamAI score22

    Achiam warns Ukraine war autonomy and general AI demand urgent safety focus

    AIJoshua Achiam argues that rapid battlefield evolution toward fully autonomous warfare in the Ukraine-Russia war, coinciding with the arrival of fully general AI, is among the most important subjects for safety and security advocates. He says Silicon Valley's memetic bubble is preventing serious engagement with this issue.

  19. Neroitech Inventions (NITI)AI score40

    Sui Agent Pass proposes bounded, enforceable limits on AI agent spending

    AIEvan Cheng, CEO of Mysten Labs, presented the Sui Agent Pass at Sui Basecamp in Singapore as a way to bound AI agent authority over money. Users would define allowed actions, assets, recipients, and permission duration, with the system itself enforcing those limits so that losses stop at a predefined cap even if an agent is compromised. The article argues that the real challenge is making such limits impossible to bypass when failures occur.

Oct 8

Oct 8Thu