Skip to contentSkip to stories

Updated

#Industry news

Showing low-relevance items too. Hide low-relevance items

Oct 8

Oct 8Thu
  1. elvisXAI score48

    Google's FlowAgent auto-repairs failing tests inside code review

    AIGoogle proposed FlowAgent, a ReAct-style agent that generates and validates fixes for pre-submit test failures and shows them in its code review tools. Two abstention filters, before and after execution, suppress weak suggestions; in a manual review of 195 real failures, 67.18% of fixes were correct. After the Google-wide launch, it suggested fixes on 295,508 changes, with developers previewing 65,069 and applying 28,554.

    Image from @omarsar0's post
  2. Sherwin WuXAI score60

    Harvey LAB-AA v1.1 adds hallucination gate; Grok 4.7 leads at 9.4%

    AISherwin Wu, an OpenAI employee, says the updated Harvey LAB-AA v1.1 benchmark, announced by Artificial Analysis with Harvey, is more useful than the original LAB results. The new Hallucination-Gated All-Pass Rate credits a task only when every rubric criterion passes and no material hallucination appears. Grok 4.7 (xhigh) leads at 9.4%, while GPT-6 Astra (max) at 8.6% has very few material hallucinations.

    Why it matters: The update adds a hallucination gate to a legal benchmark, showing that models with high all-pass rates can rank much lower once material errors count.

  3. GoodfireOfficialAI score25

    Goodfire launches a challenge to build AI models on a dataset

    AIParticipants will use a dataset to build AI models, evaluated through a series of evals ranging from general benchmarks to more complex tasks. Top teams will be shortlisted and have their experimental hypotheses tested in Prima Mente's wet lab.

  4. GoodfireOfficialAI score44

    Alzheimer's Translation Challenge Built on 150M-Cell Atlas

    AIThe Alzheimer's Translation Challenge is built on a new atlas of 150M cells, covering neurons, astrocytes, and microglia across different genetic backgrounds under combinatorial perturbations with multi-modal readouts. The data will be made available through the AD workbench and Prima Mente's modeling platform.

  5. The DecoderNewsAI score62

    Anthropic's updated usage policy bans sustained abusive behavior toward Claude

    AIAnthropic has updated Claude's usage policy for the first time in over a year, banning sustained and needless abusive or cruel behavior toward Claude. The company says ordinary frustration, pushback, dark creative themes, and model testing are not covered, and that the rule applies only in extreme cases. Violations can lead to warnings, throttling, restriction, suspension, or termination of access.

  6. Andrew CurranXAI score62

    Three fired OpenAI safety researchers publish open letter to leadership

    AIThree OpenAI safety and alignment employees, Tomek Korbak, Jasmine Wang, and Mikita Balesni, were fired last week and have published an open letter to OpenAI's safety and governance committees. The letter argues that OpenAI cannot make AI safe on its own, calls for open debate, third-party collaboration, and clear internal procedures, and says the firing and its handling bear directly on safety oversight.

    Image from @AndrewCurran_'s post
  7. ZDNet · AINewsAI score50

    Microsoft 365 Family and Premium plans will switch to shared storage and AI credit pools

    AIMicrosoft will replace per-user OneDrive allowances on Microsoft 365 Family and Premium plans with a shared 2 TB storage pool, down from 6 TB across six accounts. Heavy users could face bills that double or triple, since each extra 1 TB adds $10 a month, while family members will be able to share a single AI usage allowance. New and upgraded subscriptions get shared storage starting Oct. 8, 2026, and existing subscribers move at their first renewal on or after May 2, 2027.

  8. Gizmodo · AINewsAI score18

    REK Stages Human-Versus-Robot Fights, Then Puts Swords on Robots

    AIRobot Entertainment Kombat, founded by Cix Liv, hosted a September 18 San Francisco event where a human fought a remotely controlled robot and lost each match, leaving with a hand injury. The California State Athletic Commission sent a cease-and-desist, saying the fights were unsanctioned and require approval before any bout involving a human. REK plans a robot deathmatch later this month with remotely controlled bots equipped with weapons.

  9. 🚨 AI News | TestingCatalogXAI score36

    Gemini Agent for Business may add Claude Opus 5 and Sonnet 5.5

    AIGoogle's recently announced Gemini Agent for Gemini Business is reportedly set to offer Gemini Argon 4, Gemini Flash 3.8, Claude Opus 5, and Claude Sonnet 5.5. If accurate, it would mark the first time Claude models appear on Google's platform alongside Google's own models, which the post frames as a way for Google to compete for enterprise customers.

    Video from @testingcatalog's post
  10. CNBC · TechnologyNewsAI score50

    US suspends Microsoft, Adobe from green card labor program amid foreign worker crackdown

    AIThe U.S. Department of Labor said it suspended Microsoft and Adobe from its Permanent Labor Certification program, citing multiple active federal investigations. Labor Secretary Keith Sonderling also said no new applications will be accepted for Cognizant, Infosys, Capgemini, Tata, Wipro and HCL. Microsoft said the vast majority of its U.S. employees are Americans and that 80% of its roughly 6,000 H-1B petitions last fiscal year were to extend or change the status of existing employees.

  11. Arena.aiOfficialAI score44

    Arena raises $200M Series B led by Lightspeed, launches Alignment Index

    AIArena has secured a $200M Series B, with Lightspeed doubling down on its investment. The company is also launching the Alignment Index, which measures how closely AI behavior aligns with human values in real-world settings. Arena reports annualized revenue above $100M since its Series A, with millions of people helping evaluate frontier models through real-world use.

  12. Artificial AnalysisOfficialAI score34

    Harvey LAB-AA: Artificial Analysis benchmark for legal AI agents

    AIArtificial Analysis has released Harvey LAB-AA, an evaluation built on Harvey's LAB dataset and developed in collaboration with Harvey. Full results are published on the Artificial Analysis evaluations page, alongside Harvey's commentary on the benchmark and human expert preferences.

  13. Artificial AnalysisOfficialAI score28

    Artificial Analysis Pareto frontier: GPT-6 Luna cheapest per task at $0.22

    AIAmong models with a Hallucination-Gated All-Pass Rate above 0%, GPT-6 Luna (max), GPT-6.1 Sol (max), Muse Spark 1.3 (max), and Grok 4.7 (xhigh) set the Pareto frontier for score versus cost per task. GPT-6 Luna (max) is the cheapest at about $0.22 per task, scoring 3.3%, while Grok 4.7 (xhigh) leads at about $9.50 per task and Muse Spark 1.3 (max) costs about $4.20. The three Claude models cost about $18 to $22 per task.

    Image from @ArtificialAnlys's post
  14. Satya NadellaXAI score11

    Satya Nadella thanks Trump for honor, pledges tech collaboration

    AIMicrosoft CEO Satya Nadella thanked President Trump for an honor bestowed at an event alongside prominent American innovators. He said he looks forward to continuing work together to advance technology and drive American prosperity. The quoted remarks from Trump credit Nadella with decades of transforming Microsoft.

  15. TechCrunch · AINewsAI score36

    Ben Affleck's AI expertise goes viral as he explains neural networks and fine-tuning

    AIActor Ben Affleck drew attention this week for explaining machine learning concepts, including convolutional neural networks, tensors, and transformers, in several recent interviews. He said he fine-tuned open video models by unfreezing weights and training only the last cinematic layer, using a dataset he built over about eight months for his startup. Affleck said he worries about students and learned helplessness more than Skynet, and predicted AI will be additive to the movie business.

  16. TechCrunch · AINewsAI score46

    Arena raises $200M at $3.1B valuation, nearly doubling in 10 months

    AIArena, the crowdsourced AI model leaderboard that started as a UC Berkeley research project, raised a $200 million Series B at a $3.1 billion valuation, led by Lightspeed Venture Partners and Khosla Ventures. The company said it reached $100 million in annualized run-rate revenue in June, up from $30 million when it raised its $150 million Series A in January at a $1.7 billion post-money valuation.

  17. TechCrunch · AINewsAI score58

    OpenAI's annualized revenue reportedly about $20 billion below earlier estimates

    AIOpenAI has reportedly told investors its annualized revenue is approaching $50 billion, about $20 billion below a previously reported $70 billion figure. The Financial Times reports the earlier number came from investor attempts to compare OpenAI with Anthropic, which counts cloud partners' sales differently. OpenAI's IPO has reportedly been pushed to early 2027.

  18. The DecoderNewsAI score80

    Mathematicians call for OpenAI boycott after AI-generated proofs flood the field

    AIThe Association of Historical Mathematicians (AHM) has called for a boycott of OpenAI after the company released more than 700 AI-generated proof files at once. Fields Medalist Terence Tao, who chairs the group, argues that AI solving open problems autonomously reduces seminars, collaborations, and fertile research directions, and that the field should shift its measure of progress toward explanation and community-building.

    Why it matters: The article links the AHM boycott call to Tao's argument that AI-driven proof volume is changing how mathematicians measure progress and whether solutions remain useful.

  19. CNBC · TechnologyNewsAI score38

    Trump's August Disclosure Shows Up to $25 Million in Meta and $5 Million in SpaceX Debt

    AITrump disclosed more than 500 securities transactions in August, including a purchase of up to $25 million in Meta stock and up to $5 million in SpaceX senior unsecured notes. The filing, which reports trades in value ranges, shows total activity of roughly $74.3 million to $273.3 million according to a CNBC analysis. The SpaceX notes were bought two days before Trump signed a national space transportation policy, and the White House says the portfolio is independently managed.

  20. TechCrunch · AINewsAI score65

    OpenAI's math solutions fall short of the field's standards, mathematicians say

    AIOpenAI released hundreds of claimed solutions to hard math problems but did not fully meet guidelines from the Advisory Group on Mathematics and Artificial Intelligence. Only 10 of 719 manuscripts included chain-of-thought releases, and just 42% of proofs were formalized. A Cambridge and King's College paper found discrepancies between a natural language proof and its Lean code for a Navier-Stokes-derived problem.

  21. SiliconANGLE · AINewsAI score38

    Automation Anywhere to acquire Boost.ai to expand customer-facing voice AI

    AIAutomation Anywhere Inc. announced an agreement to acquire Boost.ai Inc., a conversational voice AI company, from Nordic Capital, to extend its autonomous enterprise platform into customer experience. Boost.ai supports more than 36 languages, serves hundreds of customers in regulated industries and Europe, and maintains more than 650 deployments and about 600 live AI agents. The deal follows Automation Anywhere's late 2025 acquisition of Aisera Inc.

  22. OpenRouterOfficialAI score18

    Ori AI sales agent cuts deal cycles 34% and raises close rate 2.6x

    AIOpenRouter reports that its Ori AI sales agent shortened deal closing from 71 to 47 days, a 34% reduction, and raised close rate 2.6x. The post says Ori automatically switches to lower-cost models while maintaining quality, so costs fall over time. Ori Slack agents are slated to roll out to all users soon.

  23. The Verge · AINewsAI score44

    USA Today Co. sues OpenAI, seeking over $250 million for copyright infringement

    AIUSA Today Co. and several local newspapers it owns sued OpenAI, alleging the company copied "hundreds of thousands" of articles to train its AI models without permission. The publisher seeks damages of more than $250 million, arguing the unauthorized use has caused real and continuing harm. OpenAI did not immediately respond to The Verge's request for comment.

  24. The Verge · AINewsAI score30

    SpaceXAI Backs Omarchy Linux Distro With $1.5 Million in Grok Tokens

    AISpaceXAI is joining the Omacom Foundation, which oversees the Omarchy Linux distribution, as a Founding Corporate Patron and donating $1.5 million worth of Grok tokens to the project. According to David Heinemeier Hansson's blog post, the tokens will primarily accelerate development, review code, and patch bugs. The partnership follows earlier controversy over Hansson's anti-immigration posts, which have drawn criticism of Omarchy's corporate contributors, including 1Password and Cloudflare.

  25. SemiAnalysisBlogAI score72

    SemiAnalysis argues China's AI safety regime is speed-first, not frontier-focused

    AISemiAnalysis argues China's real AI safety approach prioritizes rapid development, regulating AI applications and outputs rather than frontier models. Its dataset of 857 releases from nine Chinese developers found only 31 (3.6%) with any published safety result, and only 9 available at launch. The author also reports that technical experts favor binding frontier rules, but none of their demands has been adopted in binding Chinese instruments.

    Why it matters: The piece tests China's stated AI safety position against its releases, statements, and rules, offering a checkable case for how US pacing debates should read Beijing.

  26. Miles BrundageXAI score22

    Miles Brundage suspects Anthropic's Claude abuse policy aims at IPO and regulatory capture

    AIMiles Brundage speculates that Anthropic's new rule, making abusive behavior toward Claude a Usage Policy violation effective November 12, 2026, is meant to help its IPO and win favor with the administration as part of a regulatory capture strategy. The post offers this as a guess about motive rather than a confirmed fact, and it relies on the policy change flagged in the quoted post by Andrew Curran.

  27. Google Cloud TechOfficialAI score10

    Google Cloud livestreams its #GeminiAtWork event now

    AIGoogle Cloud Tech announced it is live streaming the #GeminiAtWork event. The post provides a link to tune in but includes no details about product announcements or specific features.

  28. Gizmodo · AINewsAI score13

    Trump Declares Anyone Using "Artificial Intelligence" Term "THE ENEMY" on Truth Social

    AIPresident Donald Trump posted on Truth Social that the White House considers anyone who uses the term "Artificial Intelligence," instead of "Super Intelligence," to be "THE ENEMY." The article reports that Nvidia's Jensen Huang, Meta's Mark Zuckerberg, Elon Musk, and Jeff Bezos have adopted the "super intelligence" term, and that ai.gov was changed to read "SI" while si.gov appears dormant.