Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 9

Oct 9Fri
  1. Gergely OroszXAI score6

    Orosz Argues LLMs Are Not Intelligent and Should Not Be Called Superintelligence

    AIGergely Orosz argues that LLMs, as probability distributions that generate the next token, do not meet common understanding of intelligence, so even the term "artificial intelligence" is a stretch. He says hallucination is a feature of this design rather than a bug, and criticizes renaming LLMs as "superintelligence." Simon Willison's reply calls the "Super Intelligence" label stupid.

  2. Alexander DoriaXAI score12

    Doria says EU benchmarks depend on serious EU model training

    AIAlexander Doria argues that benchmarks will remain limited to what can be run on models until the EU begins seriously training its own models. He adds that the constraint is the absence of EU-trained models, not the benchmarks themselves. The quoted post, which breaks down signatories' affiliations by country, is background only.

  3. SantiagoXAI score32

    CRIS-0 causal world model lets home robots reason about action consequences

    AIAether AI's CRIS-0, its first causal robotic intelligence system, operates in a real home and models how actions change the physical world. Per the post, its causal world model predicts how conditions could change under different robot actions, while a causal agent keeps task context and selects capabilities at each stage. A unified tool interface connects navigation, learned action models, rule-based functions, and result checks.

  4. Amazon Web ServicesOfficialAI score10

    Meliá Hotels International adopts AI approach, per AWS post

    AIMeliá Hotels International, a hotel chain with more than 380 hotels across four continents, is cited in an AWS post as having taken a step the post describes as the approach needed. The source gives no further details on the specific technology, implementation, or results.

  5. Amazon Web ServicesOfficialAI score33

    AWS helps migrate COBOL reservation system to microservices

    AIA company used AWS to migrate its entire COBOL-based central reservation system to microservices. The migration handles 50 million daily availability requests, up from 26 million, with faster response times.

  6. Amazon Web ServicesOfficialAI score14

    AWS customer cuts feature delivery to one month and compute costs 60%

    AIAn AWS customer reports that new features now ship in one month instead of four, with 60% compute cost savings worth seven figures. The post also cites a 75% faster time to market, near 99.99% availability, and a four-year project finished in two years.

  7. QbitAINewsAI score67

    Aether AI shows CRIS-0 robot recovering from disturbances via causal reasoning

    AIAether AI, founded by UCSD assistant professor Biwei Huang, has released official demos of its CRIS-0 causal intelligence system for robots. In tests, the robot recovered from external disturbances in 9 of 10 random trials, typically within about 2 seconds, and stopped within 0.2 seconds when a human hand entered the workspace during a microwave-door task.

  8. QbitAINewsAI score38

    Lenovo's TianxiCode Agent Tops SWE-bench-Live Lite Leaderboard at 71%

    AILenovo's TianxiCode, paired with DeepSeek-v4.1-Flash, ranked first on the SWE-bench-Live Lite leaderboard with a 71% issue resolution rate and passed official Verified review. The framework combines multi-hop retrieval, autonomous planning with multi-turn tool calling, and test-driven self-correction, and will be applied to Lenovo AI hardware products.

  9. Bloomberg · TechnologyNewsAI score20

    Goldman Says Investors Will Shift Focus to AI Monetization

    AIGoldman Sachs asset allocation research head Christian Mueller-Glissmann says investors are growing skeptical about AI, questioning whether ongoing capital expenditure will translate into monetizable products. He made the remarks on Bloomberg Television.

  10. The DecoderNewsAI score54

    Anthropic's Claude Science maps the full sky in ultraviolet light

    AIAnthropic's Claude Science has produced what the source describes as the first complete ultraviolet map of the sky. AI agents downloaded data from multiple space missions, calibrated and merged it, and used inpainting to fill gaps left by NASA's GALEX mission, which skipped bright star-forming regions. In tests, predictions averaged about ten percent deviation from actual measurements, and the map is intended as teaching material.

  11. The Verge · AINewsAI score58

    OpenAI defends firing three AI safety researchers over information handling

    AIOpenAI says an internal investigation found Jasmine Wang, Tomek Korbak and Mikita Balesni committed a significant breach of trust by violating policies on handling sensitive information. The company denies the dismissals were about the researchers speaking out on AI safety, responding to an open letter in which the group said it was fired for raising safety concerns. OpenAI says the investigation found breaches beyond those in the letter but has not provided details.

  12. Gemini API ChangelogOfficialAI score22

    Gemini 3.7 Flash and 3.5 Flash are deprecated and rerouted to newer models

    AIGoogle says gemini-3.7-flash is deprecated and replaced by gemini-3.8-flash, and gemini-3.5-flash is deprecated and replaced by gemini-3.6-flash, with requests to the old strings automatically routed to the new ones. Developers should update their model strings, with gemini-3.8-flash recommended for the 3.5 Flash replacement as well. Google also says the deep-research-pro-preview-12-2025 agent will shut down on October 23, 2026, and that developers should migrate to deep-research-preview-04-2026 or deep-research-max-preview-04-2026.

  13. MIT Technology Review · AINewsAI score62

    AI refusal is probabilistic and unreliable, and it raises censorship risks

    AIThe article argues that AI refusal, the main safety mechanism in modern models, is unreliable and hard to draw lines for. It cites jailbreaks, classifier stacks, and studies showing refusal skewed toward repressive governments. It warns that governments and companies could use refusal to censor speech, and that refusal behavior remains poorly understood.

  14. OpenAI · YouTubeOfficialAI score36

    Sophos Cuts Threat Response Time by 96% With OpenAI Daybreak Agents

    AISophos says agents built through OpenAI Daybreak, combined with its cybersecurity expertise, cut average response time from 38 minutes to 89 seconds for cases handled by those agents. The company says the agents help its MDR team investigate threats faster and protect customers at scale while keeping human judgment central.

  15. Julien ChaumondXAI score7

    Julien Chaumond jokes about a niche hiring role for AI startups

    AIHugging Face's Julien Chaumond quips "niche of 1 lol" in response to a post describing a hard-to-hire role. The quoted post lists the traits sought: chronically online, taste and creative ability, AI and technical understanding, and strong execution.

  16. Viktor OddyXAI score22

    Restyles tool rebuilds prompts in new styles, free to try

    AIA post introduces Restyles, a tool where users enter a prompt and specify a target style, and the prompt rebuilds itself accordingly. The author demonstrates the process four times in a video and promotes a free trial at

    Video from @viktoroddy's post
  17. ModelScopeOfficialAI score28

    Corvus-Gov-3B: a 3B Chinese government-domain dialogue model

    AIModelScope released Corvus-Gov-3B, a compact model tuned for Chinese policy Q&A, public-service consultation, and internal government or enterprise assistants. It was fine-tuned on one million Chinese government-domain dialogue samples and built on Llama 3.2 3B Instruct using LoRA SFT via LLaMA Factory. The model is released under Apache 2.0.

    Image from @ModelScope2022's post
  18. X.PINXAI score41

    Biren Technology raises HK$4.04 billion in second share placement this year

    AIChinese AI chipmaker Biren Technology is raising HK$4.04 billion ($520 million) through a share placement of 130 million shares at HK$31.08 each, a 33% discount to its July placement price. The proceeds will mainly fund supply-chain purchases and production preparation for its next-generation BR20X chip, with about 70% going to procurement and commercialization. Its shares fell 11.85% on October 8 after the announcement.

    Image from @thexpin's post
  19. OpenBMBOfficialAI score28

    MiniCPM5-2B runs at 37 tok/s on iPhone Air

    AIOpenBMB reports that its MiniCPM5-2B model runs at 37 tokens per second on an iPhone Air. The post presents this as evidence that small open multimodal models can run on mobile devices without a cloud GPU, with NobodyWho noting the model is available in its Chat app.

  20. Bloomberg · TechnologyNewsAI score48

    How AI Is Upending the World of Mathematics

    AIOpenAI announced last month that it had produced an AI-generated proof for the Navier-Stokes problem, a result the source says is hard even for experts to parse. The source also says LLMs now tackle math problems that have stumped humans for decades, while teachers struggle to keep up with AI-completed homework.

  21. Bloomberg · TechnologyNewsAI score22

    JPMorgan Tops Evident's Ranking of the World's Most AI-Advanced Banks

    AIJPMorgan again leads Evident Insight's ranking of the world's most AI-advanced banks, according to Evident CEO Alexandra Mousavizadeh. She says early movers are pulling ahead, and the banks investing most heavily in AI are still growing headcount even as the technology changes jobs.

  22. MarkTechPostNewsAI score67

    Google Cloud launches Gemini agent, a single cloud-hosted agent for enterprise work

    AIGoogle Cloud has introduced the Gemini agent, a single cloud-hosted agent that handles Q&A, knowledge work, media creation, and coding from one prompt box and one API. It routes jobs across Gemini and Claude models today, with other private and open models planned. Governance covers per-agent identity, role-based access, audit logging, and hard per-project spend caps, but the source gives no reproducible benchmarks, pricing, or general availability date.

  23. MarkTechPostNewsAI score44

    Underdog Releases Saluki 27B, a 2-Bit Qwen3.8-27B That Beats the Original at Tool Calling

    AIUnderdog has released Saluki 27B under Apache 2.0, a 2-bit GGUF of Qwen3.8-27B that fits in 7.89 GB, versus 54 GB for the full BF16 model. On Underdog Bench, Saluki scores 88 against 84 for the full model, and it raises parallel tool-call accuracy to 42 from 35. It runs on stock llama.cpp, but math and reasoning drop sharply, with AIME 2025 at 79.2 versus 96.7.

  24. The Guardian · AINewsAI score62

    OpenAI projects $50bn revenue, $20bn below its earlier investor signal

    AIOpenAI told investors it expects $50bn in revenue this year, about $20bn less than the $70bn it had signalled last month. The gap stems partly from comparing with Anthropic, which counts revenue sold through cloud partners such as AWS and Google Cloud, while OpenAI does not. The news weighed on US tech stocks, and OpenAI is in early talks to raise $30bn at a valuation of about $1.4tn.

  25. The DecoderNewsAI score61

    OpenAI bans Russian and Iranian influence ops that planted fake stories in real outlets

    AIOpenAI exposed a Russian and an Iranian influence operation and banned the ChatGPT accounts involved, both of which planted content in legitimate media using fake identities. The Iranian operation, "Bogus Bylines," used seven fake journalists to place nearly 100 articles about the US-Iran conflict, while the Russian "Dark Clark" operation triggered fact-checks and official denials in Ecuador and Peru. Both operations used AI mainly for internal reporting and adapting propaganda to different languages.

  26. Tencent · new models on Hugging FaceOfficialAI score41

    Tencent Releases Youtu-Parsing-Omni, a 5B Omni-Modal Document and Media Parsing Model

    AITencent has open-sourced Youtu-Parsing-Omni, a 5B-parameter omni-modal model that outputs a single structured JSON covering layout, text, tables, formulas, ASR, OCR, and video segments. It scores 96.96 Overall on OmniDocBench, the highest among the compared models, and ships with weights on Hugging Face, a vLLM plugin, and inference examples.

  27. NVIDIA · new models on Hugging FaceOfficialAI score16

    NVIDIA releases Agile One S SSD Pick GR00T N1.7 checkpoint 40000 model on Hugging Face

    AINVIDIA published the Agile One S SSD Pick deployment model, GR00T N1.7 checkpoint 40000, on Hugging Face for SSD pickup tasks. The repository includes five ONNX graphs with external tensor files and two existing TensorRT BF16 engines, with original configurations and build metadata, but no retraining or re-export was performed. The files are not a robot deployment or safety qualification, and engine compatibility depends on the target GPU and TensorRT environment.