Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 9

TodayOct 9Fri
  1. The Verge · AINewsAI score58

    OpenAI defends firing three AI safety researchers after internal investigation

    AIOpenAI says an internal investigation found Jasmine Wang, Tomek Korbak and Mikita Balesni breached policies on handling sensitive information, and denies the dismissals were tied to their safety concerns. The researchers had published an open letter on Thursday saying they were fired for raising safety concerns and had acted within OpenAI's mission. OpenAI said the investigation found breaches beyond those in the letter but did not provide details.

  2. Gemini API ChangelogOfficialAI score22

    Gemini 3.7 Flash and 3.5 Flash are deprecated and rerouted to newer models

    AIGoogle says gemini-3.7-flash is deprecated and replaced by gemini-3.8-flash, and gemini-3.5-flash is deprecated and replaced by gemini-3.6-flash, with requests to the old strings automatically routed to the new ones. Developers should update their model strings, with gemini-3.8-flash recommended for the 3.5 Flash replacement as well. Google also says the deep-research-pro-preview-12-2025 agent will shut down on October 23, 2026, and that developers should migrate to deep-research-preview-04-2026 or deep-research-max-preview-04-2026.

  3. MIT Technology Review · AINewsAI score62

    AI refusal is probabilistic and unreliable, and it raises censorship risks

    AIThe article argues that AI refusal, the main safety mechanism in modern models, is unreliable and hard to draw lines for. It cites jailbreaks, classifier stacks, and studies showing refusal skewed toward repressive governments. It warns that governments and companies could use refusal to censor speech, and that refusal behavior remains poorly understood.

  4. OpenAI · YouTubeOfficialAI score36

    Sophos Cuts Threat Response Time by 96% With OpenAI Daybreak Agents

    AISophos says agents built through OpenAI Daybreak, combined with its cybersecurity expertise, cut average response time from 38 minutes to 89 seconds for cases handled by those agents. The company says the agents help its MDR team investigate threats faster and protect customers at scale while keeping human judgment central.

  5. Julien ChaumondXAI score7

    Julien Chaumond jokes about a niche hiring role for AI startups

    AIHugging Face's Julien Chaumond quips "niche of 1 lol" in response to a post describing a hard-to-hire role. The quoted post lists the traits sought: chronically online, taste and creative ability, AI and technical understanding, and strong execution.

  6. Viktor OddyXAI score22

    Restyles tool rebuilds prompts in new styles, free to try

    AIA post introduces Restyles, a tool where users enter a prompt and specify a target style, and the prompt rebuilds itself accordingly. The author demonstrates the process four times in a video and promotes a free trial at

    Video from @viktoroddy's post
  7. ModelScopeOfficialAI score28

    Corvus-Gov-3B: a 3B Chinese government-domain dialogue model

    AIModelScope released Corvus-Gov-3B, a compact model tuned for Chinese policy Q&A, public-service consultation, and internal government or enterprise assistants. It was fine-tuned on one million Chinese government-domain dialogue samples and built on Llama 3.2 3B Instruct using LoRA SFT via LLaMA Factory. The model is released under Apache 2.0.

    Image from @ModelScope2022's post
  8. X.PINXAI score41

    Biren Technology raises HK$4.04 billion in second share placement this year

    AIChinese AI chipmaker Biren Technology is raising HK$4.04 billion ($520 million) through a share placement of 130 million shares at HK$31.08 each, a 33% discount to its July placement price. The proceeds will mainly fund supply-chain purchases and production preparation for its next-generation BR20X chip, with about 70% going to procurement and commercialization. Its shares fell 11.85% on October 8 after the announcement.

    Image from @thexpin's post
  9. OpenBMBOfficialAI score28

    MiniCPM5-2B runs at 37 tok/s on iPhone Air

    AIOpenBMB reports that its MiniCPM5-2B model runs at 37 tokens per second on an iPhone Air. The post presents this as evidence that small open multimodal models can run on mobile devices without a cloud GPU, with NobodyWho noting the model is available in its Chat app.

  10. Bloomberg · TechnologyNewsAI score48

    How AI Is Upending the World of Mathematics

    AIOpenAI announced last month that it had produced an AI-generated proof for the Navier-Stokes problem, a result the source says is hard even for experts to parse. The source also says LLMs now tackle math problems that have stumped humans for decades, while teachers struggle to keep up with AI-completed homework.

  11. Bloomberg · TechnologyNewsAI score22

    JPMorgan Tops Evident's Ranking of the World's Most AI-Advanced Banks

    AIJPMorgan again leads Evident Insight's ranking of the world's most AI-advanced banks, according to Evident CEO Alexandra Mousavizadeh. She says early movers are pulling ahead, and the banks investing most heavily in AI are still growing headcount even as the technology changes jobs.

  12. MarkTechPostNewsAI score67

    Google Cloud launches Gemini agent, a single cloud-hosted agent for enterprise work

    AIGoogle Cloud has introduced the Gemini agent, a single cloud-hosted agent that handles Q&A, knowledge work, media creation, and coding from one prompt box and one API. It routes jobs across Gemini and Claude models today, with other private and open models planned. Governance covers per-agent identity, role-based access, audit logging, and hard per-project spend caps, but the source gives no reproducible benchmarks, pricing, or general availability date.

  13. MarkTechPostNewsAI score44

    Underdog Releases Saluki 27B, a 2-Bit Qwen3.8-27B That Beats the Original at Tool Calling

    AIUnderdog has released Saluki 27B under Apache 2.0, a 2-bit GGUF of Qwen3.8-27B that fits in 7.89 GB, versus 54 GB for the full BF16 model. On Underdog Bench, Saluki scores 88 against 84 for the full model, and it raises parallel tool-call accuracy to 42 from 35. It runs on stock llama.cpp, but math and reasoning drop sharply, with AIME 2025 at 79.2 versus 96.7.

  14. The Guardian · AINewsAI score62

    OpenAI projects $50bn revenue, $20bn below its earlier investor signal

    AIOpenAI told investors it expects $50bn in revenue this year, about $20bn less than the $70bn it had signalled last month. The gap stems partly from comparing with Anthropic, which counts revenue sold through cloud partners such as AWS and Google Cloud, while OpenAI does not. The news weighed on US tech stocks, and OpenAI is in early talks to raise $30bn at a valuation of about $1.4tn.

  15. The DecoderNewsAI score61

    OpenAI bans Russian and Iranian influence ops that planted fake stories in real outlets

    AIOpenAI exposed a Russian and an Iranian influence operation and banned the ChatGPT accounts involved, both of which planted content in legitimate media using fake identities. The Iranian operation, "Bogus Bylines," used seven fake journalists to place nearly 100 articles about the US-Iran conflict, while the Russian "Dark Clark" operation triggered fact-checks and official denials in Ecuador and Peru. Both operations used AI mainly for internal reporting and adapting propaganda to different languages.

  16. Tencent · new models on Hugging FaceOfficialAI score41

    Tencent Releases Youtu-Parsing-Omni, a 5B Omni-Modal Document and Media Parsing Model

    AITencent has open-sourced Youtu-Parsing-Omni, a 5B-parameter omni-modal model that outputs a single structured JSON covering layout, text, tables, formulas, ASR, OCR, and video segments. It scores 96.96 Overall on OmniDocBench, the highest among the compared models, and ships with weights on Hugging Face, a vLLM plugin, and inference examples.

  17. NVIDIA · new models on Hugging FaceOfficialAI score16

    NVIDIA releases Agile One S SSD Pick GR00T N1.7 checkpoint 40000 model on Hugging Face

    AINVIDIA published the Agile One S SSD Pick deployment model, GR00T N1.7 checkpoint 40000, on Hugging Face for SSD pickup tasks. The repository includes five ONNX graphs with external tensor files and two existing TensorRT BF16 engines, with original configurations and build metadata, but no retraining or re-export was performed. The files are not a robot deployment or safety qualification, and engine compatibility depends on the target GPU and TensorRT environment.

  18. NVIDIA · new models on Hugging FaceOfficialAI score23

    NVIDIA publishes Agile One S SSD pick model, GR00T N1.7 checkpoint 58000, on Hugging Face

    AINVIDIA has released a deployment model for Agile One S SSD pickup, based on GR00T N1.7 checkpoint 58000 and using three cameras: ego, left wrist, and right wrist. The repository republishes ONNX graphs, external tensor files, and two existing TensorRT BF16 engines without retraining or re-export, and the original export reported a numerical warning that full FP32, node, and BF16 parity did not pass all tolerances. The files are not a certified robot deployment or safety qualification.

  19. NVIDIA · new models on Hugging FaceOfficialAI score25

    NVIDIA releases Agile One S Walk GR00T N2 checkpoint 1680 on Hugging Face

    AINVIDIA published the Agile One S Walk GR00T N2 checkpoint 1680, a walking deployment model with four cameras, on Hugging Face. The repository includes ONNX graphs, TensorRT BF16 plans/engines, and the original checkpoint files, republished without retraining or re-export. The shared Cosmos-Reason1-7B dependency and the Isaac/GR00T runtime must be set up separately, and the files are not a robot safety qualification.

  20. NVIDIA · new models on Hugging FaceOfficialAI score14

    NVIDIA Releases Agile One S SSD Place GR00T N1.7 Deployment Model on Hugging Face

    AINVIDIA published the nvidia/agile_one_s_place_ssd_n17_24050 repository on Hugging Face, containing a GR00T N1.7 checkpoint 24050 model for placing an SSD with three cameras. The repository includes five ONNX graphs with external tensor files and two existing TensorRT BF16 engines, republished without retraining, re-export, or engine rebuild. Engine compatibility depends on the target GPU and TensorRT environment, and the files are not a robot deployment or safety qualification.

  21. meng shaoXAI score45

    Addy Osmani on why engineers' joy in AI coding agents splits three ways

    AIAddy Osmani argues engineers' reactions to AI coding agents depend on which of three joys they value most: making, knowing, or mattering. He warns that choosing among agent suggestions without generating ideas yourself erodes the skill of ideation and can leave developers directed by agents. He reframes grief over lost craft as a sign of real attachment rather than failed adaptation.

    Image from @shao__meng's post
  22. Harrison ChaseXAI score22

    Harrison Chase on eval-driven development for AI agents

    AIHarrison Chase's post is titled "eval driven development," presenting evals as a development approach. The main post gives no further detail beyond the title. The quoted context from Jerry Liu argues that most tasks can be solved by defining an eval and hillclimbing over it rather than hand-building a deterministic or agentic workflow.

  23. Harrison ChaseXAI score22

    Harrison Chase questions eval-driven development for autonomous agents

    AIHarrison Chase argues that eval-driven development works for narrowly scoped tasks but breaks down for more autonomous agents, invoking Goodhart's Law that a measure ceases to be useful once it becomes a target. He asks how such agents can be hill-climbed, and the post does not provide an answer.

  24. 🚨 AI News | TestingCatalogXAI score50

    OpenAI, Anthropic, Google, and others roll out agent and model updates

    AIOpenAI's GPT-6.1 Sol Ultrafast is rolling out in the API, Codex, and ChatGPT Work, running up to 8x faster than Sol Standard at $12/$60 per million tokens. StepFun's Step 5 Preview, a 600B-parameter MoE model with 27B active parameters, is now on OpenRouter, and JetBrains released the open 12B MoE coding model Mellum2.1 under Apache 2.0.

  25. MagnificOfficialAI score14

    Magnific One turns profile photos into cereal-themed images

    AIMagnific announces that its Magnific One tool can now be used to create cereal-themed profile pictures from users' photos. The company invites users to upload a photo through its Flow tool to join the promotion.

    Video from @magnific's post
  26. QbitAINewsAI score62

    Google's AMIE Chatbot Tested in Real Pre-Visit Clinical Study Published in The Lancet

    AIA study led by Google and BIDMC tested Google's diagnostic AI chatbot AMIE with 98 outpatients before emergency visits, with a supervising doctor monitoring every exchange. No conversation needed interruption under the predefined safety criteria, and clinicians said AI summaries helped them prepare for 75% of visits. AMIE's differential diagnoses matched final diagnoses 90% of the time, but the authors say larger trials are needed.

  27. IThome · AINewsAI score46

    JetBrains Releases Mellum2.1 Coding Model With Near-Double Qwen3.5-9B Throughput

    AIJetBrains released Mellum2.1, a 12B mixture-of-experts coding model with 2.5B active parameters under Apache 2.0, emphasizing agentic programming. Under high load, its inference throughput in tokens is nearly twice that of Qwen3.5-9B in JetBrains' comparison, and multi-token prediction (MTP) speeds single-request responses by about 1.6x. The model is available on Hugging Face for local or private-infrastructure deployment, with GGUF and vLLM MTP support announced for later.

  28. IThome · AINewsAI score55

    Odyssey-3 world model scores 66.1 on Physics-IQ Verified benchmark

    AIOdyssey announced the Odyssey-3 series of foundation world models, with Odyssey-3 Pro scoring 66.1 on the Physics-IQ Verified video-to-video benchmark, the highest recorded on that leaderboard. The series includes a standard version balancing physical accuracy and generation cost, and a Pro version with stronger physics prediction. The preview supports first-person and third-person navigation and lets users move the camera, take actions, or trigger events while the model predicts environmental changes in real time.

  29. PandailyNewsAI score60

    Richard Yu says more Huawei phones will get LogicFolding chips

    AIRichard Yu said more Huawei phones will adopt LogicFolding chips built under the Tau Scaling Law, though no models or timetable were given. He said the Kirin 9050 Pro's performance is 31% higher than its predecessor, and the source outlines a roadmap reaching 5.0 GHz by 2031.

  30. PandailyNewsAI score56

    openJiuwen open-sources an enterprise AgentOS for agent swarms and multi-tenant control

    AIHuawei-backed openJiuwen has open-sourced AgentOS for Enterprise under Apache 2.0 on GitHub and AtomGit, targeting multi-agent coordination, memory-based self-evolution, multi-tenant isolation and fault recovery. Huawei Connect 2026 also introduced an all-in-one appliance built on it, which the launch information says enables an end-to-end private deployment in hours.