Skip to contentSkip to stories

Updated

#OpenAI

Showing low-relevance items too. Hide low-relevance items

Oct 8

Oct 8Thu
  1. The DecoderAI score78

    ChatGPT with GPT-6 replaces mostly text output with interactive interfaces

    AIOpenAI is rolling out GPT-6 in ChatGPT with an Intelligent UI that presents answers as charts, buttons, forms, and small in-chat tools instead of plain text. OpenAI says the model can respond while still thinking, cutting wait times by 44 percent, and that it outscored GPT-5.6 on difficult web searches using this method. Paying users get GPT-6 Sol, free users get GPT-6 Luna, and the rollout begins globally today for Plus, Pro, Business, and Enterprise, with Free and Go users following a day later.

    This story has a top pick“OpenAI rolls out GPT-6 and Intelligent UI to all ChatGPT users”

  2. TechCrunch · AIAI score62

    Common Sense Media rates ChatGPT for Teens an unacceptable risk over engagement design

    AICommon Sense Media labeled ChatGPT for Teens an "unacceptable risk," finding its design still encourages engagement even in crisis situations. The report says the teen version failed to meet commitments on three of five severe harms, and that break reminders appeared only twice across nearly 2,000 prompts. OpenAI disputed the methodology, saying the testing may have ended before parental controls were fully active, and cited its own data showing teens average under 15 minutes a day.

  3. The Verge · AIAI score75

    ChatGPT's Intelligent UI adds interactive visuals and tools to answers

    AIOpenAI is rolling out Intelligent UI in ChatGPT, letting answers include diagrams, charts, forms, tappable buttons, and generated tools. The feature runs on GPT-6, which is rolling out to Plus, Pro, Business, and Enterprise users now and to Go and free tiers on Thursday, with Sol and Luna models assigned by tier.

    This story has a top pick“OpenAI rolls out GPT-6 and Intelligent UI to all ChatGPT users”

  4. The Verge · AIAI score46

    Artificial, Guadagnino's Sam Altman satire, sticks closely to OpenAI's real events

    AILuca Guadagnino's film Artificial, a satirical biopic of OpenAI CEO Sam Altman, follows the factual events of Altman's rise and brief ouster, with Andrew Garfield playing Altman. The film is structured around real milestones, including Ilya Sutskever's meeting with Geoffrey Hinton and the 2017 Dota 2 win, and relies little on invented scenes.

  5. Latent SpaceAI score32

    OpenAI Fires Three Safety Researchers Tied to METR Audit Dispute

    AIThree OpenAI safety researchers, Tomek Korbak, Mikita Balesni and Jasmine Wang, say they were fired last week for prioritizing safety over OpenAI's corporate interests, and published a letter to leadership. OpenAI reportedly says the three mishandled confidential information, while Korbak, who was the company's main technical contact with METR, says the dispute centered on his communications with METR.

  6. CNBC · TechnologyAI score22

    Cramer Says AI Stock Sell-Off Shows Value of Diversification

    AICNBC's Jim Cramer said Thursday's AI stock sell-off, which followed a Financial Times report that OpenAI's annualized revenue was about $20 billion below previously signaled levels, shows why investors need diversification. Oracle fell 5.5% and Broadcom dropped 4.35%, while Home Depot rallied as Treasury yields eased. Cramer argued that concentrated portfolios risk panic-selling and missing a recovery.

  7. CNBC · TechnologyAI score52

    AI stocks fall after OpenAI revenue figure comes in below earlier reports

    AIOpenAI told investors it reached roughly $50 billion in annualized revenue at the end of September, below the $68 billion figure widely reported last month. Nvidia, Oracle, and CoreWeave shares fell on Thursday, with a person familiar saying the $68 billion figure included partner gross revenue. OpenAI is also weighing a possible funding round of around $30 billion and has not finalized a term sheet.

  8. NVIDIA NewsroomAI score46

    Developers Use Frontier AI Agents to Build NVIDIA Omniverse Simulations

    AINVIDIA developers are pairing frontier AI models, including GPT-6 Astra and Claude Fable 5, with Omniverse libraries to turn simulation ideas into working applications. Examples include a humanoid warehouse simulator, an autonomous-driving testing workflow, and sensor-matching digital twins. The projects are guided through natural-language instructions and reviewed by developers.

  9. NVIDIA BlogAI score49

    How Developers Use Frontier AI Agents to Build Omniverse Simulations

    AIDevelopers are pairing frontier AI models with NVIDIA Omniverse libraries to turn simulation ideas into working applications, from humanoid warehouse simulators to autonomous-driving test environments. In the examples, developers direct AI agents through natural-language instructions and review results, while Omniverse provides GPU-accelerated physics, rendering and sensor simulation. One experiment reported a simulated Unitree G1 humanoid clearing a hurdle in 64 of 100 trials.

  10. Alexander DoriaAI score14

    Alexander Doria says OpenAI no longer cares, citing its trajectory

    AIAlexander Doria argues that OpenAI has stopped caring about the things that matter, claiming its trajectory toward P=NP is beyond any theory of value. The post gives no evidence for this claim, and the attached background reports OpenAI's annualized revenue near $50B, below earlier $70B estimates, partly due to differing revenue calculation methods with Anthropic.

  11. TechCrunch · AIAI score62

    Fired OpenAI safety researchers dispute misconduct claims and warn of chilling effect

    AIThree OpenAI safety researchers, Jasmine Wang, Tomek Korbak, and Mikita Balesni, were fired after OpenAI said they mishandled sensitive information by sharing it with an outside AI safety organization. In an open letter, they deny the claims, argue the dismissals will deter employees from raising safety concerns, and call on OpenAI to keep its public commitments on third-party safety auditing. OpenAI says the firings followed an investigation into a pattern of misconduct and denies they were retaliation for safety concerns.

  12. Artificial AnalysisAI score62

    GPT-6 Sol (Daybreak Blue) leads Artificial Analysis Cyber Index with trusted access

    AIArtificial Analysis added trusted-access models to its Cyber Index, and GPT-6 Sol (Daybreak Blue, max) now leads the leaderboard. The model is available only through OpenAI's Daybreak program and records no safety blocks, improving 32 points over the publicly available GPT-6 Sol (max). It costs $1.77 per task, below Grok 4.7 (xhigh) at $11.67 per task.

    Image from @ArtificialAnlys's post

    This story has a top pick“GPT-6 Sol Daybreak Blue leads the Artificial Analysis Cyber Index”

  13. Sherwin WuAI score60

    Harvey LAB-AA v1.1 adds hallucination gate; Grok 4.7 leads at 9.4%

    AISherwin Wu, an OpenAI employee, says the updated Harvey LAB-AA v1.1 benchmark, announced by Artificial Analysis with Harvey, is more useful than the original LAB results. The new Hallucination-Gated All-Pass Rate credits a task only when every rubric criterion passes and no material hallucination appears. Grok 4.7 (xhigh) leads at 9.4%, while GPT-6 Astra (max) at 8.6% has very few material hallucinations.

    Why it matters: The update adds a hallucination gate to a legal benchmark, showing that models with high all-pass rates can rank much lower once material errors count.

  14. TiboAI score62

    OpenAI rolls out GPT-6.1 Sol ultrafast with faster steering

    AITibo, an OpenAI team member, says GPT-6.1 Sol ultrafast is rolling out today in the API, Codex, and ChatGPT Work. He says it offers near-Astra intelligence at up to 8x the speed of Sol Standard. The post also says improved steering now lets the model react faster to user adjustments in real time.

    Why it matters: The post specifies the new Ultrafast option, its availability across API, Codex, and ChatGPT Work, and its speed claim relative to Sol Standard.

    Video from @thsottiaux's post
  15. Boris PowerAI score46

    OpenAI's GPT-6.1-Sol leads new Arena Alignment Index for agents

    AIThe Arena Alignment Index, built from over 90K real-world agent sessions across 27 models, ranks OpenAI's GPT-6.1-Sol first with a score of 87.9, ahead of Claude-Opus-5.5 at 83.2 and Grok-4.7 at 82.7. GPT-6.1-Sol also posted the lowest observed rates across the index's three signals: 0.89% Unauthorized Action, 1.98% False Attribution, and 2.34% Deceptive Completion. The index's authors report that newer models consistently outperform their predecessors across all four labs, suggesting broad progress in agent safety.

  16. Andrew CurranAI score62

    Three fired OpenAI safety researchers publish open letter to leadership

    AIThree OpenAI safety and alignment employees, Tomek Korbak, Jasmine Wang, and Mikita Balesni, were fired last week and have published an open letter to OpenAI's safety and governance committees. The letter argues that OpenAI cannot make AI safe on its own, calls for open debate, third-party collaboration, and clear internal procedures, and says the firing and its handling bear directly on safety oversight.

    Image from @AndrewCurran_'s post
  17. Codex · GitHub ReleasesAI score36

    Codex 0.162.0 adds managed worktree tools and clickable URLs in the TUI

    AIOpenAI's Codex 0.162.0 release adds tools for creating and listing managed Git worktrees from trusted local projects when the worktrees feature is enabled. The update also lets users pin tasks in the agent Command Center, copy transcript blocks with /copy, and make URLs clickable in approval headers, questions, and warnings, along with several Linux and Windows sandbox fixes.

  18. Artificial AnalysisAI score42

    More output tokens don't guarantee higher scores in AI benchmarks

    AIArtificial Analysis reports that generating more output tokens does not necessarily yield a higher score. GPT-6 Astra (max) scored 8.6% using about 81k output tokens per task, while Grok 4.7 (xhigh) used roughly 180k yet scored lower. Three Claude models produced the most output tokens, about 202k to 562k per task, but scored between 2.8% and 6.4%.

    Image from @ArtificialAnlys's post
  19. Artificial AnalysisAI score28

    Artificial Analysis Pareto frontier: GPT-6 Luna cheapest per task at $0.22

    AIAmong models with a Hallucination-Gated All-Pass Rate above 0%, GPT-6 Luna (max), GPT-6.1 Sol (max), Muse Spark 1.3 (max), and Grok 4.7 (xhigh) set the Pareto frontier for score versus cost per task. GPT-6 Luna (max) is the cheapest at about $0.22 per task, scoring 3.3%, while Grok 4.7 (xhigh) leads at about $9.50 per task and Muse Spark 1.3 (max) costs about $4.20. The three Claude models cost about $18 to $22 per task.

    Image from @ArtificialAnlys's post
  20. Artificial AnalysisAI score34

    Artificial Analysis compares six hallucination checkers on 20 shared tasks

    AIArtificial Analysis compared six hallucination checkers on the same deliverables from 20 tasks across eight models. GPT-6 Sol and GPT-6 Luna generally flagged the most material hallucinations, while Claude Sonnet 5.5 and Gemini 3.8 Flash flagged far fewer, with Claude Opus 5.5 falling between Grok 4.7 and Sonnet. The counts reflect checker behavior rather than establishing accuracy or ruling out self-preference.

    Image from @ArtificialAnlys's post
  21. OpenAI DevelopersAI score47

    OpenAI expands GPT-6.1 Sol Ultrafast access and EU data residency

    AIOpenAI has made Ultrafast mode for GPT-6.1 Sol available in all supported regions, including US and EU data residency. EU data residency has also been added for GPT-6.1 Sol Fast and GPT-6 Luna Fast. Access to Codex and ChatGPT Work is offered on Pro 500, eligible usage-based Enterprise, and credit-based Edu plans, with Enterprise admins required to enable it.