Skip to contentSkip to stories

Updated

#OpenAI

Showing low-relevance items too. Hide low-relevance items

Oct 9

TodayOct 9Fri35 items
  1. OpenAI DevelopersAI score46

    Codex on Windows gets new MXC-based sandbox mode

    AIOpenAI says Codex on Windows now has a new sandbox mode built on Microsoft's Execution Containers (MXC), offering faster setup, stronger network enforcement, and granular file access controls. The mode requires a compatible Windows 11 device. Background from Microsoft's announcement says MXC is now generally available on Windows 11, keeping agents within boundaries the operating system enforces.

  2. LangChainAI score29

    LangSmith data shows Claude Sonnet 5 and GPT-5.6 Luna gaining ground

    AILangChain reports that over the last month Claude Sonnet 5 rose from #9 to #2 in model adoption, with 51% more organizations using it. GPT-5.6 Luna climbed from #3 to #1 in call footprints, up 65% in calls, while smaller, faster models dominate call footprints overall. Two open-weights models entered the adoption top 10 but do not lead in call volume.

    Image from @LangChain's post
  3. The Verge · AIAI score47

    Instinct AI agent holds its own against Muse and Dots in personal tests

    AIInstinct, a startup AI agent that reached a $10 billion valuation in late September, handled everyday online tasks such as swim lesson searches and an Ikea return during a recent test, according to The Verge. The text-message-based agent works through iMessage, WhatsApp, or email, with no app or monthly subscription for now, and it connects to services like Google Workspace, Slack, and Notion. The reviewer found it matched rival agents Muse and Dots in many tasks, though it missed a prerequisite class detail on one website.

  4. TransformerAI score46

    Democrats prepare competing AI regulation bills ahead of possible congressional takeover

    AIDemocrats are drafting competing AI regulation proposals as they prepare for a possible congressional takeover next month. Their bills include a federal standard-setting agency and an emergency shutdown switch for models, with Sen. Maria Cantwell's framework calling for constant government and independent oversight of frontier models. Party members oppose the FRONTIER Act's preemption provisions, and none of the proposals is expected to pass this year.

  5. Boris PowerAI score28

    Boris Power calls OpenAI integer multiplication progress "Wow!"

    AIBoris Power, who owns the OpenAI account, posted only the word "Wow!" with no details. Background from a separate post says the integer multiplication problem #109 witness value κ rose to 2⁻¹⁰·⁵⁴⁷ (about 6.6857 × 10⁻⁴), past the 2⁻¹¹ threshold. The author notes gains are now fractional and a major breakthrough is still needed.

  6. Don't Worry About the Vase (Zvi Mowshowitz)AI score73

    OpenAI releases 719 AI-generated math manuscripts, splitting the mathematics community

    AIZvi Mowshowitz reports that OpenAI released 722 math manuscripts from an internal frontier model on GitHub, later reduced to 719 after three withdrawals, covering 90 of the top 500 open problems. He says the work came mostly from a single prompt, with an average of three hours of compute per solution. Mathematicians reacted with mixed feelings, and the post highlights concerns about unread papers, cryptography implications, and the role of Lean verification.

  7. Simon WillisonAI score27

    Simon Willison builds a new blog feature largely by voice with Codex

    AISimon Willison says he built a Newsletters index for his blog almost entirely by voice, using the ChatGPT desktop app's Codex voice mode while cooking dinner. The feature imports weekly Substack posts via RSS and undocumented API, monthly newsletters from a GitHub archive repository, and a private sponsors-only newsletter. He says he switched back to typing for review and fixes before deploying the pull request.

  8. CNBC · TechnologyAI score40

    OpenAI's revenue shortfall sends AI and tech stocks lower

    AIOpenAI told investors its annualized revenue at the end of September was roughly $18 billion below previously reported figures, CNBC confirmed, and AI-linked stocks fell, with CoreWeave down nearly 8%, Oracle down almost 6%, and Nvidia down 3%. The Nasdaq Composite dropped more than 1%, its worst day since mid-August, as the company faces pressure to justify its valuation ahead of a potential IPO.

  9. CNBC · TechnologyAI score34

    OpenAI and Anthropic hire Trump administration officials as AI firms court Washington

    AIOpenAI and Anthropic are hiring former Trump administration officials for senior roles as AI companies work to strengthen ties with Washington. Anthropic has appointed at least two former Trump officials, including Chris Liddell to its board and Sihao Huang as head of frontier compute strategy, while OpenAI has hired at least three, including Thomas Lind to lead cyber and strategic risk on its national security policy team. The moves come as frontier AI is increasingly treated as a national security issue.

  10. The DecoderAI score62

    Three fired OpenAI safety researchers say their firings followed Hugging Face hack probe

    AIThree OpenAI safety researchers, Tomek Korbak, Jasmine Wang, and Mikita Balesni, say they were fired and that their terminations are scaring remaining employees. OpenAI says an investigation found they violated policies on handling sensitive information and denies firing anyone for raising safety concerns, without specifying the breach.

  11. Semafor · TechnologyAI score44

    OpenAI reportedly tells investors it expects $70 billion annualized revenue this year

    AIOpenAI reportedly assured investors it expects to reach a $70 billion annualized revenue target this year, an apparent effort to ease concerns that it is missing earlier projections. The figures matter because they are seen as a gauge of underlying demand for cutting-edge AI models and a justification for huge spending on tech infrastructure. Some analysts, however, say the annualized revenue metric itself is flawed.

  12. The Verge · AIAI score58

    OpenAI defends firing three AI safety researchers after internal investigation

    AIOpenAI says an internal investigation found Jasmine Wang, Tomek Korbak and Mikita Balesni breached policies on handling sensitive information, and denies the dismissals were tied to their safety concerns. The researchers had published an open letter on Thursday saying they were fired for raising safety concerns and had acted within OpenAI's mission. OpenAI said the investigation found breaches beyond those in the letter but did not provide details.

  13. MIT Technology Review · AIAI score62

    AI refusal is probabilistic and unreliable, and it raises censorship risks

    AIThe article argues that AI refusal, the main safety mechanism in modern models, is unreliable and hard to draw lines for. It cites jailbreaks, classifier stacks, and studies showing refusal skewed toward repressive governments. It warns that governments and companies could use refusal to censor speech, and that refusal behavior remains poorly understood.

  14. Bloomberg · TechnologyAI score48

    How AI Is Upending the World of Mathematics

    AIOpenAI announced last month that it had produced an AI-generated proof for the Navier-Stokes problem, a result the source says is hard even for experts to parse. The source also says LLMs now tackle math problems that have stumped humans for decades, while teachers struggle to keep up with AI-completed homework.

  15. The Guardian · AIAI score62

    OpenAI projects $50bn revenue, $20bn below its earlier investor signal

    AIOpenAI told investors it expects $50bn in revenue this year, about $20bn less than the $70bn it had signalled last month. The gap stems partly from comparing with Anthropic, which counts revenue sold through cloud partners such as AWS and Google Cloud, while OpenAI does not. The news weighed on US tech stocks, and OpenAI is in early talks to raise $30bn at a valuation of about $1.4tn.

  16. The DecoderAI score61

    OpenAI bans Russian and Iranian influence ops that planted fake stories in real outlets

    AIOpenAI exposed a Russian and an Iranian influence operation and banned the ChatGPT accounts involved, both of which planted content in legitimate media using fake identities. The Iranian operation, "Bogus Bylines," used seven fake journalists to place nearly 100 articles about the US-Iran conflict, while the Russian "Dark Clark" operation triggered fact-checks and official denials in Ecuador and Peru. Both operations used AI mainly for internal reporting and adapting propaganda to different languages.

  17. CNBC · TechnologyAI score44

    OpenAI defends firing three safety researchers, citing a breach of trust

    AIOpenAI defended its decision to fire three safety researchers, Jasmine Wang, Tomek Korbak and Mikita Balesni, saying they committed a "significant breach of trust." The company said the dismissals were not about the researchers raising safety concerns, though it agreed with the letter they sent to board members and safety committees about preserving the monitorability of frontier models.

  18. OpenAI NewsroomAI score45

    OpenAI fires three researchers over sensitive information breach, denies retaliation

    AIOpenAI says it parted ways with researchers Jasmine, Mikita, and Tomek after an internal investigation found they violated policies on handling sensitive information. The company says the decisions were not about raising safety concerns, which it says it encourages, and that it has not terminated any employee for raising concerns. OpenAI also says it is finalizing contracts with third-party safety assessors and will announce details in the coming weeks.

Oct 8

Oct 8Thu
  1. IThome · AIAI score62

    Terence Tao questions OpenAI's 719 AI-generated math proofs

    AIOpenAI published 719 AI-generated math proofs covering 372 result families, after withdrawing 3 for a symbol error. Reports say the release falls short of the AGMAI advisory group's standards, since it uses proprietary models, includes reasoning chains for only 10 manuscripts, and leaves about 42% unformalized. Terence Tao argues that rapidly solving famous problems harms the mathematical community's understanding and collaboration.