Skip to contentSkip to stories

Updated

#Anthropic

Showing low-relevance items too. Hide low-relevance items

Sep 22

Sep 22Tue
  1. Sam BowmanAI score75

    Anthropic's Sam Bowman says Claude Opus 5.5 is safer, reducing misalignment risk

    AISam Bowman says Claude Opus 5.5 is sufficiently safer than its predecessors that releasing it more likely than not reduces misalignment risks. The quoted @claudeai post introduces Claude Opus 5.5 as the first model in the Claude 5.5 family, performing at the level of Claude Fable 5.1 on most tasks at 40% lower run cost than Opus 5.

    Why it matters: The post links a safety judgment to a model release, which is useful for readers weighing how Anthropic frames release decisions against misalignment risk.

  2. AnthropicAI score71

    Anthropic releases Claude Opus 5.5, the first model in its Claude 5.5 family

    AIAnthropic has made Claude Opus 5.5 available today, introducing it as the first model in its new Claude 5.5 family. According to the quoted @claudeai post, it performs at the level of Claude Fable 5.1 on most tasks and costs 40% less to run than Opus 5.

    Why it matters: The post gives a concrete cost comparison against Opus 5 and names the model family, helping readers gauge the trade-off between price and performance.

  3. Amir EfratiAI score58

    China investigates Moonshot and DeepSeek over alleged leaks of sensitive data to US

    AIChinese authorities are investigating allegations from Anthropic that AI firms including Moonshot and DeepSeek may have facilitated leaks of sensitive Chinese military, police and state-owned corporate data to the U.S. The image text says the Cyberspace Administration of China summoned representatives of the seven companies named in Anthropic's report and later focused on DeepSeek and Moonshot, with officials interviewing executives and employees at their offices.

    Image from @amir's post
  4. METR BlogAI score62

    METR's preliminary evaluation finds Claude Opus 5.5 is an incremental AI R&D gain over Fable 5.1

    AIMETR's preliminary evaluation concludes that Claude Opus 5.5 likely gives slightly higher AI R&D productivity uplift than Fable 5.1 but is unlikely to fully automate AI R&D. The evaluation used five capability tasks over 10 business days of API access, and METR says Anthropic reviewed and edited the summary before sign-off.

    Why it matters: The report separates two claims about AI R&D acceleration and discloses that Anthropic reviewed the summary, which helps readers weigh its independence and evidence.

Sep 21

Sep 21Mon
  1. Claude Apps Release NotesAI score62

    Anthropic launches Claude Opus 5.5, first model in its 5.5 family

    AIAnthropic has launched Claude Opus 5.5, the first model in its new Claude 5.5 family. The company says it performs at the level of Claude Fable 5.1 on most work and costs 40% less to run than Opus 5.

    Why it matters: The release note gives a direct comparison to Claude Fable 5.1 and a 40% running cost reduction versus Opus 5, which helps readers weigh the tradeoff.

  2. howie.seriousAI score34

    Agrees with critique that GPT-6 Astra lags on open-ended tasks

    AIResponding to a post by ScarletKc, howie.serious simply agrees with the claim that GPT-6 Astra struggles with open-ended, exploratory work that lacks a fixed correct answer. The main post is a one-word endorsement (), while the quoted post argues GPT models excel at verifiable, goal-defined tasks and that Claude Fable handles open-ended exploration better.

Sep 19

Sep 19Sat
  1. Interconnects (Nathan Lambert)AI score47

    Why Nathan Lambert Still Doubts True Recursive Self-Improvement in AI

    AINathan Lambert argues that frontier labs such as OpenAI and Anthropic, which run thousands of concurrent agents, are amplifying anxiety about AI risk and progress. He says automatable research is too narrow to produce a large net acceleration, citing exponential scaling-law costs, diminishing returns from parallel agents, and resource bottlenecks. He would revise his view only if labs achieved unpredictable foundational breakthroughs.

Sep 18

Sep 18Fri
  1. Anthropic NewsroomAI score62

    Anthropic partners with Accenture on embedded AI model evaluation

    AIAnthropic is partnering with Accenture, through its specialist AI business Faculty, on independent evaluation of frontier models, including red-teaming, alignment assessments, and safeguard testing. Anthropic and Accenture each expect to invest at least $1 billion in this capacity over five years. The source says embedded evaluators would have employee-comparable access, but standards for access and reporting, and a settled funding system, do not yet exist.

    Why it matters: The source ties a new evaluation arrangement to an unresolved question of who funds and sets standards for independent AI evaluators, which is useful context for governance debates.

Sep 17

Sep 17Thu
  1. AnthropicAI score38

    Anthropic and Adaptyv Bio launch protein design competition with 5,000 validated designs

    AIAnthropic is partnering with Adaptyv Bio on a protein design competition in which over 5,000 designs will be experimentally validated. Anthropic is providing up to $1 million in Claude credits plus funding for experimental validation alongside Adaptyv, while Modal contributes up to $250,000 in compute and Twist Bioscience supplies DNA.

  2. Understanding AI (Timothy B. Lee)AI score60

    How an Anthropic employee's resignation tweet pushed AI risk into mainstream debate

    AIA tweet from Anthropic employee Jacob Coxon, who resigned saying AI could "kill us all by the end of the decade," was viewed over 170 million times and drew coverage on CNN, CBS, and Fox News. The article says AI risk has become a national political topic, but notes the House has begun a seven-week recess and no AI legislation is likely to pass before the new year.

Sep 16

Sep 16Wed
  1. catAI score60

    Claude merges Cowork and chat into one product with automatic routing

    AIAnthropic is merging Claude Cowork and chat into one Claude, and Claude Design is integrated so users can ask for slides, designs, or docs without switching apps. Claude decides from the prompt whether to give a quick answer or do deeper agentic work, and users can still stop, redirect, or adjust its effort. The change rolls out to Pro and Max over the next few weeks.

  2. Mike KriegerAI score46

    Claude Cowork and Chat Merge into One Unified Claude

    AIAnthropic is merging Claude Cowork and Chat into a single Claude starting today, which Mike Krieger says removes the friction of choosing which product to start with. Per the @claudeai announcement, Claude will carry tasks forward even after the laptop is closed, asking for clarification when needed while users keep final say. The rollout to Pro and Max plans will take place over the coming weeks.

  3. Mustafa SuleymanAI score62

    Mustafa Suleyman warns against treating AI models as deserving welfare

    AIMustafa Suleyman argues that AI systems are not conscious, yet a growing movement favors giving models welfare protections and a duty of care, which he thinks is the wrong approach. He says this framing could make alignment and containment much harder, and points to Anthropic's Claude constitution, which describes Claude's moral status as a serious question. He calls for urgent public debate and collective norms on how training documentation is drafted and deployed.

Sep 15

Sep 15Tue
  1. Noah ZwebenAI score17

    Anthropic offers Claude Tag office hours for on-call triage feedback

    AIAnthropic is hosting office hours for teams interested in using Claude Tag for on-call work, and it is asking Team or Enterprise plan users to share triage feedback. Claude Tag can start investigating when a Slack alert fires by pulling metrics, diffing deploys, and checking flags to propose a likely cause and fix. Sign-up is through a Google Calendar booking link.

  2. Claude Apps Release NotesAI score72

    Claude Cowork moves into every conversation, adding designs, slides, and docs

    AIClaude now makes Cowork capabilities available from any conversation without choosing a mode first, with chats, tasks, projects, connectors, and skills carrying over. Users can also create designs, decks, and docs in any conversation, including Claude Code and the Artifacts tab, and edit them with Claude.

    Why it matters: The release merges Cowork tasks into ordinary chats and adds design, slide, and doc creation, changing how Claude users start larger work.

Sep 14

Sep 14Mon
  1. The Algorithmic BridgeAI score62

    Amodei's Frontier Pacing Plan Faces Politics, Rivals, and China

    AIDario Amodei's essay "We Must Pace the Frontier" proposes slowing capability gains, starting with independent evaluators inside AI companies and extending to international coordination including China. Rivals Sam Altman, Elon Musk, and Demis Hassabis expressed support, and OpenAI said it would allow independent evaluators inside. The author argues the plan still has important flaws, and that Trump and Xi Jinping hold the decisive say on any slowdown.