Skip to contentSkip to stories

Updated

#Agent

Showing low-relevance items too. Hide low-relevance items

Oct 8

Oct 8Thu
  1. SiliconANGLE · AINewsAI score62

    Manus raises over $500M at reported $4B valuation after Meta deal collapsed

    AIManus, the developer of the Manus AI agent, has raised more than $500 million led by Boyu Capital, with Tencent also investing. Bloomberg had reported a $4 billion valuation, roughly double Meta's reported offer last December, which Chinese regulators blocked in April. The funding comes days after Manus 2.0 added a new harness, Cloud Computer, and Cue, which lets agents use email accounts and digital wallets.

  2. SiliconANGLE · AINewsAI score24

    CoreWeave Pitches Open Full-Stack AI Cloud With Forge Development Platform

    AICoreWeave is positioning its AI cloud around an open development loop, connecting training, inference and evaluation through its newly announced CoreWeave Forge platform. Chief marketing officer Jean English said the company wants production learnings to improve models and agents and that the loop should work across different models, frameworks and clouds. She argued that competitive differentiation extends beyond GPUs to partner tooling, infrastructure and APIs.

  3. ClaudeDevsOfficialAI score42

    Anthropic halves Sonnet 5.5 cache read prices on Claude Platform

    AIAnthropic has cut Sonnet 5.5 cache read pricing in half to $0.10 per million tokens, with input at $2 and output at $10 per million tokens. The company says this makes Sonnet 5.5 roughly 20% cheaper on most agentic work. The change applies to API usage only and does not alter Claude Code usage limits.

  4. Claude Code · GitHub ReleasesOfficialAI score56

    Claude Code v2.1.295 adds hook failure blocking and gateway controls

    AIClaude Code v2.1.295 adds onFailure: "block" for command and HTTP hooks, so a hook that cannot start, times out, or exits unexpectedly blocks the action. The release also adds an optional models list for Claude apps gateway upstreams, plus upstream_request_id in the inference audit event, and fixes a range of MCP, plugin, and terminal issues.

  5. LiveKitOfficialAI score4

    LiveKit and Modal host AI Agents Speakeasy in LA on October 14

    AILiveKit is hosting an AI Agents Speakeasy with Modal during LA Tech Week on Wednesday, October 14, from 6 to 9 p.m. in Los Angeles. The event offers cocktails, food, and conversation with people building AI agents, with RSVPs available through a Partiful link.

    Image from @livekit's post
  6. Eugene SmartsXAI score44

    Grok Bot runs named AI coworkers on one shared persistent cloud computer

    AIGrok Bot, from dot.com, lets an office roster of named AI workers such as Chief, Sales Outbound, Talent Scout, and Inbox Manager share one persistent cloud computer. Sales Outbound uses Hex and Salesforce to queue 36 personalized outreach drafts overnight, with human review before anything is sent. Isolation is set per user rather than per bot, so every worker shares the same browser cookies, files, and authenticated SaaS sessions.

    Image from @EugeneSmarts's post
  7. Harrison ChaseXAI score28

    LangChain's Sam explains decision models and using Jev in harnesses

    AISam from LangChain discusses where decision models fit inside an agent harness and how to use Jev with LangChain. The post links to a LangChain blog on building a harness with Jev, which the background post says Jev from typesafeai popularized alongside OpenAI's Decisions API and Databricks' ai_decide function.

  8. The DecoderNewsAI score62

    Anthropic's updated usage policy bans sustained abusive behavior toward Claude

    AIAnthropic has updated Claude's usage policy for the first time in over a year, banning sustained and needless abusive or cruel behavior toward Claude. The company says ordinary frustration, pushback, dark creative themes, and model testing are not covered, and that the rule applies only in extreme cases. Violations can lead to warnings, throttling, restriction, suspension, or termination of access.

  9. Codex · GitHub ReleasesOfficialAI score36

    Codex 0.162.0 adds managed worktree tools and clickable URLs in the TUI

    AIOpenAI's Codex 0.162.0 release adds tools for creating and listing managed Git worktrees from trusted local projects when the worktrees feature is enabled. The update also lets users pin tasks in the agent Command Center, copy transcript blocks with /copy, and make URLs clickable in approval headers, questions, and warnings, along with several Linux and Windows sandbox fixes.

  10. 🚨 AI News | TestingCatalogXAI score36

    Gemini Agent for Business may add Claude Opus 5 and Sonnet 5.5

    AIGoogle's recently announced Gemini Agent for Gemini Business is reportedly set to offer Gemini Argon 4, Gemini Flash 3.8, Claude Opus 5, and Claude Sonnet 5.5. If accurate, it would mark the first time Claude models appear on Google's platform alongside Google's own models, which the post frames as a way for Google to compete for enterprise customers.

    Video from @testingcatalog's post
  11. Tessl BlogOfficialAI score29

    One Brain Means Owning Your Organizational Memory

    AILeapfrog, a small team doing high-volume AI visual and production work for fashion and brand clients, is building a "one brain" system that makes company knowledge and client context searchable through natural-language agents. The starter stack described is OpenClaw in a sandbox, a GitHub repository, Obsidian on the local machine, and Telegram as the access point. The system's research structure had roughly 1,200 files at the time of the talk.

  12. Tessl BlogOfficialAI score42

    Agent Skills Should Be Treated as Supply Chain Components

    AITessl's talk at AI Native DevCon London argues that agent skills, which can be markdown files with instructions and bundled material, act as supply chain components that can shape agent behavior. The author says reading SKILL.md once is insufficient because risks can sit in supporting files, updates, and workspace trust settings. He identifies the danger as the combination of private context, untrusted content, and external communication, and cites research scanning roughly 4,000 public skills for issues including malware-like behavior.

  13. elvisXAI score42

    Voyager: an open harness for creative AI work across video and games

    AIElvis Saravia argues that creative work needs domain-specific agent harnesses rather than coding-oriented ones, and he highlights Voyager as an open harness for video, graphics, and games. According to the quoted post, Voyager lets agents work with local files and drive apps such as Blender, DaVinci Resolve, and Unity, and it is designed to work with models like Opus, Astra, and DeepSeek.

    Video from @omarsar0's post
  14. Tessl BlogOfficialAI score52

    Cisco engineer argues agent skills need a context pipeline with evals

    AIJohn Groetzinger, writing in a personal capacity rather than for Cisco, argues that enterprise skills need packaging, evaluation, syncing, and distribution rather than scattered markdown files. He describes using skills to make cheaper models viable, converting curated TAC knowledge-base articles into maintained skills, and rolling out an eval framework across teams. He also describes syncing a repository README to Confluence with a deterministic script.

  15. Artificial AnalysisOfficialAI score34

    Harvey LAB-AA: Artificial Analysis benchmark for legal AI agents

    AIArtificial Analysis has released Harvey LAB-AA, an evaluation built on Harvey's LAB dataset and developed in collaboration with Harvey. Full results are published on the Artificial Analysis evaluations page, alongside Harvey's commentary on the benchmark and human expert preferences.

  16. AWS Machine Learning BlogOfficialAI score46

    AWS Pays Per Inference for AI Agents with BlockRun and Incarna

    AIAmazon Bedrock AgentCore payments lets AI agents pay for model inference one request at a time, using x402 with USDC on the Base network. Incarna used the service to connect its agents to BlockRun, a pay-as-you-go router serving more than 90 models from more than 15 providers. Spending limits are enforced at the infrastructure layer, outside the model.

  17. 🚨 AI News | TestingCatalogXAI score49

    Voyager desktop app lets AI agents work inside creative tools on Mac

    AIVoyager has launched a Mac desktop app that lets AI agents read project files and operate creative tools such as After Effects, DaVinci Resolve, Blender, and Unity. The agents produce editable results for video edits, motion graphics, color grading, 3D scenes, and game prototypes. Built-in and custom skills, plus a memory that learns each user's workflow, are included.

    Video from @testingcatalog's post
  18. NVIDIA Technical BlogOfficialAI score29

    NVIDIA KGMON Places Second in KDD Cup 2026 Data Agents Competition

    AIThe NVIDIA KGMON team placed second in the KDD Cup 2026 Data Agents competition with a system built around a smaller, clearer, and easier-to-verify agent harness. The competition required agents to answer natural-language questions over heterogeneous sources, including databases, CSV and JSON files, prose documents, PDFs, and briefing videos.

  19. laurenXAI score29

    Omarchy seeks feedback on Grok Bot plugins and integrations

    AILauren Tan invites users of Grok Bot on Omarchy and developers building plugins for it to share feedback and feature requests. The post points to the Omarchy plugin catalog and asks what integrations could be supported. Background from DHH says SpaceXAI joined the Omacom Foundation as a Founding Corporate Patron, contributing $1,500,000 in Grok tokens for Omarchy's maintenance and development.

  20. Tessl BlogOfficialAI score42

    Tessl Proposes Executable Specs to Verify AI Coding Agent Output

    AITessl argues AI code review is slow because generated code outpaces trust, and proposes executable specs that let agents check preview environments against product intent. Its spec reviewer splits work between a planner agent that extracts requirements and parallel verifier agents that test each one against the code and base branch.

  21. TechCrunch · AINewsAI score72

    Google launches unified Gemini agent for businesses, consumers to follow

    AIGoogle announced at a Google Cloud event a unified Gemini agent that can plan and complete tasks from a single interface, starting with businesses. The agent has its own Workspace account, connects to systems including Google Workspace, Microsoft 365, Slack, and Jira through MCP, and writes an audit trail attributed to the agent. Google said consumers will get access later, after it addresses security, scale, and performance.

    Why it matters: The source details how the agent takes objectives, connects to business systems, and logs actions, showing how enterprise agent deployment is being structured.

  22. PyTorch BlogOfficialAI score62

    NVIDIA Dynamo adds session-level IDs to route and cache agentic inference

    AINVIDIA Dynamo uses a unified session-level identifier to make its inference stack aware of agent sessions, subagents, and their KV cache across turns and tool calls. On SWE-bench, two TP4 MiniMax-M2 replicas on one 8xH100 node gained roughly 12-16% throughput from program-aware scheduling over KV-aware routing alone. The post also describes experimental shared-pool indexing and a proposed KvHint interface for session-aware cache policies in vLLM and SGLang.

    Why it matters: The post explains how session identifiers let an inference stack track agent working sets, with measured throughput gains on SWE-bench and agentic RL rollouts.

  23. SCOTTY BEAMXAI score40

    Grok Bot adds email, coordination, and cheaper plans from $20/month

    AIX's Scotty Beam says Grok Bot, in public beta since August 11, 2026, now gets its own email address, a main bot that coordinates other bots, and expanded access to subscription plans starting at $20 per month. The post says the bot can organize inboxes, research topics, build software, and keep working after users close their laptops. The post also claims the entry price dropped 90%, but does not give the prior price.

  24. SantiagoXAI score42

    Voyager: open harness connecting AI models to creative apps like Blender

    AIVoyager is an open harness for creative work that connects models with applications to build videos, graphics, and games. It works with Blender, DaVinci Resolve, After Effects, Ableton, and Unity, operating similarly to Codex or Claude Code. The harness is designed to get strong creative results from models such as Opus, Astra, and DeepSeek.

    Video from @svpino's post
  25. SiliconANGLE · AINewsAI score38

    Automation Anywhere to acquire Boost.ai to expand customer-facing voice AI

    AIAutomation Anywhere Inc. announced an agreement to acquire Boost.ai Inc., a conversational voice AI company, from Nordic Capital, to extend its autonomous enterprise platform into customer experience. Boost.ai supports more than 36 languages, serves hundreds of customers in regulated industries and Europe, and maintains more than 650 deployments and about 600 live AI agents. The deal follows Automation Anywhere's late 2025 acquisition of Aisera Inc.

  26. Sundar PichaiOfficialAI score62

    Google introduces Gemini agent as a single universal agent for work

    AIGoogle introduced a new Gemini agent that combines question answering, knowledge work, image and media creation, and code writing in one prompt box. The agent connects to personal workflows, systems of record, and enterprise controls, and runs in the cloud with a shared memory and personalization graph. It can create sub-agents for multi-step tasks, act as a coworker agent with its own identity, and orchestrate across multiple models to balance quality and cost.

    Image from @sundarpichai's post
  27. OpenRouterOfficialAI score22

    Sales workflow tool cuts demo prep and CRM time by about an hour

    AIThe post says demo prep fell from 30 minutes to 5, note-taking from 15 minutes to 2, and CRM updates from 30 minutes to 7 per call. It estimates this saves about an hour per call, letting each rep take roughly two more calls a day while staying prepared.

    Image from @OpenRouter's post