Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 2

Oct 2Fri
  1. AI at MetaOfficialAI score61

    Meta shares six math papers from mathematician-AI collaborations on open problems

    AIAI at Meta says mathematicians used Muse Spark 1.1 and Muse Spark 1.2 in Thinking Mode through the standard meta.ai chat interface to find solutions to open problems. The company is sharing six resulting papers, each marking which passages were drafted primarily by humans or AI, with mathematicians guiding the work and a second group reviewing it.

    Why it matters: The post shows AI models helping mathematicians on open problems, with human guidance, peer review, and disclosure of AI-drafted passages, which clarifies how such collaborations are documented.

  2. eric zakariassonXAI score47

    xAI releases experimental TypeScript SDK with Grok models and tools

    AIxAI has released an experimental TypeScript SDK, installable via npm install @xai-official/sdk, that covers text, voice, image, and video in one package. It provides access to the latest Grok models along with server-side tools including real-time X search, web search, code execution, and remote MCP.

    Video from @ericzakariasson's post
  3. DatabricksOfficialAI score44

    Omnigent: open-source meta-harness coordinating Claude Code and Codex agents

    AIDatabricks' new open-source meta-harness, Omnigent, lets multiple coding agents such as Claude Code and Codex share sessions, rules, and security policies in one system. A walkthrough by @leonvz demonstrates forking work across agents, multi-agent review and debate with Debby, and splitting implementation across subagents with Polly.

    Video from @databricks's post
  4. ChatGPTOfficialAI score60

    ChatGPT adds Finances for subscriptions, budgets, credit, and investments

    AIChatGPT now offers Finances, which can find forgotten subscriptions, flag unfamiliar or duplicate charges, and track recurring bills that have increased. It also provides weekly updates, monthly spending breakdowns, budget building, credit score tracking, debt payoff planning, emergency fund estimates, and investment mix and concentration views. Users can access it at

    Why it matters: The post lists concrete finance features across budgeting, debt, and investments, showing how a general assistant is expanding into personal money management.

  5. ChatGPTOfficialAI score60

    Finances in ChatGPT rolls out to Free and Go users in the U.S.

    AIChatGPT's Finances feature is rolling out to Free and Go users in the U.S. Users can securely connect their accounts through Plaid and Experian to get answers based on their own financial information.

    Why it matters: The post names the rollout scope and the account connection method, which helps readers judge how the feature handles personal financial data.

    Video from @ChatGPT's post
  6. SpaceXAIOfficialAI score13

    Grok 4.7 is 50% off in Ramp Router until October 6

    AIRamp Router is offering Grok 4.7 at 50% off through October 6, in a partnership with SpaceXAI to lower token costs. Developers can get an API key at router.com to access the discounted model.

  7. Harrison ChaseXAI score53

    Google Research's Cogentic uses multi-agent proof search to produce verified results

    AIGoogle Research's Cogentic is a multi-agent harness running on Gemini that searches for proofs of open theoretical computer science problems without expert hints. It runs rounds where an orchestrator launches provers, two adversarial verifiers must both accept each draft, and shared disk documents store attempts and verified lemmas. The system produced new results on five open problems in online learning, auction theory, and mechanism design, each checked by domain experts.

  8. Harrison ChaseXAI score38

    LangSmith Custom Apps lets teams build trace review UIs in-workspace

    AILangChain's LangSmith Custom Apps lets agent teams build their own review UI over their traces and publish it directly into the workspace. Developers build the interface on their LangSmith data, while hosting, authentication, and permissions are handled by the platform.

  9. SunoOfficialAI score8

    Braylon Browner and Nina McNeely reveal creative process behind Suno music video

    AISuno's X post offers a behind-the-scenes look at how Braylon Browner and Emmy-winning choreographer and director Nina McNeely moved from dance to song to a complete creative world built around movement. The post gives no further details on tools, production steps, or release timing.

    Video from @suno's post
  10. CursorOfficialAI score42

    Cursor's Rollouts detects deployment regressions and launches cloud agent fixes

    AICursor introduced Rollouts, a tool that writes a monitoring plan and watches changes as they deploy to catch regressions before users see them. When Rollouts detects a regression, it identifies the offending PR and opens an issue, and one click starts a cloud agent to fix it. Rollouts usage credits are included through Oct 3.

  11. GitHubOfficialAI score44

    GitHub Copilot adds Project HydraFusion and new models to model picker

    AIGitHub has made the Project HydraFusion research preview available in the GitHub Copilot app and @code, where it orchestrates multiple models rather than acting as a single model. New models from Anthropic (Fable 5.1 and Opus 5.5) and OpenAI (GPT-6.1 Sol) are also now selectable in the Copilot model picker.

  12. Together AIOfficialAI score34

    Together AI shares how its team uses AI to boost collective productivity

    AITogether AI's CPO and product team outlined how they use AI to make the whole team more productive, not just individuals. The approach includes a shared context repo readable by any AI harness, cutting half a day of research to about 5 minutes, and evals that test their product the way agents actually use it.

  13. Ant LingOfficialAI score27

    Ling-3.1-flash now available free on OpenRouter

    AIAnt Ling has made Ling-3.1-flash available on OpenRouter at no cost, inviting users to try it and share feedback. The post provides no details on model size, benchmarks, context length, or pricing beyond the free access.

  14. Harrison ChaseXAI score26

    LangChain improves memory for managed Deep Agents in enterprise settings

    AIHarrison Chase says LangChain is improving memory in managed Deepagents, noting that memory is difficult to get working well in company settings. The linked background post describes user memory in Managed Deep Agents 0.8, which lets an agent remember the people it works with.

  15. ElevenLabsOfficialAI score32

    ElevenLabs earns FedRAMP 20x Class A certification for federal voice AI

    AIElevenLabs has achieved FedRAMP 20x Class A certification, covering ElevenAgents and its Text to Speech and Speech to Text APIs when run in Zero Retention Mode with US data residency. The company says the published evidence package and FedRAMP Marketplace listing should shorten security reviews for public sector and regulated buyers. The certification extends ElevenLabs for Government, its dedicated offering for federal agencies.

    Image from @ElevenLabs's post
  16. PixVerseOfficialAI score16

    PixVerse plugin turns product ideas into unboxing videos

    AIPixVerse announced a plugin that lets users generate an unboxing video from a product description, setting, and presenter reaction. According to the post, the AI agent uses PixVerse to produce the video.

    Video from @PixVerse's post
  17. Microsoft AIOfficialAI score41

    Microsoft MAI voice models now available on LiveKit for agents

    AIMicrosoft's MAI-Transcribe-2-Streaming, MAI-Voice-2.1, and MAI-Voice-2.1-Flash are now live on LiveKit for building voice agents. LiveKit says MAI-Transcribe-2-Streaming debuts at #1 on the Artificial Analysis accuracy leaderboard, and suggests pairing it with MAI-Voice-2.1-Flash for efficient, expressive voice agents.

  18. Microsoft CopilotOfficialAI score20

    Copilot Code lets more people build apps and workflows

    AIMicrosoft's Copilot Code is designed to help more people turn ideas into apps, workflows, and solutions for their work. Microsoft Copilot EVP Jacob Andreou discusses how the product expands who gets to build.

    Video from @MSFTCopilot's post
  19. Google AIOfficialAI score62

    Google launches Project Suncatcher prototype satellite to test TPUs in orbit

    AIGoogle AI announced that its Project Suncatcher prototype satellite, built with Planet, has launched into orbit on SpaceX's Transporter-18 rideshare mission. The initial mission will gather data on how Google TPUs handle the physical stress and extremes of spaceflight. The post says low Earth orbit systems could generate up to 8x more solar power than on Earth, and that future work may link multiple satellite constellations for scaled machine learning.

    Why it matters: The post explains a space-based machine learning prototype and why orbit's near-constant sunlight matters, which helps readers weigh the idea's practical potential.

    Video from @GoogleAI's post
  20. Kilo (acq. by Anaconda)OfficialAI score36

    Ling 3.1 Flash is free in Kilo Code until October 13

    AIKilo Code is offering Ling 3.1 Flash for free until October 13, with the model served by Novita Labs. Ant Ling's background post describes the model as roughly 560B total parameters with about 25B active per token and a context window of up to 1M tokens. Ant Ling says it scores 1,673 Elo on GDPVal-AA v2.1, 75.16 on FrontierSWE, and 65.35 on HealthBench Professional, and plans to open-source it soon.

  21. Microsoft AIOfficialAI score36

    Microsoft's MAI voice models now available on OpenRouter

    AIMicrosoft AI's MAI voice models are now accessible through OpenRouter, according to the announcement. The quoted OpenRouter post highlights MAI-Voice-2.1, a text-to-speech model that supports 23 languages with native accents and is priced at $22 per 1M characters.

  22. Hugging Face BlogOfficialAI score70

    Ai2 open-sources AstaBrief 8B, a fast model for generating cited research reports

    AIAi2 released AstaBrief 8B, an open-weights model that turns a research question and retrieved literature excerpts into a cited report, along with its training data. The model runs as Fast mode in Asta, averaging 51.1 seconds per report versus 178.5 seconds for Thinking mode, about 3.5x faster. The post also describes filtering synthetic training data by citation density and building DPO pairs judged by two models that agreed.

    Why it matters: The post explains how supervised fine-tuning, preference data, and citation-density filtering were used to build a cited-report model, which is useful for teams training their own models.