Skip to contentSkip to stories

Updated

#Product update

Items with an AI score under 20 are hidden. Show low-relevance items

Sep 15

Sep 15Tue
  1. Google AI StudioOfficialAI score72

    Google releases Gemini 3.8 Live and 3.5 Transcribe for real-time voice apps

    AIGoogle AI Studio released Gemini 3.8 Live, a native speech-to-speech model with an Extended Thinking variant, and made it available through the Live API. Gemini 3.5 Transcribe, released last month, supports 85+ languages with a reported 4.0% streaming and 2.6% non-streaming Word Error Rate, and accepts a custom vocabulary of up to 1,000 terms. Live API audio pricing is listed at $0.005/min for input and $0.018/min for output.

    Why it matters: The post lists concrete Live API capabilities, per-minute audio pricing, and transcription accuracy figures, helping developers weigh voice agent options against their own cascaded pipelines.

  2. Microsoft Foundry BlogOfficialAI score32

    Microsoft Launches Foundry Dev Pack to Install Foundry Development Tools in One Command

    AIMicrosoft has launched Foundry Dev Pack, an all-in-one installer that sets up tools for Microsoft Foundry development across the terminal, IDE, and coding agents. Depending on the environment, it installs Azure CLI (az), Azure Developer CLI (azd) with the Microsoft Foundry Extension for azd, the Microsoft Foundry Skill, the Microsoft Foundry Toolkit for Visual Studio Code, and Foundry Canvas (preview), with the last two conditional on VS Code or GitHub Copilot App being present.

  3. Josh WoodwardXAI score31

    Gemini Notebook adds spoken Q&A and lecture audio notes for students

    AIGoogle's Gemini Notebook now offers live spoken Q&A over class materials in about 100 languages, and lets students record lectures on the go with audio notes saved automatically to a chosen notebook. University students in 140+ countries can also still get a free Google AI Plan for higher limits and access to more Google products.

  4. Greg BrockmanXAI score46

    ChatGPT Work adds Data agent for dashboards and actions on company data

    AIOpenAI's Greg Brockman says ChatGPT Work can operate over and act on a company's data, including building dashboards, by connecting existing tools such as PowerBI, Tableau, Clickhouse, Oracle BI, and AWS Redshift. The linked ChatGPT announcement describes a Data agent with a Data Plugin that turns company data into answers, interactive dashboards, and actions through conversation.

  5. Logan KilpatrickXAI score44

    Google launches Gemini 3.8 Live audio models with 97-language support

    AIGoogle has released Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, described as new state-of-the-art live audio models available at frontier pricing and performance. The 3.8 Live model supports 97 languages with seamless switching between them, along with async tool calls.

    Image from @OfficialLoganK's post
  6. Google AIOfficialAI score72

    Google rolls out Gemini 3.8 Live and Extended Thinking across consumer, developer, and enterprise channels

    AIGoogle is rolling out Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking across several channels. Consumers get them in Search Live and Gemini Live, developers get public preview access through the Gemini API, and enterprises get private preview through Gemini Enterprise, with Customer Experience support coming soon.

    Why it matters: The post lays out where each Gemini 3.8 Live variant reaches consumers, developers, and enterprises, which clarifies access paths for a voice model release.

  7. Google DeepMindOfficialAI score33

    Gemini 3.8 Live Extended Thinking adds upgraded reasoning for real-time programming tutoring.

    AIGoogle DeepMind demonstrated 3.8 Live Extended Thinking acting as a programming tutor in Gemini Live. Both 3.8 Live models feature upgraded reasoning, near real-time visual understanding, automatic detection across 97 languages, and background tool calling that doesn't interrupt the chat. The Extended Thinking variant adds higher performance and precision for harder tasks and narrates its progress, and it is available in Gemini Live in the Gemini app or through the Gemini API via Google AI Studio.

    Video from @GoogleDeepMind's post
  8. Google · Innovation & AIOfficialAI score52

    Google says its language technology now covers over 300 languages with new speech, data, and on-device tools

    AIGoogle reports that its technologies and products now power everyday interactions in more than 300 languages used by over 7 billion people, about 86% of the global population. The post describes new speech models, including Gemini 3.5 Live Translate and Gemini 3.5 Transcribe, plus the TranslateGemma open translation models trained across 55 languages.

  9. TechNode · AINewsAI score38

    Twoo Adds AI as a Third Member to Two-Person Relationships

    AIShanghai-based Twoo is building an AI that participates in conversations between two users, sharing context, remembering experiences, and helping them plan activities together rather than serving one person. Founder Cobe Chen says the product's monetization includes in-app "shells" that unlock outfits for its octopus AI character, plus premium memberships for professional uses such as collaborative writing and teaching. The company is preparing overseas expansion into North America, Japan, and South Korea.

  10. MiniMax Design (H3)OfficialAI score26

    MiniMax Design canvas runs Astra agent and Blender to produce full scene

    AIMiniMax Design lets a single brief drive a full production workflow on one canvas, with the Astra agent working in Blender through an official connector to build the scene and camera direction. The final video is generated with MiniMax H3 from the same canvas, with outputs syncing directly onto the canvas.

  11. Gemini API ChangelogOfficialAI score62

    Google makes Gemini 3.8 Live models generally available for real-time voice

    AIGoogle has made two audio-to-audio models, Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, generally available through the Live API. Gemini 3.8 Live, model ID gemini-3.8-live, is the default for low-latency voice agents, with interleaved reasoning and asynchronous function calling. Gemini 3.8 Live Extended Thinking, model ID gemini-3.8-live-extended-thinking, supports background reasoning during live audio and is recommended when more reasoning is needed.

    Why it matters: The changelog names two model IDs and their intended use, showing how Live API developers can choose between low-latency voice and higher background reasoning.

Sep 14

Sep 14Mon
  1. Intern Large ModelsOfficialAI score23

    Intern-S2-397B, a scientific multimodal model, gets SGLang Day-0 support

    AISGLang announces Day-0 support for Intern-S2-397B from Intern Large Models, a 397B multimodal foundation model built for scientific intelligence and long-horizon agents. The model is pre-trained directly on raw scientific literature pages without parsing and uses reinforcement learning across more than 20 scientific domains, from biomolecule design to material generation. It also applies black-box agentic reinforcement learning in large-scale sandboxed environments.

  2. vLLM BlogOfficialAI score53

    Novita AI open-sources Chord, a W4A16 MoE kernel for Kimi K2.x on vLLM

    AINovita AI has open-sourced Chord, a W4A16 MoE CUDA operator with BF16 activations, INT4 weights and group-32 scales, built for Kimi K2.x serving shapes. Measured per layer against public Humming, it reports 1.11–1.20x on H200 EP8 prefill, 1.17–1.33x on H200 TP8 serving, and 1.81–2.15x on B300 EP8 decode against an untuned Humming default. Integration of the grouped operators with vLLM's Humming backend is still a work in progress.

  3. Claude Apps Release NotesOfficialAI score46

    Anthropic Launches Salesforce Plugin for Claude in Beta

    AIAnthropic has launched a Salesforce plugin for Claude that brings sellers' accounts, opportunities, and pipeline into the Claude app, with 37 pre-built sales skills. The beta is available on all paid plans for organizations Salesforce approves through its beta sign-up.

  4. MiniMax (official)OfficialAI score41

    MiniMax H3 video generation exceeds 2× real-time on 8× B200

    AIMiniMax H3 with SGLang-Diffusion and VDN-H3 generates 14.4 seconds of 768p video in 9.0 seconds end-to-end after warmup on 8× B200 GPUs. Eight-step denoising takes 6.9 seconds, exceeding 2× real-time, with no measured quality regression versus dense 50-step H3 across 103 test prompts.

  5. Google AIOfficialAI score44

    Google Labs' Dreambeans turns connected data into personalized daily stories

    AIGoogle Labs has launched Dreambeans, an opt-in experience that connects data from Gmail, Calendar, Search, the Gemini app, and Google Photos face grouping to generate personalized illustrated stories. It can spot events such as a friend's upcoming birthday and suggest gift ideas, with daily in-app notifications when stories are ready. Users can tap a story for links to next steps like movie trailers or gift purchases, and give a thumbs-down to help the system learn their preferences.

    Video from @GoogleAI's post
  6. Intern Large ModelsOfficialAI score25

    Intern-S2-397B gets Day-0 support in vLLM

    AIIntern-S2-397B, a model built for long-horizon scientific research, now has Day-0 support in vLLM. The model brings multimodal, reasoning, coding, and scientific agent capabilities, and vLLM has published a run recipe for it.

  7. Baidu Inc.OfficialAI score38

    Baidu's Miaoda upgrade expands no-code platform for enterprises and creators

    AIBaidu's Miaoda no-code platform has upgraded with enhanced AI agents for design, app generation, and testing. The update adds enterprise tools for private deployment and collaboration, plus a marketplace linking businesses with creators for templates and custom development. Baidu says Miaoda has served over 40M users and enabled 5M business apps.

    Image from @Baidu_Inc's post
  8. MiniMax (official)OfficialAI score36

    MiniMax H3 community projects speed up open-source video generation

    AIMiniMax highlighted open-source community progress on its H3 video generation model, which it built with native stereo audio and multimodal reference control. Recent highlights include FastH3's 4-step distillation running on DGX Spark and Apple Silicon, and NVIDIA's Sol-H3 generating 15 seconds of 768p video with audio in 6.6 seconds on 8×B300 in a warm-inference benchmark. Other releases include VDN's faster-inference attention work with code and weights, and 8-step Acc-LoRAs from Alibaba PAI, with LightX2V offering 4- and 8-step Turbo LoRAs.

    Image from @MiniMax_AI's post
  9. MiniMax Design (H3)OfficialAI score22

    Hailuo AI highlights AI-driven 3D workflow integration with Blender

    AIHailuo AI promotes bringing AI tools into professional 3D production workflows. A quoted post from @akiyoshisan describes connecting MiniMax Design, using GPT-6 Astra, to Blender via MCP for direct 3D creation. The creator says this lets 3D representation in Blender be built into an AI production flow rather than relying on MiniMax Design alone.

Sep 13

Sep 13Sun
  1. Satya NadellaXAI score20

    Microsoft Foundry adds security, auditability, and FinOps to long-running agents

    AISatya Nadella highlighted a Microsoft Foundry example showing how long-running, multi-agent, multi-model workflows can be built with security, safety guardrails, auditability, and FinOps included from the start. The example was shared from Jeff Hollan's post, which says Foundry's observability and governance features keep agents within user-defined bounds, including control over data access, data flow, action traceability, and cost budgets.

Sep 12

Sep 12Sat

Sep 11

Sep 11Fri
  1. VercelOfficialAI score26

    Tailscale's Aperture model router is built on Vercel AI Gateway

    AITailscale offers instant access to hundreds of models for any user in a secure tailnet through its customer-facing model router, Aperture. Aperture is built on Vercel's AI Gateway and offers zero data retention, zero markup with free BYOK, and cost and usage data on every response.

  2. Cognition Blog (Devin, Windsurf)OfficialAI score51

    Cognition introduces Fusion in Devin Desktop and CLI for lower-cost coding

    AICognition is making Fusion available in Devin Desktop and CLI, a harness where a frontier lead model plans and reviews while a cheaper sidekick executes. Across listed coding benchmarks, Cognition reports Fusion cuts cost per task by about 11% to 46% versus the lead model alone, while the sidekick does the implementation work. The post recommends pairing Fable 5.1 with SWE-2, and argues price per task matters more than price per token.