Skip to contentSkip to stories

Updated

#Product update

Items with an AI score under 20 are hidden. Show low-relevance items

Sep 24

Sep 24Thu
  1. Lewis Tunstall @ COLM 🌉XAI score42

    Hugging Face releases over 5,000 RL environments for data science tasks

    AIHugging Face released SmolDataEnvs, more than 5,000 open-source RL environments aimed at real-world data science tasks. They target the gap between simple educational games and frontier-level benchmarks, especially for improving coding in models under 10B parameters. The environments are designed as a testbed for developing new RL methods such as GRPO or OPSD.

  2. Baseten BlogOfficialAI score44

    LangSmith Fine-Tuning Trains Open Models on Agent Traces via Baseten Loops

    AILangChain launched LangSmith Fine-Tuning, which lets users fine-tune open models on their LangSmith agent traces using the open-source smithtune CLI. Training runs on Baseten Loops in the user's own workspace, and smithtune deploy places the evaluated checkpoint on a Baseten Dedicated Inference deployment. Loops is in early access, so users may need to request access for their workspace.

  3. GitHub Blog · AI & MLOfficialAI score66

    GitHub Security Lab shows an LLM agent running AI-driven fuzzing for C/C++ projects

    AIGitHub Security Lab describes the Fuzzing Taskflow, an LLM agent pipeline that identifies entrypoints, writes harnesses, runs AFL++, reads coverage reports, and triages crashes for C/C++ repositories. The agent makes decisions while MCP tools handle execution, and state is stored in a SQLite database. The post also warns that the taskflow runs AFL and build commands directly on the host, so it should be used only in disposable environments without elevated privileges.

    Why it matters: The post explains how an LLM agent automates fuzzing steps like harness writing, coverage gap chasing, and crash triage, with a runnable workflow and design tradeoffs.

  4. Azure BlogOfficialAI score67

    Microsoft Foundry adds voice agents and continuous optimization for production agents

    AIMicrosoft Foundry expands its agent platform with voice agents in public preview, long-running resilience for hosted agents, and tools for evaluating production agents. The post also says GPT-6 Sol, GPT-6 Luna, and Claude Opus 5.5 are now available in Foundry. Agent optimizer, Insights, and Rubric evaluator are described as tools for continuous improvement, with some reaching general availability later this month.

    Why it matters: The post shows how Foundry combines model choice, voice agents, long-running resilience, and production evaluation into one agent workflow, with a customer example.

  5. LiveKitOfficialAI score28

    LiveKit tests Gemini 3.8 Flash-Lite TTS in a live voice agent

    AILiveKit tested Gemini 3.8 Flash-Lite TTS inside a LiveKit agent, letting users direct a voice line by line and hear it hold up in a real conversation. The post highlights expressive speech, custom voices, and a production-ready voice library.

  6. Google DeepMindOfficialAI score62

    Google DeepMind adds Live Avatar to Gemini 3.8 Live for enterprise

    AIGoogle DeepMind has launched Gemini 3.8 Live with Live Avatar, which adds near real-time visual presence to its native live dialogue models. The feature is available today in Gemini Enterprise, supports 97 languages with adaptive lip-sync, and allows custom avatars through enterprise allowlisting. All output carries an imperceptible SynthID watermark.

    Why it matters: The post specifies the new avatar capabilities, the Gemini Enterprise access path, and the SynthID watermark, which helps readers judge its enterprise deployment fit.

  7. Google for DevelopersOfficialAI score37

    Gemma 4 now runs on-device in the Antigravity SDK

    AIGoogle says Gemma 4 can now run locally on-device within the Antigravity SDK. Developers can build fully local or hybrid multi-agent workflows that pair cloud models with Gemma 4 agents for auditing, patching, and testing code. The post emphasizes total data privacy and zero API fees, powered by LiteRT.

    Video from @googledevs's post
  8. Philipp SchmidXAI score56

    Gemini 3.8 TTS adds custom voice creation from a short recording or prompt

    AIGemini 3.8 TTS lets users replicate their own voice or design a custom voice from a text prompt. The workflow is to record about 20 seconds of speech with a consent sentence, create the voice through an API call, then use it in any request with styles set in speech_metadata. The author also points readers to a guide for setting up and testing the process with an agent.

  9. Philipp SchmidXAI score62

    Gemini 3.8 Flash TTS adds custom voice creation from recordings or a sentence

    AIGemini 3.8 Flash TTS and Flash-Lite TTS are now available in the Gemini API and AI Studio, with a new option to replicate a user's own voice from two recordings or design one from a sentence. The guide says the reusable voice ID can be passed in later requests, or an encrypted voicekey that expires after 7 days can be used if nothing is stored server-side. Prompting changed from gemini-3.1-flash-tts-preview: input text is spoken word for word, delivery goes in speech_metadata.style, and non-streaming responses are now real WAV.

  10. Google · Gemini appOfficialAI score62

    Google launches Gemini 3.8 Live with Live Avatar for enterprises

    AIGoogle introduced Gemini 3.8 Live with Live Avatar, which adds a visual persona with lip-syncing and expressions to its live dialogue models. The feature is available in Gemini Enterprise and supports 97 languages, with custom avatars available through enterprise allowlisting. Google says all output is watermarked with SynthID.

    Why it matters: The post specifies enterprise availability, custom avatar allowlisting, and 97-language support, which clarifies who can use the feature and how far it reaches.

  11. vLLMOfficialAI score34

    vLLM and RL-Kernel achieve bit-exact logprob match on AMD MI300X

    AIThe RLKernel team integrated RL-Align/RL-Kernel with vllm-project/vime, and a 200-step Qwen3-8B GRPO run on 8× AMD MI300X recorded zero logprob mismatches between Megatron training and vLLM rollout. The strict path aligns reduction order, intermediate precision, rounding points, and math primitives across both sides to achieve bit-for-bit matching on ROCm.

  12. Microsoft Foundry BlogOfficialAI score61

    Microsoft Foundry Routines reach general availability for scheduled and event-driven agents

    AIMicrosoft announced general availability of Routines in Foundry Agent Service, a managed way to run agents on a timer, on a recurring schedule, or in response to GitHub issue events and new Microsoft Teams channel messages. Routines keep the trigger, agent action, identity, connections, and run history in the Foundry project, and each routine can run under the creator's identity or the agent's own Microsoft Entra ID identity. A preview reminder tool lets a Hosted Agent schedule itself to resume later on the same conversation.

    Why it matters: The post explains how scheduled, event-based, and self-reminding agent runs are managed in one place, along with the creator versus agent identity choice for unattended tasks.

  13. Google Cloud · AI & Machine LearningOfficialAI score55

    Gemini 3.8 Live with Live Avatar becomes generally available in Gemini Enterprise

    AIGoogle says Gemini 3.8 Live with Live Avatar is now generally available in Gemini Enterprise, with US and EU endpoints, provisioned throughput, and enterprise compliance. Its video avatars use synchronized lip-syncing, custom avatars are limited to an allowlist, and generated audio and video carry SynthID watermarks. The model also understands and speaks 97 languages and can run tool calls in the background while the conversation continues.

  14. Liquid AI NewsletterOfficialAI score38

    Liquid AI's Liquid Context now optimized for Snapdragon NPUs; LFM Longevity models released

    AILiquid AI announced its on-device Liquid Context layer is now optimized for Snapdragon processors using the Qualcomm Hexagon NPU, letting edge agents learn user routines and share context across devices. Separately, Liquid AI released LFM2-1.2B-Longevity and LFM2-2.6B-Longevity, which the company says often match or outperform much larger frontier LLMs on longevity prediction tasks.

  15. OpenBMBOfficialAI score34

    FIT-GGUF enables size-targeted mixed-precision quantization of MiniCPM5-2B

    AIDeveloper @Scorp1o_117 used FIT-GGUF to build four MiniCPM5-2B GGUF variants, ranging from about 1.14 GiB to 1.46 GiB, tuned to target file sizes or fidelity tiers. Instead of fixed presets, FIT-GGUF allocates precision tensor by tensor, with Quality, Balanced, Compact, and Mini options, and its generated files matched predicted sizes. Builds are evaluated with KL Divergence and Same-top metrics and are available on Hugging Face.

    Image from @OpenBMB's post
  16. Google · Innovation & AIOfficialAI score62

    Google's Project Suncatcher will test TPUs in orbit on a prototype satellite

    AIGoogle's Project Suncatcher will launch a prototype satellite on the Transporter-18 rideshare mission with SpaceX to test how its TPUs handle spaceflight. Initial ground tests showed the Trillium TPUs survived vibration and a radiation dose greater than a five-year space mission would deliver. Google says cooling with heat pipes and radiators and laser links between satellites in 2027 remain open engineering challenges.

    Why it matters: The source reports concrete radiation, vibration, and cooling test results for TPUs, showing what space-based AI compute still has to solve.

  17. TechNode · AINewsAI score34

    H3C Shifts AI Infrastructure Focus From More GPUs to Token Efficiency

    AIH3C argued at the 2026 Apsara Conference that AI infrastructure competition is shifting from adding GPUs to maximizing useful Tokens per GPU. The company showcased its UniPoD S80000 SuperPod, supporting 32 to 1,024 GPUs and scaling to 16,384, alongside switches for Scale-Up, Scale-Out, and Scale-Across interconnects. It also pitched its UniStor X20000 storage, which it says delivers up to 200GB/s bandwidth and cuts GPU waiting time by 30%.

  18. KrASIA · Big TechNewsAI score55

    Mind Lab launches Mint Recursive, a post-training platform for companies

    AIMind Lab unveiled Mint Recursive, a post-training and inference platform for industry use, alongside Macaron-V1.1, a model post-trained entirely on it. Macaron-V1.1 is a 752-billion-parameter model built from GLM-5.3 with four two-billion-parameter LoRA expert modules for chat, agents, coding, and generation. The platform is serverless and bills by token usage, and it collects feedback from models in use to support continued training.

  19. Lovable BlogOfficialAI score44

    Lovable Now Offers Free Chat for Planning and App Work

    AILovable now lets users chat for free to explore app ideas, review existing projects, and draft business materials before making changes. The chat can connect to tools like Notion, Granola, and Linear, and Free, Pro, and Business workspaces include a daily free chat allowance. Chats that generate images or video, or hand work off to Plan or Build, use credits as usual, and current chat pricing applies through October 31, 2026.

  20. Lovable BlogOfficialAI score80

    How Lovable's Chats connect conversations to agent work on projects

    AILovable describes how its Chats feature lets a workspace-level chat agent hand work to project builder agents and receive progress back. The design records each agent's history as an append-only, forkable trajectory, and passes messages through durable inboxes that activations wake. Agents can suspend at iteration boundaries and resume on freshly deployed nodes without killing long-running runs.

    Why it matters: The post details how trajectories, inboxes, and activations let agents share work and resume after deploys, useful for designing comparable agent systems.

  21. Tencent HyOfficialAI score34

    Tencent Hy Translation launches with Hy-MT2 offline on-device translation

    AITencent Hy Translation has launched, powered by Hy-MT2, supporting 33 languages and 5 Chinese minority languages and dialects. It offers voice and photo translation with full offline, on-device operation requiring no network, and is already live in 12 countries and regions.

    Image from @TencentHunyuan's post
  22. MiniMax (official)OfficialAI score34

    MiniMax-H3 video generation accelerated on AMD MI355X by Nunchux

    AINunchux runs MiniMax-H3 on AMD MI355X GPUs, generating 5 seconds of video in 1.3 seconds with up to 26.7x faster inference than SGLang on 8 GPUs. The stack supports streaming generation, letting users change prompts while the video plays. Free access to MiniMax-H3 through Nunchux is coming soon, with a waitlist open.

  23. LangChain BlogOfficialAI score50

    LangSmith Engine v2 adds red teaming and pre-validated agent fixes

    AILangChain released LangSmith Engine v2, an in-platform agent that scans production traces to detect agent issues and validates proposed fixes before human review. Engine v2 adds Red Teaming, currently in Private Beta for LangSmith Deployment users, which tests agents for weaknesses such as hallucinations and system-prompt violations before they reach production. Engine v2 is available in SaaS deployments for LangSmith Plus and Enterprise plans, with Self-Hosted support and BYOK for Engine coming later.

  24. LangChain BlogOfficialAI score50

    LangSmith Fine-Tuning and smithtune Turn Agent Trajectories Into Custom Models

    AILangChain launched LangSmith Fine-Tuning and smithtune, a CLI that turns LangSmith agent trajectories into fine-tuned models through dataset creation, training with Fireworks or Baseten, and evaluation in LangSmith. smithtune currently supports supervised fine-tuning, training models on recorded examples of good agent behavior by updating model weights. The tool lets teams train specialized models without building the data pipeline by hand.

  25. LangChain BlogOfficialAI score44

    LangSmith Launches Trajectories for Readable, Chronological Agent Session Views

    AILangChain has launched Trajectories in LangSmith, a chronological, conversational view that aggregates human, AI, and tool messages across an agent and its subagents. Trajectories work with traces from LangChain, LangGraph, Deep Agents, OpenAI and Claude agent SDKs, and coding agents like Codex, Claude Code, and Cursor. The feature is available now on all plans in the US.

Sep 23

Sep 23Wed
  1. Midjourney UpdatesOfficialAI score31

    Midjourney Alpha changelog adds style previews, default parameters, and a new Create feed

    AIMidjourney's alpha site now lets users preview their current prompt across styles with "Live previews" in the Styles sidebar and save prompt-bar settings as defaults via Settings → Advanced → Your defaults. The Create feed received a full-width masonry redesign with hover-based prompts and buttons, and Korean is now live for all users on midjourney.com.

  2. Simon WillisonXAI score34

    Datasette blog backup now queryable by voice via ChatGPT iPhone app

    AISimon Willison had a voice conversation with his blog's Datasette backup through datasette-mcp, which is now accessible from the ChatGPT iPhone app. The quoted post says ChatGPT Voice can now use plugins and run on GPT-6 Astra, Sol and Luna in ChatGPT Work on web and mobile.