Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Sep 17

Sep 17Thu
  1. OpenBMBOfficialAI score40

    OpenMed and MiniCPM5-2B demo local agentic clinical AI workflow

    AIOpenMed paired with MiniCPM5-2B to demonstrate a local clinical AI workflow combining privacy-preserving data processing with a compact model's tool use and long-context reasoning. OpenMed masks sensitive identifiers and extracts clinical context before MiniCPM5-2B calls tools, compares lab results, and generates clinical handoffs with source references. The post presents this as an example of keeping inference on local, resource-constrained hardware.

    Image from @OpenBMB's post
  2. OpenBMBOfficialAI score29

    Kahya-TTS: Turkish speech model fine-tuned from VoxCPM2 on 100 hours

    AIDeveloper Alican Kiraz fine-tuned OpenBMB's open-source VoxCPM2 voice model on nearly 100 hours of natural Turkish speech, creating Kahya-TTS for Turkish text-to-speech. The project shows how open-source voice models can be adapted to new languages and specialized datasets. The model is available on Hugging Face.

    Image from @OpenBMB's post

Sep 16

Sep 16Wed
  1. OpenBMBOfficialAI score20

    OpenBMB praises Dubedo's VoxCPM2-based voice cloning and dubbing studio

    AIOpenBMB says Dubedo is the kind of product it hoped VoxCPM2 would enable, citing speaker-aware cloning, multilingual generation, and an editing studio. The post praises @dubedostudio's work, while the background post describes Dubedo as dubbing into 30 languages with per-speaker voice cloning and a beta open for trial.

  2. TinkerOfficialAI score32

    Sundial trains Inkling-Small to fix LaTeX errors in under a second

    AISundial fine-tuned Thinking Machines' Inkling-Small with RLVR on 3,978 verified TeX.StackExchange fixes, using rewards for compilation and PDF match and penalties for removed content. The trained model fixes 83.7% of LaTeX errors in under one second at $0.0013 per fix, according to the post. Sundial says it is rolling out the model in its editor, applying fixes as suggestions and rebuilding the PDF.

  3. Kilo (acq. by Anaconda)OfficialAI score40

    Kilo Mobile lets users run full AI agent loops from their phone

    AIKilo Mobile now lets users spawn Cloud Agents, start sessions on remote machines, and dictate prompts by voice from a phone. Users can also review and comment on pull requests and approve Security Agent remediations without a laptop. On iPhone, Live Activities show session status on the Lock Screen when an agent needs input.

    Image from @kilocode's post

Sep 15

Sep 15Tue
  1. Zed BlogOfficialAI score72

    Zed launches Delta public beta to replace pull requests with agent threads

    AIZed has launched the public beta of Delta, a multiplayer environment for coding with agents and reviewing their work, which replaces pull requests with shared threads. Delta is built on DeltaDB, which records edits and messages between Git commits, and it is free during the beta, with paid plans for individuals and teams to follow.

    Why it matters: The post explains how Delta replaces pull requests with shared agent threads and DeltaDB, showing a concrete alternative to the GitHub review workflow.

  2. VercelOfficialAI score22

    Delphi ships 100+ deploys daily on Vercel's Python backend

    AIDelphi, which turns experts' knowledge into digital minds, runs its Python backend on Vercel with a 10-person team and no dedicated infrastructure role. The stack uses Vercel Workflows for long-running agents and Vercel Queues for background jobs. The team reports more than 100 production deploys a day.

  3. Air Street PressBlogAI score39

    Air Street Capital leads $40 million Series A in Jack & Jill, an AI career agent startup

    AIAir Street Capital has led a $40 million Series A round in Jack & Jill, with Madrona joining and existing investors Creandum and Entrepreneurs First investing again. The company runs Jack, an AI agent that talks with job seekers about career moves and searches millions of job postings, and Jill, a recruiting agent that introduces candidates to hiring managers. Jack & Jill has arranged 25,000 interviews and plans 5,000 more each month, according to the article.

  4. TechNode · AINewsAI score38

    Twoo Adds AI as a Third Member to Two-Person Relationships

    AIShanghai-based Twoo is building an AI that participates in conversations between two users, sharing context, remembering experiences, and helping them plan activities together rather than serving one person. Founder Cobe Chen says the product's monetization includes in-app "shells" that unlock outfits for its octopus AI character, plus premium memberships for professional uses such as collaborative writing and teaching. The company is preparing overseas expansion into North America, Japan, and South Korea.

  5. Kilo (acq. by Anaconda)OfficialAI score22

    Kilo App launches on Product Hunt for iOS and Android

    AIKilo announces that its Kilo App is live on Product Hunt, letting users start coding agents, check sessions, and review pull requests from iOS and Android. The company asks supporters to upvote or comment on its Product Hunt listing.

    Image from @kilocode's post

Sep 14

Sep 14Mon
  1. vLLM BlogOfficialAI score53

    Novita AI open-sources Chord, a W4A16 MoE kernel for Kimi K2.x on vLLM

    AINovita AI has open-sourced Chord, a W4A16 MoE CUDA operator with BF16 activations, INT4 weights and group-32 scales, built for Kimi K2.x serving shapes. Measured per layer against public Humming, it reports 1.11–1.20x on H200 EP8 prefill, 1.17–1.33x on H200 TP8 serving, and 1.81–2.15x on B300 EP8 decode against an untuned Humming default. Integration of the grouped operators with vLLM's Humming backend is still a work in progress.

  2. InferactOfficialAI score42

    Inferact and Google Cloud partner to make TPUs first-class in vLLM

    AIInferact and Google Cloud announce a partnership to make Google TPUs a first-class platform in the vLLM open-source project. The collaboration targets production serving features, optimized kernels, a native PyTorch path via TorchTPU, and day-0 support for frontier model releases. A community program will offer shared TPU capacity and review and design help from vLLM core maintainers, with all outputs released as open source.

    Image from @inferact's post
  3. Intern Large ModelsOfficialAI score25

    Intern-S2-397B gets Day-0 support in vLLM

    AIIntern-S2-397B, a model built for long-horizon scientific research, now has Day-0 support in vLLM. The model brings multimodal, reasoning, coding, and scientific agent capabilities, and vLLM has published a run recipe for it.

Sep 13

Sep 13Sun
  1. Ian Johnson 🔬🤖XAI score34

    Flying through 30 million embeddings as a video game

    AIIan Johnson visualizes 30 million jina-v5-nano embeddings from 12 billion tokens across multilingual FineWeb, StarCoder, The Pile, and RedPajama as a flyable video game. He frames the project as making data exploration engaging rather than a chore.

    Video from @enjalot's post

Sep 11

Sep 11Fri
  1. InferactOfficialAI score34

    vLLM adds day-0 support for DeepSeek v4.1 Flash across six NVIDIA GPUs

    AIInferact says vLLM now supports DeepSeek v4.1 Flash on day zero across H100, H200, B200, B300, GB200, and GB300 GPUs. SemiAnalysis independently verified the NVIDIA support, while the post notes AMD vLLM still does not work with the model. Serving recipes are available at recipes.vllm.ai.

  2. VercelOfficialAI score26

    Tailscale's Aperture model router is built on Vercel AI Gateway

    AITailscale offers instant access to hundreds of models for any user in a secure tailnet through its customer-facing model router, Aperture. Aperture is built on Vercel's AI Gateway and offers zero data retention, zero markup with free BYOK, and cost and usage data on every response.

  3. LlamaIndex 🦙OfficialAI score22

    LlamaParse adds high-effort page-level confidence scores with explanations

    AILlamaParse's new high-effort mode provides granular page-level confidence scores with text explanations of parsing quality. The scoring also checks the original document as an extra verification step. High-effort mode costs 5 additional credits per page and is meant to be used only when needed.

    Image from @llama_index's post

Sep 10

Sep 10Thu
  1. ollamaOfficialAI score22

    Ollama adds ChatGPT support in its desktop app

    AIOllama says users can enable ChatGPT inside its app, with downloads available at The post gives no further details on how the integration works or which models it covers.

  2. Kilo (acq. by Anaconda)OfficialAI score7

    Kilo is hiring a Staff Software Engineer for its extension

    AIKilo is hiring a Staff Software Engineer to lead performance work, context optimization, and architectural patterns for its extension. The company says the extension has passed 5 million downloads and processes over 10 trillion tokens a month. The post invites applicants who can ship a first pull request on day one.

  3. Thomas DohmkeXAI score13

    Entire approves merges via iPhone Duo, Thomas Dohmke says

    AIThomas Dohmke says Entire is already approving merges this way, with the capability coming to iPhone Duo in October. The background post notes the new Trails split view was designed with iPhone Duo in mind.

  4. RadixArkOfficialAI score22

    RadixArk publishes Miles cookbook for DeepSeek V4.1 Flash

    AIRadixArk has published a cookbook on its Miles documentation site covering how to run DeepSeek V4.1 Flash. The post itself contains only a link to the cookbook page, so no further details about features, figures, or setup steps are available.

Sep 9

Sep 9Wed
  1. BAAI · new models on Hugging FaceOfficialAI score24

    BAAI open-sources EPT, UniPath, and MiSI AIDD molecular and crystal modeling resources

    AIBAAI released open-source resources for three AIDD projects on Hugging Face: EPT, an equivariant pretrained transformer for unified 3D molecular representation learning, and UniPath, a learnable-time flow matching method for crystal structure and energy prediction. The repository mirrors their GitHub source code and READMEs, with setup, preprocessing, training, and evaluation documentation. The MiSI benchmark is released separately on Hugging Face.

Sep 3

Sep 3Thu
  1. Hacker News · Launch HN, YC launches (10+ points)BlogAI score33

    Mireye launches API and MCP server giving AI agents US location data

    AIMireye, a Y Combinator S26 startup, launches an API and MCP server that supply AI agents with cited facts, property enrichment, tools, and change signals for any US location. The founder says a free tier of 5,000 credits is available with no card, and that his earlier site-screening app was dropped because customers wanted the underlying engine. Usage data cited in the post shows 311 of the 317 catalog fields are queried.

Sep 1

Sep 1Tue
  1. Hacker News · Launch HN, YC launches (10+ points)BlogAI score34

    Nori Robotics launches A3 low-cost humanoid robot for development at $1,688

    AINori Robotics, a Y Combinator S26 company, is selling the Nori A3, a humanoid robot priced at $1,688 with no deposit and shipping in fall 2026. The robot has 7+1 degree-of-freedom arms with 1.5 kg payload per arm, a 12 m lidar, four 720p RGB cameras, and 6-8 hours of battery life. Nori says it is assembled in San Francisco and offers a skills marketplace and a laptop app for training and operating the robot.

Aug 31

Aug 31Mon
  1. Hacker News · Launch HN, YC launches (10+ points)BlogAI score36

    Almanac launches AI workspace that connects tasks, projects and knowledge

    AIAlmanac, a Y Combinator S26 company, launches a personal AI workspace that brings together tasks, projects, conversations and a connected wiki. The workspace picks up follow-ups from connected email and calendar accounts, and the Mac desktop app is the starting point. A seven-day free trial requires a card, then the plan renews at $20 per month unless canceled, and model usage counts toward a connected ChatGPT subscription's Codex limits.