Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Sep 17

Sep 17Thu
  1. Gemini NotebookOfficialAI score37

    Mariposa Museum exhibit shows town in 1859, a decade after Gold Rush

    AIA Mariposa Museum exhibit photographed by writer Steven Johnson, shared by Gemini Notebook, documents the Sierra Nevada town in 1859, ten years after the Gold Rush began. Johnson says he used the Gemini Notebook mobile app's camera feature to generate a detailed report from photos of the display, which he says was 99% accurate on fact-checking.

  2. Josh WoodwardXAI score25

    Gemini Notebook's phone camera turns field photos into research documents

    AISteven B. Johnson describes how the camera feature in the Gemini Notebook mobile app let him photograph a Mariposa Museum exhibit and get a detailed text report of its contents. He says the transcription was about 99% accurate on his fact-check, though some blurry text was hard to read, and he plans to use Notebook to compare the new material against his existing sources.

  3. Google DeepMindOfficialAI score33

    Researchers use AlphaGenome Atlas to identify disease-causing DNA variants

    AIResearchers at the Broad Institute, the University of Exeter, and other institutions are already using AlphaGenome Atlas to identify potential disease-causing DNA variants and interpret their role. The post is a thread announcement, and it gives no further details on methods or results.

    Video from @GoogleDeepMind's post
  4. OpenBMBOfficialAI score40

    OpenMed and MiniCPM5-2B demo local agentic clinical AI workflow

    AIOpenMed paired with MiniCPM5-2B to demonstrate a local clinical AI workflow combining privacy-preserving data processing with a compact model's tool use and long-context reasoning. OpenMed masks sensitive identifiers and extracts clinical context before MiniCPM5-2B calls tools, compares lab results, and generates clinical handoffs with source references. The post presents this as an example of keeping inference on local, resource-constrained hardware.

    Image from @OpenBMB's post
  5. inclusionAI (Ant Ling) · new models on Hugging FaceOfficialAI score46

    Ming-Image-0.1-Design-Layer splits flattened design images into RGBA layers

    AIinclusionAI has released Ming-Image-0.1-Design-Layer on Hugging Face, a model that decomposes a flattened design image into a requested number of RGBA layers using an image and a layer plan. The model runs at 1024 resolution (512 for faster processing) with 12 sampling steps, a CFG scale of 2.0, and BF16 precision on one CUDA GPU with 80 GiB VRAM. It is released under the MIT License.

  6. TechNode · AINewsAI score53

    Huawei unveils Ascend 960 SuperPoD with NPO for AI infrastructure

    AIAt HUAWEI CONNECT 2026, Huawei announced the Ascend 960 SuperPoD, which supports up to 4,096 NPU cards and uses NPO technology. According to Huawei, Hi-ONE NPO units can replace 48,000 800G optical modules with about 5,500 units, reducing power consumption by more than 550kW, and system availability will reach 99.8%.

  7. inclusionAI (Ant Ling) · new models on Hugging FaceOfficialAI score42

    inclusionAI releases Ming-Image-0.1-Design, a 6B text-to-image model for text-rich designs

    AIinclusionAI has released Ming-Image-0.1-Design, a 6B text-to-image model for UI, infographics, and posters that outputs RGBA images with transparent backgrounds. The model is available on Hugging Face and ModelScope under the MIT License. It runs at 2048 x 2048 with 12 sampling steps and a CFG scale of 1.0, validated on one CUDA GPU with 80 GiB VRAM.

  8. WanOfficialAI score44

    Wan3.0 generates single 30-second video shots with director-level control

    AIAlibaba's Wan3.0 video model now produces a single 30-second shot directly, up from 15 seconds and a year ago's 5-second clips. It adds director-level control and omni-reference input accepting up to five videos, letting creators generate long takes instead of stitching short clips. A filmmaker used Wan3.0 in a production workflow to make Soulscape and Johnny Mai.

    Video from @Alibaba_Wan's post
  9. Gemini API ChangelogOfficialAI score38

    Antigravity Agent 09-2026 replaces 05-2026 with new built-in file and search tools

    AIGoogle released the antigravity-preview-09-2026 agent, which replaces and deprecates antigravity-preview-05-2026. Remote sandbox users reading only output_text or model_output steps need only update the agent string, while local-environment users or those parsing function_call steps must adapt to renamed tools, PascalCase parameters, and line-range file edits. The 05-2026 preview shuts down on October 5, 2026.

Sep 16

Sep 16Wed
  1. Kling AIOfficialAI score3

    Kling AI shares a link to a full film

    AIKling AI's post links to a full film hosted at watch.fountain0.com. The post itself gives no further details about the film's content, production, or release.

  2. WanOfficialAI score34

    Wan 3.0 generates a full anime-style fantasy fight sequence

    AIQwen's Wan 3.0 video model produced a complete fantasy fight between a mage and a swordswoman, with strong timing, poses, and complex layouts that the post calls anime-ready. The quoted post notes Wan is stronger than Seedance 2.5 on timing and layouts, while Seedance is better at keeping characters on-model, and that all three support Vid2Vid and Omni Reference.

  3. WanOfficialAI score7

    Wan 3.0 used to generate a magical girl video clip

    AIWan (@Alibaba_Wan) shares a magical girl clip made with Wan 3.0. The post notes it is also a test of effect generation, so more magical girl content may follow. The clip was shared alongside the TapNow tool.

  4. Amp NewsOfficialAI score50

    Amp Runners Now Serve Multiple Directories and Update Themselves

    AIAmp runners can now serve multiple directories, specified with repeated --dir flags or found automatically with --discover-dirs, which scans Git checkouts up to two levels deep by default. Runners also check for new releases about once an hour, install them, and restart into the new version once no thread is running, at most once every 12 hours, with auto-update disabled via amp.runner.autoUpdate.enabled: false.

  5. Google Developers BlogOfficialAI score38

    Google and Speakeasy open-source OpenAPI SDK generator suite under AGPLv3 license

    AISpeakeasy is open-sourcing its full OpenAPI client suite under the AGPLv3 license, including generators for seven languages (Python, TypeScript, Go, Java, C#, PHP, Ruby), an agent-native CLI generator, and a documentation MCP server generator. Google said the move followed the May 2026 shutdown of the SDK generation provider it had been using, which it cited as evidence that closed-source generators pose platform risk. Google's new Google GenAI SDKs for the Interactions, Agents, and Webhooks APIs were built with this pipeline across six targets.

  6. Midjourney UpdatesOfficialAI score34

    Midjourney Alpha Site Adds Editing, Mobile, and Folder Fixes in 9/16/26 Update

    AIMidjourney's September 16, 2026 alpha site update fixes the v8.2 editor, adds a Korean language option, and makes its mobile and tablet layouts fill the screen. Users can now drag images into folders with the sidebar collapsed, use Add to Folder without existing folders, and delete uploads from a menu. Default parameters and sref previews are listed as upcoming work.

  7. SpaceXAIOfficialAI score55

    Grok Voice is now available on fal for low-latency voice agents

    AIGrok Voice is now live on fal, and the post says it answers in 0.70 seconds and can finish tool calls before a sentence ends. Its listed capabilities include transcription with word-level timestamps, text to speech in 30+ voices across 25+ languages, and voice cloning from two minutes of audio.

  8. TinkerOfficialAI score32

    Sundial trains Inkling-Small to fix LaTeX errors in under a second

    AISundial fine-tuned Thinking Machines' Inkling-Small with RLVR on 3,978 verified TeX.StackExchange fixes, using rewards for compilation and PDF match and penalties for removed content. The trained model fixes 83.7% of LaTeX errors in under one second at $0.0013 per fix, according to the post. Sundial says it is rolling out the model in its editor, applying fixes as suggestions and rebuilding the PDF.

  9. Alex AlbertXAI score45

    Claude Cowork and chat merge into one unified interface

    AIAnthropic is merging Claude Cowork and chat into a single Claude experience, which Alex Albert says feels much better than either product alone. He also highlights the new slides, docs, and design integrations as working very well. The merged version is rolling out to Pro and Max users over the next few weeks.

  10. Kilo (acq. by Anaconda)OfficialAI score40

    Kilo Mobile lets users run full AI agent loops from their phone

    AIKilo Mobile now lets users spawn Cloud Agents, start sessions on remote machines, and dictate prompts by voice from a phone. Users can also review and comment on pull requests and approve Security Agent remediations without a laptop. On iPhone, Live Activities show session status on the Lock Screen when an agent needs input.

    Image from @kilocode's post
  11. LiveKitOfficialAI score14

    LiveKit offers startups a free year of voice stack with credits

    AILiveKit is offering startups $10K in Cloud credits, a year of its Scale plan, and $7K in partner inference credits, with $1K each for Google Gemma, Deepgram, Inworld, Rime, Fish Audio, AssemblyAI, and Gradium. The partner credits cover the voice stack in one workspace without requiring separate provider keys. Startups can apply through

    Video from @livekit's post
  12. Google for DevelopersOfficialAI score22

    Google invites developers to Gemini Audio event with new voice models

    AIGoogle is hosting Gemini Audio | At Night in San Francisco on Thursday, September 24, from 6:00 PM to 10:00 PM at The Pearl. Attendees can try Gemini 3.5 Transcribe, 3.8 Live, and 3.8 Live Extended Thinking, meet the teams that built them, and test live demos. RSVP is available at goo.gle/4xpdcBn.

    Image from @googledevs's post
  13. Baseten BlogOfficialAI score54

    Baseten launches Hosted Tools with web search for open-source models

    AIBaseten has launched Hosted Tools, starting with Baseten Grounded Inference, a server-side web search capability for models hosted on Baseten. Developers enable it by adding a hosted search tool to a Messages, Chat Completions, or Responses request, and the platform runs the search loop with partners Exa, Keenable, Parallel, and You.com. In Baseten's benchmarks, agents using the hosted tools saw a 15% reduction in end-to-end latency compared with client-side tools, and the feature is in playground preview with 25 RPM rate limits and $2 of free credits.

  14. catXAI score60

    Claude merges Cowork and chat into one product with automatic routing

    AIAnthropic is merging Claude Cowork and chat into one Claude, and Claude Design is integrated so users can ask for slides, designs, or docs without switching apps. Claude decides from the prompt whether to give a quick answer or do deeper agentic work, and users can still stop, redirect, or adjust its effort. The change rolls out to Pro and Max over the next few weeks.

  15. Mike KriegerXAI score46

    Claude Cowork and Chat Merge into One Unified Claude

    AIAnthropic is merging Claude Cowork and Chat into a single Claude starting today, which Mike Krieger says removes the friction of choosing which product to start with. Per the @claudeai announcement, Claude will carry tasks forward even after the laptop is closed, asking for clarification when needed while users keep final say. The rollout to Pro and Max plans will take place over the coming weeks.

Only the first 50 pages are available. Search or browse topics for older items.