Skip to contentSkip to stories

Updated

All AI news

Items with an AI score under 20 are hidden. Show low-relevance items

Sep 30

Sep 30Wed
  1. IdeogramOfficialAI score22

    Ideogram 4.5 keeps product details accurate across edits

    AIIdeogram 4.5 lets users change one element at a time in AI-generated images while the product stays accurate. The post says this makes building a full campaign much easier.

    Video from @ideogram_ai's post
  2. IdeogramOfficialAI score22

    Ideogram 4.5 holds up across repeated edits where rivals degrade

    AIIdeogram posted a side-by-side test of identical edits on Ideogram 4.5, GPT Image 2.5 Sunburst, Nano Banana Pro, and Nano Banana 2. The post claims GPT Image and Nano Banana outputs become unusable within a few edits, while Ideogram 4.5 remains clean across repeated edits.

    Video from @ideogram_ai's post
  3. IdeogramOfficialAI score38

    Ideogram 4.5 launches as a precise image edit model

    AIIdeogram released Ideogram 4.5, which it calls the most precise edit model, claiming it avoids the artifacts, pixel shifts, and color changes that leading models add with each edit. The company says this eliminates artifact buildup and makes multi-turn editing possible. It is live in Ideogram, via the API, and with launch partners, with open weights promised soon.

    Video from @ideogram_ai's post
  4. Tejas ManoharXAI score22

    Hightouch launches AXO to help brands win over personal agents

    AIHightouch launches AXO, a connector that helps brands make their sites easy for personal agents like Muse to discover, navigate, and understand. The company says the tool gives agents a tailored experience and gives brands insights into agent searches, and it is onboarding a select number of brands.

    Video from @tejasmanohar's post
  5. ModelScopeOfficialAI score46

    IndexTeam releases Index-Translate multilingual translation model family

    AIIndexTeam has released Index-Translate, a multilingual family covering text, speech, dubbing, and long-document translation across 150 languages. Its 9B model scores 0.8789 on FLORES, 0.8209 on instTrans, and 0.7387 on MEME, and the 2B and 9B models are released under Apache 2.0.

    Video from @ModelScope2022's post
  6. ModelScopeOfficialAI score62

    InSpatio-World 1.5 turns images and videos into real-time explorable 4D worlds

    AIInSpatio-World 1.5 from InSpatio_AI turns a single image, four images, a panorama, or a video into a navigable scene with wide viewpoint changes. The 1.3B model scores 68.72 on WorldScore-Dynamic, ranking first among evaluated real-time and interactive methods, with speeds up to 24 FPS. The post says the code is released under Apache 2.0 and that dependencies keep their own licenses.

    Why it matters: The post gives specific benchmark, speed, and input details, so readers can judge how the model handles real-time scene exploration from images or video.

    Video from @ModelScope2022's post
  7. ollamaOfficialAI score23

    Ollama adds Nimble, Tev1 4B, and Tev1 0.8B models

    AIOllama now lists three models, Nimble, Tev1 4b, and Tev1 0.8b, for use with /v1/systemone. The post provides a library link for each model, with the Tev1 variants listed at 4b and 0.8b.

  8. ollamaOfficialAI score34

    Ollama adds support for JEV-style decision models

    AIOllama now supports JEV-style decision models, with documentation for the decision capability and API. The post links to a blog announcement, decision capability docs, and API docs, but gives no further details on the models or their performance.

  9. ollamaOfficialAI score30

    Ollama adds local decision models like Nimble via new API

    AIOllama now supports decision models such as Nimble locally, usable for tasks like ticket triaging, model routing, and content moderation. The model can be installed with ollama pull nimble and accessed through the new local /v1/systemone API, as shown in a real-time Ollama racer demo.

    Video from @ollama's post
  10. Mastra BlogOfficialAI score22

    Mastra Factory adds Jira, GitLab, and incident.io work intake integrations

    AIMastra Factory now supports Jira, GitLab, and incident.io as work intake sources, joining GitHub, Linear, and Slack. The intake column can be filtered by source, and each item can be moved through the pipeline until the work is complete. The integrations ship in the @mastra/factory package, with credentials configured under Settings → Work Intake.

Sep 29

Sep 29Tue
  1. SGLangOfficialAI score36

    SGLang adds native decision API for classification and scoring models

    AISGLang says it turned Qwen3.8-27B into a multimodal decision model that beat Pokémon FireRed's Elite Four and champion with sub-100 ms decisions from live game state. It introduces a native /v1/decisions endpoint for turning LLMs and VLMs into classification and scoring models. A /v1/systemone endpoint is also added so Jev-like open models can work with the TypeSafe SDK.

    Video from @sgl_project's post
  2. OpenClaw🦞OfficialAI score70

    OpenClaw Enterprise launches as an open-source control plane for persistent agents

    AIThe OpenClaw Foundation announced OpenClaw Enterprise, an open-source enterprise control plane for persistent agents, in collaboration with Red Hat, NVIDIA, and OpenAI. The product is built to run on an organization's own infrastructure and will always be free for organizations to use.

    Why it matters: The announcement names its collaborators and deployment model, which helps organizations judge how the enterprise control plane would fit their own infrastructure.

    Image from @openclaw's post
  3. Kilo (acq. by Anaconda)OfficialAI score32

    Kilo adds Sign in with ChatGPT across all its surfaces

    AIKilo now supports Sign in with ChatGPT, letting users apply their ChatGPT plan across VS Code, JetBrains, the CLI, Cloud Agents, Code Reviews, and the mobile app. The integration enables running 10 agents in parallel, and Cloud Agents continue working after the user closes their laptop.

    Image from @kilocode's post
  4. ModelScopeOfficialAI score54

    IQuest-Q1 released as 320B MoE model for long-horizon coding agents

    AIModelScope announced IQuest-Q1, a 320B MoE model with 15B active parameters and a 512K context window for agentic coding. The post reports scores of 84.5 on CyberGym, 83.2 on Terminal-Bench 2.1, 64.6 on DeepSWE v1.1, and 63.0 on NL2Repo, and says weights are released under the IQuest-Q1 License.

    Image from @ModelScope2022's post
  5. OpenBMBOfficialAI score34

    MiniCPM-o 4.5 now runs in SGLang Omni v0.1.7 for developers

    AIOpenBMB announced that MiniCPM-o 4.5 is now supported in SGLang Omni v0.1.7, giving developers more flexibility to run and build with the model. The background release notes add that MiniCPM-o 4.5 brings multimodal input and speech output to the runtime. MiniCPM-o and MiniMax-Music3 also gained Intel XPU support in the same release.

  6. vLLMOfficialAI score58

    IQuest-Q1 320B MoE coding model gets day-0 support in vLLM

    AIvLLM announced day-0 support for IQuest-Q1, a 320B-parameter MoE model with 15B active per token, 256 experts with 8 active, and a 524,288-token context. The post credits existing vLLM features such as the hybrid KV cache coordinator, sinks attention path, and EAGLE speculative decoding with probabilistic draft sampling. The linked material includes a Docker image and vllm serve commands, with and without recursive MTP.

    Image from @vllm_project's post
  7. SGLangOfficialAI score53

    SGLang adds Day-0 support for IQuest-Q1 with a single-node serve command

    AISGLang says it has Day-0 support for IQuest-Q1, an open-source sparse MoE model with 320B total and 15B active parameters for coding and agentic tasks. The post includes a single-node serving command for H200 GPUs in BF16, using tensor parallelism of 8, EAGLE speculative decoding, and the iquest_q1 reasoning and tool-call parsers. The image marks the command as not verified.

    Image from @sgl_project's post
  8. Mastra BlogOfficialAI score42

    Mastra Adds Memory Hooks to Observe and Modify Agent Memory Cycles

    AIMastra has added memory hooks that let developers monitor or alter an agent's observational memory cycles. Lifecycle hooks such as onObservationStart and onReflectionEnd report on each cycle, including token usage for spotting cost spikes, while transform hooks like beforeObservation and afterReflection can prune, remove, or redact memory data.

Sep 28

Sep 28Mon
  1. ModelScopeOfficialAI score44

    Audio8 ASR Infinite enables unlimited-length streaming speech transcription with bounded memory

    AIAudio8 ASR Infinite transcribes Chinese and English audio of unlimited length using a rolling KV Cache that avoids accumulated drift. At a 480 ms delay, it reports 1.75 CER on AISHELL-1, 2.89 on AISHELL-4, and 3.04/6.81 WER on LibriSpeech test-clean/test-other. The preview release is under Apache 2.0, with deployment through an adapted vLLM stack.

    Video from @ModelScope2022's post
  2. KhazixXAI score38

    Khazix open-sources AIHOT, the AI news site, with its full pipeline and prompts

    AIKhazix (Shuzi Shengming Kazike) says the monthly-active-million AI news site AIHOT is now open source on GitHub, including its collection workflow, curation scoring, clustering mechanism, and all production prompts. He says the release is meant to hand the project to others, since readers have asked for versions for industries such as gaming, law, HR, and finance.

  3. François CholletXAI score32

    K3-Node: a Keras 3 GNN library running on JAX, PyTorch, and TF

    AIK3-Node is a graph neural network library built natively on Keras 3, with models that run on JAX, PyTorch, and TensorFlow with hardware acceleration including Apple Silicon and TPU. According to the post, it achieves 100% public API parity with PyG and incorporates foundation models and architectures from Spektral and StellarGraph.

  4. ReplicateOfficialAI score28

    Pruna's P-Video-2-Pro video model now runs on Replicate

    AIReplicate has added P-Video-2-Pro, the latest video model from Pruna AI, which sits on the edge of the preference-speed and preference-price Pareto frontiers. Design Arena ranks its Quality and Speed variants tied for #2 on the Image to Video leaderboard with an Elo of 1325, with the Quality version generating in 8.0 seconds and the Speed version in 4.5 seconds.

  5. Hacker News · Launch HN, YC launches (10+ points)BlogAI score54

    Vespper launches a DOCX MCP for agents editing Word documents

    AIVespper, a Y Combinator F24 startup, launches a DOCX MCP that lets agents edit Word files through HTML that a trained reconciler converts back to OOXML. On its 279-task internal benchmark, the company reports its MCP is 2.7–2.9x cheaper and 2.7–3.5x faster than Anthropic's DOCX skill, with higher pass rates. The post also lists current limitations, including no comment creation or reply, no image or video attachment, and no support for latent styles.

  6. FireworksOfficialAI score22

    Normal Factory's CAD Arena joins the Specialized Intelligence Index

    AINormal Factory joins the Specialized Intelligence Index with CAD Arena, which tests whether AI agents can turn engineering drawings into accurate, editable CAD parts. The benchmark evaluates agents across five CAD platforms, extending the SII into engineering design.

    Image from @FireworksAI_HQ's post
  7. Josh WoodwardXAI score22

    Josh Woodward praises a Yosemite photography simulator built with three.js

    AIJosh Woodward, Google's Gemini leader, called a Yosemite-themed browser project "awesome" after returning from the park. The project, by @trondw, is a three.js and WebGL photography simulator featuring real sun and Moon positions, real lidar data, and a tripod-mounted film camera for virtual shooting at Yosemite locations.

  8. RadixArkOfficialAI score46

    RadixArk releases Miles v0.1.1 with multi-LoRA and expanded model support

    AIRadixArk has released Miles v0.1.1, adding multi-LoRA with Tinker API compatibility so multiple training jobs can share one base model. The update also supports agentic training with harnesses like Claude Code and runs Harbor tasks in sandboxes including AgentENV, Daytona, E2B, and Modal. It further reduces memory needs for training larger models on validated NVIDIA and AMD GPUs and adds stable support for Qwen3.8-Flash-Next, GLM-5.3-Flash, and Kimi-K3.

    Image from @radixark's post
  9. Daniel HanXAI score29

    Unsloth Desktop serves local Laya decision models for real-time packing demo

    AIUnsloth Desktop can now serve local Laya decision models through a Jev-compatible API, shown in a real-time packing demo where suitcase items update as the user types. The demo runs through Unsloth's Decision API, and the team says more optimizations are coming to speed up local hardware performance.

    Video from @danielhanchen's post
  10. Unsloth AIOfficialAI score34

    Laya Decision models can now run locally on 4GB RAM

    AIUnsloth AI says Laya Decision models can run locally on just 4GB of RAM, on CPU, Mac, Windows, Linux, and GPU setups. The post adds that Laya can be served through a Jev-compatible API via Unsloth Desktop.

    Image from @UnslothAI's post
  11. ModelScopeOfficialAI score46

    Qwen-Image-2.1 LoRAs extract and remove layers for editing

    AIModelScope released two Qwen-Image-2.1 LoRAs, LayerExtract and LayerRemove, for layer-based image editing. LayerExtract isolates a prompt-specified subject onto a transparent background, while LayerRemove deletes the matching object from the source image and reconstructs the scene behind it. Both can be hot-swapped within the same DiffSynth-Studio pipeline, and the LoRA weights are licensed under Apache 2.0, with Qwen-Image-2.1 base-model terms also applying.

    Image from @ModelScope2022's post