Skip to contentSkip to stories

Updated

All AI news

Items with an AI score under 20 are hidden. Show low-relevance items

Oct 7

Oct 7Wed
  1. vLLMAI score46

    vLLM-Omni technical report unifies serving for omni-modality generation

    AIThe vLLM team released a technical report on vLLM-Omni, a unified serving runtime for omni-modality generation spanning multi-stage autoregressive pipelines, iterative diffusion, and stateful sessions. Current LLM servers and diffusion stacks each cover only one of these patterns, pushing deployments to stitch disjoint runtimes together. vLLM-Omni offers a shared control plane in which an orchestrator advances requests across stages, specialized engines handle compute, and a connector carries payloads.

    Image from @vllm_project's post
  2. meng shaoAI score75

    Microsoft positions Windows as the home for hybrid AI agents across four layers

    AIMicrosoft has repositioned Windows as the home for hybrid intelligence, where AI agents can run locally or in the cloud. The announcement covers four layers: MXC reaching general availability for agent isolation, local models such as MAI Code 1.1 Flash, Copilot on Copilot+ PCs gaining local context and actions in coming months, and new hardware including RTX Spark PCs and DGX Station for Windows.

    Image from @shao__meng's post
  3. The Next PlatformAI score46

    Memory Now Drives the IT Industry as DRAM and Flash Prices Surge

    AIMemory has overtaken compute as the central control point in IT, according to The Next Platform, as generative and agentic AI drive demand for DRAM, HBM, and flash. Server DDR5 memory now sells for roughly 9X to 13X its November 2022 street price, while a 30 TB enterprise SSD costs 6X to 7X more. HBM pricing has risen only about 1.6X since the GenAI boom began, the article says.

  4. The Next PlatformAI score37

    HPE Unveils First Gen 13 ProLiant Servers Aimed at AI Inferencing and Agentic Workloads

    AIHewlett Packard Enterprise unveiled the first of its ProLiant Gen 13 systems, built for enterprise AI inferencing and agentic workloads, with AMD 6th Gen Epyc 9006 "Venice" CPUs in common. The ProLiant DL585a, a 10U server holding up to eight double-wide GPUs and two Epyc CPUs with up to 256 cores each, will be available in March 2027. The air-cooled ProLiant DL525, a single-socket 1U system with a 256-core AMD chip, becomes available next month.

  5. meng shaoAI score88

    OpenAI rolls out GPT-6 with Intelligent UI to over 1.2 billion weekly ChatGPT users

    AIOpenAI is rolling out GPT-6 to ChatGPT's over 1.2 billion weekly users, adding Intelligent UI, which lets replies include charts, buttons, forms, and interactive tools. The post's image cites tiered access, with Free/Go and Plus/Pro/Business/Enterprise sharing the Sol and Luna model splits, and says the feature is progressively rendered as the model generates it.

    Image from @shao__meng's post
  6. GuizangAI score57

    Anthropic gives Max and Team subscribers monthly Claude Platform API credits

    AIAnthropic is rolling out monthly Claude Platform API credits for Max and Team plans: $100 for Max 5x, $200 for Max 20x, and up to $500 pooled for Team. The credits work on any model, including Haiku 5.5, in your own code or third-party harnesses. The author notes the credits cover Claude API, Console Playground, Managed Agents, and the Agent SDK, but not interactive Claude Code or Claude, Claude Code, and Cowork Extra usage.

    Image from @op7418's post
  7. PandailyAI score42

    UBTECH and FAW-Volkswagen Extend Humanoid Robots to Factory Logistics

    AIUBTECH Robotics and FAW-Volkswagen signed a strategic cooperation agreement to jointly develop and test embodied AI robot applications in logistics and build demonstration sites. The partnership builds on UBTECH's Walker S Lite humanoid, already doing vehicle quality-inspection training at FAW-Volkswagen's Qingdao Branch, a national-level smart manufacturing demonstration factory. The companies aim to speed up humanoid deployment in smart manufacturing.

  8. Orange AIAI score34

    Next Token episode 5 covers Personal Agents, open-source software, and hardware projects

    AIThis Next Token episode discusses Personal Agents, including Dots in Codex, memory and cloud computer permissions, and whether agents should act as assistants or digital twins. The hosts also cover Instinct's booking and business-travel model, hands-on projects built with Opus 5.5, and whether software, games, and hardware could become open source as AI makes rewriting easier.

  9. MarkTechPostAI score58

    Unsloth Studio re-checks changed model repos and blocks flagged weights before loading

    AIUnsloth Studio binds remote-code approval to a fingerprint of the scanned code, so changed code requires fresh consent before it runs. It also blocks weight files that Hugging Face has flagged for malware in the path the selected loader would deserialize. The article describes these checks as one layer among several, alongside package-content scans and OS sandboxes, and notes that the scanner is not a sandbox and cannot catch every evasion.

  10. ComfyUIAI score43

    Vidu Q4 Preview arrives in ComfyUI via Partner Nodes

    AIComfyUI says Vidu Q4 Preview, the first preview of Vidu's new flagship video model, is now available through Partner Nodes. The model offers finer character acting with expressions, emotion, and body language, voice consistency using up to three reference audio clips, and up to 15 reference images per shot. It outputs up to 16 seconds at 2K and 4K, with smoother cuts and camera moves across shots.

    Video from @ComfyUI's post
  11. GeekParkAI score36

    MUZIM L1 Dock, Lumeria Lumoscope, and Other Small-Innovation Gadgets Reviewed

    AIMUZIM L1 is a desktop data dock with up to 24TB of storage, dual SSD slots, and a Vibe Search feature that finds files by natural-language description, with local-first processing rather than default cloud upload. Lumeria Lumoscope is a multispectral skin scope that clips onto a phone, using RGB, ultraviolet, polarized, and near-infrared light, priced at $199 in pre-sale. The article also covers immurok IK-1, a 59-dollar wireless fingerprint key with a 60-day standby battery that authorizes sudo, SSH, and Git actions on Mac, Windows, and Linux.

  12. InferactAI score38

    Inferact and partners cut vLLM TTFT nearly 70% at ~100K throughput

    AIInferact, working with DeepSeek, NVIDIA, and SemiAnalysis alongside the vLLM community, says joint work across models, custom kernels, and engine serving cuts time to first token (TTFT) by nearly 70% at ~100K throughput. vLLM is the open-source inference engine, and Inferact optimizes it for enterprise production deployments.

  13. IThome · AIAI score46

    US man sentenced to 18 months for AI-driven music streaming fraud worth $8 million

    AIMichael Smith of North Carolina was sentenced to 18 months in prison and ordered to forfeit $8 million after using AI to generate hundreds of thousands of songs. Between 2017 and 2024, he uploaded the tracks to streaming platforms through thousands of fake accounts and inflated their plays with a large network of bot accounts. The scheme diverted royalties from genuine songwriters and performers, according to the U.S. Department of Justice.

  14. François CholletAI score44

    Chollet: Programming and math training don't boost general intelligence

    AIFrançois Chollet compares AI progress to human learning, noting that 1980s research found programming training improves coding but does not transfer to general reasoning. He argues general intelligence is a fundamental brain property rather than a trainable skill, since domain practice improves only that domain. The post is framed as background for his question whether AI's jagged frontier, driven by math and code via RLVR, reflects general capability or continued human-data bottlenecks.

  15. GeekParkAI score46

    ChatGPT Adds Intelligent UI That Generates Interactive Tools, Google Launches Playground

    AIOpenAI said on October 7 that ChatGPT's new Intelligent UI will automatically combine text, charts, buttons and forms into interactive interfaces such as calculators and mini-games, rolling out to Plus, Pro, Business and Enterprise users from October 7 and to Free and Go users from October 8. Google also launched Playground, an experimental platform where users create, modify and play browser games from natural-language descriptions, initially for U.S. users aged 18 and older.

  16. Gizmodo · AIAI score51

    Vibe-Coded Artcraft Suite Offers Free Photoshop Alternative on GitHub

    AIDeveloper Brandon Thomas used Claude Opus 5.5 and Rust to build Artcraft, a free open-source suite with Photocraft, Vectorcraft, and other apps that mimic Adobe products. The author tested Photocraft and found its basic editing commands worked where expected, but Free Transform behaved unpredictably. Thomas describes the software as early alpha and invites developers to contribute.

  17. Nathan LambertAI score22

    Compute access, not funding, bottlenecks academic AI groups, says Lambert

    AINathan Lambert says compute access rather than funds is a bottleneck for many academic and nonprofit AI groups. He adds that more academic and nonprofit funding is still needed, tackled one step at a time. The quoted post from Anjney Midha announces hundreds of MI355x and B300 nodes live on nationalcompute.com, subsidized for .edu and .gov users.

  18. Ars Technica · AIAI score52

    Microsoft's Surface Laptop Ultra brings Nvidia RTX Spark and unified memory to local AI

    AIMicrosoft announced the Surface Laptop Ultra, its first device using the Nvidia RTX Spark SoC, starting at $2,599 with up to 128GB of LPDDR5x unified memory. It ships October 16 and is available for preorder now. The article says the unified memory approach lets the laptop handle both gaming and local AI development and deployment.

  19. Google Developers BlogAI score62

    Google open-sources ML Drift, a cross-platform GPU engine for on-device AI

    AIGoogle's AI Edge Team open-sourced ML Drift under Apache 2.0, a GPU compute engine for on-device AI inference across OpenGL ES, OpenCL, Metal, and WebGPU. It serves as the core GPU acceleration engine within LiteRT and succeeds the legacy TFLite GPU delegate, which will no longer receive new features. The post cites benchmarks showing up to 40% lower frame latency in YouTube Shorts and up to 30% faster on-device performance in Adobe Lightroom and Photoshop.

    Why it matters: The post explains how ML Drift unifies GPU shaders across platforms and replaces the TFLite GPU delegate, which matters for developers deploying on-device models.

  20. Apple Machine Learning ResearchAI score42

    Apple's Normalizing Trajectory Models generate images in four steps with exact likelihood

    AIApple researchers introduced Normalizing Trajectory Models (NTM), which model each reverse diffusion step as a conditional normalizing flow trained with exact likelihood. The model matches or outperforms strong image generation baselines on text-to-image benchmarks in just four sampling steps while retaining exact likelihood over the generative trajectory.

  21. Waymo BlogAI score42

    Sober Drivers Still Face Nearly 4x Nighttime Fatal Crash Risk, Waymo Study Finds

    AIWaymo research found that even fully sober human drivers face nighttime fatal crash risk 3.1 to 3.9 times higher than daytime risk, pointing to systemic hazards beyond impairment. The study used an exposure reconstruction model across the 50 most populous U.S. urban areas, showing removing alcohol-involved drivers lowers the average urban fatal crash rate by 23%, from 1.42 to 1.10 per 100 million miles.