Skip to contentSkip to stories

Updated

#Video

Showing low-relevance items too. Hide low-relevance items

Sep 29

Sep 29Tue
  1. Google Developers BlogAI score47

    Google Details Sparse Attention Speedup for Video Diffusion on TPUs

    AIGoogle Developers Blog describes how Sparse VideoGen (SVG) routes video diffusion attention heads into spatial or temporal sparse masks and implements them as custom JAX and Pallas Splash Attention kernels on TPU v6e. In isolated single-chip tests with 75.6K tokens and 10 heads, the sparse variants retain about 38.87% of query-key pairs. The article argues that theoretical sparsity must be converted into hardware tile skipping to yield real speedups.

Sep 28

Sep 28Mon
  1. ReplicateAI score28

    Pruna's P-Video-2-Pro video model now runs on Replicate

    AIReplicate has added P-Video-2-Pro, the latest video model from Pruna AI, which sits on the edge of the preference-speed and preference-price Pareto frontiers. Design Arena ranks its Quality and Speed variants tied for #2 on the Image to Video leaderboard with an Elo of 1325, with the Quality version generating in 8.0 seconds and the Speed version in 4.5 seconds.

  2. Kling AIAI score42

    Kling 4.0 Flash launches now for Ultra Yearly subscribers; Kling 4.0 arrives October

    AIKling AI says its Kling 4.0 Flash is live now for Ultra Yearly subscribers, with the full Kling 4.0 coming this October. The update advertises up to 4K resolution, 10-bit HDR output, stereo audio, and native 30-second generation. It also adds Omni Reference supporting up to 15 multimodal references and multi-keyframe control with up to 10 keyframes.

    Video from @Kling_ai's post
  3. TechNode · AIAI score60

    Sanxingdui: Future Past, China's AI-produced theatrical film, releases October 23

    AIBona Film Group announced that Sanxingdui: Future Past, a 100-minute film using AI throughout production, will screen nationwide on October 23, 2026. The production team says AI handled tasks like image generation while over 100 professionals kept creative control, and it took two years and over 1.2 million source images to maintain consistency across the film.

  4. Manus BlogAI score60

    Manus 2.0 adds Cascade agent harness, Manus Studio, and Cue app

    AIManus 2.0 introduces a new agent harness called Cascade, Manus Studio with Video Editor and Game Dev environments, and a standalone Cue app for personal agents. In one tested configuration, Cascade used 23.2% fewer tokens, completed tasks 28.2% faster, and cost 32% less to run than the previous system. Cue is in early access and available with an invite code.

    Why it matters: The post separates the new agent harness, Studio, and Cue, and its Cascade chart gives measured token, time, and cost comparisons against the previous system.

Sep 26

Sep 26Sat
  1. Higgsfield AI 🧩AI score14

    Higgsfield API launches native 1080p Seedance 2.5 with cashback promotion

    AIHiggsfield AI has introduced Seedance 2.5 in native 1080p on its API, positioned as a US-based option for commercial and large-scale productions. The company is offering 100% instant cashback on API spend from an $18M pool, with caps of up to $200,000 per business and $1,000 per individual, and unused cashback expires September 30.

    Video from @higgsfield's post

Sep 25

Sep 25Fri
  1. Amazon ScienceAI score38

    Amazon and Reactor build kernel path to real-time video generation on Trainium

    AIUsing the Neuron Kernel Interface, Reactor and Amazon's Neuron Science team built a kernel-centric path to real-time autoregressive diffusion video generation on Trainium. They addressed dynamic shapes, memory access patterns, and cache management, which are hard for generic compilers, and developed techniques intended to generalize across models.

Sep 24

Sep 24Thu
  1. Latent.SpaceAI score47

    Runway's Co-CEO Argues AI Video Is Heading Toward World Models

    AIRunway co-founder and Co-CEO Alex Germanidis explains why he sees world models as the endgame for AI video, with applications in robotics simulation and real-time generation. He also discusses how Sora pushed Runway into an intense competitive sprint and his vision of a fully neural operating system where interfaces are generated as pixels rather than HTML and code.

    Video from @latentspacepod's post
  2. Google DeepMindAI score62

    Google DeepMind adds Live Avatar to Gemini 3.8 Live for enterprise

    AIGoogle DeepMind has launched Gemini 3.8 Live with Live Avatar, which adds near real-time visual presence to its native live dialogue models. The feature is available today in Gemini Enterprise, supports 97 languages with adaptive lip-sync, and allows custom avatars through enterprise allowlisting. All output carries an imperceptible SynthID watermark.

    Why it matters: The post specifies the new avatar capabilities, the Gemini Enterprise access path, and the SynthID watermark, which helps readers judge its enterprise deployment fit.

  3. Google · Gemini appAI score62

    Google launches Gemini 3.8 Live with Live Avatar for enterprises

    AIGoogle introduced Gemini 3.8 Live with Live Avatar, which adds a visual persona with lip-syncing and expressions to its live dialogue models. The feature is available in Gemini Enterprise and supports 97 languages, with custom avatars available through enterprise allowlisting. Google says all output is watermarked with SynthID.

    Why it matters: The post specifies enterprise availability, custom avatar allowlisting, and 97-language support, which clarifies who can use the feature and how far it reaches.

  4. Google Cloud · AI & Machine LearningAI score55

    Gemini 3.8 Live with Live Avatar becomes generally available in Gemini Enterprise

    AIGoogle says Gemini 3.8 Live with Live Avatar is now generally available in Gemini Enterprise, with US and EU endpoints, provisioned throughput, and enterprise compliance. Its video avatars use synchronized lip-syncing, custom avatars are limited to an allowlist, and generated audio and video carry SynthID watermarks. The model also understands and speaks 97 languages and can run tool calls in the background while the conversation continues.