Skip to contentSkip to stories

Updated

#Image generation

Showing low-relevance items too. Hide low-relevance items

Oct 5

Oct 5Mon
  1. Liquid AI · new models on Hugging FaceAI score44

    LiquidAI releases d1-omni-600M, a 600M decision model for text, image and audio

    AILiquidAI has released d1-omni-600M on Hugging Face, a 587M-parameter model that answers named yes/no, choice and score questions over text, images or up to 30 seconds of speech in a single forward pass. It returns typed answers with zero output tokens by reading the model's distribution over options, and is built on LFM2.5-Encoder-350M with a 16,384-token context length. The model is not a chat model and does not generate text.

  2. GeekParkAI score46

    Why AI keeps generating beautiful women: a feedback loop of data, taste, and profit

    AIAI image models default to attractive women because training data, averaged-face aesthetics, and user preference feedback reinforce one another. A 1973 test image from Playboy, later widely used in image processing, shows how such defaults form early. Reward models trained on user choices can increase NSFW output even when prompts are unrelated.

Oct 4

Oct 4Sun

Oct 3

Oct 3Sat

Oct 2

Oct 2Fri
  1. SenseTimeAI score9

    SenseTime's SenseNova 6.8 Flash Preview animates still landscape images with dynamic design

    AISenseTime says its SenseNova 6.8 Flash Preview powers Dynamic Design, which adds motion to a static landscape image so sand swirls, a stream flows, ducks move, and snow drifts while buildings stay still. The demonstration, created by Xiaohongshu creator Mr.Muzi Lizi (@Muzlz8888), shows the feature turning a still picture into a scene with rhythm.

    Video from @SenseTime_AI's post
  2. Kling AIAI score52

    Kling 4.0 Enters Closed Beta With Stable Motion and 30-Second Takes

    AIKling 4.0 is in closed beta, with an official launch planned for October, and supports video up to 30 seconds long with up to 4K resolution and 10-bit HDR output. Creator Johnson Sheng reports stable dynamic motion, consistent characters and props across shots, and an unedited 30-second fight sequence, while noting that Kling 4.0 is aimed at commercial production.

Oct 1

Oct 1Thu
  1. NVIDIA · new models on Hugging FaceAI score44

    NVIDIA releases PixelUMM, an encoder-free model for pixel-space image and video tasks

    AINVIDIA has released PixelUMM, an encoder-free unified multimodal model with 15,199,672,064 parameters that handles text, image, and video understanding and generation directly in pixel space. It represents images as 16-by-16 RGB pixel patches on a Qwen3-8B language backbone, with iterative denoising for generation. The checkpoint is licensed for non-commercial research or evaluation only, while the source code is under Apache License 2.0.