Skip to contentSkip to stories

Updated

#Image generation

Showing low-relevance items too. Hide low-relevance items

Sep 30

Sep 30Wed
  1. SenseTimeAI score23

    SenseTime previews Dynamic Design, animating static images with SenseNova 6.8 Flash

    AISenseTime previewed Dynamic Design, powered by SenseNova 6.8 Flash, which turns static images into animated visuals. The system decides which elements stay static and which to animate, chooses HTML/CSS, SVG, transparent images, or video for each element, and choreographs text reveals and subject motion. SenseNova 6.8 Flash is coming soon, and SenseTime is offering a limited beta.

    Video from @SenseTime_AI's post
  2. Tencent HyAI score13

    Tencent Hunyuan launches three-track image challenge with cash prizes for winners.

    AITencent Hunyuan announced a three-track creative challenge with three winners sharing prizes. Background from GMI Cloud says the Hy Image 3.5 preview is free there for seven days, supporting image generation and editing up to 2K with up to five reference images. The post also points to $1,800 in cash and credits for winners.

  3. Kling AI BlogAI score62

    Kling 4.0 enters early access with 30-second native video generation

    AIKling 4.0 is entering early access, with a wider rollout planned for October, and Kling 4.0 Flash opens to Ultra Yearly subscribers on September 28. The update generates videos up to 30 seconds in a single pass, accepts up to 15 reference assets, and supports up to 10 keyframe images. Upcoming features include 10-bit HDR output at 4K and 1080p and video extension up to 2 minutes.

    Why it matters: The post specifies concrete capability limits such as 30-second native generation, up to 15 references, and 10 keyframes, which help users judge fit for production workflows.

Sep 29

Sep 29Tue
  1. Google ResearchAI score35

    Google Research unveils Diffusion Controller for steering AI image generation

    AIGoogle Research introduced Diffusion Controller, a framework that treats image generation as a continuous control problem rather than separate inference-time guidance and fine-tuning fixes. Its lightweight add-on "steering damper" network keeps the base model frozen and works on black-box or gray-box models, and it outperformed the industry standard on human preference matching. In a Stable Diffusion v1.4 test, the fully unlocked version achieved a 90% win rate over the baseline.

  2. Luma AI NewsAI score22

    AI Photo Editing Prompt Formula Preserves Color, Light, and Skin in Campaign Edits

    AIThe article presents a four-part prompt structure (action verb, target element, desired result, protection instructions) for AI photo editing, saying it preserves approved work across platforms. It identifies three common failure causes: unmatched light direction, stacked edits in one prompt, and vague visual language. It states that simple skin retouching takes 2-3 minutes versus 15-30 minutes manually.

Sep 28

Sep 28Mon
  1. ModelScopeAI score46

    Qwen-Image-2.1 LoRAs extract and remove layers for editing

    AIModelScope released two Qwen-Image-2.1 LoRAs, LayerExtract and LayerRemove, for layer-based image editing. LayerExtract isolates a prompt-specified subject onto a transparent background, while LayerRemove deletes the matching object from the source image and reconstructs the scene behind it. Both can be hot-swapped within the same DiffSynth-Studio pipeline, and the LoRA weights are licensed under Apache 2.0, with Qwen-Image-2.1 base-model terms also applying.

    Image from @ModelScope2022's post
  2. Kling AI BlogAI score9

    Kling IMAGE 3.0 Generates Basketball League Logo Concepts From Written Prompts

    AIKling AI's blog outlines a structured prompt method for basketball league logos, covering league identity, basketball symbol, style, colours, and composition. It provides six example prompts for professional, modern, youth, retro, minimal, and street styles, and shows how to generate concepts with Kling IMAGE 3.0 from text or reference images.

Sep 26

Sep 26Sat

Sep 25

Sep 25Fri

Sep 24

Sep 24Thu
  1. Midjourney UpdatesAI score34

    Midjourney Adds Styles Live Previews and Improved Edit Model Inpainting

    AIMidjourney is testing fast models on alpha.midjourney.com and has added live previews in the Styles sidebar, showing how the latest prompt looks across liked and featured styles. Its edit model now changes only the selected pixels during inpainting and outpainting, allowing repeated edits without degrading image quality. Tiling with --tile also works better for V8.1/8.2, with seams between tiles now blending invisibly.

  2. ModelScopeAI score38

    Qwen-Image-2.1-Fun-Controlnet-Union adds eight controls and inpainting

    AIModelScope released Qwen-Image-2.1-Fun-Controlnet-Union, a single checkpoint adding eight structural controls, including Canny, Depth, Pose, and Scribble, plus inpainting to Qwen-Image 2.1. Control and inpainting share one branch with 16 injection points across every second Transformer block, keeping the base model frozen and requiring no checkpoint switching. It runs at guidance scale 1.0 with CFG-distilled sampling and prefix KV caching, and is available under the Qwen Research License with base Qwen-Image 2.1 weights required.

    Image from @ModelScope2022's post

Sep 23

Sep 23Wed
  1. Google LabsAI score29

    Google Labs Releases Six Flow Tools Built by Creatives in Sound, Design, and Content

    AIGoogle Labs released six new Google Flow Tools built by creatives across architecture, sound design, and digital content, including Mondo Sónico, CaptionCast, ThumbnailForge, Surface, CollageMotion Pro, and SwissFlow Studio. Each tool targets a specific workflow, such as generating synchronized audio stems, transcribing and styling captions, or producing animated collages from text prompts. Users can try the tools, duplicate and remix them, or build their own by describing a task in Google Flow.

  2. Comfy BlogAI score62

    Comfy Router launches one API for frontier image, video, 3D, and audio models

    AIComfy Router is now live on the Comfy Developer Platform, giving developers one API to call frontier image, video, 3D, and audio models. Day one models include Seedance 2.5, MiniMax H3, Nano Banana Pro, GPT Image 2, Kling, and Black Forest Labs, and the provider for each job is selectable. Requests fail rather than silently switching providers, and inputs and outputs are deleted after 24 hours.

    Why it matters: The post shows how one API key and a provider parameter let developers swap routes for media models without rewriting calls, with failed requests reporting the provider.

  3. QwenAI score60

    Qwen Intelligence launches three mobile agents and opens its benchmark suite

    AIAlibaba's Qwen launched Qwen Intelligence with three mobile agents: a Mobile Planner Agent, a Mobile-Use Agent, and a Mobile Creative Agent. The post reports benchmark results including MobileWorld 82.1, MobileWorld-Real 92.2, and AndroidDaily 97.2, plus a 90% end-to-end success rate, and says the MobilePA-Bench, MobileWorld, MobileWorld-Real, and MobileWorld-Safety benchmarks are open.

    Image from @Alibaba_Qwen's post
  4. KrASIA · Big TechAI score46

    Tencent Hy Image 3.5 preview refined through its consumer and business products

    AITencent has released a preview of its Hy Image 3.5 image generation model, which product teams across Yuanbao, WorkRally, Ima, and other services are helping refine through co-design. Tencent Cloud prices the model at USD 0.024 per 2K output image, and it supports text-to-image and image-to-image generation with up to five reference images. Tencent said an internal blind evaluation found it on par with ByteDance's Seedream 5.0 Pro and slightly better than Nano-Banana Pro and Qwen-Image-3.0 Pro.

  5. Tencent HyAI score22

    Tencent Hunyuan's Hy Image3.5 preview free for two weeks on OnSolo

    AITencent Hunyuan has made its Hy Image3.5 preview available on the OnSolo platform, free for two weeks. The model targets short drama character sheets, full-motion video game assets, and keyframes, with characters kept consistent across episodes and edits that refine rather than regenerate images. OnSolo's background post says the preview supports 5 references at 2K resolution and is free for Members during the two-week window.

Sep 22

Sep 22Tue
  1. Tencent HyAI score43

    Tencent Hunyuan previews Hy Image3.5 in ComfyUI

    AITencent Hunyuan's Hy Image3.5 preview is now available in ComfyUI, with a claimed 30% higher win rate in human evaluation than Hy Image3.0. The model handles text-to-image and image-to-image in one model at up to 2K resolution, with multilingual text and small print rendering correctly. It also keeps identity and product features consistent across scene, outfit, and style changes.

  2. ModelScopeAI score62

    inclusionAI open-sources Ming-Image-0.1-Design models for visual design

    AIinclusionAI open-sources the Ming-Image-0.1-Design family, two complementary 6B models for visual-design workflows, under an MIT License. Design generates complete UIs, dashboards, infographics, and posters up to 2048×2048 with native transparent RGBA output, and Layer decomposes flattened graphics into independently editable RGBA layers.

    Image from @ModelScope2022's post