Tencent Hunyuan releases open-source AuK audio model for speech generation and editing
Original title🚀 AuK is officially here. Nano banana🍌 for audio
AISummary
Tencent Hunyuan has released AuK, an open-source foundation model for unified speech generation and editing that takes natural-language instructions and reference audio.
It supports tasks including zero-shot TTS, timbre, style and emotion editing, denoising, and music separation.
A companion AuK-Flash variant runs 4-step inference and is about 4.5 times faster under matched conditions, with code, weights, and a demo now available.
Source: Tencent Hunyuan · x.comPublished · added here