# kandinskylab/kandinsky-6 > Kandinsky 6.0 is a family of diffusion models for synchronized text-to-audio-video and image-to-audio-video generation: 5-second clips with 44 kHz audio and lip-sync, upscalable to Full-HD. - Magnitude: 3.7 out of 10 — Early signal - Stars: 122 total · +25 stars measured 2026-10-06, 15:03–17:36 UTC, ≈ 32 by evening - Star trust: star growth looks organic - Category: Generative media · Language: Python · License: MIT · Created: 2026-10-05 · Last push: 2026-10-06 - GitHub: https://github.com/kandinskylab/kandinsky-6 · Page: https://gitnova.dev/en/r/kandinskylab/kandinsky-6 ## Useful for - Generate a 5-second clip with audio and lip-sync from a text prompt - Animate a still image into video with synchronized audio - Install kandinsky6 and kandinsky6-sr in ComfyUI via ComfyUI Manager ## Why it’s here - Star-counter measurements on 2026-10-06 (UTC), 15:03–17:36: 97 → 122 stars (+25). This is the change over that interval. - Estimated end-of-day forecast: about +32 stars, using observed gains and the previous day. - The repository is 1 day old and already has 122 stars. With less than two weeks of history, there's no usual pace to compare the spike against yet. - Top new repositories this week: #121. - About 18 forks a day — people are taking the code. - Recent forks include notable developers: @leffff (245 followers). ## Star trust Star growth looks organic. Star-trust labels are heuristics based on the repository’s behavior, not a check of every stargazer. ## Numbers - Forks: 9 - Issues and pull requests: 7 - Watchers: 1 - Average over the last week: 23 per day - Usual pace: too little history (under two weeks) - Stars in the last hour (measured): 1 ## Stars per day, last 3 days (oldest → newest, today is partial) 2026-10-04 … 2026-10-06: 0, 13, 25 ## Spotted in now - Top new repositories this week: #121 ## Similar by description 1. **meituan-longcat/LongCat-Video** — 2.3 · Steady · Generative media · Python · +18 stars measured 2026-10-06, 00:19–17:40 UTC, ≈ 24 by evening Meituan's open 13.6B foundational video generation model unifying Text-to-Video, Image-to-Video and Video-Continuation, including long videos and audio-driven character animation. Full card: https://gitnova.dev/en/r/meituan-longcat/LongCat-Video.md 2. **MiniMax-AI/MiniMax-H3** — 2.1 · Steady · Generative media · Python · +32 stars measured 2026-10-06, 00:22–17:41 UTC, ≈ 43 by evening MiniMax H3 is an omni-modal generative system that understands text, images, video and audio and generates video with native stereo audio up to 2K and 15 seconds. The repository bundles prompt-writing and style-specific video generation… Full card: https://gitnova.dev/en/r/MiniMax-AI/MiniMax-H3.md 3. **QwenLM/Qwen-Image-2.1** — 1.4 · Cooling · Generative media · Python · +10 stars measured 2026-10-06, 00:23–17:45 UTC, ≈ 13 by evening Qwen's open-source text-to-image generation and image editing model, including transparent RGBA output and up to 10 reference images. 7B generation component, 2K support, integrations with Diffusers, ComfyUI, vLLM, SGLang. Full card: https://gitnova.dev/en/r/QwenLM/Qwen-Image-2.1.md 4. **gongnyang/awesome-ai-motion** — 0.4 · Steady · Generative media · HTML · +0 stars measured 2026-10-06, 00:27–17:49 UTC, ≈ 0 by evening A field guide to 637 motion techniques for AI video, infographics and scroll decks: effect cards with parameters, examples, recipes and prompts for Claude Code and Codex. Connects as an agent skill and renders clips locally. Full card: https://gitnova.dev/en/r/gongnyang/awesome-ai-motion.md --- Magnitude (0–10) measures how fast and how unusually interest in a repository is growing right now. It is not a quality score. Days are UTC. “So far today” is a fact; “expected by the end of the day” is a forecast. Summaries and use cases are written by an LLM (DeepSeek V4.1 Flash) from the README and may be inaccurate: verify specific claims (benchmarks, speed, hardware) in the repository itself. Data as of 2026-10-06 17:51 UTC, updated every 30 minutes.