LiveTalking: a streaming digital human that talks back live
LiveTalking by lipku is a real-time streaming digital human: synced audio-video, interruptible speech, multi-session, low-latency WebRTC, Apache-2.0.
Published entries across all sections carrying the “Video generation” tag, newest first by publication date on this site.
12 entries
LiveTalking by lipku is a real-time streaming digital human: synced audio-video, interruptible speech, multi-session, low-latency WebRTC, Apache-2.0.
Meituan's LongCat team open-sourced a model that turns one photo and one audio track into a talking video — Whisper lip sync, 8-step distillation, MIT.
HeyGen turns a clip of you into a lip-synced AI avatar for videos in 175+ languages. Free: 3 videos a month, 1 minute each, watermarked. Paid starts at $29/mo.
OpenMontage turns your AI coding assistant into a video studio: 12 pipelines, a storyboard approval gate with per-asset cost, free without any API key.
Anthropic never advertised video for Opus 5.5. Within a week users made it a category: 1,401 videos, four public bills, one fight over whether it is video.
geeklee's MIT agent skill turns SRT subtitles into whiteboard animation: 25–35 s scenes, strokes synced to the narration, built for explainer videos.
AiToEarn is an open-source AI agent that publishes your content to 14 platforms, auto-replies to comments, and settles brand deals — via web, MCP, or Docker.
Open-source SparkDiffusion pushes 720P-14B video generation to 265x on one RTX 5090 via sparse attention, distillation and FP8; VBench-2.0 drops 0.5 points.
html-explainer is an open-source Agent Skill for Claude Code, Codex and other coding agents: give it a topic and it runs research, narration, voiceover, subtitles, frame-by-frame rendering and covers end to end, producing a hard-subtitled MP4. Scenes are written in HTML/CSS/GSAP and rendered with a deterministic seek renderer, so audio and animation stay locked in sync. The default edge-tts voiceover is free and everything runs locally.
A copy-ready motion showreel prompt: an identity challenge with time-driven hard constraints. We ran it on GLM-5.3; coding models like Opus 5.5 take it well.
Skillry's gallery has 389 viral Claude Opus 5.5 videos with original prompts and the site's own remakes; the most reused is a one-line showreel challenge.
The open-source tool Story-Flicks takes nothing but a story theme and automatically produces a complete short video with AI illustrations, narration, and subtitles; the text model can connect to DeepSeek, Alibaba Cloud, OpenAI, and other providers, with one-command Docker deployment and multi-language output.