2 links tagged with all of: video-generation + text-to-video
Click any tag below to further narrow down your results
Links
This repo introduces LongCat-Video, a 13.6B-parameter model that handles text-to-video, image-to-video, and video continuation within a single framework. It uses block sparse attention and a coarse-to-fine strategy to produce minutes-long 720p/30fps videos without quality drift. The project also includes an audio-driven avatar extension with Whisper-based lip sync and distillation-accelerated inference.
This article covers a service that turns plain text into HD cartoon videos in just a few minutes. It uses AI to generate characters, compose scenes, animate movements, and export the final clip—all without any animation skills.