Train a real-time, playable Video World Models on 8 GPUs — keyboard & mouse control & gamepad, fully open and reproducible.
-
Updated
Aug 19, 2026 - Python
Train a real-time, playable Video World Models on 8 GPUs — keyboard & mouse control & gamepad, fully open and reproducible.
Quartz — custom CUDA/Vulkan GGUF inference engine + SD1.5/SDXL image and video generation (Rust, no llama.cpp)
Unofficial ComfyUI integration of Causal Forcing / Causal Forcing++ (official frame-wise 2-step and 4-step checkpoints, official inference path, separate WSL2/Linux runtime)
Unofficial ComfyUI integration of MonarchRT (training-free, on the public Self-Forcing / Wan2.1-T2V-1.3B weights) with a dense A/B baseline, run in a separate WSL2/Linux runtime.
Make AI trailers 100% locally on a 6 GB GPU: video (Wan 2.1), music (ACE-Step), narration (OmniVoice) and ffmpeg mix, driven by a storyboard.json, by hand or by an LLM agent.
Local text to video pipeline that turns a text prompt into a lip synced talking video. Wan2.1 handles video generation, Qwen3 TTS creates the voice, and MuseTalk 1.5 syncs the lips. Everything runs offline after the models are downloaded.
InfernoBench: open-source AI model installers, local inference benchmarks, and reproducible browser labs. Wan2.1 and Sulphur-2 included.
🎥 AI Video Generation Studio — Text-to-Video & Image-to-Video engine with Gradio & Free Cloud GPU Acceleration
InfiniteTalk (Wan2.1) talking-avatar pipeline: ComfyUI workflows, RunPod pod/serverless setup, A/B matrix and forensic evaluation vs HeyGen
To associate your repository with the wan2-1 topic, visit your repo's landing page and select "manage topics."