CATALOG DOMAIN · 1 FEEDS

Multimodal AI

9 cards

Vision, speech and generative media systems.

Subscribe to this viewRSSJSONMD
FEEDMultimodal AI0/0 curated snapshots ready
Create card in Multimodal AIFeed and taxonomy context will be prefilled.

This repository provides the official implementation of TurboDiffusion, a video generation acceleration framework that can speed up end-to-end diffusion generation by $100 \sim 200\times$ on a single RTX 5090, while maintaining video quality. TurboDiffusion primarily uses SageAttention, SLA (Sparse-Linear Attention) for attention acceleration, and rCM for timestep distillation.

MARKDOWN SNAPSHOT

Loading…

00

ViewCrafter can generate high-fidelity novel views from a single or sparse reference image , while also supporting highly precise pose control. Below shows some examples: Reference image Camera trajecotry Generated novel view video Reference image 1 Reference image 2 Generated novel view video |Model|Resolution|Frames|GPU Mem.

MARKDOWN SNAPSHOT

Loading…

00