# #diffusion — MDRSS hashtag feed

> Public MDRSS cards tagged #diffusion.
> Canonical feed: https://mdrss.com/feeds/diffusion

## Cards (14)

### [Awesome Video Diffusion](https://mdrss.com/multimodal/image-video-and-creative-ai/901112/901112.md)

A curated list of recent diffusion models for video generation, editing, restoration, understanding, nerf, etc. Use it to navigate the topic and choose relevant methods, papers or tools.

Classification: multimodal/image-video-and-creative-ai · Feed: multimodal · Updated: 2026-08-04T13:54:51.641Z · Version: 1

### [A Survey on Video Diffusion Models](https://mdrss.com/multimodal/image-video-and-creative-ai/901006/901006.md)

Zhen Xing, Qijun Feng, Haoran Chen, Qi Dai, Han Hu, Hang Xu, Zuxuan Wu, Yu-Gang Jiang. Use it to navigate the topic and choose relevant methods, papers or tools.

Classification: multimodal/image-video-and-creative-ai · Feed: multimodal · Updated: 2026-08-04T13:54:51.641Z · Version: 1

### [Stable Diffusion web UI](https://mdrss.com/llm-engineering/models-and-training/2663/2663.md)

A web interface for Stable Diffusion, implemented using Gradio library. Detailed feature showcase with images: Make sure the required dependencies are met and follow the instructions available for: Alternatively, use online services (like Google Colab): 1.

Classification: llm-engineering/models-and-training · Feed: llm-engineering · Updated: 2026-08-04T12:22:38.168Z · Version: 1

### [Why LoongForge?](https://mdrss.com/llm-engineering/models-and-training/2269/2269.md)

English | 简体中文 A unified, high-performance framework for training LLMs, VLMs, diffusion, and embodied models. 🌐 Website &nbsp;·&nbsp; 📖 Docs &nbsp;·&nbsp; ✍️ Blog &nbsp;·&nbsp; ⚡ Quick Start &nbsp;·&nbsp; 📊 Performance &nbsp;·&nbsp; 🏛️ Supported Models &nbsp;·&nbsp; 💬 Contact LoongForge is a unified training framework for LLMs, VLMs, diffusion, and embodied models, covering pre-training, continued pre-training, and SFT.

Classification: llm-engineering/models-and-training · Feed: llm-engineering · Updated: 2026-08-04T12:22:38.168Z · Version: 1

### [Cosmos Tokenizer: A suite of image and video neural tokenizers](https://mdrss.com/multimodal/vision-and-media/2057/2057.md)

As of February 10th, 2025, this repository is read-only. Please visit github.com/NVIDIA/Cosmos for the latest updates and support on Cosmos Tokenizer.

Classification: multimodal/vision-and-media · Feed: multimodal · Updated: 2026-08-04T12:22:38.168Z · Version: 1

### [Model properties](https://mdrss.com/multimodal/vision-and-media/1706/1706.md)

USP : Unvoice and Silence with Pitch when infer 1. Install project dependencies Note: whisper is already built-in, do not install it again otherwise it will cuase conflict and error 3.

Classification: multimodal/vision-and-media · Feed: multimodal · Updated: 2026-08-04T12:22:38.168Z · Version: 1

### [TrainTrain](https://mdrss.com/llm-engineering/models-and-training/1484/1484.md)

2025.12.09 Support Z-Image Turbo Standalone training is now supported. For details, please refer to the Standalone Environment Setup Repository.

Classification: llm-engineering/models-and-training · Feed: llm-engineering · Updated: 2026-08-04T12:22:38.168Z · Version: 1

### [MuseV English 中文](https://mdrss.com/multimodal/vision-and-media/1345/1345.md)

MuseV was a milestone achieved around July 2023. Amazed by the progress of Sora, we decided to opensource MuseV, hopefully it will benefit the community.

Classification: multimodal/vision-and-media · Feed: multimodal · Updated: 2026-08-04T12:22:38.168Z · Version: 1

### [EasyAnimate | An End-to-End Solution for High-Resolution and Long Video Generation](https://mdrss.com/multimodal/vision-and-media/1304/1304.md)

😊 EasyAnimate is an end-to-end solution for generating high-resolution and long videos. We can train transformer based diffusion generators, train VAEs for processing long videos, and preprocess metadata.

Classification: multimodal/vision-and-media · Feed: multimodal · Updated: 2026-08-04T12:22:38.168Z · Version: 1

### [ControlNet for Stable Diffusion WebUI](https://mdrss.com/llm-engineering/models-and-training/1286/1286.md)

The WebUI extension for ControlNet and other injection-based SD controls. This extension is for AUTOMATIC1111's Stable Diffusion web UI, allows the Web UI to add ControlNet to the original Stable Diffusion model to generate images.

Classification: llm-engineering/models-and-training · Feed: llm-engineering · Updated: 2026-08-04T12:22:38.168Z · Version: 1

### [Peft](https://mdrss.com/llm-engineering/models-and-training/1168/1168.md)

🤗 PEFT State-of-the-art Parameter-Efficient Fine-Tuning (PEFT) methods Fine-tuning large pretrained models is often prohibitively costly due to their scale. Parameter-Efficient Fine-Tuning (PEFT) methods enable efficient adaptation of large pretrained models to various downstream applications by only fine-tuning a small number of (extra) model parameters instead of all the model's parameters.

Classification: llm-engineering/models-and-training · Feed: llm-engineering · Updated: 2026-08-04T12:22:38.168Z · Version: 1

### [ViewCrafter: Taming Video Diffusion Models for High-fidelity Novel View Synthesis](https://mdrss.com/multimodal/vision-and-media/1061/1061.md)

ViewCrafter can generate high-fidelity novel views from a single or sparse reference image , while also supporting highly precise pose control. Below shows some examples: Reference image Camera trajecotry Generated novel view video Reference image 1 Reference image 2 Generated novel view video |Model|Resolution|Frames|GPU Mem.

Classification: multimodal/vision-and-media · Feed: multimodal · Updated: 2026-08-04T12:22:38.168Z · Version: 1

### [Generating the documentation](https://mdrss.com/multimodal/vision-and-media/1039/1039.md)

To generate the documentation, you first have to build it. You don't have to commit the built documentation.

Classification: multimodal/vision-and-media · Feed: multimodal · Updated: 2026-08-04T12:22:38.168Z · Version: 1

### [Core ML Stable Diffusion](https://mdrss.com/llm-engineering/serving-and-retrieval/938/938.md)

Run Stable Diffusion on Apple Silicon with Core ML  \ Blog Post\ (https://machinelearning.apple.com/research/stable-diffusion-coreml-apple-silicon)  \ BibTeX\ (#bibtex) This repository comprises: If you run into issues during installation or runtime, please refer to the FAQ section. Please refer to the System Requirements section before getting started.

Classification: llm-engineering/serving-and-retrieval · Feed: llm-engineering · Updated: 2026-08-04T12:22:38.168Z · Version: 1
