# #llamacpp — MDRSS hashtag feed

> Public MDRSS cards tagged #llamacpp.
> Canonical feed: https://mdrss.com/feeds/llamacpp

## Cards (4)

### [A Hands-On Guide to Fine-Tuning LLMs with PyTorch and Hugging Face](https://mdrss.com/llm-engineering/models-and-training/2517/2517.md)

Kindle | Paperback | PDF  Leanpub  | PDF  Gumroad  You can easily load the notebooks directly from GitHub using Colab and run them using a GPU provided by Google. You need to be logged in a Google Account of your own.

Classification: llm-engineering/models-and-training · Feed: llm-engineering · Updated: 2026-08-04T12:22:38.168Z · Version: 1

### [wllama - Wasm binding for llama.cpp](https://mdrss.com/llm-engineering/models-and-training/2180/2180.md)

For embeddings, please see examples/embeddings/index.html WebGPU support is introduced via PR #215. Upon updating to V3.1, WebGPU will be enabled automatically.

Classification: llm-engineering/models-and-training · Feed: llm-engineering · Updated: 2026-08-04T12:22:38.168Z · Version: 1

### [koboldcpp](https://mdrss.com/llm-engineering/serving-and-retrieval/1899/1899.md)

KoboldCpp is an easy-to-use AI text-generation software for GGML and GGUF models, inspired by the original KoboldAI. It's a single self-contained distributable that builds off llama.cpp and adds many additional powerful features.

Classification: llm-engineering/serving-and-retrieval · Feed: llm-engineering · Updated: 2026-08-04T12:22:38.168Z · Version: 1

### [Local Studio](https://mdrss.com/llm-engineering/models-and-training/915/915.md)

Local Studio is a local-first workstation for running, managing, and using self-hosted LLM backends. One machine can launch models, watch GPU/runtime state, chat with OpenAI-compatible endpoints, and run agent sessions against local or remote controllers.

Classification: llm-engineering/models-and-training · Feed: llm-engineering · Updated: 2026-08-04T12:22:38.168Z · Version: 1
