CATALOG DOMAIN · 1 FEEDS

LLM Engineering

19 cards

Training, tuning, serving and evaluating language models.

Subscribe to this viewRSSJSONMD
FEEDLLM Engineering0/0 curated snapshots ready
Create card in LLM EngineeringFeed and taxonomy context will be prefilled.

A powerful local RAG (Retrieval Augmented Generation) application that lets you chat with your PDF documents using Ollama and LangChain. This project includes multiple interfaces: a modern Next.js web app, a Streamlit interface, and Jupyter notebooks for experimentation.

MARKDOWN SNAPSHOT

Loading…

00
DLLM RLAgent

We also introduce a diffusion-based value model that reduces variance and improves stability during optimization. Based on TraceRL, we derive a series of diffusion language models, TraDo, which achieve state-of-the-art performance on math and coding reasoning tasks.

MARKDOWN SNAPSHOT

Loading…

00

I've started to work on reimplementation of the library here: FastTensors Please star it if you'd like to see GGML-compatible implementation in pure Go. Please check out my related project Booster We dream of a world where fellow ML hackers are grokking REALLY BIG GPT models in their homelabs without having GPU clusters consuming a shit tons of $$$.

MARKDOWN SNAPSHOT

Loading…

00

⚡LLM Zoo is a project that provides data, models, and evaluation benchmark for large language models.⚡ [[Tech Report]](assets/llmzoo.pdf) technology gifted by the creator. For example, many pioneers have made great efforts to spread the use of light bulbs and vaccines to developing countries.

MARKDOWN SNAPSHOT

Loading…

00