This repo contains the source code of the Python package loralib and several examples of how to integrate it with PyTorch models, such as those in Hugging Face. See our paper for a detailed description of LoRA.
This repo contains the source code of the Python package loralib and several examples of how to integrate it with PyTorch models, such as those in Hugging Face. See our paper for a detailed description of LoRA.
PowerInfer is a CPU/GPU LLM inference engine leveraging activation locality for your device. Project Kanban https://github.com/SJTU-IPADS/PowerInfer/assets/34213478/fe441a42-5fce-448b-a3e5-ea4abb43ba23 PowerInfer v.s.
A research and development fork of llama.cpp, providing unique KV cache codecs, inference techniques, and bleeding edge features. Why pay 3-bit or 4-bit quality for a context length you only sometimes reach?
🚅 LiteLLM LiteLLM AI Gateway Open Source AI Gateway for 100+ LLMs. LiteLLM Proxy Server (AI Gateway) | Hosted Proxy | Enterprise Tier | Website --- LiteLLM is an open source AI Gateway that gives you a single, unified interface to call 100+ LLM providers — OpenAI, Anthropic, Gemini, Bedrock, Azure, and more — using the OpenAI format.
A Swift client library for interacting with the Ollama API. Pass "json" to get back a JSON string, or specify a full JSON Schema: The format parameter works with both chat and generate methods.
🔥 Large Language Models(LLM) have taken the ~~NLP community~~ ~~AI community~~ the Whole World by storm. Here is a curated list of papers about large language models, especially relating to ChatGPT.
🤗 PEFT State-of-the-art Parameter-Efficient Fine-Tuning (PEFT) methods Fine-tuning large pretrained models is often prohibitively costly due to their scale. Parameter-Efficient Fine-Tuning (PEFT) methods enable efficient adaptation of large pretrained models to various downstream applications by only fine-tuning a small number of (extra) model parameters instead of all the model's parameters.
Deep Lake: Database for AI Docs • Get Started • API Reference • LangChain & VectorDBs Course • Blog • Whitepaper • Slack • Twitter Deep Lake is a Database for AI powered by a storage format optimized for deep-learning applications. Storing and searching data plus vectors while building LLM applications 2.
𝕏 Follow me on X • 🤗 Hugging Face • 💻 Blog • 📙 LLM Engineer's Handbook The LLM course is divided into three parts: 1. 🧩 LLM Fundamentals is optional and covers fundamental knowledge about mathematics, Python, and neural networks.
This newer video covers the an updated 2024 version of the state of MLOps. You can join the Machine Learning Engineer newsletter.
Just like a compass guides us on our journey, OpenCompass will guide you through the complex landscape of evaluating large language models. With its powerful algorithms and intuitive interface, OpenCompass makes it easy to assess the quality and effectiveness of your NLP models.
Unsloth Studio lets you run and train models locally. Features • News • Quickstart • Notebooks • Documentation Unsloth Studio (Beta) lets you run and train text, audio, embedding, vision models on Windows, Linux and macOS.
Osmantic Deployment System Turn your PC, Mac, or Linux box into a private AI server. AI server and homelab setup is rapidly becoming a solved problem.
This is a hands-on guide to machine learning for programmers with no background in AI. Using a neural network doesn’t require a PhD, and you don’t need to be the person who makes the next breakthrough in AI in order to use what exists today.
I will be updating this tutorials site on a daily basis adding all relevant topcis for 2022 - 2024 especially pertaining to GPU programming, Data Centric AI, Emerging topics like Sustainable AI with Web3AI.js (DeFI, DAO, NFT) and much more. NOTE: All these tutorials are supported and accelerated on NVIDIA GPUs More importantly the applications of ML/DL/AI into industry areas such as Transportation, Medicine/Healthcare etc.
Run Stable Diffusion on Apple Silicon with Core ML [\[Blog Post\]](https://machinelearning.apple.com/research/stable-diffusion-coreml-apple-silicon) [\[BibTeX\]](#bibtex) This repository comprises: If you run into issues during installation or runtime, please refer to the FAQ section. Please refer to the System Requirements section before getting started.
Local Studio is a local-first workstation for running, managing, and using self-hosted LLM backends. One machine can launch models, watch GPU/runtime state, chat with OpenAI-compatible endpoints, and run agent sessions against local or remote controllers.
Geospatial analysis, or just spatial analysis, is an approach to applying statistical analysis and other analytic techniques to data which has a geographical or spatial aspect. with GNSS (global navigation satellite system).
--- 🔥 Large Language Models(LLM) have taken the ~~NLP community~~ ~~AI community~~ the Whole World by storm. Here is a curated list of papers about large language models, especially relating to ChatGPT.