Llama.cpp is a GitHub repository or organization with source code, releases, documentation, or project resources.
Directory
Search results
Published directory entries matching your search.
The llama-cpp-agent framework is a tool designed for easy interaction with Large Language Models.
Cpp Core Guidelines is a GitHub repository or organization with source code, releases, documentation, or project resources.
modern-cpp-tricks is a GitHub repository or organization with source code, releases, documentation, or project resources.
Pure-Rust tokenizer for GGUF models, compatible with llama.cpp tokenization.
A self-hosted copilot clone that uses the library behind llama.cpp to run the 6 billion parameter Salesforce Codegen model in 4 GB of RAM.
Agentic components of the Llama Stack APIs.
UI tool for fine-tuning and testing your own LoRA models base on LLaMA, GPT-J and more. One-click run on Google Colab. + A Gradio ChatGPT-like Chat UI to demonstrate your language models.
LlamaIndex is a data framework for your LLM applications.
Official repository to use and implement LLaMA 3.1 models, Meta's state-of-the-art large language models.
Llama3 implementation one matrix multiplication at a time.
Official inference framework for 1-bit LLMs, by Microsoft. opensource.
Port of OpenAI's Whisper model in C/C++. It can be executed locally.
Port of OpenAI's Whisper model in C/C++. opensource.
Instruct-tune LLaMA on consumer hardware.
7B Large Language Model fine-tune by 34B Chinese Character Corpus, based on LLaMA and Alpaca.
A project to make it easier to use large external knowledge bases with LLMs.
Open-source project that supports AI agents, automated workflows, orchestration, or delegated tasks.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Open-source project that provides conversational AI for questions, tasks, and general assistance.
Provides a central interface to connect your LLM's with external data.
Inspired on Private GPT with the GPT4ALL model replaced with the Vicuna-7B model and using the InstructorEmbeddings instead of LlamaEmbeddings.
Chinese LLM, Based on LLaMA and fine tune by Stanford Alpaca, Alpaca LoRA, Japanese-Alpaca-LoRA.
OpenAI compatible API for LLMs and embeddings (LLaMA, Vicuna, ChatGLM and many others).