Advanced search

GitHub projects

3 440 entries

Robocorp

github.com

Create, deploy and operate Actions using Python anywhere to enhance your AI agents and assistants. Batteries included with an extensive set of libraries, helpers and logging.

Repository GitHub projects Self-hosted open source

Search with Lepton

github.com

Build your own conversational search engine using less than 500 lines of code by LeptonAI.

Repository GitHub projects Self-hosted open source

Langchain-Chatchat

github.com

Formerly langchain-ChatGLM, local knowledge based LLM (like ChatGLM) QA app with langchain.

Repository GitHub projects Self-hosted open source

IntelliServer

github.com

simplifies the evaluation of LLMs by providing a unified microservice to access and test multiple AI models.

Repository GitHub projects Self-hosted open source

Embedchain

github.com

Framework to create ChatGPT like bots over your dataset.

Repository GitHub projects Self-hosted open source

Serge

github.com

a chat interface crafted with llama.cpp for running Alpaca models. No API keys, entirely self-hosted!

Repository GitHub projects Self-hosted open source

Agenta

github.com

Easily build, version, evaluate and deploy your LLM-powered apps.

Repository GitHub projects Self-hosted open source

promptfoo

github.com

Test your prompts. Evaluate and compare LLM outputs, catch regressions, and improve prompt quality.

Repository GitHub projects Self-hosted open source

wechat-chatgpt

github.com

Use ChatGPT On Wechat via wechaty

Repository GitHub projects Self-hosted open source

magentic

github.com

Seamlessly integrate LLMs as Python functions

Repository GitHub projects Self-hosted open source

LiteChain

github.com

Lightweight alternative to LangChain for composing LLMs

Repository GitHub projects Self-hosted open source

Swiss Army Llama

github.com

Comprehensive set of tools for working with local LLMs for various tasks.

Repository GitHub projects Self-hosted open source

LlamaIndex

github.com

A Python library for augmenting LLM apps with data.

Repository GitHub projects Self-hosted open source

LangChain

github.com

A popular Python/JavaScript library for chaining sequences of language model prompts.

Repository GitHub projects Self-hosted open source

dspy

github.com

DSPy: The framework for programming—not prompting—foundation models.

Repository GitHub projects Self-hosted open source

Easily deploy any LLM on a VM with minimal configuration, using Ansible.

Repository GitHub projects Self-hosted open source

prima.cpp

github.com

A distributed implementation of llama.cpp that lets you run 70B-level LLMs on your everyday devices.

Repository GitHub projects Self-hosted open source

Liger-Kernel

github.com

Efficient Triton Kernels for LLM Training.

Repository GitHub projects Self-hosted open source

LMDeploy

github.com

A high-throughput and low-latency inference and serving framework for LLMs and VLs

Repository GitHub projects Self-hosted open source

Infinity

github.com

Inference for text-embeddings in Python

Repository GitHub projects Self-hosted open source

Inference for text-embeddings in Rust, HFOIL Licence.

Repository GitHub projects Self-hosted open source

DeepSpeed-Mii

github.com

MII makes low-latency and high-throughput inference, similar to vLLM powered by DeepSpeed.

Repository GitHub projects Self-hosted open source

OpenLLM

github.com

Fine-tune, serve, deploy, and monitor any open-source LLMs in production. Used in production at BentoML for LLMs-based applications.

Repository GitHub projects Self-hosted open source

SkyPilot

github.com

Run LLMs and batch jobs on any cloud. Get maximum cost savings, highest GPU availability, and managed execution -- all with a simple interface.

Repository GitHub projects Self-hosted open source

mistral.rs

github.com

Blazingly fast LLM inference.

Repository GitHub projects Self-hosted open source

exllama

github.com

A more memory-efficient rewrite of the HF transformers implementation of Llama for use with quantized weights.

Repository GitHub projects Self-hosted open source

MInference

github.com

To speed up Long-context LLMs' inference, approximate and dynamic sparse calculate the attention, which reduces inference latency by up to 10x for pre-filling on an A100 while maintaining accuracy.

Repository GitHub projects Self-hosted open source

FasterTransformer

github.com

NVIDIA Framework for LLM Inference(Transitioned to TensorRT-LLM)

Repository GitHub projects Self-hosted open source

TensorRT-LLM

github.com

Nvidia Framework for LLM Inference

Repository GitHub projects Self-hosted open source

SGLang

github.com

SGLang is a fast serving framework for large language models and vision language models.

Repository GitHub projects Self-hosted open source

Axolotl

github.com

Open-source framework for fine-tuning and evaluating LLMs. It simplifies the process of experimenting with different training configurations and makes it easy to reproduce and share results, supporting features like LoRA, QLoRA, DeepSpeed, PEFT, and multi-GPU setups.

Repository GitHub projects Self-hosted open source

OpenRLHF

github.com

An Easy-to-use, Scalable and High-performance RLHF Framework (70B+ PPO Full Tuning & Iterative DPO & LoRA & RingAttention & RFT).

Repository GitHub projects Self-hosted open source

Transformer Engine

github.com

A library for accelerating Transformer model training on NVIDIA GPUs.

Repository GitHub projects Self-hosted open source

GPT-NeoX

github.com

An implementation of model parallel autoregressive transformers on GPUs, based on the DeepSpeed library.

Repository GitHub projects Self-hosted open source

maxtext

github.com

A simple, performant and scalable Jax LLM!

Repository GitHub projects Self-hosted open source

Mesh Tensorflow

github.com

Mesh TensorFlow: Model Parallelism Made Easier.

Repository GitHub projects Self-hosted open source

BMTrain

github.com

Efficient Training for Big Models.

Repository GitHub projects Self-hosted open source

veRL

github.com

veRL is a flexible and efficient RL framework for LLMs.

Repository GitHub projects Self-hosted open source

ROLL

github.com

An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models.

Repository GitHub projects Self-hosted open source

torchtune

github.com

A Native-PyTorch Library for LLM Fine-tuning.

Repository GitHub projects Self-hosted open source

Megatron-DeepSpeed

github.com

DeepSpeed version of NVIDIA's Megatron-LM that adds additional support for several features such as MoE model training, Curriculum Learning, 3D Parallelism, and others.

Repository GitHub projects Self-hosted open source

torchtitan

github.com

A native PyTorch Library for large model training.

Repository GitHub projects Self-hosted open source

Megatron-LM

github.com

Ongoing research training transformer models at scale.

Repository GitHub projects Self-hosted open source

DeepSpeed

github.com

DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.

Repository GitHub projects Self-hosted open source

nanotron

github.com

Minimalistic large language model 3D-parallelism training.

Repository GitHub projects Self-hosted open source

Litgpt

github.com

20+ high-performance LLMs with recipes to pretrain, finetune and deploy at scale.

Repository GitHub projects Self-hosted open source

Ragas

github.com

a framework that helps you evaluate your Retrieval Augmented Generation (RAG) pipelines.

Repository GitHub projects Self-hosted open source

Giskard

github.com

Testing & evaluation library for LLM applications, in particular RAGs

Repository GitHub projects Self-hosted open source

Submit a site to the catalog

Just send the link — we will work out the rest.

We will review what you send and add it to the catalog if it fits.

Not sure how to implement it? We can help

Tell us about your task — we will pick the tools and suggest where to start.

0 / 5000
Verification code

Fields marked with an asterisk are required. Your data is used only to reply.