Тип

Подборка — 3 881

Megatron-DeepSpeed

github.com

DeepSpeed version of NVIDIA's Megatron-LM that adds additional support for several features such as MoE model training, Curriculum Learning, 3D Parallelism, and others.

Репозиторий GitHub-проекты Своё развёртывание

torchtune

github.com

A Native-PyTorch Library for LLM Fine-tuning.

Репозиторий GitHub-проекты Своё развёртывание

ROLL

github.com

An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models.

Репозиторий GitHub-проекты Своё развёртывание

veRL

github.com

veRL is a flexible and efficient RL framework for LLMs.

Репозиторий GitHub-проекты Своё развёртывание

BMTrain

github.com

Efficient Training for Big Models.

Репозиторий GitHub-проекты Своё развёртывание

Mesh Tensorflow

github.com

Mesh TensorFlow: Model Parallelism Made Easier.

Репозиторий GitHub-проекты Своё развёртывание

maxtext

github.com

A simple, performant and scalable Jax LLM!

Репозиторий GitHub-проекты Своё развёртывание

GPT-NeoX

github.com

An implementation of model parallel autoregressive transformers on GPUs, based on the DeepSpeed library.

Репозиторий GitHub-проекты Своё развёртывание

Transformer Engine

github.com

A library for accelerating Transformer model training on NVIDIA GPUs.

Репозиторий GitHub-проекты Своё развёртывание

OpenRLHF

github.com

An Easy-to-use, Scalable and High-performance RLHF Framework (70B+ PPO Full Tuning & Iterative DPO & LoRA & RingAttention & RFT).

Репозиторий GitHub-проекты Своё развёртывание

Axolotl

github.com

Open-source framework for fine-tuning and evaluating LLMs. It simplifies the process of experimenting with different training configurations and makes it easy to reproduce and share results, supporting features like LoRA, QLoRA, DeepSpeed, PEFT, and multi-GPU setups.

Репозиторий GitHub-проекты Своё развёртывание

TensorRT-LLM

github.com

Nvidia Framework for LLM Inference

Репозиторий GitHub-проекты Своё развёртывание

FasterTransformer

github.com

NVIDIA Framework for LLM Inference(Transitioned to TensorRT-LLM)

Репозиторий GitHub-проекты Своё развёртывание

MInference

github.com

To speed up Long-context LLMs' inference, approximate and dynamic sparse calculate the attention, which reduces inference latency by up to 10x for pre-filling on an A100 while maintaining accuracy.

Репозиторий GitHub-проекты Своё развёртывание

exllama

github.com

A more memory-efficient rewrite of the HF transformers implementation of Llama for use with quantized weights.

Репозиторий GitHub-проекты Своё развёртывание

mistral.rs

github.com

Blazingly fast LLM inference.

Репозиторий GitHub-проекты Своё развёртывание

SkyPilot

github.com

Run LLMs and batch jobs on any cloud. Get maximum cost savings, highest GPU availability, and managed execution -- all with a simple interface.

Репозиторий GitHub-проекты Своё развёртывание

OpenLLM

github.com

Fine-tune, serve, deploy, and monitor any open-source LLMs in production. Used in production at BentoML for LLMs-based applications.

Репозиторий GitHub-проекты Своё развёртывание

DeepSpeed-Mii

github.com

MII makes low-latency and high-throughput inference, similar to vLLM powered by DeepSpeed.

Репозиторий GitHub-проекты Своё развёртывание

Inference for text-embeddings in Rust, HFOIL Licence.

Репозиторий GitHub-проекты Своё развёртывание

Infinity

github.com

Inference for text-embeddings in Python

Репозиторий GitHub-проекты Своё развёртывание

LMDeploy

github.com

A high-throughput and low-latency inference and serving framework for LLMs and VLs

Репозиторий GitHub-проекты Своё развёртывание

Liger-Kernel

github.com

Efficient Triton Kernels for LLM Training.

Репозиторий GitHub-проекты Своё развёртывание

prima.cpp

github.com

A distributed implementation of llama.cpp that lets you run 70B-level LLMs on your everyday devices.

Репозиторий GitHub-проекты Своё развёртывание

Easily deploy any LLM on a VM with minimal configuration, using Ansible.

Репозиторий GitHub-проекты Своё развёртывание

dspy

github.com

DSPy: The framework for programming—not prompting—foundation models.

Репозиторий GitHub-проекты Своё развёртывание

LangChain

github.com

A popular Python/JavaScript library for chaining sequences of language model prompts.

Репозиторий GitHub-проекты Своё развёртывание

LlamaIndex

github.com

A Python library for augmenting LLM apps with data.

Репозиторий GitHub-проекты Своё развёртывание

Swiss Army Llama

github.com

Comprehensive set of tools for working with local LLMs for various tasks.

Репозиторий GitHub-проекты Своё развёртывание

LiteChain

github.com

Lightweight alternative to LangChain for composing LLMs

Репозиторий GitHub-проекты Своё развёртывание

magentic

github.com

Seamlessly integrate LLMs as Python functions

Репозиторий GitHub-проекты Своё развёртывание

wechat-chatgpt

github.com

Use ChatGPT On Wechat via wechaty

Репозиторий GitHub-проекты Своё развёртывание

promptfoo

github.com

Test your prompts. Evaluate and compare LLM outputs, catch regressions, and improve prompt quality.

Репозиторий GitHub-проекты Своё развёртывание

Agenta

github.com

Easily build, version, evaluate and deploy your LLM-powered apps.

Репозиторий GitHub-проекты Своё развёртывание

Serge

github.com

a chat interface crafted with llama.cpp for running Alpaca models. No API keys, entirely self-hosted!

Репозиторий GitHub-проекты Своё развёртывание

Embedchain

github.com

Framework to create ChatGPT like bots over your dataset.

Репозиторий GitHub-проекты Своё развёртывание

Предложить сайт в каталог

Пришлите ссылку — остальное мы выясним сами.

Мы рассмотрим, что вы прислали, и добавим в каталог, если подойдёт.

Не знаете, как внедрить? Мы поможем

Расскажите про задачу — подберём инструменты и подскажем, с чего начать.

0 / 5000
Проверочный код

Поля со звёздочкой обязательны. Данные используются только для ответа.