Расширенный поиск

GitHub-проекты

3 440 записей

LLMApp

github.com

LLM App is a Python library that helps you build real-time LLM-enabled data pipelines with few lines of code.

Репозиторий GitHub-проекты открытый код

LangKit

github.com

Out-of-the-box LLM telemetry collection library that extracts features and profiles prompts, responses and metadata about how your LLM is performing over time to find problems at scale.

Репозиторий GitHub-проекты открытый код

LangFlow

github.com

An effortless way to experiment and prototype LangChain flows with drag-and-drop components and a chat interface.

Репозиторий GitHub-проекты открытый код

Laminar

github.com

Open-source all-in-one platform for engineering AI products. Traces, Evals, Datasets, Labels.

Репозиторий GitHub-проекты открытый код

Hypersigil

github.com

Open-source prompt lifecycle management and gateway with a Web UI.

Репозиторий GitHub-проекты открытый код

Hive

github.com

Open-source AI agent framework for building goal-driven, self-improving autonomous agents with auto-generated graphs, evolution loops, and MCP integration.

Репозиторий GitHub-проекты открытый код

GPTCache

github.com

Creating semantic cache to store responses from LLM queries.

Репозиторий GitHub-проекты открытый код

Glide

github.com

Cloud-Native LLM Routing Engine. Improve LLM app resilience and speed.

Репозиторий GitHub-проекты открытый код

Dstack

github.com

Cost-effective LLM development in any cloud (AWS, GCP, Azure, Lambda, etc).

Репозиторий GitHub-проекты открытый код

deeplake

github.com

Stream large multimodal datasets to achieve near 100% GPU utilization. Query, visualize, & version control data. Access data w/o the need to recompute the embeddings for the model finetuning.

Репозиторий GitHub-проекты открытый код

Contexto

github.com

Self-hosted context engine for AI agents with persistent conversation memory and recall. Works as a drop-in OpenAI-compatible proxy, OpenClaw plugin, or memory SDK — no code changes required.

Репозиторий GitHub-проекты открытый код

Cheshire Cat AI

github.com

Web framework to create vertical AI agents. FastAPI based, plugin system inspired to WordPress, admin panel, vector DB included

Репозиторий GitHub-проекты открытый код

BudgetML

github.com

Deploy a ML inference service on a budget in less than 10 lines of code.

Репозиторий GitHub-проекты открытый код

AI studio

github.com

A Reliable Open Source AI studio to build core infrastructure stack for your LLM Applications. It allows you to gain visibility, make your application reliable, and prepare it for production with features such as caching, rate limiting, exponential retry, model fallback, and more.

Репозиторий GitHub-проекты открытый код

AgentMark

github.com

Type-Safe Markdown-based Agents

Репозиторий GitHub-проекты открытый код

semantic-coverage

github.com

Visualizes RAG knowledge gaps and "blind spots" using 2D UMAP clustering and density detection.

Репозиторий GitHub-проекты открытый код

Future AGI

github.com

Production-grade SDK for observability, automated evaluations and prompt management with sub-100ms guardrails for LLM/agent workflows.

Репозиторий GitHub-проекты открытый код

traceAI

github.com

Open-source AI tracing framework built on OpenTelemetry for deep observability across agentic and LLM workflows.

Репозиторий GitHub-проекты открытый код

RagTune

github.com

CLI tool for debugging and benchmarking RAG retrieval. EXPLAIN ANALYZE for your retrieval layer.

Репозиторий GitHub-проекты открытый код

onWatch

github.com

Lightweight Go CLI that tracks AI API quota usage across 7 providers (Anthropic, OpenAI, GitHub Copilot, MiniMax, and more). Background daemon, <50MB RAM, zero telemetry, SQLite storage.

Репозиторий GitHub-проекты открытый код

whylogs

github.com

The open standard for data logging

Репозиторий GitHub-проекты открытый код

OpenTelemetry-based observability and monitoring for LLM and agents workflows.

Репозиторий GitHub-проекты открытый код

Helicone

github.com

Open source LLM observability platform. One line of code to monitor, evaluate, and experiment with features like prompt management, agent tracing, and evaluations.

Репозиторий GitHub-проекты открытый код

Great Expectations

github.com

Always know what to expect from your data.

Репозиторий GitHub-проекты открытый код

QWED

github.com

Deterministic verification protocol for LLM outputs using 8 formal verification engines (SymPy, Z3, AST, SQLGlot). Prevents hallucinations through mathematical proofs rather than statistical methods.

Репозиторий GitHub-проекты открытый код

Fiddler AI

github.com

Evaluate, monitor, analyze, and improve machine learning and generative models from pre-production to production. Ship more ML and LLMs into production, and monitor ML and LLM metrics like hallucination, PII, and toxicity.

Репозиторий GitHub-проекты открытый код

EvalView

github.com

Regression testing for AI agents. Snapshot behavior, detect tool-call and output regressions, with golden-baseline diffing and LLM-as-judge scoring. Supports LangGraph, CrewAI, OpenAI, Claude, and any HTTP API.

Репозиторий GitHub-проекты открытый код

Azure OpenAI Logger

github.com

"Batteries included" logging solution for your Azure OpenAI instance.

Репозиторий GitHub-проекты открытый код

Plexiglass

github.com

A Python Machine Learning Pentesting Toolbox for Adversarial Attacks. Works with LLMs, DNNs, and other machine learning algorithms.

Репозиторий GitHub-проекты открытый код

dstack

github.com

Open-source confidential AI framework for secure LLM deployment with data privacy, providing hardware-enforced isolation using Intel TDX and NVIDIA Confidential Computing.

Репозиторий GitHub-проекты открытый код

brood-box

github.com

CLI tool for running coding agents inside hardware-isolated microVMs with snapshot isolation, egress control, and MCP authorization.

Репозиторий GitHub-проекты открытый код

Kaito

github.com

A Kubernetes operator that simplifies serving and tuning large AI models (e.g. Falcon or phi-3) using container images and GPU auto-provisioning. Includes an OpenAI-compatible server for inference and preset configurations for popular runtimes such as vLLM and transformers.

Репозиторий GitHub-проекты открытый код

KubeAI

github.com

Deploy and scale machine learning models on Kubernetes. Built for LLMs, embeddings, and speech-to-text.

Репозиторий GitHub-проекты открытый код

Xinference

github.com

Replace OpenAI GPT with another LLM in your app by changing a single line of code. Xinference gives you the freedom to use any LLM you need. With Xinference, you're empowered to run inference with any open-source language models, speech recognition models, and multimodal models, whether in the clou

Репозиторий GitHub-проекты открытый код

ray-llm

github.com

LLMs on Ray - RayLLM (Archived)

Репозиторий GitHub-проекты открытый код

lanarky

github.com

FastAPI framework to build production-grade LLM applications

Репозиторий GitHub-проекты открытый код

langchain-serve

github.com

Serverless LLM apps on Production with Jina AI Cloud (Archived)

Репозиторий GitHub-проекты открытый код

The Triton Inference Server provides an optimized cloud and edge inferencing solution.

Репозиторий GitHub-проекты открытый код

Torchserve

github.com

Serve, optimize and scale PyTorch models in production (Archived)

Репозиторий GitHub-проекты открытый код

TFServing

github.com

A flexible, high-performance serving system for machine learning models.

Репозиторий GitHub-проекты открытый код

Mosec

github.com

A machine learning model serving framework with dynamic batching and pipelined stages, provides an easy-to-use Python interface.

Репозиторий GitHub-проекты открытый код

Jina

github.com

Build multimodal AI services via cloud native technologies · Model Serving · Generative AI · Neural Search · Cloud Native

Репозиторий GitHub-проекты открытый код

x-stable-diffusion

github.com

Real-time inference for Stable Diffusion - 0.88s latency. Covers AITemplate, nvFuser, TensorRT, FlashAttention. (Archived)

Репозиторий GitHub-проекты открытый код

whisper.cpp

github.com

Port of OpenAI's Whisper model in C/C++

Репозиторий GitHub-проекты открытый код

whisper-ctranslate2

github.com

is a 4x faster and low-memory usage drop-in cli replacement that supports word-level timestamps and VAD filter

Репозиторий GitHub-проекты открытый код

Large Language Model Text Generation Inference

Репозиторий GitHub-проекты открытый код

Rapid-MLX

github.com

OpenAI-compatible LLM inference server for Apple Silicon using MLX. 2-4x faster than Ollama with tool calling and prompt caching.

Репозиторий GitHub-проекты открытый код

Off Grid

github.com

Open-source iOS/Android app running LLMs on-device via llama.cpp. Voice (Whisper), vision, image gen, tool calling — fully offline.

Репозиторий GitHub-проекты открытый код

Предложить сайт в каталог

Пришлите ссылку — остальное мы выясним сами.

Мы рассмотрим, что вы прислали, и добавим в каталог, если подойдёт.

Не знаете, как внедрить? Мы поможем

Расскажите про задачу — подберём инструменты и подскажем, с чего начать.

0 / 5000
Проверочный код

Поля со звёздочкой обязательны. Данные используются только для ответа.