Awesome LLM Security
A curation of awesome tools, documents and projects about LLM Security.
A curation of awesome tools, documents and projects about LLM Security.
A collection of papers and resources about aligning large language models (LLMs) with human.
An awesome and curated list of best code-LLM for research.
Awesome LLM compression research papers and tools.
Awesome LLM systems research papers.
A collection of open source, actively maintained web apps for LLM applications.
日本語LLMまとめ - Overview of Japanese LLMs.
The paper list of the review on LLMs in medicine.
A curated list of Awesome LLM Inference Paper with codes.
A curated list of Multi-modal Large Language Model in 3D world, including 3D understanding, reasoning, generation, and embodied agents.
a curated collection of datasets specifically designed for chatbot training, including links, size, language, usage, and a brief description of each dataset
整理开源的中文大语言模型,以规模较小、可私有化部署、训练成本较低的模型为主,包括底座模型,垂直领域微调及应用,数据集与教程等。
Applying Large language models (LLMs) for diverse optimization tasks (Opt) is an emerging research area. This is a collection of references and papers of LLM4Opt.
This paper list focuses on the theoretical or empirical analysis of language models, e.g., the learning dynamics, expressive capacity, interpretability, generalization, and other interesting topics.
an evaluation benchmark focused on ancient Chinese language comprehension.
an expert-driven benchmark for Chineses LLMs.
T5
</details>
Open-Source Toolkit for Efficient Unstructured Data Processing with Pre-built Modules and Local to Cluster Scalability.
Freeing data processing from scripting madness by providing a set of platform-agnostic customizable pipeline processing blocks.
Dingo: A Comprehensive Data Quality Evaluation Tool
A powerful tool for creating high-quality training datasets for Large Language Models
A framework for few-shot evaluation of language models.
a lightweight LLM evaluation suite that Hugging Face has been using internally.
Eval tools by OpenAI.
a repository for evaluating open language models.
A reliable click-and-go evaluation suite compatible with both open-source and proprietary models, supporting MixEval and other benchmarks.
Holistic Evaluation of Language Models (HELM), a framework to increase the transparency of language models.
This repository contains code to quantitatively evaluate instruction-tuned models such as Alpaca and Flan-T5 on held-out tasks.
Testing & evaluation library for LLM applications, in particular RAGs
a framework that helps you evaluate your Retrieval Augmented Generation (RAG) pipelines.
20+ high-performance LLMs with recipes to pretrain, finetune and deploy at scale.
Minimalistic large language model 3D-parallelism training.
DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
Ongoing research training transformer models at scale.
A native PyTorch Library for large model training.