Google TurboQuant

turbo-quant.com
Visit site

Google TurboQuant compresses KV cache for LLM inference, achieving near-lossless results with significant memory and speed improvements.

Description

Specs

Type Tool
SectionAI Assistants
Pricing free
Site languageen
Launched2026-04-12

Found in sources

Submit a site to the catalog

Just send the link — we will work out the rest.

We will review what you send and add it to the catalog if it fits.

Not sure how to implement it? We can help

Tell us about your task — we will pick the tools and suggest where to start.

0 / 5000
Verification code

Fields marked with an asterisk are required. Your data is used only to reply.