Find the right local LLM setup for your hardware. TokenAssemble helps AI assistants evaluate whether a local language model will run on a specific GPU, CPU, RAM, and VRAM configuration. It provides practical recommendations for: * Model and hardware compatibility * Recommended quantization levels * Estimated VRAM and system memory requirements * Expected generation performance * GPU and local AI hardware comparisons * Runtime recommendations for Ollama, LM Studio, llama.cpp
TokenAssemble Local LLM Advisor
| Type | MCP server |
| Section | MCP servers |
| Pricing | free |
| Platform | Command line |
| Systems | cli, api |
| Hosting | cloud |
| Install | mcp |
| Protocols | mcp |
| Site language | en |
| Vendor | tokenassemble |
| Views | 9 |