What is a #llama-cpp MCP server?
An MCP server tagged #llama-cpp implements the Model Context Protocol so AI assistants like Claude, Cursor, and VS Code can access llama-cpp-related tools, data, or APIs.
Every MCP server and client below is tagged #llama-cpp — install one to give Claude, Cursor, VS Code, or any other MCP-compatible client access to llama-cpp tools.
qso-graph
Named after the Q-signal QSP ("Will you relay?"), qsp-mcp relays tool calls between a local LLM and MCP servers. Any model with function calling capability gains access to the full qso-graph tool ecosystem — 71+ tools across 12 servers — from local weights, not from cloud.
JoniMartin27
InferBench's MCP server lets coding agents run, serve and benchmark local LLMs (text + image, llama.cpp + Stable Diffusion) on your own hardware on demand. Measures real tokens/sec, picks the optimal quant for your GPU, and exposes a 124-model catalog. Local-first, no cloud requi
Common questions about MCP servers and clients tagged #llama-cpp
An MCP server tagged #llama-cpp implements the Model Context Protocol so AI assistants like Claude, Cursor, and VS Code can access llama-cpp-related tools, data, or APIs.
mcp.so currently lists 2 MCP servers and clients tagged #llama-cpp.
Open any server below and copy its install snippet into Claude Desktop, Cursor, VS Code, or another MCP client's configuration — remote servers need no separate download.