← Top 100 / Full directory (6,010)

LocalLLM-MCP

AI & Memory Local

MCP server that exposes local llama.cpp LLM models to IBM Bob in VS Code via STDIO, enabling natural language interaction with models like Granite, Nemotron, Gemma, Qwen, and Llama. Configuration is managed through a single models.json file, making it easy to add or disable models without code changes.

How to connect

1
Glama registry
View https://glama.ai/mcp/servers/wyt6q1zhc2 for deploy options, or install from https://github.com/amit11-ibm/LocalLLM-MCP-GITHUB (see README for MCP config).
2
GitHub
Install from https://github.com/amit11-ibm/LocalLLM-MCP-GITHUB and add the server to your MCP client configuration (see repository README).

Tools

Tool names are not listed in our registry for this server. Connect it in your MCP client to see the live tool list.

Related servers