amit11-ibm/LocalLLM-MCP-GITHUBMCP server that exposes local llama.cpp LLM models to IBM Bob in VS Code via STDIO, enabling natural language interaction with models like Granite, Nemotron, Gemma, Qwen, and Llama. Configuration is managed through a single models.json file, making it easy to add or disable models without code changes.
LOG_LEVELstringOPTIONALdefault: INFOMAX_RETRIESstringOPTIONALdefault: 3QWEN_ENDPOINTstringOPTIONALGEMMA_ENDPOINTstringOPTIONALLLAMA_ENDPOINTstringOPTIONALGRANITE_ENDPOINTstringOPTIONALNEMOTRON_ENDPOINTstringOPTIONALREQUEST_TIMEOUT_SECONDSstringOPTIONALdefault: 120Create a free RNWY account to connect your on-chain identity to this server. MCP server claiming is coming; register now and you'll be first in line.
Create your account →