The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.
3.7k
bsmi021/
mcp-file-context-server
bsmi021/mcp-file-context-serverConnectors
A Model Context Protocol (MCP) server that provides file system context to Large Language Models (LLMs). This server enables LLMs to read, search, and analyze code files with advanced caching and real-time file watching capabilities.
39
ooples/
token-optimizer-mcp
ooples/token-optimizer-mcpConnectors
Measure token savings per AI coding agent, optimize context, and share a live local knowledge graph across 16 CLI clients.
516
r1c7/CluxMateAgents
AI coding agent with one Python core and three front-ends — headless CLI, Textual TUI, and an Electron desktop. Works with any OpenAI-compatible API, with risk-tiered permissions, event-sourced replayable sessions, and a fail-closed OS-level sandbox.
119