Vendor-agnostic orchestration for training, inference and agentic workloads across NVIDIA, AMD, TPU, and Tenstorrent on clouds, Kubernetes, and bare metal.
dstackai/
dstack
dstackai/dstackTools
2.2k
pleaseai/
shunt
pleaseai/shuntTools
Shunt Claude Code agents to any model — selective, per-agent inference-layer routing proxy
98
ddalcu/
mlx-serve
ddalcu/mlx-serveTools
Native LLM inference server for Apple Silicon. OpenAI + Anthropic API compatible. No Python. Includes MLX Core macOS app with chat, agent mode, and tool calling.
1.2k
ashuiGordon/
stata-cli
Agent-native Stata CLI for AI coding agents (Codex, Claude Code, Cursor). Run .do files, inspect data, and return econometric results as JSON via PyStata.
56
The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.
3.7k