Magnitude
by Magnitude
Runs open-weight models locally on kernels tuned to your hardware and connects them to coding agents through the magnitude CLI and an OpenAI- and Anthropic-compatible API.
Skills
Hardware Model Recommendations
Inspects local hardware and ranks compatible catalog models by preference with magnitude catalog recommendations, then pulls one to run.
Agent Connections
Configures Pi, OpenCode, Hermes, OpenClaw, Codex, Claude Code, Oh My Pi, or Cline to use local models with magnitude connections add.
Local Inference API
Serves OpenAI Chat Completions, OpenAI Responses, and Anthropic Messages routes on port 10100, as JSON or server-sent event streams.
Remote Inference Server
Runs magnitude serve on a dedicated machine behind an API key so agents on another computer can use its models over the network.
Related Agents
AgentOps for Apify Builders Bundle
Apify actor bundle for agent builders: normalize run traces for QA, control costs, and guard tool calls with a firewall…
Colibri
Runs large Mixture-of-Experts models such as GLM, DeepSeek V4, and Kimi on consumer hardware by streaming routed expert…
CubeSandbox
Runs self-hosted, KVM-isolated sandboxes for agent code execution behind an E2B-compatible API, with snapshots, cloning…
ds4
Runs DeepSeek V4 Flash, GLM 5.x, and Qwen models locally on Metal, CUDA, or ROCm with a native engine, a CLI, a coding…