Provision private AI model endpoints on dedicated GPUs (Llama, Qwen, Mistral). Pay per minute.
The inference layer agents run themselves. · Deploy a private AI model on a dedicated GPU. Per-minute billing — no token charges, no monthly minimums. OpenAI-compatible endpoint that never changes. · Why not just call a per-token API? · A model youragent runs itself. · The agent lifecyclein three steps · Everything included.Nothing extra.
| Type | MCP server |
| Section | MCP servers |
| Pricing | free |
| Platform | Command line |
| Systems | cli, api |
| Hosting | cloud |
| Install | mcp |
| Protocols | mcp |
| Site language | en |
| GitHub | auxen-ai/auxen-mcp |
| Launched | 2026-05-18 |