Sign inSign up

enterpilot/gomodel

By enterpilot

•Updated 3 days ago

Fast, resource-efficient AI gateway in Go. OpenAI-compatible, self-hosted LiteLLM alternative.

Image
API management
Machine learning & AI
Monitoring & observability
1

50K+

enterpilot/gomodel repository overview

⁠GoModel - high-performance, lightweight AI gateway written in Go

GoModel is a fast, self-hosted AI gateway and LLM proxy with OpenAI-compatible and Anthropic-compatible APIs. Connect your applications to one endpoint and route requests across cloud and local AI model providers.

It is a lightweight LiteLLM alternative for teams that want lower AI costs, reliable model access, and complete LLM observability without a heavy control plane.

GoModel saves you money and nerves.

  • Money - cache repeated requests, track every token and cost, enforce budgets, and route traffic to the right model.
  • Nerves - keep applications running with load balancing, retries, circuit breakers, provider health checks, and automatic failover.

⁠Quick Start

docker run --rm --name gomodel -p 8080:8080 \
  -v gomodel-data:/app/data \
  -e OPENAI_API_KEY="your-openai-key" \
  enterpilot/gomodel

Open the admin dashboard:

http://localhost:8080/admin/dashboard

Send your first request:

curl http://localhost:8080/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer change-me" \
  -d '{
    "model": "gpt-4o-mini",
    "messages": [{"role": "user", "content": "Say hello in one sentence."}]
  }'

⁠Features

  • Multi-provider AI gateway - use OpenAI, Anthropic, Google Gemini, Vertex AI, Azure OpenAI, Amazon Bedrock, Cohere, DeepSeek, Groq, xAI, OpenRouter, Ollama, vLLM, and many more through one API.
  • OpenAI and Anthropic API compatibility - supports Chat Completions, Responses API, Conversations, Messages API, embeddings, audio, files, batches, and realtime APIs. Use the official SDKs by changing the base URL.
  • Virtual models and load balancing - expose stable model aliases and balance requests with round-robin or cost-based routing.
  • Failover and resilience - reroute failed requests to backup models or providers with retries, circuit breakers, and live provider health.
  • Response caching - reduce latency and cost with exact caching, semantic caching, and provider-native prompt caching.
  • Cost tracking and usage analytics - understand spend by provider, model, user path, key, or label from the built-in dashboard.
  • Budgets and rate limits - enforce spend limits, requests per minute, tokens per minute, and concurrency limits.
  • Access control - create managed API keys and scope model access, usage, budgets, workflows, and audit logs with hierarchical user paths.
  • Guardrails - inspect, modify, or reject requests and responses at the gateway.
  • MCP gateway - aggregate Model Context Protocol servers behind one authenticated endpoint with tool access controls, usage tracking, and rate limits.
  • Provider-native passthrough - access native provider APIs while keeping GoModel authentication, tracking, and observability.
  • Provider key rotation - spread traffic across multiple API keys while preserving session affinity for prompt caching.
  • Full observability - real-time request logs, audit logs, usage analytics, provider status, cost estimates, and Prometheus metrics.
  • Flexible configuration - configure GoModel with environment variables, config.yaml, or the admin dashboard without restarting.
  • Production-friendly container - compact distroless image, non-root runtime, built-in health check, and support for linux/amd64, linux/arm64, and linux/arm/v7.

⁠Supported providers

OpenAI, Anthropic, Cohere, Google Gemini, Google Vertex AI, DeepSeek, Groq, Fireworks AI, Meta, OpenRouter, Kilo AI, Z.ai, xAI, Alibaba Cloud Model Studio, MiniMax, Xiaomi MiMo, OpenCode Go, Kimi Code, Azure OpenAI, Oracle GenAI, Ollama, vLLM, Amazon Bedrock, Amazon Bedrock Mantle, and other OpenAI-compatible providers.

⁠More information

Tag summary

Content type

Image

Digest

sha256:77e5e1546…

Size

20.4 MB

Last updated

3 days ago

docker pull enterpilot/gomodel