#77th of 13 in Observability
Future AGI
Tracing, evals, simulation, guardrails and an LLM gateway for AI agents
Live demo ↗
(opens in a new tab)Documentation ↗
(opens in a new tab)Website ↗
(opens in a new tab)Repository on GitHub ↗
(opens in a new tab)
- Stars
- 2.1k
- License
- Apache-2.0
- Last commit
- Oct 2026
- Last release
- Oct 2026
- Language
- Python
Overview
Future AGI is a Django and Go platform that traces agents over OpenTelemetry, scores outputs with 50+ evaluators, simulates multi-turn conversations, and applies guardrail scanners. It also ships an OpenAI-compatible gateway with routing, caching and virtual keys, plus prompt-optimization algorithms. The default install runs one app container with Postgres and ClickHouse; a distributed Compose setup and a Helm chart cover larger deployments.
Who it is for: Teams running LLM agents who want tracing, evals and a gateway in one stack
Strengths
- OTel tracing with instrumentors for 50+ frameworks in Python, TypeScript, Java and C#
- Gateway is OpenAI-compatible with 100+ providers, semantic caching and virtual keys
- Standalone install is one command and needs 2 vCPUs and 4 GB for Docker
- Apache-2.0 core; Compose, production overlay, Helm and air-gapped modes documented
Weaknesses
- No supported path to move Standalone data to Distributed or Helm later
- Distributed setup needs 4+ vCPUs and 12-16 GB, plus Kafka and PeerDB
- Many components (Postgres, ClickHouse, Redis, Temporal) make it heavy to operate
- Benchmark figures come from the README; independent verification is unknown
What it needs
- RAM ≥ 4 GB
- no GPU
- Docker + Compose
- Needs PostgreSQL, ClickHouse, Redis, Temporal, Docker Compose v2.24+
- Models: OpenAI-compatible providers (100+ via gateway)
- port 3000
Also in Observability
See all 13| Rank | Project | Score |
|---|---|---|
| 1 | LangfuseTracing, prompt management and evals for LLM apps on ClickHouse | 74 out of 100 |
| 2 | PhoenixLLM tracing, evals, datasets and prompt playground built on OpenTelemetry | 74 out of 100 |
| 3 | promptfooCLI for evaluating and red-teaming prompts, agents and RAG | 71 out of 100 |
| #4 | MLflowTracing, evals, prompt registry and AI gateway plus classic ML tracking | 70 out of 100 |
| #5 | OpikTrace, evaluate and monitor LLM apps and agents, Apache-2.0 end to end | 64 out of 100 |