Phoenix
Arize Phoenix · Open-source tracing, evaluation and experimentation for AI agents
Best for: AI engineers who need to trace, evaluate and iterate on LLM agents on their own infrastructure
Running and serving models, hosted or self-managed.
12 tools ranked · updated 2026-08
Ranked 1
Arize Phoenix · Open-source tracing, evaluation and experimentation for AI agents
Best for: AI engineers who need to trace, evaluate and iterate on LLM agents on their own infrastructure
Ranked 2
Helicone · AI gateway and LLM observability for routing, debugging and analysing apps
Best for: AI engineering teams routing, debugging and monitoring LLM calls across many providers
Ranked 3
Replicate · Run, fine-tune and deploy AI models through a cloud API
Best for: Developers who want to run, fine-tune or deploy AI models via an API without managing GPUs
Ranked 4
Langfuse · Open-source tracing, evaluation and prompt management for LLM apps
Best for: Engineering teams tracing, evaluating and improving LLM apps who want open source and self-hosting.
Ranked 5
DigitalOcean · Cloud platform for AI inference, GPUs and general app hosting
Best for: Developers and small teams wanting per-product cloud pricing for AI inference and app hosting
Ranked 6
Ollama · Run open models locally, with cloud capacity for larger ones
Best for: Developers who want to run open models on their own hardware, with cloud capacity for larger models.
Ranked 7
Paperspace · Cloud GPU notebooks, training machines and model deployments
Best for: ML engineers and small teams wanting on-demand cloud GPUs for notebooks, training and deployment
Ranked 8
OpenLLMetry · OpenTelemetry-based tracing for LLM applications
Best for: Python or TypeScript teams who want LLM traces in an observability tool they already run
Ranked 9
Jan · Open-source ChatGPT alternative that runs models locally
Best for: Individuals who want to run open-source AI models locally or connect to cloud models.
Ranked 10
GPT4ALL · Local chatbot running open-source language models on your own device
Best for: Developers, teams and AI power-users who want to run open-source models locally, off the cloud.
Ranked 11
DeepSeek · Chat and API access to the DeepSeek V4 models, billed per token
Best for: Developers wanting a token-billed LLM API with OpenAI- and Anthropic-compatible endpoints
Ranked 12
Open-Notebook · Open-source AI note-taking and research platform
Best for: Researchers, students and professionals who want AI note-taking with control over models and data.
Every score is measured, not opinion: we read each product’s own site, documentation, pricing and changelog, and combine what is there into weighted dimensions covering integration, documentation, transparency and reliability. Scores refresh as products change — see how the rankings work.