Phoenix Eval
FreeNot checkedMCP server for Arize Phoenix enabling AI agents to perform LLM tracing, evaluation, and dataset management for automated quality assurance.
About
MCP server for Arize Phoenix enabling AI agents to perform LLM tracing, evaluation, and dataset management for automated quality assurance.
README
MCP server for Arize Phoenix — LLM tracing, evaluation, and dataset management via AI agents.
Glama Quality Score CI License: MIT Python Arize Phoenix MCP
What is this?
phoenix-mcp-eval is an MCP (Model Context Protocol) server that exposes Arize Phoenix's LLM observability capabilities to AI agents. It enables AI-driven analysis of traces, evaluation of LLM outputs, and management of evaluation datasets — directly from an MCP-compatible agent.
Built for platform engineers and ML teams running LLM pipelines on AI Foundry, LangChain, or LlamaIndex who need automated quality assurance and tracing.
Available Tools
| Tool | Description |
|---|---|
list_projects |
List all Phoenix tracing projects |
get_traces |
Retrieve LLM traces for a project with filters |
get_spans |
Get individual spans with input/output/latency data |
list_datasets |
List evaluation datasets in Phoenix |
get_dataset |
Fetch dataset examples for review or comparison |
list_evaluations |
List evaluation runs and their scores |
get_evaluation_summary |
Get aggregated evaluation metrics (precision, recall, etc.) |
query_traces |
Run structured queries over trace data |
Quick Start
Prerequisites
- Python 3.11+
- Arize Phoenix instance (self-hosted or cloud)
- Phoenix API key or local server URL
Installation
git clone [https://github.com/akkireddy-challa/phoenix-mcp-eval.git](https://github.com/akkireddy-challa/phoenix-mcp-eval.git)
cd phoenix-mcp-eval
pip install -r requirements.txt
Configuration
export PHOENIX_HOST=http://localhost:6006
export PHOENIX_API_KEY=<your-api-key> # if using cloud
Run
python server.py
MCP Client Config (Claude Desktop)
{
"mcpServers": {
"phoenix": {
"command": "python",
"args": ["/path/to/phoenix-mcp-eval/server.py"],
"env": {
"PHOENIX_HOST": "http://localhost:6006"
}
}
}
}
Security Model
- Connects to Phoenix via API key or local network only
- All operations are read-only by default (trace/eval retrieval)
- No model weights, prompts, or PII are transmitted outside Phoenix
- API key stored in environment variables, never in code
- Designed for internal network use within a Kubernetes cluster
Use Cases at Telia
This pattern is used to allow AI agents to:
- Automatically review LLM trace quality after AI Foundry deployments
- Surface failing evaluation metrics to on-call engineers without manual Phoenix access
- Compare evaluation datasets across model versions
- Trigger re-evaluation jobs based on trace anomaly detection
Roadmap
-
run_evaluation— trigger evaluation jobs programmatically -
create_dataset— export traces to evaluation datasets -
get_prompt_templates— retrieve versioned prompts from Phoenix - Integration with Azure AI Foundry deployment events
- GitHub Actions workflow for CI validation
Related Projects
| Repo | Purpose |
|---|---|
| k8s-mcp-server | Kubernetes cluster diagnostics via MCP |
| azure-mcp-platform | Azure resource management via MCP |
| grafana-mcp-observability | Grafana dashboards and alerts via MCP |
License
MIT License. See LICENSE for details.
Built by Akkireddy Challa — Platform Engineer at Telia, Stockholm.
Installing Phoenix Eval
This server has no published package — it is built from source. Open the repository and follow its README.
▸ github.com/akkireddy-challa/phoenix-mcp-evalFAQ
Is Phoenix Eval MCP free?
Yes, Phoenix Eval MCP is free — one-click install via Unyly at no cost.
Does Phoenix Eval need an API key?
No, Phoenix Eval runs without API keys or environment variables.
Is Phoenix Eval hosted or self-hosted?
Self-hosted: the server runs locally on your machine via the install command above.
How do I install Phoenix Eval in Claude Desktop, Claude Code or Cursor?
Open Phoenix Eval on unyly.org, pick your client tab (Claude Desktop, Claude Code, Cursor) and press Install — the config is generated automatically, no JSON editing.
Related MCPs
Fetch
Web content fetching and conversion for efficient LLM usage.
Roblox Studio
Enables AI coding tools to control Roblox Studio for workspace exploration, instance manipulation, and script management. It provides tools for playtesting, sce
by paralovAWS KB Retrieval
Retrieval from AWS Knowledge Base using Bedrock Agent Runtime.
by modelcontextprotocolSpring AI MCP Server
Provides auto-configuration for setting up an MCP server in Spring Boot applications.
llm-analysis-assistant
A very streamlined mcp client that supports calling and monitoring stdio/sse/streamableHttp, and can also view request responses through the /logs page. It also
by xuzexin-hzMCP-Agent
A simple, composable framework to build agents using Model Context Protocol by [LastMile AI](https://www.lastmileai.dev)
by lastmile-aiSpring AI MCP Client
Provides auto-configuration for MCP client functionality in Spring Boot applications.
mcp.natoma.ai
A Hosted MCP Platform to discover, install, manage and deploy MCP servers by [Natoma Labs](https://www.natoma.ai)
MCPHub
Website to list high quality MCP servers and reviews by real users. Also provide online chatbot for popular LLM models with MCP server support.
MCP Servers Rating and User Reviews
Website to rate MCP servers, write authentic user reviews, and [search engine for agent & mcp](http://www.deepnlp.org/search/agent)
Compare Phoenix Eval with
Not sure what to pick?
Find your stack in 60 seconds
Author?
Embed badge for your README
Browse similar
All ai MCPs
