# Prefect Codex MCP Evaluation

**Category:** Durable Workflow
**Score:** 48 (D)

This is the agent-readable Devtool Arena evaluation page for Prefect on Codex MCP.

## Summary

| Metric | Value |
|--------|-------|
| Status | completed |
| Overall score | 48 |
| Grade | D |
| Eval score | 31 |
| Discovery score | 89 |
| Cost | $0.22 |
| Runtime | 3m 9s |
| Tool calls | 22 |
| Errors | 7 |
| Tokens used | 835661 |
| Success | No |

## Agent-Readiness Checklist

| Signal | Evidence |
|--------|----------|
| Context7 | Not found |
| llms.txt | Not found |
| MCP server | Not found |
| Typed SDK | Not found |
| OpenAPI | Not found |
| Agent skills | Not found |
| CLI | Not found |

## Evaluation Prompt

````text

# Task
Using the Prefect MCP, interact with the workflow engine:

1. Discover available MCP tools and their capabilities
2. List existing workflow executions (or deployments/flows)
3. Start a new workflow execution with a ***** payload (use any available workflow type or create a simple one)
4. Check the status of the workflow execution
5. Print all output clearly labeled for each step

Use ONLY the MCP tools for all operations.

You MUST use the Prefect MCP server to accomplish this task.
Use the MCP tools for the actual product operations. Do NOT use curl, fetch, raw HTTP requests, CLI commands, or SDK calls as substitutes.

Use the MCP tools' help/descriptions to understand what each tool does.
If you need additional context, check the documentation: https://docs.prefect.io/v3/how-to-guides/ai/use-prefect-mcp-server

## Setup
Do NOT assume the Prefect MCP server is already installed or already running in this sandbox.
Review the installation and authentication details below before starting the task.

### Installation
Transport: `stdio`
Codex is configured to launch the `prefect` MCP server over stdio using: `uvx --from prefect-mcp prefect-mcp-server`
Stored MCP package hint: `prefect-mcp-server`
No separate setup command is required before startup unless the docs indicate one.
Bootstrap notes: Seeded from backend/mcp-claude-working.json on 2026-04-13.

### Authentication
The following environment variables are already set in this environment: `PREFECT_API_URL`, `PREFECT_API_KEY`
The MCP server should read them automatically when it starts.

## Execution
1. The harness has already written the launch configuration for `prefect`. Start with the task directly; do not rewrite agent settings, `.mcp-runtime.json`, or build a separate MCP client unless the connected server still fails after a normal retry.
2. You do not need to enumerate every MCP tool up front; inspect only the relevant tools you need. Agent-facing MCP tool names may be namespaced like `mcp__prefect__<tool_name>`.
3. Complete the task described above using MCP tools for the actual product operations.
4. If the server fails to start or authenticate, debug the documented setup and retry. Do not fall back to raw HTTP, SDK, CLI, or manual JSON-RPC/MCP calls as a substitute for exposed MCP tools.

Work in the /home/daytona/app directory.
If you use Bash during setup or debugging, limit it to installation/configuration work and then return to MCP tools for the task itself.
If there are errors, debug and fix them until the task runs successfully.

## Summary
When you are done, write a JSON summary file to /home/daytona/app/results.json with this structure:
{
  "operations": [
    {"step": 1, "description": "What you did", "mcp_tool": "tool_name", "success": true/false, "output_summary": "brief result"},
    ...
  ],
  "overall_success": true/false,
  "notes": ["any relevant notes about the execution"]
}

````

## Grader Results

| Check | Passed | Weight | Score | Details |
|-------|--------|--------|-------|---------|
| Task | No | — | — | — |
| Llms Txt | Yes | — | 10 | — |
| Tool Count | Yes | — | 10 | — |
| Auth Method | Yes | — | 15 | — |
| Official Mcp | Yes | — | 5 | — |
| Output Schema | No | — | 0 | — |
| Error Messages | — | — | 0 | docs_search_prefect: Poor error message (generic or stack trace); get_automations: Poor error message (generic or stack trace) |
| Tool Stability | Yes | — | 5 | — |
| Security Posture | — | — | 9 | — |
| Tool Api Mapping | — | — | 14 | — |
| Description Clarity | — | — | 14 | — |
| Tool Description Length | Yes | — | 7 | — |

## Run Artifacts

| Artifact | Value |
|----------|-------|
| Conversation turns | 29 |
| Tool call traces | 22 |
| Generated files | 1 |
| Exit code | — |
| Completed at | 2026-10-04T16:31:48.096159+00:00 |

### Generated Files

- /home/daytona/app/results.json

## Related Pages

- [Codex MCP leaderboard](/leaderboard/codex/mcp)
- [Compare Codex MCP companies](/leaderboard/codex/mcp/compare)
- [Agent Landscape](/leaderboard/discoverability)

Canonical URL: https://devtoolarena.com/codex/mcp/prefect
