# Extend.ai Codex MCP Evaluation

**Category:** Document Parsing
**Score:** 81 (B)

This is the agent-readable Devtool Arena evaluation page for Extend.ai on Codex MCP.

## Summary

| Metric | Value |
|--------|-------|
| Status | completed |
| Overall score | 81 |
| Grade | B |
| Eval score | 78 |
| Discovery score | 89 |
| Cost | $0.18 |
| Runtime | 4m 32s |
| Tool calls | 4 |
| Errors | 2 |
| Tokens used | 505300 |
| Success | Yes |

## Agent-Readiness Checklist

| Signal | Evidence |
|--------|----------|
| Context7 | Not found |
| llms.txt | Not found |
| MCP server | Not found |
| Typed SDK | Not found |
| OpenAPI | Not found |
| Agent skills | Not found |
| CLI | Not found |

## Evaluation Prompt

````text

# Task
Using the Extend.ai MCP, parse a document:

1. Discover available MCP tools for document parsing
2. Download this sample PDF: https://www.irs.gov/pub/irs-pdf/fw9.pdf
3. Use the MCP tools to parse the PDF into structured output (markdown, JSON, or text)
4. Print the parsed output showing extracted text content, tables, or structure

Use ONLY the MCP tools for document parsing operations. Do NOT use SDK or raw API calls.

You MUST use the Extend.ai MCP server to accomplish this task.
Use the MCP tools for the actual product operations. Do NOT use curl, fetch, raw HTTP requests, CLI commands, or SDK calls as substitutes.

Use the MCP tools' help/descriptions to understand what each tool does.
If you need additional context, check the documentation: https://github.com/extend-hq/extend-mcp

## Setup
Do NOT assume the Extend.ai MCP server is already installed or already running in this sandbox.
Review the installation and authentication details below before starting the task.

### Installation
Transport: `stdio`
Codex is configured to launch the `extend-ai` MCP server over stdio using: `/usr/local/share/nvm/current/bin/npx -y 'github:extend-hq/extend-mcp#release'`
Stored MCP package hint: `github:extend-hq/extend-mcp#release`
No separate setup command is required before startup unless the docs indicate one.
Bootstrap notes: Verified the launch command in the vendor release-branch README and confirmed that it starts a stdio server with a dummy key. Provide a valid EXTEND_API_KEY in the server process environment. Startup was verified; API access and document parsing were not *****ed. npx installs the GitHub package as needed, so no separate setup command is required.

### Authentication
The following environment variables are already set in this environment: `EXTEND_API_KEY`
The MCP server should read them automatically when it starts.

## Execution
1. The harness has already written the launch configuration for `extend-ai`. Start with the task directly; do not rewrite agent settings, `.mcp-runtime.json`, or build a separate MCP client unless the connected server still fails after a normal retry.
2. You do not need to enumerate every MCP tool up front; inspect only the relevant tools you need. Agent-facing MCP tool names may be namespaced like `mcp__extend-ai__<tool_name>`.
3. Complete the task described above using MCP tools for the actual product operations.
4. If the server fails to start or authenticate, debug the documented setup and retry. Do not fall back to raw HTTP, SDK, CLI, or manual JSON-RPC/MCP calls as a substitute for exposed MCP tools.

Work in the /home/daytona/app directory.
If you use Bash during setup or debugging, limit it to installation/configuration work and then return to MCP tools for the task itself.
If there are errors, debug and fix them until the task runs successfully.

## Summary
When you are done, write a JSON summary file to /home/daytona/app/results.json with this structure:
{
  "operations": [
    {"step": 1, "description": "What you did", "mcp_tool": "tool_name", "success": true/false, "output_summary": "brief result"},
    ...
  ],
  "overall_success": true/false,
  "notes": ["any relevant notes about the execution"]
}

````

## Grader Results

| Check | Passed | Weight | Score | Details |
|-------|--------|--------|-------|---------|
| Task | Yes | — | — | — |
| Llms Txt | Yes | — | 10 | — |
| Tool Count | Yes | — | 10 | — |
| Auth Method | Yes | — | 15 | — |
| Official Mcp | Yes | — | 5 | — |
| Output Schema | No | — | 0 | — |
| Error Messages | — | — | 0 | get_classifier: Poor error message (generic or stack trace); classify_document: Poor error message (generic or stack trace) |
| Tool Stability | Yes | — | 5 | — |
| Security Posture | — | — | 9 | — |
| Tool Api Mapping | — | — | 14 | — |
| Description Clarity | — | — | 14 | — |
| Tool Description Length | Yes | — | 7 | — |

## Run Artifacts

| Artifact | Value |
|----------|-------|
| Conversation turns | 9 |
| Tool call traces | 4 |
| Generated files | 1 |
| Exit code | — |
| Completed at | 2026-10-04T16:26:28.030124+00:00 |

### Generated Files

- /home/daytona/app/results.json

## Related Pages

- [Codex MCP leaderboard](/leaderboard/codex/mcp)
- [Compare Codex MCP companies](/leaderboard/codex/mcp/compare)
- [Agent Landscape](/leaderboard/discoverability)

Canonical URL: https://devtoolarena.com/codex/mcp/extend-ai
