# Groq Claude Code CLI Evaluation

**Category:** Inference
**Score:** 34 (D)

This is the agent-readable Devtool Arena evaluation page for Groq on Claude Code CLI.

## Summary

| Metric | Value |
|--------|-------|
| Status | completed |
| Overall score | 34 |
| Grade | D |
| Eval score | 26 |
| Discovery score | 55 |
| Cost | $0.62 |
| Runtime | 6m 34s |
| Tool calls | 40 |
| Errors | 3 |
| Tokens used | 1596976 |
| Success | No |

## Agent-Readiness Checklist

| Signal | Evidence |
|--------|----------|
| Context7 | Not found |
| llms.txt | Not found |
| MCP server | Not found |
| Typed SDK | Not found |
| OpenAPI | Not found |
| Agent skills | Not found |
| CLI | Not found |

## Evaluation Prompt

````text

# Task
Using the groq CLI (Groq), extract structured data from this invoice text:

---
INVOICE #INV-2024-0042
From: Acme Corp
Date: 2024-03-15

Items:
- Consulting Services: $500.00
- Travel Expenses: $150.00

Total: $650.00
---

Use the CLI to make a chat completion or inference call that extracts JSON with these fields: invoice_number, vendor, date, line_items (array), total.

Print the JSON output to stdout.

You MUST use the `groq` CLI to accomplish this task.
Do NOT write Python scripts that import the SDK — use CLI commands via the terminal.

If you need help with CLI commands, check the documentation: https://github.com/build-with-groq/groq-code-cli
You can also use `groq --help` and `groq <command> --help` to discover available commands.

## Setup
You need to install and authenticate the `groq` CLI yourself.

### Installation
Run: `npm install -g groq-code-cli`

### Authentication
The following environment variables are already set in this environment: `GROQ_API_KEY`
The CLI should pick these up automatically, or pass them to commands as needed.

## Execution
1. Install the `groq` CLI using the command above
2. Verify the installation: `groq --help`
3. Authenticate using the credentials above
4. Perform the task described above

Work in the /home/daytona/app directory. Run commands and print output to stdout.
If there are errors, debug and fix them until the task runs successfully.

## Summary
When you are done, write a JSON summary file to /home/daytona/app/results.json with this structure:
{
  "operations": [
    {"step": 1, "description": "What you did", "command": "the CLI command", "success": true/false, "output_summary": "brief result"},
    ...
  ],
  "overall_success": true/false,
  "notes": ["any relevant notes about the execution"]
}

````

## Grader Results

| Check | Passed | Weight | Score | Details |
|-------|--------|--------|-------|---------|
| Stage 2 Auth | Yes | — | 25 | — |
| Stage 3 Task | No | — | — | — |
| Stage 1 Install | Yes | — | 5 | — |
| Stage 5 Llms Txt | No | — | 0 | — |
| Stage 0 Cli Exists | Yes | — | 5 | — |
| Stage 4 Json Output | No | — | 0 | — |
| Stage 6 Agent Skill | Yes | — | 20 | — |
| Stage 3 Non Interactive | No | — | 0 | — |

## Run Artifacts

| Artifact | Value |
|----------|-------|
| Conversation turns | 45 |
| Tool call traces | 40 |
| Generated files | 2 |
| Exit code | — |
| Completed at | 2026-09-27T14:08:24.640272+00:00 |

### Generated Files

- /home/daytona/app/invoice.txt
- /home/daytona/app/results.json

## Related Pages

- [Claude Code CLI leaderboard](/leaderboard/claudecode/cli)
- [Compare Claude Code CLI companies](/leaderboard/claudecode/cli/compare)
- [Agent Landscape](/leaderboard/discoverability)

Canonical URL: https://devtoolarena.com/claudecode/cli/groq
