> ## Documentation Index
> Fetch the complete documentation index at: https://docs.veadk.xyz/llms.txt
> Use this file to discover all available pages before exploring further.

# View experiments and results

The `eval experiment` command group (alias `exp`) is used to view experiments submitted by [`eval run`](/productions/agentkit-cli/archives/0.50.1/en/commands/eval/run): list experiments, view an experiment's status and aggregate scores, and fetch per-case results.

<Note>
  All subcommands share `-p, --project <name>` (project name, default `default`) and `--json` (output raw JSON).
</Note>

## experiment list

List the experiments in a project: id, name, status, start time, and aggregate score.

```bash lines theme={null}
agentkit eval experiment list
```

## experiment get

Show details of an experiment: status, dataset, target, weighted overall score, and each evaluator's average score and score distribution.

| Flag / Argument | Description | Default |
| - | - | - |
| `<id>` | Experiment id (required) | None |

```bash lines theme={null}
agentkit eval experiment get 75901xxxxxxxxxxxxx
```

```text lines theme={null}
ID              75901xxxxxxxxxxxxx
Name            qa-set-2026-07-03-12-16-21
Status          Success
Dataset         qa-set (75900xxxxxxxxxxxxx)
Target          75901xxxxxxxxxxxxx
Weighted score  enabled
Overall score   1.00

Evaluators:
  relevance @0.0.1  avg 1.00
    1.00: 2 (100%)
```

## experiment results

Fetch per-case results: each dataset case, the target output, and each evaluator's score and reasoning.

| Flag / Argument | Description | Default |
| - | - | - |
| `<id>` | Experiment id (required) | None |
| `--limit <n>` | Items per page (max 20) | `20` |
| `--page <n>` | Page number | `1` |

```bash lines theme={null}
agentkit eval experiment results 75901xxxxxxxxxxxxx --json
```

## Experiment status

| Status | Meaning |
| - | - |
| `Success` | All cases finished executing |
| `Failed` | Some cases failed to execute (e.g. the target did not respond) |
| `Draining` / `Processing` | Still running |

<Tip>
  When a case fails but there is no experiment-level error message, it is usually because the **target runtime did not respond correctly** (not deployed, not ready, or authentication failed), rather than an evaluation configuration issue — first confirm that the runtime pointed to by `--target` is running.
</Tip>
