Overview
Eval case generator for an agent or LLM task: describe the task and get back a draft llm eval dataset of {input, expected, rubric} cases spanning typical, edge, and adversarial scenarios. Pass example cases to steer style, and set n for how many to generate (default 8, max 20). Every case is labeled a candidate needing human review, not validated ground truth: this drafts a first pass at agent test cases, it doesn't certify them. Use it as an agent regression tests starter, an eval dataset gener
Health
x402 Payment Validation
Recent Health Checks
| Time | Status | HTTP | Latency | Error |
|---|---|---|---|---|
| 2026-09-23 04:15:36 | healthy | 402 | 2449ms | |
| 2026-09-23 02:33:04 | healthy | 402 | 127ms | |
| 2026-09-22 16:01:47 | healthy | 402 | 495ms | |
| 2026-09-22 11:10:04 | timeout | — | — | timeout |
| 2026-09-21 22:33:20 | healthy | 402 | 60ms | |
| 2026-09-21 16:01:27 | healthy | 402 | 629ms | |
| 2026-09-21 08:04:47 | healthy | 402 | 83ms | |
| 2026-09-20 19:15:20 | healthy | 402 | 6789ms | |
| 2026-09-20 09:38:57 | timeout | — | — | timeout |