Overview
Eval case generator for an agent or LLM task: describe the task and get back a draft llm eval dataset of {input, expected, rubric} cases spanning typical, edge, and adversarial scenarios. Pass example cases to steer style, and set n for how many to generate (default 8, max 20). Every case is labeled a candidate needing human review, not validated ground truth: this drafts a first pass at agent test cases, it doesn't certify them. Use it as an agent regression tests starter, an eval dataset gener
Health
x402 Payment Validation
Recent Health Checks
| Time | Status | HTTP | Latency | Error |
|---|---|---|---|---|
| 2026-08-07 13:44:37 | healthy | 402 | 4168ms | |
| 2026-08-07 08:46:44 | healthy | 402 | 32ms | |
| 2026-08-07 04:34:34 | healthy | 402 | 3184ms | |
| 2026-08-06 22:19:41 | healthy | 402 | 70ms | |
| 2026-08-06 17:58:51 | healthy | 402 | 4396ms | |
| 2026-08-06 12:38:10 | healthy | 402 | 14ms | |
| 2026-08-06 10:11:56 | healthy | 402 | 36ms | |
| 2026-08-06 05:01:05 | healthy | 402 | 67ms | |
| 2026-08-05 22:08:00 | healthy | 402 | 23ms |