Overview
Get a vision model's read on an image: a description, alt text, extracted text, tags, or a caption. Send an image_url and a mode (describe, alt_text, ocr, tags, or caption), or override with a custom prompt, and it returns the generated text along with the mode used. OCR mode pulls out any visible text verbatim. Use it as an image captioning API, alt-text generator, image OCR tool, or vision-based tagging endpoint.
Health
x402 Payment Validation
Recent Health Checks
| Time | Status | HTTP | Latency | Error |
|---|---|---|---|---|
| 2026-09-05 08:50:50 | healthy | 402 | 47ms | |
| 2026-09-05 03:29:55 | healthy | 402 | 528ms | |
| 2026-09-04 21:47:20 | healthy | 402 | 12ms | |
| 2026-09-04 14:36:38 | healthy | 402 | 6113ms | |
| 2026-09-04 04:27:15 | healthy | 402 | 3402ms | |
| 2026-09-03 18:23:35 | healthy | 402 | 87ms | |
| 2026-09-03 12:20:49 | healthy | 402 | 63ms | |
| 2026-09-03 00:16:46 | degraded | 402 | 5630ms | |
| 2026-09-02 17:51:50 | healthy | 402 | 5865ms | |
| 2026-09-02 13:52:05 | healthy | 402 | 4951ms |