Overview
Captions an image and translates the caption into any of 100+ languages in one call. Composite: one call runs describe-image + translate. A vision LLM writes a single-sentence caption, then it is translated into the target language. Returns the source caption, translated caption, and per-component telemetry. Use it for translated image captions, multilingual alt text, or vision caption plus translation.
Health
x402 Payment Validation
Recent Health Checks
| Time | Status | HTTP | Latency | Error |
|---|---|---|---|---|
| 2026-08-14 09:33:48 | healthy | 402 | 4537ms | |
| 2026-08-14 04:02:38 | healthy | 402 | 3010ms | |
| 2026-08-13 20:46:34 | healthy | 402 | 21ms | |
| 2026-08-13 13:45:45 | healthy | 402 | 276ms | |
| 2026-08-13 07:40:11 | healthy | 402 | 3820ms | |
| 2026-08-13 03:59:32 | healthy | 402 | 28ms | |
| 2026-08-12 21:07:52 | healthy | 402 | 80ms | |
| 2026-08-12 15:55:10 | healthy | 402 | 58ms | |
| 2026-08-12 09:54:00 | healthy | 402 | 33ms | |
| 2026-08-12 05:18:48 | healthy | 402 | 65ms | |
| 2026-08-12 01:38:48 | healthy | 402 | 37ms |