Overview
Captions an image and translates the caption into any of 100+ languages in one call. Composite: one call runs describe-image + translate. A vision LLM writes a single-sentence caption, then it is translated into the target language. Returns the source caption, translated caption, and per-component telemetry. Use it for translated image captions, multilingual alt text, or vision caption plus translation.
Health
x402 Payment Validation
Recent Health Checks
| Time | Status | HTTP | Latency | Error |
|---|---|---|---|---|
| 2026-10-06 07:15:38 | timeout | — | — | timeout |
| 2026-10-05 18:49:43 | healthy | 402 | 3945ms | |
| 2026-10-05 04:42:58 | timeout | — | — | timeout |
| 2026-10-04 20:49:59 | healthy | 402 | 3676ms | |
| 2026-10-04 10:12:24 | timeout | — | — | timeout |
| 2026-10-04 03:14:04 | healthy | 402 | 2218ms | |
| 2026-10-03 19:30:40 | healthy | 402 | 84ms | |
| 2026-10-03 10:56:40 | healthy | 402 | 103ms |