$ man image-description
/image-description
PRICE / CALL
$0.02
USDC · base mainnet · scheme: exact
METHOD
POST
CLUSTER
wordmintCATEGORY
uncategorized
STATUS
● live
NAME
image-description — takes a public image url and returns an ai vision description, alt text, ocr text, tags, or caption depending on mode
synonym alias of describe-image — reuses the canonical handler.
SYNOPSIS
POST https://x402.agentutility.ai/image-description
Content-Type: application/json
X-PAYMENT: <signed-transferWithAuthorization>
{ ... }↳ first call →
402 Payment Required. Sign USDCtransferWithAuthorization, retry with theX-PAYMENT header.DESCRIPTION
Takes a public image URL and returns an AI vision description, alt text, OCR text, tags, or caption depending on mode. Use it as an image description API, AI image captioner, or image-to-text endpoint.
INPUT — request schema
| property | type | description | req? |
|---|---|---|---|
| image_url | string | URL of the image to analyze. Must be http or https. | required |
| mode | string | Analysis mode: 'describe' (default), 'alt_text', 'ocr', 'tags', or 'caption'. enum: describe · alt_text · ocr · tags · caption | optional |
| prompt | string | Optional. Custom instruction that overrides the mode's default prompt. | optional |
OUTPUT — response shape
| field | type | description |
|---|---|---|
| text | string | — |
| mode | string | — |
EXAMPLES — two ways to call
EXAMPLE 1 · curl
curl -X POST https://x402.agentutility.ai/image-description \
-H 'Content-Type: application/json' \
-d '{ }'first response =
402 Payment Required with payment requirements; sign + retry with X-PAYMENT.EXAMPLE 2 · mcp
# Install the MCP package for this endpoint's cluster npx -y @agentutility/mcp-<cluster> # Required: EVM private key with USDC on Base export X402_PRIVATE_KEY=0x... # Then call the image-description tool from your MCP-aware agent.
MCP server handles payment automatically — your coding agent just calls the tool by name.
METADATA
- tags
- wordmintimagedescriptionimage-description
- methods
- POST
- cluster
- wordmint
- price
- $0.02 USDC per call
ADJACENT — other endpoints in wordmint
| endpoint | description | price |
|---|---|---|
| alt-text-generator | Turns a public image URL into ready-to-use text: concise alt text for screen readers, a natural-language description, OCR text pulled fro… | $0.02 |
| classify | Sort text into categories you define on the spot, no training run required. | $0.02 |
| classify-text | Sorts a piece of text into categories you define on the fly, no training or fixed label set required. | $0.02 |
| describe-image | Describes images with a vision LLM across five modes: describe, alt_text (accessibility, <=125 chars), OCR (extract visible text), tags (… | $0.02 |
| detect-pii | Detects PII in text: emails, phones, SSNs, credit cards, addresses, names, IPs, and API tokens. | $0.02 |
| email-draft | Writes emails with AI: subject, body, salutation, and sign-off. | $0.02 |
| extract | Pull structured entities out of raw text instead of hand-parsing it yourself. | $0.02 |
| image-describe | Get a vision model's read on an image: a description, alt text, extracted text, tags, or a caption. | $0.02 |
SEE ALSO