cloud-api.near.ai) serves specialized TEE-hosted models for embeddings, reranking, scoring, image generation, audio transcription, and privacy classification.
Gateway responses include an X-Request-Id response header. Support may ask you for that opaque value when debugging a request. It is support/debugging metadata, not W3C traceparent or distributed trace context. If you send your own X-Request-Id, use a non-sensitive UUID value; X-Request-Id values must not contain secrets or PII. X-Org-Id and X-Workspace-Id are internal tenant headers; public clients cannot set or override them.
Embeddings
Generate vector embeddings withQwen/Qwen3-Embedding-0.6B:
data[0].embedding is the vector). input also accepts an array of strings for batch embedding.
Reranking
Score documents against a query withQwen/Qwen3-Reranker-0.6B:
relevance_score; index refers to the position in your input documents array.
Image Generation
Generate images withblack-forest-labs/FLUX.2-klein-4B (billed per image — see the Models page):
data[0].b64_json. An /v1/images/edits endpoint is also available for image-to-image editing.
Audio Transcription
Transcribe audio withopenai/whisper-large-v3 (multipart form upload, OpenAI-compatible):
PII Detection & Redaction
Theopenai/privacy-filter model detects personally identifiable information (names, emails, phone numbers, …) in text. It is not a chat model — it exposes dedicated privacy endpoints:
Classify returns labeled character spans with confidence scores:
input accepts a string or an array of strings for both endpoints.
/v1/privacy/redact is gateway-only and is served at https://cloud-api.near.ai/v1.
Vision (Image Input)
Qwen/Qwen3-VL-30B-A3B-Instruct accepts images in standard OpenAI chat format (image_url content parts) on /v1/chat/completions. Check input_modalities in /v1/models to see which models accept images.