Skip to main content
NEAR AI Cloud is not limited to chat completions. The NEAR AI Cloud Gateway (cloud-api.near.ai) serves specialized TEE-hosted models for embeddings, reranking, scoring, image generation, audio transcription, and privacy classification. Gateway responses include an X-Request-Id response header. Support may ask you for that opaque value when debugging a request. It is support/debugging metadata, not W3C traceparent or distributed trace context. If you send your own X-Request-Id, use a non-sensitive UUID value; X-Request-Id values must not contain secrets or PII. X-Org-Id and X-Workspace-Id are internal tenant headers; public clients cannot set or override them.

Embeddings

Generate vector embeddings with Qwen/Qwen3-Embedding-0.6B:
The response follows the OpenAI embeddings format (data[0].embedding is the vector). input also accepts an array of strings for batch embedding.

Reranking

Score documents against a query with Qwen/Qwen3-Reranker-0.6B:
Results are ordered by relevance_score; index refers to the position in your input documents array.

Image Generation

Generate images with black-forest-labs/FLUX.2-klein-4B (billed per image — see the Models page):
The response returns the generated image as base64 in data[0].b64_json. An /v1/images/edits endpoint is also available for image-to-image editing.

Audio Transcription

Transcribe audio with openai/whisper-large-v3 (multipart form upload, OpenAI-compatible):

PII Detection & Redaction

The openai/privacy-filter model detects personally identifiable information (names, emails, phone numbers, …) in text. It is not a chat model — it exposes dedicated privacy endpoints: Classify returns labeled character spans with confidence scores:
Redact (gateway only) returns the text with PII replaced:
input accepts a string or an array of strings for both endpoints. /v1/privacy/redact is gateway-only and is served at https://cloud-api.near.ai/v1.

Vision (Image Input)

Qwen/Qwen3-VL-30B-A3B-Instruct accepts images in standard OpenAI chat format (image_url content parts) on /v1/chat/completions. Check input_modalities in /v1/models to see which models accept images.