Skip to model directory
SilixarCloudIncSingapore
post@silixarcloudinc.com
SilixarCloudIncModel API platform
Corporate homeModel API catalog · Public pricing

45 models. One account. Clear usage pricing.

Browse proprietary commercial models from providers in North America and Europe, compare their current public API rates, and choose the right endpoint for each workload.

Browse the directory
Current catalog

Compare every currently listed model and its public rate.

Prices are shown in USD using the billing unit published by each model provider. Search by model or provider, or narrow the directory by workload type.

Catalog pricing snapshot · 2026-09-18
45 models
15

Vision

Visual understanding and multimodal reasoning

ModelType & contextPublic priceActions
Claude Fable 5.1Anthropic
Vision · Text Generation

Commercial multimodal understanding for text, images, and visual prompts.

Context 1M · Max output 128K
Input$10/ M Tokens
Cached input$0.25/ M Tokens
Output$50/ M Tokens
Model details
Claude Opus 5Anthropic
Vision · Text Generation

Commercial multimodal understanding for text, images, and visual prompts.

Context 1M · Max output 128K
Input$5/ M Tokens
Cached input$0.50/ M Tokens
Output$25/ M Tokens
Model details
Claude Sonnet 5Anthropic
Vision · Text Generation

Commercial multimodal understanding for text, images, and visual prompts.

Context 1M · Max output 128K
Input$2/ M Tokens
Cached input$0.20/ M Tokens
Output$10/ M Tokens
Model details
Claude Haiku 4.5Anthropic
Vision · Text Generation

Anthropic's fastest current Claude model for interactive, high-throughput, and cost-sensitive workloads.

Context 200K · Max output 64K
Input$1/ M Tokens
Cached input$0.10/ M Tokens
Output$5/ M Tokens
Model details
Amazon Nova 2 LiteAWS
Vision · Text Generation

Commercial multimodal understanding for text, images, and visual prompts.

Context 1M · Max output 64K
Input$0.30/ M Tokens
Cached input$0.075/ M Tokens
Output$2.50/ M Tokens
Model details
Gemini 3.8 FlashGoogle
Vision · Text Generation

Commercial multimodal understanding for text, images, and visual prompts.

Context 1.05M · Max output 64K
Input$0.75/ M Tokens
Cached input$0.075/ M Tokens
Output$3.75/ M Tokens
Model details
Gemini 3.5 Flash-LiteGoogle
Vision · Text Generation

Google's stable, low-latency multimodal model for high-volume automation, document parsing, and tool use.

Context 1.05M · Max output 64K
Input$0.30/ M Tokens
Cached input$0.03/ M Tokens
Output$2.50/ M Tokens
Model details
Gemini 3.1 Pro PreviewGoogle
Vision · Text Generation

Google's preview multimodal reasoning model for demanding software engineering and agent workflows.

Context 1.05M · Max output 64K
Input$2/ M Tokens
Cached input$0.20/ M Tokens
Output$12/ M Tokens
Model details
Muse Spark 1.3Meta
Vision · Text Generation

Meta's multimodal reasoning model for long-horizon agentic and coding work, with text, image, video, audio, and PDF input.

Context 1.05M · Max output 128K
Input$1.25/ M Tokens
Cached input$0.15/ M Tokens
Output$4.25/ M Tokens
Model details
GPT-6 AstraOpenAI
Vision · Text Generation

OpenAI's flagship reasoning model with text and image input, billed at double the input rate above 272K input tokens.

Context 1.05M · Max output 128K
Input$10/ M Tokens
Cached input$1/ M Tokens
Output$50/ M Tokens
Model details
GPT-5.6 SolOpenAI
Vision · Text Generation

Commercial multimodal understanding for text, images, and visual prompts.

Context 1.05M · Max output 128K
Input$4/ M Tokens
Cached input$0.40/ M Tokens
Output$20/ M Tokens
Model details
GPT-5.6 TerraOpenAI
Vision · Text Generation

Commercial multimodal understanding for text, images, and visual prompts.

Context 1.05M · Max output 128K
Input$2/ M Tokens
Cached input$0.20/ M Tokens
Output$12/ M Tokens
Model details
GPT-5.6 LunaOpenAI
Vision · Text Generation

The GPT-5.6 reasoning stack for cost-sensitive, high-volume text and image workloads.

Context 1.05M · Max output 128K
Input$0.20/ M Tokens
Cached input$0.02/ M Tokens
Output$1.20/ M Tokens
Model details
GPT-5.3-CodexOpenAI
Vision · Code Generation

OpenAI's model for long-running agentic software engineering in Codex-style environments.

Context 400K · Max output 128K
Input$1.75/ M Tokens
Cached input$0.175/ M Tokens
Output$14/ M Tokens
Model details
Grok 4.6xAI
Vision · Text Generation

Commercial multimodal understanding for text, images, and visual prompts.

Context 500K · Max output 450K
Input$2/ M Tokens
Cached input$0.50/ M Tokens
Output$6/ M Tokens
Model details
08

Image

Text-led image generation

ModelType & contextPublic priceActions
FLUX.2 [pro]Black Forest Labs
Text-to-Image

Commercial image generation and editing through a dedicated media endpoint.

Generation$0.03/ Megapixel
Model details
Nano Banana 2Google
Text-to-Image

Google's fast image generation and editing model, billed per output image at 1K and 2K resolutions.

Context 128K · Max output 32K
Generation$0.067/ Image (1K)
Model details
Nano Banana ProGoogle
Text-to-Image

Google's flagship image generation and editing model with 2K and 4K output, multi-image fusion, and web-search grounding.

Context 64K · Max output 32K
Generation$0.134/ Image (1K)
Model details
Luma UNI-1.1 MaxLuma AI
Text-to-Image

Commercial image generation and editing through a dedicated media endpoint.

Generation$0.1000/ Image (2K)
Model details
GPT-Image-2.5 FlareOpenAI
Text-to-Image

OpenAI's fast image generation and editing model for high-volume production work.

Generation$0.05268/ Image (1K)
Generation$0.05529/ Image (2K)
Generation$0.10008/ Image (4K)
Model details
GPT-Image-2.5 SunburstOpenAI
Text-to-Image

OpenAI's most capable image generation and editing model, tuned for detailed instructions and reference-guided edits.

Generation$0.05268/ Image (1K)
Generation$0.05529/ Image (2K)
Generation$0.10008/ Image (4K)
Model details
GPT-Image-2OpenAI
Text-to-Image

Commercial image generation and editing through a dedicated media endpoint.

Generation$0.053/ Image (1K)
Model details
Grok Imagine Image 2.0xAI
Text-to-Image

xAI's commercial image generation and editing model with 1K and 2K output and multi-image references.

Generation$0.04/ Image (1K)
Model details
06

Video

Text- and image-led video generation

ModelType & contextPublic priceActions
FLUX 3 VideoBlack Forest Labs
Text-to-Video

Commercial video generation from text or image inputs.

Generation$0.17/ Second (HD)
Model details
Gemini Omni FlashGoogle
Text-to-Video

Google's stable video model for fast generation and conversational editing from text, images, or short video inputs.

Context 1.05M
Generation$0.10/ Second (720p)
Model details
Veo 3.1Google
Text-to-Video

Commercial video generation from text or image inputs.

Generation$0.40/ Second (720p)
Model details
Luma Ray3.2Luma AI
Text-to-Video

Commercial video generation from text or image inputs.

Generation$0.30/ 5-second clip (720p)
Model details
Runway Gen-4.5Runway
Text-to-Video

Commercial video generation from text or image inputs.

Generation$0.12/ Second
Model details
Grok Imagine Video 1.5xAI
Text-to-Video

xAI's commercial model for text-to-video, image-to-video, and reference-guided generation up to 1080p.

Generation$0.080/ Second (720p)
Model details
11

Audio

Speech recognition, transcription, and synthesis

ModelType & contextPublic priceActions
Eleven v3ElevenLabs
Text-to-Speech

Commercial speech recognition, transcription, and text-to-speech generation.

Speech$0.10/ 1K Characters
Model details
Scribe v2ElevenLabs
Speech-to-Text

Commercial speech recognition, transcription, and text-to-speech generation.

Transcription$0.22/ Audio hour
Model details
Gemini 3.5 TranscribeGoogle
Speech-to-Text

Google's stable speech-to-text model for accurate, low-latency transcription across multilingual audio.

Transcription$0.005/ Audio minute
Model details
Gemini 3.1 Flash Live PreviewGoogle
Speech-to-Speech

Google's preview audio-to-audio model for low-latency, voice-first applications with multimodal input and tool use.

Context 128K · Max output 64K
Input$0.005/ Audio minute
Output$0.018/ Audio minute
Model details
Voxtral Mini Transcribe 2Mistral AI
Speech-to-Text

Commercial speech recognition, transcription, and text-to-speech generation.

Transcription$0.003/ Audio minute
Model details
GPT-TranscribeOpenAI
Speech-to-Text

OpenAI's high-accuracy speech-to-text model for completed files, streamed files, and committed Realtime turns.

Transcription$0.0045/ Audio minute
Model details
GPT-Realtime-2.1OpenAI
Speech-to-Speech

OpenAI's realtime reasoning model for interactive voice agents, image input, and live tool use.

Context 128K · Max output 32K
Input$32/ M Audio Tokens
Cached input$0.40/ M Audio Tokens
Output$64/ M Audio Tokens
Model details
GPT-Audio-1.5OpenAI
Speech-to-Speech

OpenAI's generally available audio model for Chat Completions workflows that accept and return audio.

Context 128K · Max output 16K
Input$32/ M Audio Tokens
Output$64/ M Audio Tokens
Model details
GPT-4o Mini TTSOpenAI
Text-to-Speech

Commercial speech recognition, transcription, and text-to-speech generation.

Context 2K
Input$0.60/ M Tokens
Output$12/ M Tokens
Model details
Grok VoicexAI
Speech-to-Speech

xAI's real-time speech-to-speech service for bidirectional voice agents with reasoning and tool use.

Speech$0.08/ Audio minute
Model details
Grok TTSxAI
Text-to-Speech

xAI's streamed and batch text-to-speech model with expressive multilingual voices and configurable delivery.

Speech$0.015/ 1K Characters
Model details
04

Embedding

Semantic representation for search and retrieval

ModelType & contextPublic priceActions
Cohere Embed 4Cohere
Embedding

Hosted vector embeddings for semantic search and retrieval.

Context 128K
Embedding$0.12/ M Tokens
Model details
Gemini Embedding 2Google
Embedding

Hosted vector embeddings for semantic search and retrieval.

Context 8K
Embedding$0.20/ M Tokens
Model details
Text Embedding 3 LargeOpenAI
Embedding

OpenAI's highest-capability text embedding model for English and multilingual retrieval.

Context 8K
Embedding$0.13/ M Tokens
Model details
Text Embedding 3 SmallOpenAI
Embedding

OpenAI's lower-cost third-generation embedding model for high-volume search and classification.

Context 8K
Embedding$0.02/ M Tokens
Model details
01

Reranker

Relevance scoring for retrieval pipelines

ModelType & contextPublic priceActions
Cohere Rerank 4 ProCohere
Reranker

Hosted relevance scoring for improving retrieved result ordering.

Context 32K
Rerank$2.50/ 1K Searches
Model details
Ready to build?

Fund your account and make the first API call.

Buy tokens, create an API key, and call the selected model by its model ID.