EllamiModels & pricing

ELLAMI / CATALOG

Models & pricing

Choose a model for text, images or video. Compare capabilities and pricing.

Model catalog

Model developer rates · USD · 2026-10-07

Select a model type to compare prices.

Models: 64

OpenAI

GPT 6.1 Sol

Code, documents and complex tasks in a single workflow.

TextContext: 1,050,000
USD / 1M tokens
Input
$2
Output
$10
Model details

Model ID: gpt-6.1-sol

Long prompts above 272,000 input tokens: input $4 · output $15 / 1M tokens.

Above 272,000 input tokens, the entire request is charged at 2× input and 1.5× output. Standard prices, excluding Batch, Flex and Fast.

Developer rate
OpenAI

GPT 6 Sol

Development, codebase work and autonomous agents.

TextContext: 1,050,000
USD / 1M tokens
Input
$2
Output
$10
Model details

Model ID: gpt-6-sol

Long prompts above 272,000 input tokens: input $4 · output $15 / 1M tokens.

Above 272,000 input tokens, the entire request is charged at 2× input and 1.5× output. Standard prices, excluding Batch, Flex and Fast.

Developer rate
OpenAI

GPT 6 Astra

Deep analysis, research and demanding engineering tasks.

TextContext: 1,050,000
USD / 1M tokens
Input
$10
Output
$50
Model details

Model ID: gpt-6-astra

Long prompts above 272,000 input tokens: input $20 · output $75 / 1M tokens.

Above 272,000 input tokens, the entire request is charged at 2× input and 1.5× output. Standard prices, excluding Batch, Flex and Fast.

Developer rate
OpenAI

GPT Image 2.5 Sunburst

Image generation and editing from prompts and references.

Images
USD / image
Image output
from $0.00588

1024×1024 · low · Input billed separately

Model details

Model ID: gpt-image-2.5-sunburst

USD / image
1024×1024 · low
$0.00588
1024×1024 · medium
$0.01317
1024×1024 · high
$0.05268
1024×1024 · xhigh
$0.09366
1024×1024 · max
$0.21072
USD / 1M tokens
Text input
$5
Image input
$8
Image output
$30

The card shows the minimum image output cost among the listed sizes and qualities. In the token table, input refers to text and output to image tokens. Input images are priced separately. Total cost depends on size, quality and token usage, rather than a fixed price per image.

Developer rate
OpenAI

GPT Image 2.5 Flare

Visual generation, variations and targeted image edits.

Images
USD / image
Image output
from $0.00588

1024×1024 · low · Input billed separately

Model details

Model ID: gpt-image-2.5-flare

USD / image
1024×1024 · low
$0.00588
1024×1024 · medium
$0.01317
1024×1024 · high
$0.05268
1024×1024 · xhigh
$0.09366
1024×1024 · max
$0.21072
USD / 1M tokens
Text input
$5
Image input
$8
Image output
$30

The card shows the minimum image output cost among the listed sizes and qualities. In the token table, input refers to text and output to image tokens. Input images are priced separately. Total cost depends on size, quality and token usage, rather than a fixed price per image.

Developer rate
OpenAI

GPT Image 2

A versatile model for image generation and editing.

Images
USD / image
Image output
from $0.00588

1024×1024 · low · Input billed separately

Model details

Model ID: gpt-image-2

USD / image
1024×1024 · low
$0.00588
1024×1024 · medium
$0.05268
1024×1024 · high
$0.21072
USD / 1M tokens
Text input
$5
Image input
$8
Image output
$30

The card shows the minimum image output cost among the listed sizes and qualities. In the token table, input refers to text and output to image tokens. Input images are priced separately. Total cost depends on size, quality and token usage, rather than a fixed price per image.

Developer rate
Anthropic

Claude Sonnet 5.5

Feature development, bug fixing and document work.

TextContext: 1,000,000
USD / 1M tokens
Input
$2
Output
$10
Model details

Model ID: claude-sonnet-5-5

Developer rate
Anthropic

Claude Opus 5.5

Complex code, reviews and long-running agentic tasks.

TextContext: 1,000,000
USD / 1M tokens
Input
$4
Output
$20
Model details

Model ID: claude-opus-5-5

Developer rate
Anthropic

Claude Fable 5.1

Demanding reasoning and long-horizon autonomous work.

TextContext: 1,000,000
USD / 1M tokens
Input
$10
Output
$50
Model details

Model ID: claude-fable-5.1

Developer rate
Anthropic

Claude Opus 5

Complex reasoning, architecture and agentic development.

TextContext: 1,000,000
USD / 1M tokens
Input
$5
Output
$25
Model details

Model ID: claude-opus-5

Developer rate
Anthropic

Claude Sonnet 5

Code, analysis and everyday professional work.

TextContext: 1,000,000
USD / 1M tokens
Input
$2
Output
$10
Model details

Model ID: claude-sonnet-5

Developer rate
Anthropic

Claude Sonnet 4.6

An established model for programming, writing and analysis.

TextContext: 1,000,000
USD / 1M tokens
Input
$3
Output
$15
Model details

Model ID: claude-sonnet-4-6

Developer rate
Google

Gemini 3.8 Flash

Multimodal analysis, code and multi-step tasks.

TextContext: 1,048,576
USD / 1M tokens
Input
$1.5
Output
$7.5
Model details

Model ID: gemini-3.8-flash

Regular rates: $1.5 input / $7.5 output. A promotion through December 31, 2026 reduces these to $0.75 / $3.75. The catalog shows regular rates. Reasoning tokens count as output; search is billed separately.

Developer rate
Google

Gemini 3.7 Flash

Fast agents, long documents and multimedia analysis.

TextContext: 1,048,576
USD / 1M tokens
Input
$1.5
Output
$7.5
Model details

Model ID: gemini-3.7-flash

Regular rates: $1.5 input / $7.5 output. A promotion through December 31, 2026 reduces these to $0.75 / $3.75. The catalog shows regular rates. Reasoning tokens count as output; search is billed separately.

Developer rate
Google

Nano Banana 2

Image generation and editing with output up to 4K.

Images
USD / image
Image output
from $0.045

0.5K · Input billed separately

Model details

Model ID: nano-banana-2

USD / image
0.5K
$0.045
1K · 1024×1024
$0.067
2K
$0.101
4K
$0.151
USD / 1M tokens
Text input
$0.5
Image input
$0.5
Image output
$60
Text output
$3

CLODEX uses nano-banana-2; the original Google API ID is gemini-3.1-flash-image. The card shows output cost for one 1K image. Estimated output costs: $0.045 (0.5K), $0.067 (1K), $0.101 (2K), $0.151 (4K); input is charged separately.

Developer rate
DeepSeek

DeepSeek V4.1 Flash

Code, reasoning and visual analysis with a large context.

TextContext: 1,048,576
USD / 1M tokens
Input
$0.3
Output
$1.2
Model details

Model ID: deepseek-v4.1-flash

Off-peak: input $0.15 · output $0.6 / 1M tokens.

The catalog shows peak rates. Off-peak prices are 50% lower. Peak hours: Monday–Friday, 01:00–04:00 and 06:00–10:00 UTC.

Developer rate
DeepSeek

DeepSeek V4 Pro

A large model for complex reasoning and development.

TextContext: 1,048,576
USD / 1M tokens
Input
$1.32
Output
$3.96
Model details

Model ID: deepseek-v4-pro

Off-peak: input $0.66 · output $1.98 / 1M tokens.

The catalog shows peak rates. Off-peak prices are 50% lower. Peak hours: Monday–Friday, 01:00–04:00 and 06:00–10:00 UTC.

Developer rate
xAI

Grok 4.7

Development, knowledge work and agents with visual input.

TextContext: 500,000
USD / 1M tokens
Input
$2
Output
$6
Model details

Model ID: grok-4.7

Long prompts above 200,000 input tokens: input $4 · output $12 / 1M tokens.

Base rates for up to 200,000 input tokens. Above 200,000: $4 input / $12 output per 1M. Built-in tools are billed separately.

Developer rate
Moonshot AI

Kimi K3

Multimodal reasoning and code with up to 1M context.

TextContext: 1,048,576
USD / 1M tokens
Input
$3
Output
$15
Model details

Model ID: kimi-k3

Developer rate
Z.ai

GLM 5.3

Complex development and long agentic workflows.

TextContext: 1,048,576
USD / 1M tokens
Input
$1.4
Output
$4.4
Model details

Model ID: glm-5.3

Developer rate
Alibaba Cloud

Qwen 3.8 Flash

Code, visual understanding and long-video analysis.

TextContext: 1,000,000
USD / 1M tokens
Input
$0.15
Output
$0.47
Model details

Model ID: qwen3.8-flash

Regular Singapore / International rates: $0.15 / $0.47. Global regions use different rates: $0.113 / $0.382 per 1M tokens.

Developer rate
MiniMax

MiniMax M3

Multimodal development and agentic tasks.

TextContext: 1,000,000
USD / 1M tokens
Input
$0.3
Output
$1.2
Model details

Model ID: MiniMax-M3

Long prompts above 512,000 input tokens: input $0.6 · output $2.4 / 1M tokens.

Current Standard rates after the permanent 50% price reduction. For M3 above 512,000 input tokens: $0.6 input / $2.4 output. Priority is priced separately.

Developer rate
MiniMax

MiniMax M2.7

Programming and workflow automation.

TextContext: 204,800
USD / 1M tokens
Input
$0.3
Output
$1.2
Model details

Model ID: MiniMax-M2.7

Developer rate
OpenAI

GPT 5.6 Sol

Professional development, deep analysis and agentic work.

TextContext: 1,050,000
USD / 1M tokens
Input
$4
Output
$20
Model details

Model ID: gpt-5.6-sol

Long prompts above 272,000 input tokens: input $8 · output $30 / 1M tokens.

Published GPT 5.6 Sol rate: $4 / $20 per 1M; OpenAI identifies it as promotional pricing available at least through November 21, 2026. A later regular rate is not published. Above 272,000 input tokens: $8 / $30 per 1M for the entire request.

Developer rate
OpenAI

GPT 5.6 Terra

Code and reasoning with a balance of quality and cost.

TextContext: 1,050,000
USD / 1M tokens
Input
$2
Output
$12
Model details

Model ID: gpt-5.6-terra

Long prompts above 272,000 input tokens: input $4 · output $18 / 1M tokens.

Above 272,000 input tokens, the entire request is charged at 2× input and 1.5× output. Standard prices, excluding Batch, Flex and Fast.

Developer rate
Anthropic

Claude Opus 4.8

Complex development and long-running tool-based tasks.

TextContext: 1,000,000
USD / 1M tokens
Input
$5
Output
$25
Model details

Model ID: claude-opus-4-8

Developer rate
Anthropic

Claude Opus 4.7

Agentic development, architecture and visual tasks.

TextContext: 1,000,000
USD / 1M tokens
Input
$5
Output
$25
Model details

Model ID: claude-opus-4-7

Developer rate
Anthropic

Claude Fable 5

Deep reasoning, research and autonomous tasks.

TextContext: 1,000,000
USD / 1M tokens
Input
$10
Output
$50
Model details

Model ID: claude-fable-5

Developer rate
Z.ai

GLM 5.3 Flash

Fast multimodal model for code and visual analysis.

TextContext: 1,048,576
USD / 1M tokens
Input
$0.15
Output
$0.5
Model details

Model ID: glm-5.3-flash

Developer rate
Z.ai

GLM 5.2

Complex reasoning, programming and agent workflows.

TextContext: 1,000,000
USD / 1M tokens
Input
$1.4
Output
$4.4
Model details

Model ID: glm-5.2

Developer rate
Z.ai

GLM 5.1

Software development and multi-step tool use.

TextContext: 202,752
USD / 1M tokens
Input
$1.4
Output
$4.4
Model details

Model ID: glm-5.1

Developer rate
xAI

Grok 4.6

Multimodal reasoning, code and agentic workflows.

TextContext: 500,000
USD / 1M tokens
Input
$2
Output
$6
Model details

Model ID: grok-4.6

Long prompts above 200,000 input tokens: input $4 · output $12 / 1M tokens.

Base rates for up to 200,000 input tokens. Above 200,000: $4 input / $12 output per 1M. Built-in tools are billed separately.

Developer rate
MiniMax

MiniMax M2.5

App development and tool-based task automation.

TextContext: 196,608
USD / 1M tokens
Input
$0.3
Output
$1.2
Model details

Model ID: MiniMax-M2.5

Developer rate
MiniMax

MiniMax M2.7 Highspeed

Faster M2.7 variant for development and work agents.

TextContext: 204,800
USD / 1M tokens
Input
$0.6
Output
$2.4
Model details

Model ID: MiniMax-M2.7-highspeed

Developer rate
MiniMax

MiniMax M2.5 Highspeed

Faster code generation and tool use.

TextContext: 196,608
USD / 1M tokens
Input
$0.6
Output
$2.4
Model details

Model ID: MiniMax-M2.5-highspeed

Developer rate
OpenAI

GPT 6 Luna

A lightweight model for code, fast responses and high-volume tasks.

TextContext: 1,050,000
USD / 1M tokens
Input
$0.1
Output
$0.5
Model details

Model ID: gpt-6-luna

Long prompts above 272,000 input tokens: input $0.2 · output $0.75 / 1M tokens.

Above 272,000 input tokens, the entire request is charged at 2× input and 1.5× output. Standard prices, excluding Batch, Flex and Fast.

Developer rate
OpenAI

GPT 5.6 Luna

A fast, economical variant for everyday work.

TextContext: 1,050,000
USD / 1M tokens
Input
$0.2
Output
$1.2
Model details

Model ID: gpt-5.6-luna

Long prompts above 272,000 input tokens: input $0.4 · output $1.8 / 1M tokens.

Above 272,000 input tokens, the entire request is charged at 2× input and 1.5× output. Standard prices, excluding Batch, Flex and Fast.

Developer rate
OpenAI

GPT 5.5

Complex reasoning, development and tool use.

TextContext: 1,050,000
USD / 1M tokens
Input
$5
Output
$30
Model details

Model ID: gpt-5.5

Long prompts above 272,000 input tokens: input $10 · output $45 / 1M tokens.

Above 272,000 input tokens, the entire request is charged at 2× input and 1.5× output. Standard prices, excluding Batch, Flex and Fast.

Developer rate
OpenAI

GPT 5.4

Code, analysis and professional tasks with a large context.

TextContext: 1,050,000
USD / 1M tokens
Input
$2.5
Output
$15
Model details

Model ID: gpt-5.4

Long prompts above 272,000 input tokens: input $5 · output $22.5 / 1M tokens.

Above 272,000 input tokens, the entire request is charged at 2× input and 1.5× output. Standard prices, excluding Batch, Flex and Fast.

Developer rate
OpenAI

GPT 5.4 Mini

A compact model for programming and work agents.

TextContext: 400,000
USD / 1M tokens
Input
$0.75
Output
$4.5
Model details

Model ID: gpt-5.4-mini

Developer rate
OpenAI

GPT 5.4 Nano

Fast short tasks, classification and data extraction.

TextContext: 400,000
USD / 1M tokens
Input
$0.2
Output
$1.25
Model details

Model ID: gpt-5.4-nano

OpenAI has announced API retirement on 2027-04-01.

Developer rate
OpenAI

GPT 5.3 Codex

Agentic development, bug fixing and repository work.

TextContext: 400,000
USD / 1M tokens
Input
$1.75
Output
$14
Model details

Model ID: gpt-5.3-codex

OpenAI has announced API retirement on 2027-04-01.

Developer rate
OpenAI

GPT 5 Mini

An economical model for code and well-defined tasks.

TextContext: 400,000
USD / 1M tokens
Input
$0.25
Output
$2
Model details

Model ID: gpt-5-mini

Developer rate
OpenAI

GPT 5 Nano

The lightest GPT 5 variant for brief responses and data processing.

TextContext: 400,000
USD / 1M tokens
Input
$0.05
Output
$0.4
Model details

Model ID: gpt-5-nano

Developer rate
OpenAI

GPT 4.1 Mini

A compact model for code, instructions and long documents.

TextContext: 1,047,576
USD / 1M tokens
Input
$0.4
Output
$1.6
Model details

Model ID: gpt-4.1-mini

Developer rate
OpenAI

GPT 4o Mini

Fast conversations, text tasks and image analysis.

TextContext: 128,000
USD / 1M tokens
Input
$0.15
Output
$0.6
Model details

Model ID: gpt-4o-mini

Developer rate
Anthropic

Claude Opus 4.6

Complex code, deep analysis and long-running agentic work.

TextContext: 1,000,000
USD / 1M tokens
Input
$5
Output
$25
Model details

Model ID: claude-opus-4-6

Developer rate
Anthropic

Claude Opus 4.5

Architecture, programming and tool-based tasks.

TextContext: 200,000
USD / 1M tokens
Input
$5
Output
$25
Model details

Model ID: claude-opus-4-5

Developer rate
Anthropic

Claude Sonnet 4.5

App development, analysis and document work.

TextContext: 200,000
USD / 1M tokens
Input
$3
Output
$15
Model details

Model ID: claude-sonnet-4-5

Developer rate
Anthropic

Claude Haiku 4.5

Fast responses, compact agents and everyday code.

TextContext: 200,000
USD / 1M tokens
Input
$1
Output
$5
Model details

Model ID: claude-haiku-4-5

Developer rate
Google

Gemini 3.6 Flash

A fast multimodal model for code and analysis.

TextContext: 1,048,576
USD / 1M tokens
Input
$1.5
Output
$7.5
Model details

Model ID: gemini-3.6-flash

Regular rates: $1.5 input / $7.5 output. A promotion through December 31, 2026 reduces these to $0.75 / $3.75. The catalog shows regular rates. Reasoning tokens count as output; search is billed separately.

Developer rate
Google

Gemini 3.5 Flash-Lite

A lightweight multimodal model for fast, high-volume tasks.

TextContext: 1,048,576
USD / 1M tokens
Input
$0.3
Output
$2.5
Model details

Model ID: gemini-3.5-flash-lite

Developer rate
Google

Gemini 3.1 Flash-Lite

Economical responses, data extraction and multimedia processing.

TextContext: 1,048,576
USD / 1M tokens
Input
$0.25
Output
$1.5
Model details

Model ID: gemini-3.1-flash-lite

The $0.25 input rate applies to text, images and video; audio input costs $0.5 per 1M tokens. Reasoning tokens count as output.

Developer rate
Google

Gemini 3.1 Pro Preview

Complex reasoning, long documents and multimodal analysis.

TextContext: 1,048,576
USD / 1M tokens
Input
$2
Output
$12
Model details

Model ID: gemini-3.1-pro-preview

Long prompts above 200,000 input tokens: input $4 · output $18 / 1M tokens.

Base rates for up to 200,000 input tokens. Above 200,000: $4 input / $18 output per 1M tokens. Reasoning tokens count as output; search is billed separately.

Developer rate
xAI

Grok 4.5

Programming, knowledge work and agentic tasks.

TextContext: 500,000
USD / 1M tokens
Input
$2
Output
$6
Model details

Model ID: grok-4.5

Long prompts above 200,000 input tokens: input $4 · output $12 / 1M tokens.

Base rates for up to 200,000 input tokens. Above 200,000: $4 input / $12 output per 1M. Built-in tools are billed separately.

Developer rate
MiniMax

MiniMax M2.1

Multilingual code and multi-step tool use.

TextContext: 196,608
USD / 1M tokens
Input
$0.3
Output
$1.2
Model details

Model ID: MiniMax-M2.1

Developer rate
MiniMax

MiniMax M2.1 Highspeed

A faster M2.1 variant for code and automation.

TextContext: 196,608
USD / 1M tokens
Input
$0.6
Output
$2.4
Model details

Model ID: MiniMax-M2.1-highspeed

Developer rate
MiniMax

MiniMax M2

An economical model for programming and work agents.

TextContext: 196,608
USD / 1M tokens
Input
$0.3
Output
$1.2
Model details

Model ID: MiniMax-M2

Developer rate
Google

Veo 3.1

Video generation with audio from text and images.

Video
USD / second
Video output
$0.4

720p / 1080p

Model details

Model ID: clodex-video/veo-3.1

USD / second
720p / 1080p
$0.4
4K
$0.6

Billed per second of successfully generated video with audio. Veo 3.1: $0.4/s for 720p and 1080p, $0.6/s for 4K. Fast: $0.1/s (720p), $0.12/s (1080p), $0.3/s (4K). Lite: $0.05/s (720p), $0.08/s (1080p).

Developer rate
Google

Veo 3.1 Fast

Fast video generation with audio for iteration and finished content.

Video
USD / second
Video output
$0.1

720p

Model details

Model ID: clodex-video/veo-3.1-fast

USD / second
720p
$0.1
1080p
$0.12
4K
$0.3

Billed per second of successfully generated video with audio. Veo 3.1: $0.4/s for 720p and 1080p, $0.6/s for 4K. Fast: $0.1/s (720p), $0.12/s (1080p), $0.3/s (4K). Lite: $0.05/s (720p), $0.08/s (1080p).

Developer rate
Google

Veo 3.1 Lite

Video generation with audio in 720p and 1080p.

Video
USD / second
Video output
$0.05

720p

Model details

Model ID: clodex-video/veo-3.1-lite

USD / second
720p
$0.05
1080p
$0.08

Billed per second of successfully generated video with audio. Veo 3.1: $0.4/s for 720p and 1080p, $0.6/s for 4K. Fast: $0.1/s (720p), $0.12/s (1080p), $0.3/s (4K). Lite: $0.05/s (720p), $0.08/s (1080p).

Developer rate
MiniMax

MiniMax H3

Video generation from text and multimodal references, up to 2K.

Video
USD / second
Video output
$0.08

768P

Model details

Model ID: clodex-video/minimax-h3

USD / second
768P
$0.08
2K
$0.13

Output: $0.08/s at 768P, $0.13/s at 2K. Audio input is free; the first 5 input images are free, then $0.04 per image. Input video is charged separately at $0.08/s or $0.13/s according to output resolution.

Developer rate
xAI

Grok Imagine Video 1.5

Video generation with audio, with image and audio reference support.

Video
USD / second
Video output
$0.08

480p

Model details

Model ID: grok-imagine-video-1.5

USD / second
480p
$0.08
720p
$0.14
1080p
$0.25

Video output: $0.08/s (480p), $0.14/s (720p), $0.25/s (1080p). Input images are billed separately at $0.01 per image. These are original xAI API rates.

Developer rate
MiniMax

Hailuo 2.3

Expressive motion video from text and images.

Video
USD / clip
Video output
$0.28

768P · 6s

Model details

Model ID: clodex-video/hailuo-2.3

USD / clip
768P · 6s
$0.28
768P · 10s
$0.56
1080P · 6s
$0.49

Fixed output price: $0.28 per 6-second 768P video; $0.56 per 10-second video; $0.49 per 6-second 1080P video. The card uses the equivalent per-second price ($0.28 / 6) for comparison; the actual charge is per full clip.

Developer rate

These are reference model developer rates. Available models and current Ellami prices for your API key: GET /api/v1/models.

How costs are calculated

Text is priced per 1 million input and output tokens. Images start at the listed price per image. Video is priced per second or clip. Size, quality and billing conditions are under “Model details”.

Services & payment terms ↗

Connect models to your app

Create an API key in Console and set base_url / baseURL: https://www.ellami.pro/api/v1. Use a model ID from the public API catalog in an OpenAI-compatible client. Request examples & documentation ↗