Models

Pick a model. Pay with your wallet only when you use it.

24 models ready to try

Recently Used

Connect a wallet to see models you’ve used lately.

All models(24)

0GM-1.0-35B-A3B

Zero Gravity

Chat

0G.AI in-house model optimized for agentic coding and tool use; thinking enabled by default. Works with tools. Understands images.

0.02 USDC/ est. use

0GM-1.0-35B-A3B-SIA

Zero Gravity

Chat

A 35B hybrid MoE model enhanced with per-token reward guidance. At each decoding step, a 4B Value Model scores candidate tokens and steers generation toward higher-quality outputs, with improvements in harmlessness, helpfulness, and honesty.

0.02 USDC/ est. use

DeepSeek-V4-Flash

DeepSeek

Chat

Lightweight MoE (284B total / 13B active) with native 1M context; low-latency, low-cost. Works with tools.

0.01 USDC/ est. use

DeepSeek-V4-Pro

DeepSeek

Chat

DeepSeek flagship for agentic coding, multi-step workflows, and complex reasoning; 1M context, up to 384K output. Works with tools.

0.025 USDC/ est. use

GLM-5

Zhipu

Chat

Next-generation model purpose-built for coding and agent workflows; 744B foundation. Works with tools.

0.02 USDC/ est. use

GLM-5.1

Zhipu

Chat

Zhipu AI flagship purpose-built for long-horizon tasks; 744B foundation, 200K context. Works with tools.

0.02 USDC/ est. use

GLM-5.2

Zhipu

Chat

Zhipu AI next-generation open-source flagship purpose-built for long-horizon tasks; 1M lossless context. Strong coding and engineering: autonomous task decomposition, architecture design, full-stack development, integration testing, and multi-platform deployment. Works with tools.

0.02 USDC/ est. use

Hunyuan-3

Hunyuan

Chat

Tencent Hunyuan 3 (hy3) flagship; 295B total / 21B active MoE, native 256K context. Text in / text out. Function calling and implicit prompt caching supported. Served via the Tencent Cloud MaaS (TokenHub international) OpenAI-compatible gateway. Works with tools.

0.02 USDC/ est. use

Kimi-K2.7-Code

Moonshot

Chat

Moonshot AI coding model for agentic coding and tool use; multimodal input (text, image, video), thinking always on. 256K context. Works with tools. Understands images.

0.025 USDC/ est. use

Kimi-K3

Moonshot

Chat

Moonshot AI Kimi K3 flagship; multimodal input (text, image, video), text output. 1M context, deep thinking always on. Function calling and implicit prompt caching supported. Works with tools. Understands images.

0.025 USDC/ est. use

MiniMax-M3

MiniMax

Chat

Frontier natively-multimodal model on MiniMax Sparse Attention (MSA); agentic coding, native tool use, and long-horizon tasks. 1M context, thinking on by default. Works with tools. Understands images.

0.02 USDC/ est. use

Whisper Large v3

OpenAI

Audio

Multilingual automatic speech recognition (ASR); transcription and translation.

0.01 USDC/ est. use

Qwen3-VL-30B-A3B-Instruct

Qwen

Chat

Alibaba multimodal vision-language model; strong at visual reasoning, OCR, and document understanding. Works with tools. Understands images.

0.025 USDC/ est. use

Qwen3.6-Plus

Qwen

Chat

Alibaba flagship with hybrid linear attention and sparse MoE; 1M context, 119 languages. Works with tools.

0.025 USDC/ est. use

Qwen3.7-Max

Qwen

Chat

Alibaba flagship with native function calling and web search; 1M context. Works with tools.

0.025 USDC/ est. use

Qwen3.7-Plus

Qwen

Chat

Alibaba multimodal model with vision and video understanding; native function calling, 1M context. Works with tools. Understands images.

0.025 USDC/ est. use

Z-Image-Turbo

Z.ai

Image Gen

Asynchronous text-to-image model with Base64 output. Generates at most 2 images per request — requesting more (n > 2) returns 2 images, not an error.

0.03 USDC/ est. use

Claude Fable 5

Anthropic

Chat

Anthropic Claude Fable 5; text and image input, text output, with a 1M-token context window. Extended thinking and tool use supported. Works with tools. Understands images.

0.027 USDC/ est. use

Claude Opus 4.8

Anthropic

Chat

Anthropic Claude Opus 4.8; text and image input, text output, with a 1M-token context window. Extended thinking and tool use supported. Works with tools. Understands images.

0.03 USDC/ est. use

Claude Opus 5

Anthropic

Chat

Anthropic Claude Opus 5; text and image input, text output, with a 1M-token context window. Extended thinking and tool use supported. Works with tools. Understands images.

0.03 USDC/ est. use

Claude Sonnet 5

Anthropic

Chat

Anthropic Claude Sonnet 5; text and image input, text output, with a 1M-token context window. Extended thinking and tool use supported. Works with tools. Understands images.

0.025 USDC/ est. use

GPT-5.6 Luna

OpenAI

Chat

Fast, cost-efficient model in the GPT-5.6 family, optimized for high-volume, cost-sensitive workloads: responsive chat, classification, extraction, lightweight coding, and agentic workflows at lower latency and cost. Text and image input, text output, 1M-token context. Works with tools. Understands images.

0.03 USDC/ est. use

GPT-5.6 Sol

OpenAI

Chat

Flagship of the GPT-5.6 series, built for advanced reasoning, complex coding, and agentic workflows: multi-step software engineering, long-horizon problem solving, and autonomous tool use. Text and image input, text output, 1M-token context. Works with tools. Understands images.

0.03 USDC/ est. use

GPT-5.6 Terra

OpenAI

Chat

Balanced model in the GPT-5.6 family, tuned for workloads that need strong reasoning, coding, and agentic capability at lower cost than the flagship tier. Text and image input, text output, 1M-token context. Works with tools. Understands images.

0.03 USDC/ est. use