Lumipat sa pangunahing nilalaman

AI models on Zeplik

289 models from 50 labs, all in one chat. Pricing and capabilities below come straight from the live registry Zeplik routes and bills against, synced continuously. Pick a model to see its full profile, or start typing on any model page to try it.

GPT-6 AstraFrontier

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work.

1.1M ctx600 cr/M outSeptember 2026
GPT-6 SolFrontier

GPT-6 Sol is the cost-efficient high-end model in OpenAI's GPT-6 series, positioned below the flagship GPT-6 Astra and above the fast GPT-6 Luna tier.

1.1M ctx120 cr/M outSeptember 2026
Claude Opus 5.5Frontier

Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5.

1M ctx240 cr/M outSeptember 2026
Claude Sonnet 5Frontier

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work.

1M ctx120 cr/M outJune 2026
Claude Fable 5.1Frontier

1M ctx600 cr/M outSeptember 2026
GPT-5.6 LunaGeneral purpose

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series.

1.1M ctx14.4 cr/M outJuly 2026
GPT-5.6 TerraFrontier

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier.

1.1M ctx144 cr/M outJuly 2026
Gemini 3.8 FlashFast and light

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

1.0M ctx45.0 cr/M outSeptember 2026
Grok 4.7Frontier

Grok 4.7 is SpaceXAI's flagship model for coding, agentic tasks, and knowledge work, succeeding Grok 4.6.

500K ctx72.0 cr/M outSeptember 2026
DeepSeek V4 Pro 0423Open weight

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window.

1.0M ctx5.0 cr/M outApril 2026

Media generation

Beyond text, Zeplik generates images, video, music and speech — and transcribes audio — right in the chat. You never pick a media model; just ask, and the product routes to the right one. Costs below are the exact credit basis the ledger charges.

Video

Fast (default)

Kling 2.5

~0.84 credit / second

a 5-second clip ≈ 4.2 credits (up to ~12.6 at the 15s max)

Video

Premium (cinematic/HD)

Kling 2.6

~1.68 credit / second

a 10-second clip ≈ 16.8 credits (up to ~25.2 at the 15s max) — asks you to confirm the spend first

Confirms the spend before generating

Text-to-speech

xAI TTS

~0.18 credit / 1,000 characters

about 0.22 credits for a reply-length passage

Music

CassetteAI

~0.24 credit / minute (rounded up)

a 30-second track ≈ 0.24 credits

3D model (from text)

Rodin

~4.8 credits / mesh

describe an object and get a rotatable GLB

3D model (from an image)

Hunyuan3D 2.1

~3.6 credits / mesh

turn a photo in the conversation into a rotatable GLB

Image (generate + edit)

Nano Banana 2

~0.96 credit / image

generate a new image or edit one in the conversation

Upscale

Clarity Upscaler

~0.36 credit / image

increase the resolution of an image in the conversation

Background removal

BiRefNet

~0.01 credit / image (a fraction of a cent)

remove the background, leaving a transparent PNG

Vectorize

Recraft

~0.12 credit / image

convert an image into a scalable SVG vector

Vector/SVG from text

Recraft

~0.96 credit / image

generate a brand-new SVG (logo, icon) from a description

Transcription

Whisper / Wizper

Free

upload an audio file and its speech is transcribed to text

OpenAI

OpenAI ships the GPT line, the most widely used family of AI models in the world. Zeplik leads with GPT-6 Astra and its deeper Astra Pro serving, and now carries the rest of the GPT-6 generation beneath it: GPT-6 Sol for demanding professional work and GPT-6 Luna for high-volume, latency-sensitive work, each with a Pro serving and each on the same 1M-token window as Astra. Behind them sit the GPT-5.6 tiers (Sol, Luna and Terra), GPT-5 and GPT-4 era releases, mini and nano tiers for fast inexpensive work, the Codex models for software engineering, GPT Audio for speech, and the GPT Image models for generation and editing.

All 100 OpenAI models

Qwen

Alibaba's Qwen team ships one of the broadest open catalogs in AI: Max and Plus capability tiers, Flash speed tiers, dedicated coder and reasoning variants, vision-language models, and dozens of open-weight sizes. Zeplik carries the Qwen3.x generation led by Qwen3.8 Max and its higher-throughput Max Prime serving, with Qwen3.8 Flash and the multimodal Omni Flash on the speed tier and the long tail of open-weight sizes behind them.

All 53 Qwen models

Google

Google DeepMind's Gemini models pair strong multimodal understanding with some of the largest context windows available, and the open-weight Gemma line brings the same research to smaller, cheaper models. Zeplik carries the Gemini 3.x generation led by Gemini 3.8 Flash, the Flash-Lite speed tiers, the Nano Banana image models, Lyria for music, and Gemma 4.

All 42 Google models

Anthropic

Anthropic builds the Claude family, known for careful reasoning, strong writing and reliable tool use. Zeplik leads with Claude Opus 5.5, Anthropic's flagship for demanding reasoning, coding and long-horizon agentic work, which succeeds Claude Opus 5 at a lower price on both input and output. Alongside it Zeplik carries Claude Fable 5.1 and the rest of the current generation — Claude Opus 5, Claude Sonnet 5 and Claude Fable 5 — with the Opus 4.x line and earlier Sonnet and Haiku releases behind them.

All 29 Anthropic models

Mistral

Mistral AI, the leading European lab, ships efficient models across every size. Zeplik carries Mistral Medium 3.5 at the head of the catalog, plus Mistral Large 3 and Mistral Small 4, the Ministral 3 sizes for edge deployment, Devstral 2 and Codestral for code, and Voxtral for audio.

All 26 Mistral models

Z.ai

Z.ai (formerly Zhipu) ships the GLM series, open-weight models with strong coding and agent performance. Zeplik carries the GLM 5.x generation led by GLM 5.3 and its high-throughput 5.3 Prime serving, with the 5.3 Flash and FlashX speed tiers, the Turbo servings and the vision-capable 5V behind them.

All 18 Z.ai models

DeepSeek

DeepSeek publishes open-weight models that compete with far more expensive closed models, particularly on reasoning and code. Zeplik leads with DeepSeek V4.1 Flash — the first DeepSeek release on the company's Causal Encoder-Decoder architecture, with vision and a 1M-token window — and carries the V4 generation behind it, including Pro, Flash and the vision-capable V4 Flash Vision, along with the V3.x and R1 lines.

All 15 DeepSeek models

NVIDIA

NVIDIA's Nemotron models are open-weight releases tuned for helpfulness and agentic tasks, built on top of leading open architectures and optimized for NVIDIA hardware. Zeplik carries Nemotron 3.5 Lightning at the head of the line, plus the Nemotron 3 Ultra, Super and Nano sizes and a content-safety classifier.

All 11 NVIDIA models

SpaceXAI

SpaceXAI — the lab formerly known as xAI — trains the Grok models with an emphasis on reasoning and up-to-date knowledge. Zeplik carries the Grok 4.x generation led by Grok 4.7, which succeeds Grok 4.6 on long-running software engineering and agentic work while costing less on both input and output, with Grok 4.6, Grok 4.5 and Grok 4.3 behind it, the Grok 4.20 Multi-Agent serving, and the coding-focused Grok Build.

All 8 SpaceXAI models

MoonshotAI

Moonshot AI builds the Kimi models, open-weight releases with standout agentic and coding ability. Zeplik leads with Kimi K3 and carries the K2.x line behind it, including the code-focused K2.7 Code and the deliberate K2 Thinking.

All 8 MoonshotAI models

MiniMax

MiniMax builds the M-series of chat models, iterating quickly on conversational quality and long-context handling. Zeplik carries MiniMax M3 at the head of the line, with the M2.x releases and the original MiniMax-01 behind it.

All 8 MiniMax models

Meta

Meta's Llama models are the most widely adopted open-weight family in the industry. Zeplik carries the Llama 4 generation (Maverick and Scout), the Llama Guard 4 safety classifier, and the proven Llama 3.x releases.

All 8 Meta models

Tencent

Tencent's Hunyuan models span general chat and dedicated translation. Zeplik carries the Hy4 preview at the head of the line, plus Hy3, the Hy-MT2 translation models in three sizes, and the open-weight Hunyuan A13B.

All 7 Tencent models

AionLabs

AionLabs ships the Aion models, multi-model roleplaying and storytelling systems in which several specialised models generate collaboratively. Zeplik carries Aion-3.5 and Aion-3.5-Mini at the head of the line, with Aion-3.0, Aion-3.0-Mini, Aion-2.0 and the role-play focused Aion-RP behind them.

Cohere

Cohere focuses on enterprise text work: retrieval-augmented generation, tool use and multilingual business writing. Zeplik leads with Command A+, Cohere's flagship for enterprise agentic workflows, which takes text and image input with strict tool schemas over a 192K window, and carries North Mini Code alongside the rest of the Command family — Command A, Command R+ and Command R.

ByteDance Seed

ByteDance's Seed team ships the Seed series, models trained for strong general reasoning with a dedicated code variant. Zeplik carries Seed 2.1 Turbo and Seed-2.0-Code at the head of the line, with the Seed-2.0 Lite and Mini sizes and the Seed 1.6 releases behind them.

OpenRouter

OpenRouter's own entries are ROUTERS, not models: each one reads the request and forwards it to whichever underlying model fits best, so its price and capabilities depend on where it lands. Zeplik carries the Auto Router and its beta, the Fusion and Pareto Code routers, and the free-models router.

inclusionAI

inclusionAI publishes the open-weight Ling models. Zeplik carries the Ling 3.0 Flash line led by Ling 3.0 Flash VL, which adds native visual perception to the base Flash model, along with the domain-tuned Sante (health) and Fin (finance) variants.

Xiaomi

Xiaomi's MiMo models are open-weight releases from the company's AI lab. Zeplik carries the MiMo-V2.6 generation — the 1T-parameter MiMo-V2.6-Pro, its roughly 10x faster Pro UltraSpeed serving, and the lighter MiMo-V2.6-Flash — with MiMo-V2.5 and MiMo-V2.5-Pro behind them.

Perplexity

Perplexity's Sonar models are built for search-grounded answering: they retrieve from the live web and cite sources as part of the response. Zeplik carries Sonar Pro Search at the head of the line, with Sonar Reasoning Pro, Sonar Deep Research and the base Sonar model behind it.

Thinking Machines

Thinking Machines ships the Inkling models. Zeplik carries both sizes — Inkling and the lighter Inkling Small — each with batch and free servings.

Poolside

Poolside builds models for software engineering. Zeplik carries the Laguna line in two sizes — Laguna S 2.1 and the smaller Laguna XS 2.1 — each with a free serving.

Amazon

Amazon's Nova models are the in-house family behind AWS Bedrock, tuned for practical business tasks at aggressive price points. Zeplik carries Nova 2 Lite and Nova Premier alongside the original Nova Pro, Lite and Micro tiers.

Upstage

Upstage builds the Solar models, compact releases that punch above their size on reasoning and document work. Zeplik carries Solar Mini 4 — a 35B mixture-of-experts with 3B active parameters and a 524K window, built for agentic work where response cost matters — alongside Solar Pro 4 and Solar Pro 3.

Sakana

Sakana AI, the Tokyo lab, researches nature-inspired approaches to model design. Its Fugu models are not single monolithic models but a learned multi-agent orchestration system that routes work across specialised models. Zeplik carries Fugu Ultra v2 and the cost-performance Fugu Max, with the original Fugu Ultra behind them.

Nous Research

Nous Research fine-tunes open-weight base models into the Hermes series, known for instruction following and a neutral, steerable voice. Zeplik carries Hermes 4 405B and the Hermes 3 releases.

Unbiased

Unbiased builds Pareto, a multimodal composite model for research, coding and agentic workflows. Zeplik carries it with vision, tool calling and a 262K-token window.

Perceptron

Perceptron builds models for grounded perception — reading what a scene contains and pointing at it, rather than describing it in prose. Zeplik carries Perceptron Mk1.5, an embodied-reasoning model for physical agents that takes text, images, video and audio and answers with text plus structured annotations — points, boxes, polygons and tracks — with tool calling and adjustable reasoning effort, and Perceptron Mk1 behind it.

Inference.net

Inference.net builds the Schematron models, small models trained for one job: turning HTML into structured JSON against a caller-supplied schema. Zeplik carries both V2 servings — Turbo, tuned for throughput on high-volume extraction, and Small, tuned for extraction quality on complex schemas and long pages.

Inception

Inception builds diffusion language models (dLLMs), which produce and refine many tokens in parallel rather than one after another — a different generation mechanism from every other lab in this catalog. Zeplik carries Mercury 2.5, the newest of them, with a 260K window and tool calling, alongside Mercury 2.

IBM

IBM's Granite models are open-weight releases aimed at enterprise deployment, published with clear licensing and small enough to run close to the data. Zeplik carries Granite 4.2 8B and the compact Granite 4.0 Micro.

StepFun

StepFun ships the Step series. Zeplik carries the Flash speed tiers — Step 3.7 Flash and Step 3.5 Flash.

Reka

Reka builds compact multimodal models designed to run efficiently. Zeplik carries Reka Edge and Reka Flash 3.

Morph

Morph builds models for applying code edits fast — taking a proposed change and merging it into a file, rather than writing prose about it. Zeplik carries Morph V3 Large and Morph V3 Fast.

Microsoft

Microsoft Research publishes the Phi models, small releases trained on carefully curated data that compete with much larger ones on reasoning. Zeplik carries Phi 4 and the WizardLM-2 8x22B mixture-of-experts model.

Fireworks

Fireworks Research builds models tuned to make each token count. Zeplik carries Ember-1, a reasoning model built on Kimi K3 that reaches its answer with substantially shorter reasoning traces, over a 1M-token window with vision and tool calling.

PrismML

PrismML works on making capable models small enough to run cheaply. Zeplik carries Ternary Bonsai 2 27B, a reasoning model derived from Qwen3.8-27B and shrunk by ternary compression, which handles code, mathematics, tool calling and image understanding over a 262K window.

Dots Studio

Dots Studio publishes the Dots models. Zeplik carries the Dots3-Note preview with a free serving.

LiquidAI

Liquid AI builds the LFM models on a non-transformer architecture designed for efficiency at small scale. Zeplik carries LFM2.5-2.6B with a free serving.

Meta

Beyond Llama, Meta publishes the Muse line under its own namespace. Zeplik carries Muse Glimmer 30B, with a batch serving.

Meituan

Meituan publishes the open-weight LongCat models. Zeplik carries LongCat 2.0.

Arcee AI

Arcee AI builds open-weight models through model merging and distillation. Zeplik carries Trinity Large Thinking, its reasoning release.

Writer

Writer builds the Palmyra models for enterprise content work — brand-consistent writing, structured documents and business analysis. Zeplik carries Palmyra X5.

Relace

Relace builds models for codebase retrieval — finding the files and symbols a change touches. Zeplik carries Relace Search.

ByteDance

ByteDance publishes UI-TARS, an open-weight model trained specifically to operate graphical user interfaces — reading a screenshot and deciding what to click, type or scroll.

Baidu

Baidu's ERNIE models are a long-running Chinese-language family with strong multilingual and multimodal coverage. Zeplik carries the vision-language ERNIE 4.5 VL, a 424B mixture-of-experts release.

More providers

Compare models side by side

Head-to-head pricing, context and capability comparisons for the pairings people actually decide between.

Browse all comparisons