APIMart
Chat API

LLM API for Every Leading Model

Call GPT, Claude, Gemini, DeepSeek, Qwen, Grok, Kimi, GLM and MiniMax models through one OpenAI-compatible API. Switch models by changing one parameter, and every model is 20% off the list price.

LLM API Model Families

Each family page lists every available version with its context window, token pricing and a playground to test prompts before you integrate.

How to Choose an LLM API

Complex reasoning and coding

Claude Opus 5.5, GPT-6.1 and Qwen 3.8 Max handle multi-step reasoning, agents and large codebases.

Fast and low-cost chat

Gemini 3.8 Flash, DeepSeek V4.1 Flash and GLM-5.3 Flash keep latency and token cost low for high-volume chat and extraction.

Long context

Several models support very long context windows for document analysis and retrieval; check each model's context limit on its family page.

Drop-in OpenAI compatibility

Use the OpenAI SDK you already have: point the base URL to APIMart and change the model name to switch providers.

LLM API FAQ

Is the APIMart LLM API OpenAI-compatible?

Yes. Chat models use the OpenAI-compatible Chat Completions format, so existing OpenAI SDK code works after changing the base URL, API key and model name. Claude models are also available through the Messages format.

How is LLM API pricing calculated?

LLM models are billed per million input and output tokens, and some models also price cached input. Each family page shows the token pricing per model, and every model is 20% off the list price, with membership tiers saving up to 28%.

Which LLM APIs are available?

APIMart offers 140+ chat models from OpenAI, Anthropic, Google, DeepSeek, Alibaba Qwen, xAI, Moonshot Kimi, Zhipu GLM and MiniMax, with new versions added as they launch.

Can I use one API key for multiple LLM providers?

Yes. A single APIMart key works for every chat, image and video model, with one balance and one invoice.

Do you support streaming and function calling?

Yes. Streaming responses, function or tool calling and vision input are supported on models whose providers offer those features.