Skip to content
Đang kiểm tra

Hợp nhất mọi model một gateway.

Một API tương thích OpenAI, CLI và MCP gateway đứng trước mọi provider — và cả các router khác. Model miễn phí, $4 credit mỗi tháng trên Go, và key của bạn không bị cắt phí.

Miễn phí để bắt đầu — $4 credit mỗi tháng trên Go, không cần thẻ nếu bạn đóng góp một key. Mở khóa pool chung với $2/tháng hoặc góp một key đang dùng được.

Dùng vớiOpenAIxAIAnthropic+ 184 model qua 17 provider
anyrouter ~ openai (python)
import os
from openai import OpenAI

client = OpenAI(
    base_url="https://anyrouter.dev/api/v1",
    api_key=os.environ["ANYROUTER_API_KEY"],
)

resp = client.chat.completions.create(
    model="anthropic/claude-sonnet-4.6",
    messages=[{"role": "user", "content": "Hi"}],
)
ANYROUTER_API_KEYlấy key

Token miễn phí do cộng đồng góp

Chúng tôi gộp key free & trial của mọi người vào một pool chung — token miễn phí cho tất cả. Góp một key, pool lớn hơn, dùng miễn phí.

01

Một endpoint. Mọi model.

Giữ nguyên SDK. Một upstream bị rate-limit — request của bạn không hề hay biết.

Trỏ SDK tới một base URL và gọi model dạng provider/model. Mọi call chạy qua cùng một pipeline: xác thực, chọn upstream, thử upstream kế tiếp khi lỗi, cache phần lặp lại, và ghi log từng attempt.

Giữa call và model · mỗi request

Cân bằng key

Phân tải giữa các key và tự cách ly key bị burn — không sửa client.

Failover tự động

Thử lại và chuyển provider/key ngay khi lỗi hoặc rate-limit.

Prompt caching

Tái sử dụng context đã cache để giảm chi phí token lặp và TTFT.

Log từng request

Mọi hop fallback, token và chi phí theo từng attempt — truy vết và truy vấn được.

03

Mọi thứ trong một gateway

Một key. Mọi khả năng bạn phải tự nối trước đây.

Không add-on, không tier khóa tính năng cơ bản. Mọi khả năng dưới đây dùng chung base URL và API key.

04

Vì sao team chuyển sang

Ba điều về cấu trúc khó sao chép.

Không bị cắt

Dùng key của bạn, giữ trọn từng xu.

Route qua key provider của bạn và AnyRouter không lấy phần — không phí nạp, không markup theo token, không cap. Trả upstream bao nhiêu thì trả đúng bấy nhiêu.

Quan sát được

Debug mọi request.

Mọi hop fallback, status, latency, token và chi phí — ghi theo từng attempt. Trace “debug request này” mà tier này ít ai ship.

Di động

Config đi theo bạn.

Keys, preset và skills sống trong AnyRouter, inject local khi cần, xóa sạch khi thoát. Cùng setup trên laptop, server, CI hay máy đồng đội.

05

Model miễn phí, do mọi người góp

Pool key chung

Mỗi thành viên đăng ký gói free/trial của provider — NVIDIA NIM, Gemini free, Groq free, và hơn nữa — rồi góp key đó. Quota cộng thành một pool mọi thành viên — kể cả bạn — đều gọi được. Góp một key là gói Go của bạn miễn phí.

06

Mở Go — hai cửa, cùng một phòng

Cách tham gia

AnyRouter chạy trên Go$2/tháng, hoặc góp một provider key gói free (người góp dùng miễn phí). Mỗi Go có $4 credit mỗi tháng.

Tạo tài khoản

Đăng ký trong vài giây — không cần thẻ. Bạn vào dashboard với API key đầu tiên sẵn sàng copy.

07

Cách vào rẻ nhất

Ba cách bắt đầu

$4 credit hàng tháng và model miễn phí đi kèm Go — $2/tháng, hoặc miễn phí khi góp một provider key. Hoặc dùng key của bạn, không markup, không cần thẻ.

Gói Go

$4 credit / tháng

Trên Go — $2/tháng hoặc một key đóng góp. $4 đi được bao xa trên model nhanh.

Token cho $4, trộn 75% input / 25% output.

Gói Go

Model miễn phí, 1000/ngày

Định tuyến qua anyrouter/free trên Go — $0 mỗi token, tối đa 1000 request/ngày.

Apple Foundation Model (on-device)
$0 in · $0 out
DeepSeek V4.1 Flash
$0 in · $0 out
Dots3-Note Preview
$0 in · $0 out
Gemma 4 31B
$0 in · $0 out
Ling-3.0-flash-Fin
$0 in · $0 out
Ling-3.0-flash-Sante
$0 in · $0 out
Ling-3.0-flash-VL
$0 in · $0 out
Ling-3.0-tiny
$0 in · $0 out
LFM 2.5 2.6B
$0 in · $0 out
Nex-N2.5-Mini
$0 in · $0 out
Nex-N2.5-Pro
$0 in · $0 out
Ising Calibration 1.5 31B
$0 in · $0 out
Xem model miễn phí
Dùng key của bạn

Key của bạn, không markup

12 provider có gói free — chỉ provider tính phí, chúng tôi không.

Thêm key của bạn
08

Đăng nhập với AnyRouter

Thêm AI vào app — người dùng mang AnyRouter của họ

Một nút OAuth đưa app của bạn một key tạm, scoped theo từng user. Mọi request tính vào tài khoản của họ — bạn không thu, lưu, hay trả API key.

Màn hình đăng nhập của app

Click thử — xem người dùng của bạn thấy gì.

  • OAuth 2.1 + PKCE

    Flow authorization-code chuẩn, public client, không có secret để lộ. Chạy được từ app chỉ có browser.

  • Key chỉ scoped inference

    Token chỉ chạy được AI request và đọc profile cơ bản — không chạm key, billing hay cài đặt tài khoản.

  • Billing & thu hồi theo từng user

    Mỗi user trả từ credit hoặc gói free của chính họ, và thu hồi app của bạn bất cứ lúc nào từ dashboard.

09

184+ model · 17 provider · 21 tháng 9, 2026

Một catalog đang lớn, luôn cập nhật

xAI's Grok 4.7 joins the catalog — 500K context, adjustable reasoning, vision in, and six BYOK upstream routes.

262K

Doubao Seed 2.1 Pro (build 260628) is a ByteDance Seed reasoning model with strong performance on coding, mathematics, and multilingual tasks. Serves via AIHubMix BYOK.

$1.60 in
$4.80 out
500K

SpaceXAI's frontier model for coding, agentic tasks, and knowledge work — succeeds Grok 4.6 with mandatory adjustable reasoning (low/medium/high/xhigh, default high).

$0.04 in
$0 out
128KZDR

TypeSafe Jev is a System One decision model: send application state and typed questions (noul / choice / score) and get structured answers with probabilities — not generated chat text. Call POST /api/v1/decisions (POST /api/v1/systemone is the same handler). Requires a BYOK key for TypeSafe, LLM Gateway, Vercel AI Gateway, or AIHubMix (dashboard → BYOK). TypeSafe list price is $0.042 per million input tokens; output tokens are free upstream. These hops are user-key only (0/0 on AnyRouter credits).

262K

Union Alpha is a stealth preview multimodal model for research, coding, and agentic workflows. It offers frontier-level general-purpose performance with vision in and text out. OpenRouter lists tools and structured output; it does not advertise reasoning on this SKU.

512K

Agnes 3.0 Flash is designed for real-world agent tasks and development workflows, covering the full execution chain from task understanding and planning to tool invocation and final delivery. The model focuses on improving stability, instruction-following, factual grounding, and output completeness in complex tasks, with vision in and text out via an OpenAI-compatible Chat Completions interface.

8KFree

NVIDIA Nemotron Parse 2.0 is a document-understanding VLM: page image in, structured text out (markdown, layout classes, bounding boxes, reading order). Served via NVIDIA NIM OpenAI-compatible /v1/chat/completions. Not image generation.

262K

Ling-3.0-flash-Fin is inclusionAI's finance-focused mixture-of-experts model, built on Ling-3.0-flash (124B total / 5.1B active). It is designed for real-world investment research, market analysis, and financial reasoning, while retaining general reasoning, coding, and agentic skills. Distinct from text-only inclusionai/ling-3.0-flash, the VL listing, and the health/medicine Sante SKU. Served free via the platform free pool, with BYOK as a fallback.

262K

Ling-3.0-flash-Sante is inclusionAI's health and medicine-focused mixture-of-experts model, built on Ling-3.0-flash (124B total / 5.1B active). It is designed for medical knowledge reasoning, clinical safety, evidence-based retrieval, and long-horizon medical tasks, while retaining general reasoning, coding, and agentic skills. Distinct from text-only inclusionai/ling-3.0-flash and the VL listing. Served free via the platform free pool, with BYOK as a fallback. Command Code lists the same product as a free-while-it-lasts promo — that is the upstream's credit, not AnyRouter Free monthly credits.

262K

Ling-3.0-flash-VL is inclusionAI's native multimodal instruct model, a 124B-parameter Mixture-of-Experts model with roughly 5.5B activated parameters per token. Built on Ling-3.0-flash, it adds native image and video understanding (up to 256K context) for visual reasoning, document and chart reading, and agentic GUI tasks. Distinct from the text-only inclusionai/ling-3.0-flash listing. Served free via the platform free pool, with BYOK as a fallback.

262K

Nex-N2.5-Mini is Nex AGI's smaller agentic coding model (35B total / 3B active MoE), built for high-speed instruction following, real-time tool execution, and visually assisted computer use. Distinct from Nex-N2.5-Pro and the retired Nex-N2-Pro listing. OpenRouter lists a :free SKU only — that wire is an alias, not the catalog id. Served free via the platform free pool, with BYOK as a fallback.

262K

Nex-N2.5-Pro is Nex AGI's larger agentic coding model (397B total / 17B active MoE), built to turn goals into working, verified outcomes with a visual feedback loop. Distinct from Nex-N2.5-Mini and the retired Nex-N2-Pro listing. OpenRouter lists a :free SKU only — that wire is an alias, not the catalog id. Served free via the platform free pool, with BYOK as a fallback.

1MFree

DeepSeek V4.1 Flash is the cost-efficient sparse MoE tier of the V4.1 family (1M context, vision in, text out). Official API id is `deepseek-flash`; OpenRouter lists the same product as `deepseek/deepseek-v4.1-flash`.

1.1M

OpenAI GPT-6 Astra reasoning/chat model with ~1.05M context; Experiential Cloud promotional free daily tier.

1M

Meta's Muse Spark 1.3 is a multimodal reasoning model for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context window.

1M

Google's Gemini 3.8 Flash is a multimodal model for fast agentic workflows, coding, and complex multi-step reasoning, with a 1,048,576-token context window.

260K

Inception Labs Mercury 2.5 Preview is a text model with thinking and tool use and a 260,000-token context window.

1M

Anthropic's Claude Fable 5.1 is the successor to Claude Fable 5 for long-running agentic coding, knowledge work, and research. Same $10/$50 list rates as Fable 5, with cache reads at $0.25 per million tokens and a 1M-token context window.

1M

Alibaba Qwen native vision-language model for coding, office, long-context reasoning, and agents. 1M context.

1M

Hy4 Preview is Tencent Hunyuan's 770B/49B-active MoE for agents, coding, office automation, and complex tool use, with a 1M-token context window.

512K

Agnes 2.5 Flash is Agnes AI's fast and efficient language model, optimized for coding tasks, agent workflows, tool calls, multi-turn dialogue, reasoning, and image understanding via an OpenAI-compatible Chat Completions interface.

1M

Agnes 2.5 Pro Alpha is Agnes AI's paid inference model for advanced coding, scientific reasoning, long-context analysis, agent workflows, and multimodal understanding via an OpenAI-compatible Chat Completions API.

1M

Agnes 2.5 Pro is Agnes AI's paid inference model and the commercially stable version of the Agnes 2.5 Pro Alpha ranking model, suited for advanced coding, scientific reasoning, long-context analysis, agent workflows, and multimodal understanding.

256K

MAI-Thinking-1 is Microsoft's first inference model in the MAI series, built for enterprise-scale workloads with strong reasoning, mathematical, and general intelligence capabilities at high throughput.

1M

GLM-5.3-Flash is a native multimodal model from Z.ai (320B-A18B, MIT License, 1M-token context). It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while reducing compute overhead.

10

FAQ

Câu hỏi thường gặp

AnyRouter là gì?

Một API tương thích OpenAI, CLI và MCP gateway đứng trước mọi provider — và cả các router khác. Model miễn phí, $4 credit mỗi tháng trên Go, và key của bạn không bị cắt phí.

Gọi model qua AnyRouter như thế nào?

Trỏ SDK tới một base URL và gọi model dạng provider/model. Mọi call chạy qua cùng một pipeline: xác thực, chọn upstream, thử upstream kế tiếp khi lỗi, cache phần lặp lại, và ghi log từng attempt.

Nếu tôi dùng key của mình thì có bị cắt phí không?

Route qua key provider của bạn và AnyRouter không lấy phần — không phí nạp, không markup theo token, không cap. Trả upstream bao nhiêu thì trả đúng bấy nhiêu.

Bắt đầu như thế nào?

Đăng ký trong vài giây — không cần thẻ. Go là gói mở AnyRouter: trả $2 / tháng, hoặc góp 1 key miễn phí. Go kèm $4 credit mỗi tháng, model miễn phí không giới hạn, và truy cập shared key pool.