Private AI
Browse and discover the best AI language models for conversations, coding, and creative writing.
Tencent Hy3
Hy3 is Tencent's 295B-parameter Mixture-of-Experts model with 21B active parameters, native 256K context, and configurable reasoning modes. It is built for coding, long-context comprehension, multi-turn dialogue, and agentic task execution with a focus on high-throughput production workloads.
Features
Context
262.1K
Max Output
262.1K
Date Added
Jul 6, 2026
Pricing
Input:
$0.07/1M
Output:
$0.26/1M
Cache:
Read $0.03/1M
Est./msg:
$0.0002
Subscription
Included in subscription
Claude Sonnet 5
Claude Sonnet 5 via Anthropic's native API.
Features
Context
1.0M
Max Output
128.0K
Date Added
Jun 30, 2026
Pricing
Input:
$2.00/1M
Output:
$10.00/1M
Cache:
Read $0.20/1M · Write $2.50/1M (5m) / $4.00/1M (1h)
Est./msg:
$0.0070
Subscription
Not included in subscription
Claude Sonnet 5 Thinking
Claude Sonnet 5 with adaptive thinking enabled for complex coding, planning, and agentic work.
LongCat 2.0
Meituan's LongCat 2.0 is an open-weight agentic model built for coding, tool use, multi-step reasoning, and long-context workflow automation.
Nex N2 Mini
Nex AGI's open-source agentic mixture-of-experts model in the Nex N2 family. It accepts text and image input and is built for coding, tool use, structured outputs, and optional reasoning with a 256K context window.
Doubao Seed 2.1 Pro
Higher-capability model in the Doubao Seed 2.1 family for agentic coding, long-context analysis, complex instruction following, and productivity workflows. Supports a 256k context window and up to 128k output tokens. Note: privacy and logging guarantees may be limited.
Doubao Seed 2.1 Turbo
Fast, lower-cost model in the Doubao Seed 2.1 family for everyday chat, coding assistance, document work, and high-throughput productivity tasks. Supports a 256k context window and up to 128k output tokens. Note: privacy and logging guarantees may be limited.
Fugu Ultra
Sakana AI's higher-quality Fugu model. It coordinates a deeper pool of expert agents for hard, high-stakes reasoning and coding tasks.
Greg 2 Super
Greg 2 Super is CrofAI's balanced Greg 2 model for strong UI design, frontend iteration, coding, writing, and everyday agent tasks at a lower cost than Ultra.
Greg 2 Ultra
Greg 2 Ultra is CrofAI's most capable Greg 2 model, tuned for premium UI design, agentic coding, creative writing, and higher-end general reasoning tasks.
Sofya Research
Sofya Research runs a web research agent and returns a structured report with sources, sub-queries, and token usage metadata.
Cohere North Mini Code 1.0
Cohere's compact coding model for fast code generation, code editing, and agentic coding prompts. It supports a 256K-token input context, up to 64K output tokens, and configurable thinking.
GLM 5.2 TEE
GLM-5.2 is Z.AI's flagship model for long-horizon tasks with a 1M-token context window. Served as a text-only TEE deployment via Phala, with provider attestation support.
GLM 5.2 Thinking TEE
GLM-5.2 with thinking enabled for harder long-horizon coding, autonomous agent workflows, complex engineering optimization, and real-world development tasks. Served through TEE providers, with provider attestation support.
GLM 5.2
GLM-5.2 is Z.AI's flagship model for long-horizon autonomous coding and engineering workflows. It is built to plan, execute, iterate, and optimize complex development tasks over extended runs. This variant keeps thinking disabled for faster direct responses.
GLM 5.2 Thinking
GLM-5.2 with thinking enabled for harder long-horizon coding, autonomous agent workflows, complex engineering optimization, and real-world development tasks.
Kimi K2.7 Code High-Speed
Kimi K2.7 Code High-Speed is the accelerated coding-focused variant tuned for fast agentic software engineering. It targets roughly 180 tokens per second, with short-context responses reaching up to about 260 tokens per second for rapid coding iterations.
Kimi K2.7 Code
Kimi K2.7 Code is Moonshot AI's coding-focused agentic model built for long-horizon software engineering workflows. It supports native image input, tool calling, and forced thinking mode; instant/non-thinking mode is not supported.