WiseOne AI logo
WiseOneAI

Private AI

Explore Text Models

Browse and discover the best AI language models for conversations, coding, and creative writing.

tencent

Tencent Hy3

Hy3 is Tencent's 295B-parameter Mixture-of-Experts model with 21B active parameters, native 256K context, and configurable reasoning modes. It is built for coding, long-context comprehension, multi-turn dialogue, and agentic task execution with a focus on high-throughput production workloads.

Features

Context

262.1K

Max Output

262.1K

Date Added

Jul 6, 2026

Pricing

Input:

$0.07/1M

Output:

$0.26/1M

Cache:

Read $0.03/1M

Est./msg:

$0.0002

Subscription

Included in subscription

Try Tencent Hy3Details
anthropic

Claude Sonnet 5

Claude Sonnet 5 via Anthropic's native API.

Features

Context

1.0M

Max Output

128.0K

Date Added

Jun 30, 2026

Pricing

Input:

$2.00/1M

Output:

$10.00/1M

Cache:

Read $0.20/1M · Write $2.50/1M (5m) / $4.00/1M (1h)

Est./msg:

$0.0070

Subscription

Not included in subscription

Try Claude Sonnet 5Details
anthropic

Claude Sonnet 5 Thinking

Claude Sonnet 5 with adaptive thinking enabled for complex coding, planning, and agentic work.

Try Claude Sonnet 5 ThinkingDetails
longcat

LongCat 2.0

Meituan's LongCat 2.0 is an open-weight agentic model built for coding, tool use, multi-step reasoning, and long-context workflow automation.

Try LongCat 2.0Details

Nex N2 Mini

Nex AGI's open-source agentic mixture-of-experts model in the Nex N2 family. It accepts text and image input and is built for coding, tool use, structured outputs, and optional reasoning with a 256K context window.

Try Nex N2 MiniDetails
doubao

Doubao Seed 2.1 Pro

Higher-capability model in the Doubao Seed 2.1 family for agentic coding, long-context analysis, complex instruction following, and productivity workflows. Supports a 256k context window and up to 128k output tokens. Note: privacy and logging guarantees may be limited.

Try Doubao Seed 2.1 ProDetails
doubao

Doubao Seed 2.1 Turbo

Fast, lower-cost model in the Doubao Seed 2.1 family for everyday chat, coding assistance, document work, and high-throughput productivity tasks. Supports a 256k context window and up to 128k output tokens. Note: privacy and logging guarantees may be limited.

Try Doubao Seed 2.1 TurboDetails
sakana

Fugu Ultra

Sakana AI's higher-quality Fugu model. It coordinates a deeper pool of expert agents for hard, high-stakes reasoning and coding tasks.

Try Fugu UltraDetails
crofai

Greg 2 Super

Greg 2 Super is CrofAI's balanced Greg 2 model for strong UI design, frontend iteration, coding, writing, and everyday agent tasks at a lower cost than Ultra.

Try Greg 2 SuperDetails
crofai

Greg 2 Ultra

Greg 2 Ultra is CrofAI's most capable Greg 2 model, tuned for premium UI design, agentic coding, creative writing, and higher-end general reasoning tasks.

Try Greg 2 UltraDetails
sofya

Sofya Research

Sofya Research runs a web research agent and returns a structured report with sources, sub-queries, and token usage metadata.

Try Sofya ResearchDetails
cohere

Cohere North Mini Code 1.0

Cohere's compact coding model for fast code generation, code editing, and agentic coding prompts. It supports a 256K-token input context, up to 64K output tokens, and configurable thinking.

Try Cohere North Mini Code 1.0Details
zhipu

GLM 5.2 TEE

GLM-5.2 is Z.AI's flagship model for long-horizon tasks with a 1M-token context window. Served as a text-only TEE deployment via Phala, with provider attestation support.

Try GLM 5.2 TEEDetails
zhipu

GLM 5.2 Thinking TEE

GLM-5.2 with thinking enabled for harder long-horizon coding, autonomous agent workflows, complex engineering optimization, and real-world development tasks. Served through TEE providers, with provider attestation support.

Try GLM 5.2 Thinking TEEDetails
zhipu

GLM 5.2

GLM-5.2 is Z.AI's flagship model for long-horizon autonomous coding and engineering workflows. It is built to plan, execute, iterate, and optimize complex development tasks over extended runs. This variant keeps thinking disabled for faster direct responses.

Try GLM 5.2Details
zhipu

GLM 5.2 Thinking

GLM-5.2 with thinking enabled for harder long-horizon coding, autonomous agent workflows, complex engineering optimization, and real-world development tasks.

Try GLM 5.2 ThinkingDetails
moonshot

Kimi K2.7 Code High-Speed

Kimi K2.7 Code High-Speed is the accelerated coding-focused variant tuned for fast agentic software engineering. It targets roughly 180 tokens per second, with short-context responses reaching up to about 260 tokens per second for rapid coding iterations.

Try Kimi K2.7 Code High-SpeedDetails
moonshot

Kimi K2.7 Code

Kimi K2.7 Code is Moonshot AI's coding-focused agentic model built for long-horizon software engineering workflows. It supports native image input, tool calling, and forced thinking mode; instant/non-thinking mode is not supported.

Try Kimi K2.7 CodeDetails