tencent logo
tencent

Tencent: Hy3

Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...

Overview

Model specifications

Context
262,144 tokens
Maximum output
128,000 tokens
Architecture
text->text
Tokenizer
Other
Knowledge cutoff
Not provided
Moderated
No
Capabilities

InputOutput

Input
text
Output
text
Reasoning

Yes · high · low · none

API

Supported API parameters

frequency_penaltyinclude_reasoninglogit_biasmax_completion_tokensmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Available providers

7 providers

GMICloud

bf16
Context
262K
Maximum output
236K
Input
$0.126
Output
$0.522
Cache read
$0.032
Cache write
Not provided

Baidu

fp8
Context
262K
Maximum output
236K
Input
$0.132
Output
$0.527
Cache read
$0.033
Cache write
Not provided

Tencent

fp8
Context
262K
Maximum output
128K
Input
$0.132
Output
$0.528
Cache read
$0.033
Cache write
Not provided

DeepInfra

fp8
Context
262K
Maximum output
131K
Input
$0.140
Output
$0.580
Cache read
$0.035
Cache write
Not provided

Novita

unknown
Context
262K
Maximum output
236K
Input
$0.140
Output
$0.580
Cache read
$0.035
Cache write
Not provided

Phala

unknown
Context
262K
Maximum output
236K
Input
$0.150
Output
$0.640
Cache read
$0.040
Cache write
Not provided

AtlasCloud

fp8
Context
262K
Maximum output
131K
Input
$0.200
Output
$0.800
Cache read
$0.050
Cache write
Not provided
tencent

models

All models