tencent/Hy3-FP8 neural network architecture graph

hfviewer renders an interactive architecture graph for the Hugging Face model tencent/Hy3-FP8. The graph is built from the model's real structure: nodes carry source-faithful module, class, and operation names (embeddings, attention blocks, feed-forward layers, normalization, task heads), edges follow the actual forward dataflow, and repeated blocks are grouped with their true repeat counts. Where the model could be executed, the graph is trace-backed; otherwise it is derived from the reviewed configuration and source code.

tencent/Hy3-FP8 architecture at a glance

tencent/Hy3-FP8 is a text-generation model (hy_v3 architecture). It uses 80 transformer layers, a hidden size of 4,096, and 64 query heads with 8 key/value heads (grouped-query attention). The feed-forward layers use a mixture-of-experts design with 192 experts (8 active per token). It has a vocabulary of 120,832 tokens and a context window of up to 262,144 tokens, in bfloat16 precision.

Model type
hy_v3
Layers
80
Hidden size
4,096
Attention
grouped-query attention (GQA), 64 query / 8 key-value heads
Feed-forward
Mixture-of-experts design with 192 experts (8 active per token)
Context length
262,144 tokens
Vocabulary
120,832 tokens
Precision
bfloat16

Task: text-generation.

Architecture components in this graph, each explained in the hfviewer glossary: Gated MLP (SwiGLU) · Grouped-query attention (GQA) · LM head (output projection) · MoE experts · MoE router · Multi-token prediction (MTP) · QK-Norm · Residual (skip) connection · RMSNorm · Rotary position embedding (RoPE) · Shared expert · Transformer block.

Related architecture graphs on hfviewer: tencent/Hy3-preview · tencent/Hy3 · tencent/Hy-MT2-30B-A3B · tencent/Hy-MT2-30B-A3B-FP8 · cyankiwi/Hy3-AWQ-NVFP4 · cyankiwi/Hy3-AWQ-INT4.

Browse more graphs from tencent or explore other models on the hfviewer home page.

Interactive model architecture

Architecture graph for tencent/Hy3-FP8.

Interactive architecture graph for tencent/Hy3-FP8, visualized from Hugging Face model metadata.

Paste a Hugging Face link to visualize it
No export step, no config hunt, no model surgery. Paste the link and inspect the graph.
Graph structure Understand the high-level graph structure of different transformer models. Quickstart guide
URL magic You can replace huggingface.co with hfviewer.com in the url to view it.
Chrome Extension! View each model directly on Hugging Face! Install extension
Embed in model card Embed the architecture graph directly in your Hugging Face model card. Add to your model card!
Granularity Block
Zoom into
Community showcase

Featureyour model

Embed the visualization in your model card (README.md) and get it featured in the Community showcase.

Article moderation

Review reports.

Triage reported model articles and comments, hide abusive content, and resolve cases.

Sign in with the HannesVonEssen Hugging Face account to review reports.

No moderation reports match this filter.

Interactive article

Choose a model for your article.

Search ready hfviewer graphs or pick one of your public Hugging Face model repos. Ready models open the editor immediately; owned models that are not viewable yet can be generated first.

Owner article If the model belongs to your Hugging Face user or organization namespace, it is labeled as an owner article.
Community article If you write about another public model, it is labeled as a community article next to the graph.

Loading models...

Release watchlist.

Watch your favorite Hugging Face orgs and get an email the moment a new release's architecture graph is ready in hfviewer.

Editor - interactive article

Loading editor...

MODEL PAGES WITH HFVIEWER

Community showcase.

Add to your model card!

Hugging Face authors are adding the hfviewer model card directly to their READMEs.

If you are interested in deploying these models to edge devices, check out our other products: