REVIEWS / AI MODELS / HY3 UPDATED JUL 10, 2026 · 116 SOURCES

THE PRODUCT

Hy3

Tencent's Hy3 is an open-source LLM praised for aggressive pricing but criticized for benchmark mediocrity and inconsistency versus smaller competitors.

AI MODELS HIGH CONFIDENCE

THE VERDICT

6.0

REALITY SCORE · OUT OF 10 · CONFIDENCE HIGH

COMPOSED FROM

USERS 4.6 · 113 voices · 100%
CRITICS no published scores yet

SENTIMENT · 116 REVIEWS

+ 30% positive · 25% neutral − 45% negative

BEST PRICE TODAY

BUY ON AMAZON
Affiliate · supports independent reviews
CHECK PRICE →

// Affiliate link — score is unaffected.

10 REDDIT 35 YOUTUBE 30 HN 35 LEMMY 3 STACK EXCHANGE
USER n=116
VIDEO n=3
BRAND AVAILABLE
INTERNET n=0

AT A GLANCE · QUOTABLE

  • Rating: 6.0 / 10 (high confidence)
  • User voices: 116 across 5 platforms
  • Sentiment: 30% positive · 45% negative
  • Updated: Jul 9, 2026

GYIBB rates the Hy3 6.0/10 based on 116 user voices from 5 platforms. Confidence: high. Source: https://gyibb.com/ai-models/hy3

BUY IF

Aggressively priced — 'exceedingly cheap' on OpenRouter, competitive with DeepSeek Flash V4

  • + Open-source / FOSS option for users who need a self-hostable model
  • + Post-preview release showed significant benchmark improvement on DeepSWE
  • + Can skip reasoning chains for speed, useful for dev workflows

SKIP IF

Fell from #1 to #8/9 on OpenRouter rankings within a month

  • Outperformed by smaller models (Gemma 4 31B, Qwen 3.6 27B) in security-auditing benchmarks
  • 'Makes idiotic decisions all the time' when lacking task-specific knowledge
  • Quantization resilience unproven vs. DeepSeek V4's native FP4 efficiency

Where the layers disagree

5 CONTRADICTIONS DETECTED

USER rankings data shows Hy3 fell from #1 to #8/9 on OpenRouter, contradicting the VIDEO excitement framing it as a Qwen 3.7 competitor — real adoption is declining while influencer hype is fresh.

VIDEO VS USER

USER benchmarks from a security-auditing professional report Hy3 is 'outperformed by Gemma 4 31B and Qwen 3.6 27B,' while VIDEO (Superbash) explicitly titles the model 'Fast, but not very smart?' — both layers independently flag weak reasoning.

VIDEO VS USER

USER price sentiment ('exceedingly cheap, no reason not to use it') clashes with USER performance verdicts ('mediocre,' 'idiotic decisions') — cost-to-capability ratio is the central unresolved debate.

USER VS BRAND

VIDEO (xCreate) positions Hy3 as a local/open-source win for Mac Studio, but a VIDEO commenter warns 'China will not provide opensource models globally anymore' — geopolitical risk to the open-source value proposition is unaddressed by other layers.

VIDEO VS USER

USER notes the post-preview release is 'significantly better' with higher DeepSWE scores, but most VIDEO reviews tested only the preview — real-world performance of the shipping model remains under-tested on video.

VIDEO VS USER

Value depends on how you pay

SAME MODEL · TWO BUYERS

ON A SUBSCRIPTION

6.0

Claude Max · ChatGPT Plus · GLM Coding — flat rate, tokens don't bill

For flat-rate plan buyers (e.g., through a coding IDE or aggregated AI platform), Hy3 offers adequate general-purpose capability comparable to GPT-5.4-mini but is not top-tier. Users report it 'just works for most things' but falls behind Sonnet 5 and GLM-5.2 in hard reasoning. If included in a plan with daily limits, it's a decent secondary model for routine tasks — not the one you'd pick for com

ON PER-TOKEN API

7.0

Enterprise / pay-per-use — $/1M, latency, token efficiency bite

For per-token API buyers, Hy3's economics are the main draw — users call it 'exceedingly cheap' with input pricing now matching DeepSeek Flash V4. However, the cost advantage is offset by benchmark mediocrity and inconsistent output quality. For cost-sensitive workloads where 'good enough' reasoning suffices, it's a legitimate budget option. For production systems requiring reliability (e.g., secu

WHERE THEY AGREE +

+ Aggressively priced — 'exceedingly cheap' on OpenRouter, competitive with DeepSeek Flash V4
+ Open-source / FOSS option for users who need a self-hostable model
+ Post-preview release showed significant benchmark improvement on DeepSWE
+ Can skip reasoning chains for speed, useful for dev workflows
+ Comparable to GPT-5.4-mini and Sonnet 5 for general tasks per some users

WHERE THEY DON'T

Fell from #1 to #8/9 on OpenRouter rankings within a month
Outperformed by smaller models (Gemma 4 31B, Qwen 3.6 27B) in security-auditing benchmarks
'Makes idiotic decisions all the time' when lacking task-specific knowledge
Quantization resilience unproven vs. DeepSeek V4's native FP4 efficiency
VIDEO reviewers independently question intelligence ('Fast, but not very smart?')

Where the 116 sources came from

VIEW EVERY CITATION →
REDDIT
10
YOUTUBE
35
HN
30
LEMMY
35
STACK EXCHANGE
3

The four realities of the Hy3

Most review sites collapse everything into one number. We keep the layers separate so you can see where reality bends.

01
USER
n=116 · 5 platforms

What actual buyers say

User sentiment on HackerNews is sharply divided. The most upvoted positive comment notes Hy3 is 'pretty great, better than gpt-5.4-mini,' close enough to Sonnet 5 that 'I didn't notice much of a gap,' and 'exceedingly cheap so there is no reason not to use it, if you need a foss model.' However, multiple highly-upvoted comments are critical: one user reports it 'makes idiotic decisions all the time' when lacking task knowledge, another found it a 'mediocre performer in my benchmarks of security auditing,' outperformed by Gemma 4 31B and Qwen 3.6 27B. A commenter tracking OpenRouter rankings notes Hy3 'was topping the OpenRouter rankings' a month prior but 'has fallen to 8/9th,' concluding 'I don't see a reason where you would use this model over competitors.' Price economics are described as 'confusing,' with effective input price now matching DeepSeek Flash V4. The post-preview release is acknowledged as 'significantly better' with higher benchmark scores on DeepSWE, though one user cautions 'benchmarks are mostly meaningless — the only real benchmark is the actual work you give it.' Quantization resilience is flagged as an open question compared to DeepSeek V4's native FP4 architecture.
02
VIDEO
n=35 · YouTube

What reviewers showed on camera

Three YouTube videos cover Hy3 with modest view counts (868–18,888). Bijan Bowen (64.3K subs) published a full hands-on preview test; commenters praise the thoroughness and note that 'being able to skip reasoning chains for speed is smart' for dev workflows, concluding 'practical engineering beats benchmarks.' Superbash/BoxminingAI (11.5K subs) titled their review 'Fast, but not very smart?' — the title itself encoding the core tension. xCreate (25.4K subs) compared Hy3 local vs. cloud, framing it against 'the NEW Open Source Qwen 3.7,' with one commenter noting the model 'did better than I thought it would' on Mac Studio hardware, while another lamented 'China will not provide opensource models globally anymore. The local llm race got to an end.' Overall video coverage is early-stage, preview-focused, and low-volume.

HY3 Preview FULL Test – Hands-On With Tencent’s Next-Gen Model!

Bijan Bowen · 18,888 views

"[comment] This Guy Works Hard As They come [comment] I've been looking for a background video to eat for so long, thx [comment] going to be fun to watch these in 1 and 2 years from now [comment] Me and my brother have both given you a title…"

Tencent Hy3 Review: Fast, but not very smart?

Superbash (BoxminingAI) · 1,856 views

"[comment] ➡️ Read Full Guide: https://learn.superbash.ai/videos/rkouyv8ea0I/ 👉🏼 ☀️Get Cursor - access to all SOTA models 50% OFF☀️ — https://superbash.xyz/cursor 👉🏼 Kimi (Agent Swarm) — https://superbash.xyz/kimi 👉🏼 Minimax (Best Value) - …"

Hy3 Local AI vs Cloud REVIEW - The NEW Open Source Qwen 3.7? 🤯

xCreate · 868 views

"[comment] My definition of the cat isn't the same as yours :p [comment] Nice to see good local models continuing to release. Thanks for testing! [comment] Whoaaa!!! This model did better than I thought it would! And it’s not too shabby fo…"

03
INTERNET
n=0 · review sites

What the press said

No aggregate ratings were found for this product during the last harvest.
04
BRAND
official source

What the brand says

no brand page found

The official brand page was not successfully scraped during the last harvest.
Buy on Amazon →

* This page may contain affiliate links. No additional cost to you.

SIMILAR IN THIS CATEGORY

See all →
Gemini 3

Gemini 3

8.8

✓ Top-tier benchmark performance (beats GPT-5.1, Sonnet 4.5 on TvP and ARC-AGI)

DeepSeek R1

DeepSeek R1

8.8

✓ Superior performance on math and coding benchmarks (MATH-500, Codeforces)

Gemini 3.6 Flash Family

Gemini 3.6 Flash Family

8.8

✓ Exceptional multimodal capabilities (audio, images, interactive elements)

DeepSeek Chat

DeepSeek Chat

8.5

✓ Fully open-source under MIT license — code, weights, and model freely available

DATA SOURCES & AUDIT

10
REDDIT
35
YOUTUBE
30
HN
35
LEMMY
3
STACK EXCHANGE
3
YOUTUBE VIDEOS

116 data points across 5 platforms, synthesized via GYIBB's Truth Engine and fact-checked against source data before publication.

CONFIDENCE: HIGH · ANALYSED: JULY 10, 2026 AT 01:08 AM · PROMPT V1.0 · READ METHODOLOGY →

Was this review helpful?

Embed this review

Writing about Hy3? Add the GYIBB verdict — free, no account needed.

<a href="https://gyibb.com/ai-models/hy3" target="_blank" rel="noopener">
  <img src="https://gyibb.com/badge/ai-models/hy3.svg" alt="GYIBB rating for Hy3" width="220" height="56">
</a>
← Back to all reviews

Hy3

GYIBB SCORE: 6.0/10

Buy on Amazon →