HOME REVIEWS AI MODELS

82 reviews. AI Models.

AVG SCORE: 6.9/10

82 of 1403 reviews AI Models × Clear all
21
GLM-5.3-Flash G
GLM-5.3-Flash
AI Models 52 sources · 60-89%

"Offers high intelligence per dollar, but suffers from high latency and strict content censorship, making it less idea…"

7.2
/10
27 AUG
22
Qwen3.8 Flash Q
Qwen3.8 Flash
AI Models 49 sources · 60-89%

"Alibaba's open-weight MoE with strong user-reported benchmarks, but overthinking, a 75GB+ local footprint and immatur…"

7.5
/10
27 AUG
23
Microsoft Phi-4 M
Microsoft Phi-4
AI Models 30 sources · <60%

"Efficient 14B SLM praised for math reasoning and local performance but criticized for limited context and coding reli…"

8.0
/10
27 AUG
24
DeepSeek V4 Flash Vision Exp D
DeepSeek V4 Flash Vision Exp
AI Models 87 sources · 60-89%

"Experimental vision variant of DeepSeek's speed-optimized V4 Flash: strong daily-driver coding reports, but image acc…"

7.5
/10
22 AUG
25
Kimi K2 K
Kimi K2
AI Models 134 sources · 90%+

"Users rank Moonshot's open-weights model near Claude for agentic coding, but hard refusals on sensitive topics and pr…"

6.5
/10
21 AUG
26
Claude Watermark C
Claude Watermark
AI Models 126 sources · 90%+

"Anthropic adds SynthID-style statistical watermarking to Claude output; users debate defeatability, false negatives a…"

7.0
/10
19 AUG
27
GLM 5.3 G
GLM 5.3
AI Models 62 sources · <60%

"Z.ai's post-trained GLM-5.2 base scores near-frontier coding results at lower cost per task, but is the token-hungrie…"

8.6
/10
19 AUG
28
GLM-5.2 G
GLM-5.2
AI Models 91 sources · <60%

"Z.ai's 744B open-weight MoE wows in YouTube agentic tests, but Hacker News users call it good-not-great and brutal to…"

8.0
/10
17 AUG
29
Al GLM-5 A
Al GLM-5
AI Models 56 sources · 60-89%

"Coding-focused LLM from Z.ai. Reviewers crown it 'Coding King' and users rate it near Opus/Gemini Pro, but speed and …"

7.5
/10
17 AUG
30
nenspace n
nenspace
AI Models 12 sources · <60%

"A fine-tuned LLM built to question assumptions and answer concisely. Launch-day data only: no independent reviews, be…"

2.5
/10
15 AUG
31
Mistral Large 3 M
Mistral Large 3
AI Models 105 sources · <60%

"675B MoE open-weights release pitched as a DeepSeek rival; local users report ~3 t/s at Q4 and doubt Mistral's edge."

6.5
/10
15 AUG
32
Gemini 3.7 Flash G
Gemini 3.7 Flash
AI Models 26 sources · <60%

"Google's workhorse LLM posts big benchmark jumps at half-price tokens, but early users report hallucinations and a be…"

7.0
/10
14 AUG
33
Grok Bot G
Grok Bot
AI Models 451 sources · 90%+

"One credible month-long user praises the multi-bot agent paradigm; cost burn, a privacy allegation and anti-Musk nois…"

6.0
/10
14 AUG
34
Nemotron 3.5 Lightning N
Nemotron 3.5 Lightning
AI Models 43 sources · 60-89%

"Fast open-source MoE LLM (~100-135 TPS on consumer Macs) with strong generalization but inconsistent coding and bench…"

6.5
/10
13 AUG
35
Grok 4.6 G
Grok 4.6
AI Models 93 sources · 90%+

"xAI's LLM offers genuine model diversity and competitive benchmarks, but trust issues and API reliability concerns dr…"

6.5
/10
13 AUG
36
DeepSeek V4 Pro 0813 D
DeepSeek V4 Pro 0813
AI Models 95 sources · 60-89%

"Open-weight frontier model praised for reasoning and cost-efficiency vs GLM-5.2/Opus-4.8, but pending API price hike …"

8.2
/10
13 AUG
37
Qwen3.8 2.4T A95B Q
Qwen3.8 2.4T A95B
AI Models 53 sources · <60%

"Massive MoE LLM rivaling Claude Opus and DeepSeek v4 Pro. Open weights are a landmark, but 4.9TB BF16 means only well…"

8.5
/10
13 AUG
38
Qwen3 Q
Qwen3
AI Models 251 sources · <60%

"Alibaba's open-weights LLM series with strong coding benchmarks and flexible quantization, but hampered by political …"

7.0
/10
10 AUG
39
Grok 4 G
Grok 4
AI Models 515 sources · <60%

"xAI's Grok 4 impresses in video coding tests, but $300/mo Heavy plan and training-data ethics raise real questions."

6.0
/10
7 AUG
40
Muse Code M
Muse Code
AI Models 386 sources · 90%+

"Meta's coding-focused LLM offers aggressive pricing but users flag trust issues, login friction, and benchmark vs. re…"

5.0
/10
6 AUG