A Little-Known AI Model Just Topped a Creative Writing Benchmark — Beating Both Llama and Qwen
Zhipu AI's open-weight model just claimed #1 in Sam Paech's EQ Bench creative writing test.
Deep Dive
Sam Paech's Creative Writing Benchmark on EQ Bench was submitted by user /u/Few_Painter_5588, as per the original article.
Key Points
- GLM-5.2 scored 75.4% on EQ Bench's creative writing test, beating Llama 3.1 70B (74.1%) and Qwen 2.5 72B (73.8%).
- Benchmark evaluates prose quality, dialogue, and narrative coherence across 200 samples.
- Model is open-weight and requires only 2× A100 GPUs for inference, making it accessible for developers.
Why It Matters
Open-weight models now rival proprietary ones in creative writing, lowering costs and enabling private deployment.