Altworld/Hemmingway-1 — new model trending #29 on Hugging Face
Altworld releases Hemmingway-1, a 27B open-weight Apache-2.0 model built on Qwen3.8-27B targeting everyday writing tasks.
Altworld released Hemmingway-1, a 27B-parameter open-weight model under Apache-2.0, built on Qwen3.8-27B with a 262,144-token context, tuned for everyday writing such as messages and emails. It claims first place against Fable 5.1, GPT-6 Astra, Kimi K3, GLM-5.3, Grok 4.6 and DeepSeek V4 Pro on its self-built CommunicationBench, and third place on the independent EQ-Bench 4. The model card discloses that CommunicationBench, Human-Likeness and StoryBench are the vendor's own benchmarks, run blind and order-balanced with a third-party judge.
- 27B open-weight Apache-2.0 model built on Qwen3.8-27B with 262,144-token context
- Claims top spot on self-built CommunicationBench, beating Fable 5.1 and GPT-6 Astra
- Placed third on independent EQ-Bench 4, within twelve points of the best model
- Vendor discloses its own benchmarks but claims blind, order-balanced judging
Full article594 words · extracted from huggingface.co · click to collapse
# Hemmingway-1
**The AI that writes like a person.** 27B parameters, open weights, Apache-2.0.
**[Weights →](https://huggingface.co/Altworld/Hemmingway-1)** · **[Try it →](https://hemmingway.io)** · **[Mac and Android apps →](https://hemmingway.io/download)** · **[Code →](https://github.com/lukeckprobierts/Hemmingway-1)**
Ask most models for a text to your landlord and you get three options, a
preamble, and a paragraph explaining the options. Hemmingway-1 just gives you
the text.
We built it for the writing people actually do every day: messages, emails, the
awkward note to a colleague, the thing you have been putting off. Then we
tested it against the biggest models in the world at exactly that, and it came
first.
## It writes the best everyday messages of any model we tested
We tested on eighty real requests. Every answer went head to head against
another model's answer to the same request, shuffled so the judge never knew
which was which.

It beats Fable 5.1, and it beats GPT-6 Astra by fifty points. Kimi K3, GLM-5.3,
Grok 4.6 and DeepSeek V4 Pro all come in behind it. That is a 27B model.
## And it is the one that sounds like a person
We ran the same matchups again with one question: which of these two did a
person write?

It finished twenty-six points clear of the next model.
## Where it wins
We broke it down by what you asked for. Higher means the judge more often took
its version for the one a person wrote.

It wins on money and admin, work, the hard asks you keep rewriting, and talking
someone round. Most of those by a wide margin. On hard asks, GPT-6 Astra gets
9%. Hemmingway-1 gets 72%.
It loses on hostile storytelling and long story turns. The story models are
better at those, and that is fair.
## You get the message, not a memo
One more thing we measured: how often a model buries the actual text in
commentary, options and notes you have to read past.

Fable 5, GLM-5.3 and Kimi K3 bury it in more than nine replies out of ten.
## It reads the room
EQ-Bench 4 is not ours. It is the public emotional-intelligence benchmark,