16.0%
Requested-length error · rank 127 of 370
- Unit
- % off the requested word count, lower is better
- Configuration
- Qwen3.6 35B A3B (think-prefill reasoning)
- Measured
- 25 Apr 2026
- Not shown
- Not other format limits such as character counts or bullet counts.
10.0%
Requested-length error · rank 81 of 370
- Unit
- % off the requested word count, lower is better
- Configuration
- Qwen3.6 35B A3B (no reasoning)
- Measured
- 22 Apr 2026
- Not shown
- Not other format limits such as character counts or bullet counts.
0.28
Style adherence · rank 361 of 370
- Unit
- score from 0 to 1, higher is better
- Configuration
- Qwen3.6 35B A3B (think-prefill reasoning)
- Measured
- 25 Apr 2026
- Not shown
- Not brand voice on your own examples: UGI's prompts are private and lean towards creative writing.
0.32
Style adherence · rank 285 of 370
- Unit
- score from 0 to 1, higher is better
- Configuration
- Qwen3.6 35B A3B (no reasoning)
- Measured
- 22 Apr 2026
- Not shown
- Not brand voice on your own examples: UGI's prompts are private and lean towards creative writing.
45.4
Writing score · rank 176 of 370
- Unit
- score out of 100, higher is better
- Configuration
- Qwen3.6 35B A3B (think-prefill reasoning)
- Measured
- 25 Apr 2026
- Not shown
- Not business copy quality: UGI's prompts are private and lean towards creative writing, and models that often refuse get no score.
35.8
Writing score · rank 235 of 370
- Unit
- score out of 100, higher is better
- Configuration
- Qwen3.6 35B A3B (no reasoning)
- Measured
- 22 Apr 2026
- Not shown
- Not business copy quality: UGI's prompts are private and lean towards creative writing, and models that often refuse get no score.