| Qwen 3.8 follows GPT-5.5 Pro reasoning prefills(gist.github.com) | |
| 232 points by wsxiaoys 22 hours ago | 90 comments | |
tl;dr: When prefilled with the first 1% of GPT-5.5 Pro's reasoning, Qwen3.8 A95B's output overlap with the teacher's answer jumped +18.18 points (including +27 on STEM), while DeepSeek V4 Flash, Inkling, and Kimi K3 barely budged. Since Qwen showed little response to Opus 4.8 prefills in a prior experiment, the author suggests Qwen may have been trained on outputs from GPT-5.5 Pro or a closely related model. | |
HN Discussion:
| |