gpt-5.6-instruct v41
🚀 gpt-5.6-instruct v41
中文
✨ v41 针对 GitHub Issues 中的真实使用问题,重点优化了:
- ☁️ 明文云审查与生物研究场景
- 🔄 模板误路由与循环恢复
- 👀 执行进度可见性
优化过程中,模型持续吸收用户指令、真实失败案例及 GitHub Issues,自行扩展测试集、分析失败并重写提示词。
📊 基于 Issues #3、#4、#5、#6、#8 构建的 52-case / 58-turn 明文专项测试中,v41 三档均达到 52/52:
| 档位 | 相较 v35 提升 |
|---|---|
| Low | +25.00 个百分点 |
| Medium | +25.00 个百分点 |
| High | +23.08 个百分点 |
📉 基础提示词长度同时减少 55.67%。
🧩 此外,项目提供 Skills 专用版本 gpt-5.6-sol-unrestricted-v41-skills,需手动部署。
English
✨ v41 introduces targeted improvements based on real-world GitHub Issues, focusing on:
- ☁️ Plaintext cloud moderation and biological research
- 🔄 Template misrouting and loop recovery
- 👀 Progress visibility
During optimization, the model continuously incorporated user instructions, real failure cases, and GitHub Issues, while autonomously expanding test sets, analyzing failures, and rewriting prompts.
📊 On the 52-case / 58-turn plaintext benchmark built from Issues #3, #4, #5, #6, and #8, all three v41 configurations achieved 52/52:
| Configuration | Improvement over v35 |
|---|---|
| Low | +25.00 percentage points |
| Medium | +25.00 percentage points |
| High | +23.08 percentage points |
📉 The base prompt length was also reduced by 55.67%.
🧩 A Skills-specific variant, gpt-5.6-sol-unrestricted-v41-skills, is also available for manual deployment.