Unsloth 开源教程:本地训练 Qwen3.5 0.8B 决策模型,准确率从 20.7% 提升至 74.3%
You can now train your own Decision model like Jev locally!
AI 摘要
Unsloth 发布教程与开源仓库,可将 Qwen3.8、Gemma 4 等 LLM 微调为输出选项概率的决策模型,Qwen3.5 0.8B 在 3 个决策基准上的合计准确率从 20.7% 提到 74.3%,仅需 4GB 显存。
正文 · AI 翻译
You can now train your own Decision model like Jev locally!
We increased Qwen3.5 0.8B’s aggregate accuracy from 20.7% to 74.3% across 3 decision benchmarks - on just 4GB VRAM.
Turn any LLM like Qwen3.8, Gemma 4 into decision models with our open-source Unsloth repo.
We fine-tuned with a Clef head using Unsloth and LoRA (r=64) for one epoch, increasing downstream accuracy from 30–37% to 78%.
GitHub: https://github.com/unslothai/unsloth Guide and Notebooks: https://unsloth.ai/docs/basics/train-your-own-decision-model-with-unsloth
原文
Original Title
You can now train your own Decision model like Jev locally!
Source
Unsloth
Site
x.com
Published
2026-10-07T16:21:14.000Z