Suggestion

#11
by chenmingchang - opened

Very good success, the situation of overthinking has been alleviated to some extent, but there should still be room for improvement? In addition, I found that the MTP reception rate of this version is a bit low. I don't know if there will be any special optimization done in the future

Thank you:

RE: MTP;
Based on my own testing and user feedback (Cold Fusion) there seems to be wide range here.
Qwens (all of them) generally have 60% ish MTP average - this is both tuned and un-tuned.

However the following factors affect this:

1 - Use cases , more creative => LOWER MTP, less creative => HIGHER.
2 - Quant(s) do play a factor.
3 - Rep pen and temp => Lower temp => HIGHER MTP acceptance ; temp over 1 OR rep pen on / above 1 => Kills MTP acceptance like a rock.
4 - Special factor -> Reasoning level

Very good success, the situation of overthinking has been alleviated to some extent, but there should still be room for improvement? In addition, I found that the MTP reception rate of this version is a bit low. I don't know if there will be any special optimization done in the future

"--spec-draft-p-min", "0.85",
use this bro,this is top -p for mtp draft

Sign up or log in to comment