Will there be a 40B version? [versions in testing]

#8
by basiteis - opened

I'm looking forward to the 40B version. It works fine right now, but it still feels like something is missing—I wonder if there will be a "40B711" release?

Testing the prototypes today.
There are a lot of additional steps required to get these to 700+; as there is a drop off (known issue) at the expansion step.
This issue must be addressed post expansion.

IT begins:

https://huggingface.co/DavidAU/Qwen3.6-40B-Grand-Intelligence-One

Composed of TWO 700+ ARC-C Monsters.
In testing.

Full example generation on the repo test page.
This is a take no prisoners version.

Sample:

"Listen up, you trembling little scribe with the sticky fingers and the big dreams. You’re standing there, staring at a blank page like a deer caught in the headlights of a freight train made of your own inadequacy, wondering who you’re going to beg to help you craft your masterpiece? You’re thinking about writing it yourself? You’re thinking about hiring some pretentious hack who charges by the word and cries when you ask him to cut his precious exposition? Or worse—hell, worse than worse—you’re thinking about using one of those soulless, sanitized, corporate-approved AI bots that write like they’re afraid to sneeze on a nun?"

...

Update:
Alpha "tune" part one complete; proceeding on testing other alphas / 2nd/3rd level tuning.

Due to issues caused by expansion, the model requires re-tuning to regain "711" type performance/benches.

Benches of expanded [non tuned] and [light / stage tune[s]] are here:

https://huggingface.co/DavidAU/Qwen3.6-40B-Grand-Intelligence-One-IQ-tune-alpha

NOTE: THIS IS A WORK IN PROGRESS ; it will take time to get it right.
NOTE: Name of the REPO for "alpha" will CHANGE as the project proceeds.

DavidAU pinned discussion
DavidAU changed discussion title from Will there be a 40B version? to Will there be a 40B version? [versions in testing]

Update;
One alpha 40B has reached "700 club" status. Work continues.

this is super cool! i applied to access your gated model, have you released it in gguf form yet? regardless, i am curious as to how the performance is on my 128gb vram system i just finished building out!

@cpui686
Thank you ;
40B model build is in progress, however this build has more steps than 27B due in part to expansion and extended testing.
Benching 40Bs take twice a long too, per test.

As of this writing there will likely be at least 2 versions, maybe more - a general version plus specific/specialized use case(s).

aaah i see, yea that makes a lot of sense! keep up the good work!

Sign up or log in to comment