Commit History

Repoint build instructions to charlie12345/ROCmFPX (ROCmFPX FP3/4/6/8 repo)
bf44fb3
verified

plunderstruck commited on

Upload README.md with huggingface_hub
2e42117
verified

plunderstruck commited on

consolidate to single best-balance rocmfp4 quant
62a67a6
verified

plunderstruck commited on

consolidate to single best-balance rocmfp4 quant
d89404f
verified

plunderstruck commited on

consolidate to single best-balance rocmfp4 quant
00b6fd8
verified

plunderstruck commited on

consolidate to single best-balance rocmfp4 quant
774e9f5
verified

plunderstruck commited on

Upload README.md with huggingface_hub
db0798f
verified

plunderstruck commited on

Upload README.md with huggingface_hub
a93b186
verified

plunderstruck commited on

Add mmproj-F32.gguf (best-of-both / vision)
bf22776
verified

plunderstruck commited on

Add Qwen3.6-40B-Deckard-MTP-ROCmFP4-COHERENT-embQ8-headQ6.gguf (best-of-both / vision)
6a826b0
verified

plunderstruck commited on

Add Qwen ASCII logo header (centered logo+title, full-width spec table)
2201c1f
verified

plunderstruck commited on

Deckard: neutralize last 'off-thinking coding' flag wording (config description only)
ca6f9d1
verified

plunderstruck commited on

Claim audit: remove cross-model generalizations (Q6-head %, f16-draft-head, decode comparisons) — state per-model config only, no borrowed sibling measurements
5e243c8
verified

plunderstruck commited on

Deckard: soften unverified 'does better off-thinking' claim — present off-thinking as an option, attribute the trend to the Qwen3.6 coder family (cited), note it's not measured on Deckard
c64ee80
verified

plunderstruck commited on

Card: datasheet/neo-brutalist restyle (monospace, hard borders, spec grid, numbered sections); claims audited (f16 KV stated as config only)
e5a9ff3
verified

plunderstruck commited on

Card: note bundled chat_template.jinja (froggeric unified template) for tool calls + inline think-toggle + vision
38f0c22
verified

plunderstruck commited on

Add froggeric v20 fixed chat template (vision + inline think-toggle + tool anti-stall); use via --chat-template-file
8e68a84
verified

plunderstruck commited on

Card: add Vision section (Qwen3-VL mmproj + --image-min-tokens 1024 fix for incorrect image reads)
8a8bad0
verified

plunderstruck commited on

provenance: set base_model_relation=quantized + repoint base_model at actual source
873c99b
verified

plunderstruck commited on

Card fix: imatrix variant sizes are 21.8GB (4-bit head) / 22.2GB (Q6 head), not both 21GB
7d3e25e
verified

plunderstruck commited on

Drop mtpF16 (f16 MTP draft head measured as a wash vs 4-bit)
98c0a97
verified

plunderstruck commited on

Card: add imatrix variants + measured KL/PPL; mark imatrix recommended; drop mtpF16
399eb86
verified

plunderstruck commited on

Add imatrix headQ6 variant (NEO-style imatrix + Q6_K output head)
38901e5
verified

plunderstruck commited on

Add imatrix base variant (NEO-style general+code imatrix; measured -10% median KL, +0.8pp top-token vs non-imatrix)
e0bd769
verified

plunderstruck commited on

Card: controlled test — f16 MTP head is a wash (~0.83 both); mtpF16 redundant, use base
4652708
verified

plunderstruck commited on

Card: f16-head warm acceptance ~0.85 sustained (cold 0.77 undersold it); not an underperformer
9677f59
verified

plunderstruck commited on

Card: honest f16-head result (did NOT beat 4-bit Qwopus-donor head; donor match > precision)
bfc7f3a
verified

plunderstruck commited on

Card: document the f16-MTP-head variant (3 variants)
e341888
verified

plunderstruck commited on

Add f16-MTP-head variant (BF16 27B donor head kept at f16; draft acceptance rate = 0.76852)
3c15a6e
verified

plunderstruck commited on

Add model card (lineage, functionally-verified status, block_count gotcha)
89096dd
verified

plunderstruck commited on

Add Q6-head variant
1b6c7bb
verified

plunderstruck commited on

Add base variant (ROCmFP4 STRIX + Q8 emb; transplanted MTP head)
ebc4309
verified

plunderstruck commited on