Repoint build instructions to charlie12345/ROCmFPX (ROCmFPX FP3/4/6/8 repo) bf44fb3 verified plunderstruck commited on Jun 21
Add Qwen3.6-40B-Deckard-MTP-ROCmFP4-COHERENT-embQ8-headQ6.gguf (best-of-both / vision) 6a826b0 verified plunderstruck commited on Jun 15
Add Qwen ASCII logo header (centered logo+title, full-width spec table) 2201c1f verified plunderstruck commited on Jun 15
Deckard: neutralize last 'off-thinking coding' flag wording (config description only) ca6f9d1 verified plunderstruck commited on Jun 14
Claim audit: remove cross-model generalizations (Q6-head %, f16-draft-head, decode comparisons) — state per-model config only, no borrowed sibling measurements 5e243c8 verified plunderstruck commited on Jun 14
Deckard: soften unverified 'does better off-thinking' claim — present off-thinking as an option, attribute the trend to the Qwen3.6 coder family (cited), note it's not measured on Deckard c64ee80 verified plunderstruck commited on Jun 14
Card: datasheet/neo-brutalist restyle (monospace, hard borders, spec grid, numbered sections); claims audited (f16 KV stated as config only) e5a9ff3 verified plunderstruck commited on Jun 14
Card: note bundled chat_template.jinja (froggeric unified template) for tool calls + inline think-toggle + vision 38f0c22 verified plunderstruck commited on Jun 14
Add froggeric v20 fixed chat template (vision + inline think-toggle + tool anti-stall); use via --chat-template-file 8e68a84 verified plunderstruck commited on Jun 14
Card: add Vision section (Qwen3-VL mmproj + --image-min-tokens 1024 fix for incorrect image reads) 8a8bad0 verified plunderstruck commited on Jun 14
provenance: set base_model_relation=quantized + repoint base_model at actual source 873c99b verified plunderstruck commited on Jun 13
Card fix: imatrix variant sizes are 21.8GB (4-bit head) / 22.2GB (Q6 head), not both 21GB 7d3e25e verified plunderstruck commited on Jun 13
Drop mtpF16 (f16 MTP draft head measured as a wash vs 4-bit) 98c0a97 verified plunderstruck commited on Jun 12
Card: add imatrix variants + measured KL/PPL; mark imatrix recommended; drop mtpF16 399eb86 verified plunderstruck commited on Jun 12
Add imatrix headQ6 variant (NEO-style imatrix + Q6_K output head) 38901e5 verified plunderstruck commited on Jun 12
Add imatrix base variant (NEO-style general+code imatrix; measured -10% median KL, +0.8pp top-token vs non-imatrix) e0bd769 verified plunderstruck commited on Jun 12
Card: controlled test — f16 MTP head is a wash (~0.83 both); mtpF16 redundant, use base 4652708 verified plunderstruck commited on Jun 12
Card: f16-head warm acceptance ~0.85 sustained (cold 0.77 undersold it); not an underperformer 9677f59 verified plunderstruck commited on Jun 12
Card: honest f16-head result (did NOT beat 4-bit Qwopus-donor head; donor match > precision) bfc7f3a verified plunderstruck commited on Jun 12
Card: document the f16-MTP-head variant (3 variants) e341888 verified plunderstruck commited on Jun 12
Add f16-MTP-head variant (BF16 27B donor head kept at f16; draft acceptance rate = 0.76852) 3c15a6e verified plunderstruck commited on Jun 12
Add model card (lineage, functionally-verified status, block_count gotcha) 89096dd verified plunderstruck commited on Jun 12
Add base variant (ROCmFP4 STRIX + Q8 emb; transplanted MTP head) ebc4309 verified plunderstruck commited on Jun 12