avlp12 commited on
Commit
f6b0c61
·
verified ·
1 Parent(s): 56c3da0

Withhold KL-vs-Q8 figures pending re-measurement: the 8-bit reference used as the baseline is defective (mx.split >2^31 silent corruption). Model quality and the port-parity KL~1e-7 (vs the fixed HF torch reference) are unaffected.

Browse files
Files changed (1) hide show
  1. README.md +24 -0
README.md CHANGED
@@ -24,6 +24,30 @@ tags:
24
  - conversational
25
  ---
26
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
27
  # Motif-3-Beta-Alis-MLX-Dynamic-2.3bpw · ⚠️ experimental floor build
28
 
29
  Apple Silicon (MLX) **mixed-precision** quantization of
 
24
  - conversational
25
  ---
26
 
27
+ > [!NOTE]
28
+ > ## 📌 KL 수치 보류 중 / KL numbers under review — the model itself is fine
29
+ >
30
+ > **2026-07-26.** 이 카드의 **"KL vs Q8(8bit 레퍼런스)" 수치는 현재 보류**합니다.
31
+ > 레퍼런스로 쓰인 [8bit 빌드에 결함](https://huggingface.co/avlp12/Motif-3-Beta-Alis-MLX-8bit/discussions/1)이
32
+ > 확인되어(`mx.split` 2³¹ 초과 텐서 침묵 손상), **손상된 레퍼런스를 기준으로 측정된 값**이기 때문입니다.
33
+ >
34
+ > The **"KL vs Q8" figures on this card are withheld pending re-measurement**: the 8-bit
35
+ > build used as the reference has been found defective, so those numbers were measured
36
+ > against a corrupted baseline.
37
+ >
38
+ > **보류되는 것 / withheld:** 슬라이스별 `KL vs Q8` 표 · **"DWQ가 KL을 58–72% 감소"** 주장 ·
39
+ > `eval_ladder.png`의 KL 축
40
+ >
41
+ > **영향 없는 것 / unaffected — 이 빌드 자체는 정상입니다:**
42
+ > - 모델 품질(한국어·영어·코드 생성)은 직접 생성으로 검증되었고 그대로 유효합니다.
43
+ > Generation quality is verified directly and stands.
44
+ > - **포팅 패리티 `KL ≈1e-7 / token`** 은 수정된 **HF torch 레퍼런스** 기준이라 8bit와 무관하며 유효합니다.
45
+ > The port-parity figure is measured against the fixed HF reference, not the 8-bit build.
46
+ > - 루프 프로브(distinct-4gram) 등 **레퍼런스가 필요 없는 측정**은 유효합니다.
47
+ >
48
+ > 8bit 레퍼런스를 재빌드한 뒤 재측정하여 이 카드의 수치를 갱신하고 이 주석을 내리겠습니다.
49
+ > The reference is being rebuilt; these figures will be re-measured and this note removed.
50
+
51
  # Motif-3-Beta-Alis-MLX-Dynamic-2.3bpw · ⚠️ experimental floor build
52
 
53
  Apple Silicon (MLX) **mixed-precision** quantization of