Repository navigation
Commit 88f7d04
Support quantized Gemma 4 per-layer input projections (#762)
Use the decoder quantization factory for PLE gates and projections.
Cover GPTQ and Olive packed weights, component layouts, and
floating-point behavior.
---------
Co-authored-by: Copilot <223556219+Copilot@users.noreply.github.com>1 parent 26a712b commit 88f7d04
1 file changed
Lines changed: 5 additions & 3 deletions
| Original file line number | Diff line number | Diff line change | |
|---|---|---|---|
| |||
13 | 13 | | |
14 | 14 | | |
15 | 15 | | |
16 | | - | |
| 16 | + | |
| 17 | + | |
17 | 18 | | |
18 | 19 | | |
19 | 20 | | |
| |||
1577 | 1578 | | |
1578 | 1579 | | |
1579 | 1580 | | |
1580 | | - | |
| 1581 | + | |
| 1582 | + | |
1581 | 1583 | | |
1582 | 1584 | | |
1583 | | - | |
| 1585 | + | |
1584 | 1586 | | |
1585 | 1587 | | |
1586 | 1588 | | |
| |||
0 commit comments