TQ4
Stores weights in 4 bits with a learned scale and codebook while retaining trainable controls above the compact substrate.
- Measured
- 870 MB and 0.9998 cosine.
- Missing
- Full method publication.
One internal run cut linear-weight memory from 2.7 GB to 870 MB while the top answer matched on five prompts.
Qwen3-1.7B output comparison
0.85 out of 1.00
Measured cosine scores: 0.85, 0.91, and 0.9998.
Each method below separates what ran from what remains unpublished.
Stores weights in 4 bits with a learned scale and codebook while retaining trainable controls above the compact substrate.
Uses hardware 4-bit math without expanding every operation back to 16 bits.
Places inference and training in one small program.
Rewards a correct model attempt instead of training only by copying a completed answer.