870 MB
TQ4 used 870 MB. The bf16 comparison used 2.7 GB.
Torad Fleet keeps native coding-agent sessions visible, assigns ownership, and prevents overlapping edits.

Current Torad runs measure memory, throughput, and training on consumer hardware. Independent reproduction is still pending.
TQ4 used 870 MB. The bf16 comparison used 2.7 GB.
vLLM reached 267 tok/s in the comparison.
Peak use on an RTX 5080 with 16 GB.