Extreme low-bit inference lab: pack weights to 1–1.58 bits and run real XNOR/popcount kernels on CPU/edge — honest STE training, not fake GPU 32× from sign().

binary-neural-networks cpu-inference edge-ai machine-learning model-compression pytorch quantization ternary xnor
2 Open Issues Need Help Last updated: Jul 27, 2026

Open Issues Need Help

View All on GitHub

Extreme low-bit inference lab: pack weights to 1–1.58 bits and run real XNOR/popcount kernels on CPU/edge — honest STE training, not fake GPU 32× from sign().

Python
#binary-neural-networks#cpu-inference#edge-ai#machine-learning#model-compression#pytorch#quantization#ternary#xnor
documentation good first issue

Extreme low-bit inference lab: pack weights to 1–1.58 bits and run real XNOR/popcount kernels on CPU/edge — honest STE training, not fake GPU 32× from sign().

Python
#binary-neural-networks#cpu-inference#edge-ai#machine-learning#model-compression#pytorch#quantization#ternary#xnor