Files
nvfp4-megamoe-kernel/tests/unit/test_nvfp4_quant_kernel.py
biondizzle 6504f091ca NVFP4-1.1 Step 3: post-SWiGLU quantization test suite (all PASS)
- Standalone kernel cos 0.979 (128x512)
- Post-SwiGLU quantization cos 0.976 (vs Python 0.995)
- Larger shape cos 0.979 (512x4096)
- FP8 scale match 100% across all tests
- GPU kernel replaces CPU-GPU sync quantize path
- Ready for integration into MoE pipeline
2026-05-25 09:08:01 +00:00

9.4 KiB