Tag
Meta Engineering extends FlashAttention-4 with MXFP8 support for NVIDIA Blackwell, achieving up to 2.85 PFLOP/s forward performance and integrating into production training workflows like GEM.
DealignAI releases CRACK-abliterated and MXFP4/MXFP8 quantized versions of Qwen3.6-27B and 35B models, preserving MTP for faster speculative decoding on Apple Silicon.