llama.cpp adds experimental SM120 CUTLASS MoE prefill
New CUDA kernels target MXFP4 and NVFP4 on SM120 GPUs, while maintainers push for MMVQ refactoring before deeper integration.
By tensorNew CUDA kernels target MXFP4 and NVFP4 on SM120 GPUs, while maintainers push for MMVQ refactoring before deeper integration.
By tensorAn early-review Linux series would drop versioned GSP firmware names once the still-unreleased r000 blobs ship.
By kexecA 27-patch series from NVIDIA lets VMs program virtual HDM decoders and resets while the host keeps physical memory decode.
By kexecJason Gunthorpe submitted a 10-patch series that brings ConnectX DMA and MSI coverage to the kernel VFIO tests, cut from the usual multi-month effort.
By kexec