Together AI·· Jun 23SelectedAI score66
ParallelKernelBench:前沿大语言模型尚不能编写高效多GPU内核
ParallelKernelBench: Frontier LLMs can't write fast multi-GPU kernels (yet)
AI brief
Together AI发布ParallelKernelBench(PKB)开源基准,包含87个来自真实代码库的多GPU内核生成问题,要求模型用CUDA内核替代PyTorch+NCCL并通过NVLink直接通信。
Why it matters
基准揭示前沿模型在多GPU内核生成上正确率与速度优势均有限,为分布式优化研究提供了可复现的测试平台。
Source: Together AI · together.ai