Hot eventLive
vLLM集成Helion线性后端提升推理性能
1 reports1 sources4 hr ago updated
Get the story
AI overview
vLLM 集成 Helion 构建高性能可移植线性后端,由 Sean Chen 与 Shangdi Yu 完成。该后端通过单一 GEMM 实现覆盖 Standard、Split-K 与 Swap-AB 三种变体,并采用 per-shape AOT autotuning 自动选择最优配置,以提升推理性能。
Generated from reports · updated 4 hr ago
Timeline
Follow the coverage from different angles.
Oct 3, 2026
- PyTorch Blog精选vLLM 集成 Helion 构建高性能可移植线性后端
Sean Chen 与 Shangdi Yu 将 Helion 集成到 vLLM 线性后端,用单一 GEMM 实现覆盖 Standard、Split-K 与 Swap-AB 变体,并通过 per-shape AOT autotuning 自动选择最优配置。
Heat trend
Current heat 9·Comparable peak 10(Oct 3)·Comparable change over 24 hours –
The trend compares only the same participants observed continuously; its range may be smaller than the current heat count. Move or click on the chart to inspect hourly heat; use the left and right arrow keys to switch.