Skip to content
Hot eventLive

vLLM集成Helion线性后端提升推理性能

1 reports1 sources4 hr ago updated

Get the story

AI overview

vLLM 集成 Helion 构建高性能可移植线性后端,由 Sean Chen 与 Shangdi Yu 完成。该后端通过单一 GEMM 实现覆盖 Standard、Split-K 与 Swap-AB 三种变体,并采用 per-shape AOT autotuning 自动选择最优配置,以提升推理性能。

Generated from reports · updated 4 hr ago

Timeline

Follow the coverage from different angles.

Oct 3, 2026
  1. PyTorch Blog精选
    vLLM 集成 Helion 构建高性能可移植线性后端

    Sean Chen 与 Shangdi Yu 将 Helion 集成到 vLLM 线性后端,用单一 GEMM 实现覆盖 Standard、Split-K 与 Swap-AB 变体,并通过 per-shape AOT autotuning 自动选择最优配置。

Heat trend

Current heat 9·Comparable peak 10(Oct 3)·Comparable change over 24 hours –

02.557.510Oct3Oct3Oct3Oct3

The trend compares only the same participants observed continuously; its range may be smaller than the current heat count. Move or click on the chart to inspect hourly heat; use the left and right arrow keys to switch.