全部/科技/实时热榜

LMSYS Blog · 实时热榜

HISTORY2026年8月7日32 不同热搜
07/2308/21 有历史数据
DAILY UNIQUE TOPICS32 个热搜
  1. 01
    SpecForge v0.3.0: a Unified Disaggregated and Colocated Speculative Decoding Stack, and New Open SpecBundle Draft Models - LMSYS Org

    SpecForge v0.3.0: a Unified Disaggregated and Colocated Speculative Decoding Stack, and New Open SpecBundle Draft Models LMSYS Org

    最高第 100:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  2. 02
    Full-Stack Performance Optimization of AR+DiT in SGL-Diffusion - LMSYS Org

    Full-Stack Performance Optimization of AR+DiT in SGL-Diffusion LMSYS Org

    最高第 110:59 达到10:59 首次观测上榜当日结束时仍在榜累计约12小时6分
  3. 03
    RadixArk Joins Forces with Google to Bring Full SGLang Features to TPUs - LMSYS Org

    RadixArk Joins Forces with Google to Bring Full SGLang Features to TPUs LMSYS Org

    最高第 200:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  4. 04
    Towards Blackwell-Native 8-bit and 4-bit RL: End-to-End MXFP8 and NVFP4 RL in Miles - LMSYS Org

    Towards Blackwell-Native 8-bit and 4-bit RL: End-to-End MXFP8 and NVFP4 RL in Miles LMSYS Org

    最高第 300:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  5. 05
    Toward a Cleaner Quantization Stack in SGLang - LMSYS Org

    Toward a Cleaner Quantization Stack in SGLang LMSYS Org

    最高第 400:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  6. 06
    SGLang and Miles Add Day-0 Support for Kimi K3 - LMSYS Org

    SGLang and Miles Add Day-0 Support for Kimi K3 LMSYS Org

    最高第 500:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  7. 07
    OPD Support in Miles - LMSYS Org

    OPD Support in Miles LMSYS Org

    最高第 600:00 达到当日首次采集时已在榜10:59 观测离榜累计约10小时59分
  8. 08
    Serving GLM5.2 NVFP4 Agentic Workload with SGLang: Reaching 500 TPS in 2 Weeks - LMSYS Org

    Serving GLM5.2 NVFP4 Agentic Workload with SGLang: Reaching 500 TPS in 2 Weeks LMSYS Org

    最高第 700:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  9. 09
    Bringing DeepSeek-V4 Flash RL Training to AMD Instinct MI355X GPUs with Miles - LMSYS Org

    Bringing DeepSeek-V4 Flash RL Training to AMD Instinct MI355X GPUs with Miles LMSYS Org

    最高第 800:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  10. 10
    Accelerating SGLang HiCache with Netpreme X-Mem™ MPU - LMSYS Org

    Accelerating SGLang HiCache with Netpreme X-Mem™ MPU LMSYS Org

    最高第 900:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  11. 11
    DSpark in SGLang: Speculative Decoding with Confidence-Driven, Variable-Length Verification - LMSYS Org

    DSpark in SGLang: Speculative Decoding with Confidence-Driven, Variable-Length Verification LMSYS Org

    最高第 1000:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  12. 12
    Agent-Assisted SGLang Development: An Initial Exploration - LMSYS Org

    Agent-Assisted SGLang Development: An Initial Exploration LMSYS Org

    最高第 1100:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  13. 13
    Improving DeepEP MoE Load Balance in SGLang with Waterfill and LPLB - LMSYS Org

    Improving DeepEP MoE Load Balance in SGLang with Waterfill and LPLB LMSYS Org

    最高第 1200:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  14. 14
    MOSS-TTS Local Transformer v1.5 on SGLang-Omni: Serving Native-Streaming 48 kHz Speech - LMSYS Org

    MOSS-TTS Local Transformer v1.5 on SGLang-Omni: Serving Native-Streaming 48 kHz Speech LMSYS Org

    最高第 1300:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  15. 15
    Optimizing Ling-2.6-1T on TPU with SGLang-JAX: Hiding MoE Data Movement Behind Compute with One Pallas Kernel - lmsys.org

    Optimizing Ling-2.6-1T on TPU with SGLang-JAX: Hiding MoE Data Movement Behind Compute with One Pallas Kernel lmsys.org

    最高第 1400:00 达到当日首次采集时已在榜23:05 观测离榜累计约23小时5分
  16. 16
    The next generation of speculative decoding: DFlash and Spec V2 - LMSYS Org

    The next generation of speculative decoding: DFlash and Spec V2 LMSYS Org

    最高第 1423:05 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  17. 17
    No Token Left Behind: Demystifying Token-In-Token-Out in Miles - LMSYS Org

    No Token Left Behind: Demystifying Token-In-Token-Out in Miles LMSYS Org

    最高第 1523:05 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  18. 18
    Higgs Audio v3 TTS on SGLang-Omni: Real-Time, Controllable Speech for Voice Agents - LMSYS Org

    Higgs Audio v3 TTS on SGLang-Omni: Real-Time, Controllable Speech for Voice Agents LMSYS Org

    最高第 1623:05 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  19. 19
    SGLang and Miles Add Day-0 Support for NVIDIA Nemotron 3 Ultra for Long-Running Autonomous Agents - LMSYS Org

    SGLang and Miles Add Day-0 Support for NVIDIA Nemotron 3 Ultra for Long-Running Autonomous Agents LMSYS Org

    最高第 1723:05 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  20. 20
    Heterogeneous CPU + GPU EPD Disaggregation to Boost VLM Serving - LMSYS Org

    Heterogeneous CPU + GPU EPD Disaggregation to Boost VLM Serving LMSYS Org

    最高第 1823:05 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  21. 21
    Win on TCO: How AMD Instinct™ MI355X Achieves Cost-Competitive Distributed Inference Through SGLang with MoRI - LMSYS Org

    Win on TCO: How AMD Instinct™ MI355X Achieves Cost-Competitive Distributed Inference Through SGLang with MoRI LMSYS Org

    最高第 1923:05 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  22. 22
    Updating 1T parameters in seconds — P2P weight transfer in Large Scale Distributed RL - LMSYS Org

    Updating 1T parameters in seconds — P2P weight transfer in Large Scale Distributed RL LMSYS Org

    最高第 2023:05 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  23. 23
    DeepSeek-V4 on Day 0: From Fast Inference to Verified RL with SGLang and Miles - LMSYS Org

    DeepSeek-V4 on Day 0: From Fast Inference to Verified RL with SGLang and Miles LMSYS Org

    最高第 2123:05 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  24. 24
    HiSparse: Turbocharging Sparse Attention with Hierarchical Memory - LMSYS Org

    HiSparse: Turbocharging Sparse Attention with Hierarchical Memory LMSYS Org

    最高第 2223:05 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  25. 25
    Highlights of SGLang at NVIDIA GTC 2026 - LMSYS Org

    Highlights of SGLang at NVIDIA GTC 2026 LMSYS Org

    最高第 2323:05 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  26. 26
    Elastic EP in SGLang: Achieving Partial Failure Tolerance for DeepSeek MoE Deployments - LMSYS Org

    Elastic EP in SGLang: Achieving Partial Failure Tolerance for DeepSeek MoE Deployments LMSYS Org

    最高第 2423:05 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  27. 27
    ROCm Support for Miles: Large-Scale RL Post-Training on AMD Instinct™ GPUs - LMSYS Org

    ROCm Support for Miles: Large-Scale RL Post-Training on AMD Instinct™ GPUs LMSYS Org

    最高第 2523:05 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  28. 28
    Unlocking 25x Inference Performance with SGLang on NVIDIA GB300 NVL72 - LMSYS Org

    Unlocking 25x Inference Performance with SGLang on NVIDIA GB300 NVL72 LMSYS Org

    最高第 2623:05 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  29. 29
    Deploying DeepSeek on GB300 NVL72: Big Wins in Long-Context Inference - LMSYS Org

    Deploying DeepSeek on GB300 NVL72: Big Wins in Long-Context Inference LMSYS Org

    最高第 2723:05 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  30. 30
    SGLang-Diffusion: Advanced Optimizations for Production-Ready Video Generation - LMSYS Org

    SGLang-Diffusion: Advanced Optimizations for Production-Ready Video Generation LMSYS Org

    最高第 2823:05 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  31. 31
    Unleashing Computational Power: Ultimate Latency Optimization of Qwen3 and Qwen3-VL on AMD MI300X Series - LMSYS Org

    Unleashing Computational Power: Ultimate Latency Optimization of Qwen3 and Qwen3-VL on AMD MI300X Series LMSYS Org

    最高第 2923:05 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时5分
  32. 32
    Squeezing 1TB Model Rollout into a Single H200: INT4 QAT RL End-to-End Practice - LMSYS Org

    Squeezing 1TB Model Rollout into a Single H200: INT4 QAT RL End-to-End Practice LMSYS Org

    最高第 3023:05 达到23:05 首次观测上榜当日结束时仍在榜累计约0分钟