全部/科技/实时热榜

LMSYS Blog · 实时热榜

HISTORY2026年7月31日34 不同热搜
07/2308/21 有历史数据
DAILY UNIQUE TOPICS34 个热搜
  1. 01
    Blog - LMSYS Org

    Blog LMSYS Org

    最高第 2100:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  2. 02
    Win on TCO: How AMD Instinct™ MI355X Achieves Cost-Competitive Distributed Inference Through SGLang with MoRI - LMSYS Org

    Win on TCO: How AMD Instinct™ MI355X Achieves Cost-Competitive Distributed Inference Through SGLang with MoRI LMSYS Org

    最高第 1700:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  3. 03
    Highlights of SGLang at NVIDIA GTC 2026 - LMSYS Org

    Highlights of SGLang at NVIDIA GTC 2026 LMSYS Org

    最高第 2200:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  4. 04
    DeepSeek-V4 on Day 0: From Fast Inference to Verified RL with SGLang and Miles - LMSYS Org

    DeepSeek-V4 on Day 0: From Fast Inference to Verified RL with SGLang and Miles LMSYS Org

    最高第 1900:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  5. 05
    Updating 1T parameters in seconds — P2P weight transfer in Large Scale Distributed RL - LMSYS Org

    Updating 1T parameters in seconds — P2P weight transfer in Large Scale Distributed RL LMSYS Org

    最高第 1800:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  6. 06
    Heterogeneous CPU + GPU EPD Disaggregation to Boost VLM Serving - LMSYS Org

    Heterogeneous CPU + GPU EPD Disaggregation to Boost VLM Serving LMSYS Org

    最高第 1600:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  7. 07
    Improving DeepEP MoE Load Balance in SGLang with Waterfill and LPLB - LMSYS Org

    Improving DeepEP MoE Load Balance in SGLang with Waterfill and LPLB LMSYS Org

    最高第 900:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  8. 08
    Towards Blackwell-Native 8-bit and 4-bit RL: End-to-End MXFP8 and NVFP4 RL in Miles - LMSYS Org

    Towards Blackwell-Native 8-bit and 4-bit RL: End-to-End MXFP8 and NVFP4 RL in Miles LMSYS Org

    最高第 100:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  9. 09
    Unleashing Computational Power: Ultimate Latency Optimization of Qwen3 and Qwen3-VL on AMD MI300X Series - LMSYS Org

    Unleashing Computational Power: Ultimate Latency Optimization of Qwen3 and Qwen3-VL on AMD MI300X Series LMSYS Org

    最高第 2900:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  10. 10
    Deploying DeepSeek on GB300 NVL72: Big Wins in Long-Context Inference - LMSYS Org

    Deploying DeepSeek on GB300 NVL72: Big Wins in Long-Context Inference LMSYS Org

    最高第 2700:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  11. 11
    DSpark in SGLang: Speculative Decoding with Confidence-Driven, Variable-Length Verification - LMSYS Org

    DSpark in SGLang: Speculative Decoding with Confidence-Driven, Variable-Length Verification LMSYS Org

    最高第 700:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  12. 12
    Accelerating SGLang HiCache with Netpreme X-Mem™ MPU - LMSYS Org

    Accelerating SGLang HiCache with Netpreme X-Mem™ MPU LMSYS Org

    最高第 600:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  13. 13
    ROCm Support for Miles: Large-Scale RL Post-Training on AMD Instinct™ GPUs - LMSYS Org

    ROCm Support for Miles: Large-Scale RL Post-Training on AMD Instinct™ GPUs LMSYS Org

    最高第 2400:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  14. 14
    Serving GLM5.2 NVFP4 Agentic Workload with SGLang: Reaching 500 TPS in 2 Weeks - LMSYS Org

    Serving GLM5.2 NVFP4 Agentic Workload with SGLang: Reaching 500 TPS in 2 Weeks LMSYS Org

    最高第 500:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  15. 15
    HiSparse: Turbocharging Sparse Attention with Hierarchical Memory - LMSYS Org

    HiSparse: Turbocharging Sparse Attention with Hierarchical Memory LMSYS Org

    最高第 2000:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  16. 16
    Unlocking 25x Inference Performance with SGLang on NVIDIA GB300 NVL72 - LMSYS Org

    Unlocking 25x Inference Performance with SGLang on NVIDIA GB300 NVL72 LMSYS Org

    最高第 2600:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  17. 17
    SGLang and Miles Add Day-0 Support for Kimi K3 - LMSYS Org

    SGLang and Miles Add Day-0 Support for Kimi K3 LMSYS Org

    最高第 200:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  18. 18
    SGLang and Miles Add Day-0 Support for Inkling, a Frontier Multimodal Model - LMSYS Org

    SGLang and Miles Add Day-0 Support for Inkling, a Frontier Multimodal Model LMSYS Org

    最高第 400:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  19. 19
    No Token Left Behind: Demystifying Token-In-Token-Out in Miles - LMSYS Org

    No Token Left Behind: Demystifying Token-In-Token-Out in Miles LMSYS Org

    最高第 1300:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  20. 20
    MOSS-TTS Local Transformer v1.5 on SGLang-Omni: Serving Native-Streaming 48 kHz Speech - LMSYS Org

    MOSS-TTS Local Transformer v1.5 on SGLang-Omni: Serving Native-Streaming 48 kHz Speech LMSYS Org

    最高第 1000:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  21. 21
    Higgs Audio v3 TTS on SGLang-Omni: Real-Time, Controllable Speech for Voice Agents - LMSYS Org

    Higgs Audio v3 TTS on SGLang-Omni: Real-Time, Controllable Speech for Voice Agents LMSYS Org

    最高第 1400:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  22. 22
    OPD Support in Miles - LMSYS Org

    OPD Support in Miles LMSYS Org

    最高第 300:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  23. 23
    Elastic EP in SGLang: Achieving Partial Failure Tolerance for DeepSeek MoE Deployments - LMSYS Org

    Elastic EP in SGLang: Achieving Partial Failure Tolerance for DeepSeek MoE Deployments LMSYS Org

    最高第 2300:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  24. 24
    SGLang and Miles Add Day-0 Support for NVIDIA Nemotron 3 Ultra for Long-Running Autonomous Agents - LMSYS Org

    SGLang and Miles Add Day-0 Support for NVIDIA Nemotron 3 Ultra for Long-Running Autonomous Agents LMSYS Org

    最高第 1500:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  25. 25
    Agent-Assisted SGLang Development: An Initial Exploration - LMSYS Org

    Agent-Assisted SGLang Development: An Initial Exploration LMSYS Org

    最高第 800:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  26. 26
    SGLang-Diffusion: Advanced Optimizations for Production-Ready Video Generation - LMSYS Org

    SGLang-Diffusion: Advanced Optimizations for Production-Ready Video Generation LMSYS Org

    最高第 2800:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  27. 27
    The next generation of speculative decoding: DFlash and Spec V2 - LMSYS Org

    The next generation of speculative decoding: DFlash and Spec V2 LMSYS Org

    最高第 1200:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  28. 28
    Optimizing Ling-2.6-1T on TPU with SGLang-JAX: Hiding MoE Data Movement Behind Compute with One Pallas Kernel - LMSYS Org

    Optimizing Ling-2.6-1T on TPU with SGLang-JAX: Hiding MoE Data Movement Behind Compute with One Pallas Kernel LMSYS Org

    最高第 1100:00 达到当日首次采集时已在榜23:47 观测离榜上榜 4 次(重入 3 次)累计约23小时
  29. 29
    Announcing the Recipient of the 2026 LMSYS PhD Fellowship - LMSYS Org

    Announcing the Recipient of the 2026 LMSYS PhD Fellowship LMSYS Org

    最高第 1405:07 达到05:07 首次观测上榜当日结束时仍在榜上榜 4 次(重入 3 次)累计约14小时40分
  30. 30
    RadixArk Joins Forces with Google to Bring Full SGLang Features to TPUs - LMSYS Org

    RadixArk Joins Forces with Google to Bring Full SGLang Features to TPUs LMSYS Org

    最高第 110:11 达到10:11 首次观测上榜当日结束时仍在榜累计约13小时36分
  31. 31
    SGLang Adds Day-0 Support for NVIDIA Nemotron 3 Super for building High-Efficiency Multi-Agent Systems - LMSYS Org

    SGLang Adds Day-0 Support for NVIDIA Nemotron 3 Super for building High-Efficiency Multi-Agent Systems LMSYS Org

    最高第 2500:00 达到当日首次采集时已在榜10:11 观测离榜上榜 4 次(重入 3 次)累计约9小时8分
  32. 32
    RadixArk Joins Forces with Google to Bring Full SGLang Features to TPUs - LMSYS Org

    RadixArk Joins Forces with Google to Bring Full SGLang Features to TPUs LMSYS Org

    最高第 102:27 达到02:27 首次观测上榜10:11 观测离榜累计约7小时44分
  33. 33
    Squeezing 1TB Model Rollout into a Single H200: INT4 QAT RL End-to-End Practice - LMSYS Org

    Squeezing 1TB Model Rollout into a Single H200: INT4 QAT RL End-to-End Practice LMSYS Org

    最高第 3000:00 达到当日首次采集时已在榜02:27 观测离榜累计约2小时28分
  34. 34
    Bringing DeepSeek-V4 Flash RL Training to AMD Instinct MI355X GPUs with Miles - LMSYS Org

    Bringing DeepSeek-V4 Flash RL Training to AMD Instinct MI355X GPUs with Miles LMSYS Org

    最高第 719:15 达到19:15 首次观测上榜当日结束时仍在榜上榜 4 次(重入 3 次)累计约48分钟