全部/科技/实时热榜

LMSYS Blog · 实时热榜

HISTORY2026年8月1日32 不同热搜
07/2308/21 有历史数据
DAILY UNIQUE TOPICS32 个热搜
  1. 01
    Blog - lmsys.org

    Blog lmsys.org

    最高第 2300:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  2. 02
    Win on TCO: How AMD Instinct™ MI355X Achieves Cost-Competitive Distributed Inference Through SGLang with MoRI - lmsys.org

    Win on TCO: How AMD Instinct™ MI355X Achieves Cost-Competitive Distributed Inference Through SGLang with MoRI lmsys.org

    最高第 1900:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  3. 03
    Highlights of SGLang at NVIDIA GTC 2026 - lmsys.org

    Highlights of SGLang at NVIDIA GTC 2026 lmsys.org

    最高第 2400:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  4. 04
    DeepSeek-V4 on Day 0: From Fast Inference to Verified RL with SGLang and Miles - lmsys.org

    DeepSeek-V4 on Day 0: From Fast Inference to Verified RL with SGLang and Miles lmsys.org

    最高第 2100:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  5. 05
    Updating 1T parameters in seconds — P2P weight transfer in Large Scale Distributed RL - lmsys.org

    Updating 1T parameters in seconds — P2P weight transfer in Large Scale Distributed RL lmsys.org

    最高第 2000:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  6. 06
    Heterogeneous CPU + GPU EPD Disaggregation to Boost VLM Serving - lmsys.org

    Heterogeneous CPU + GPU EPD Disaggregation to Boost VLM Serving lmsys.org

    最高第 1800:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  7. 07
    Improving DeepEP MoE Load Balance in SGLang with Waterfill and LPLB - lmsys.org

    Improving DeepEP MoE Load Balance in SGLang with Waterfill and LPLB lmsys.org

    最高第 1000:03 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  8. 08
    Towards Blackwell-Native 8-bit and 4-bit RL: End-to-End MXFP8 and NVFP4 RL in Miles - lmsys.org

    Towards Blackwell-Native 8-bit and 4-bit RL: End-to-End MXFP8 and NVFP4 RL in Miles lmsys.org

    最高第 200:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  9. 09
    Deploying DeepSeek on GB300 NVL72: Big Wins in Long-Context Inference - lmsys.org

    Deploying DeepSeek on GB300 NVL72: Big Wins in Long-Context Inference lmsys.org

    最高第 2800:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  10. 10
    DSpark in SGLang: Speculative Decoding with Confidence-Driven, Variable-Length Verification - lmsys.org

    DSpark in SGLang: Speculative Decoding with Confidence-Driven, Variable-Length Verification lmsys.org

    最高第 800:03 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  11. 11
    Accelerating SGLang HiCache with Netpreme X-Mem™ MPU - lmsys.org

    Accelerating SGLang HiCache with Netpreme X-Mem™ MPU lmsys.org

    最高第 700:03 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  12. 12
    ROCm Support for Miles: Large-Scale RL Post-Training on AMD Instinct™ GPUs - lmsys.org

    ROCm Support for Miles: Large-Scale RL Post-Training on AMD Instinct™ GPUs lmsys.org

    最高第 2600:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  13. 13
    Serving GLM5.2 NVFP4 Agentic Workload with SGLang: Reaching 500 TPS in 2 Weeks - lmsys.org

    Serving GLM5.2 NVFP4 Agentic Workload with SGLang: Reaching 500 TPS in 2 Weeks lmsys.org

    最高第 600:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  14. 14
    RadixArk Joins Forces with Google to Bring Full SGLang Features to TPUs - lmsys.org

    RadixArk Joins Forces with Google to Bring Full SGLang Features to TPUs lmsys.org

    最高第 100:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  15. 15
    HiSparse: Turbocharging Sparse Attention with Hierarchical Memory - lmsys.org

    HiSparse: Turbocharging Sparse Attention with Hierarchical Memory lmsys.org

    最高第 2200:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  16. 16
    Unlocking 25x Inference Performance with SGLang on NVIDIA GB300 NVL72 - lmsys.org

    Unlocking 25x Inference Performance with SGLang on NVIDIA GB300 NVL72 lmsys.org

    最高第 2700:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  17. 17
    SGLang and Miles Add Day-0 Support for Kimi K3 - lmsys.org

    SGLang and Miles Add Day-0 Support for Kimi K3 lmsys.org

    最高第 300:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  18. 18
    SGLang and Miles Add Day-0 Support for Inkling, a Frontier Multimodal Model - lmsys.org

    SGLang and Miles Add Day-0 Support for Inkling, a Frontier Multimodal Model lmsys.org

    最高第 500:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  19. 19
    Announcing the Recipient of the 2026 LMSYS PhD Fellowship - lmsys.org

    Announcing the Recipient of the 2026 LMSYS PhD Fellowship lmsys.org

    最高第 1400:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  20. 20
    No Token Left Behind: Demystifying Token-In-Token-Out in Miles - lmsys.org

    No Token Left Behind: Demystifying Token-In-Token-Out in Miles lmsys.org

    最高第 1500:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  21. 21
    MOSS-TTS Local Transformer v1.5 on SGLang-Omni: Serving Native-Streaming 48 kHz Speech - lmsys.org

    MOSS-TTS Local Transformer v1.5 on SGLang-Omni: Serving Native-Streaming 48 kHz Speech lmsys.org

    最高第 1100:03 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  22. 22
    Higgs Audio v3 TTS on SGLang-Omni: Real-Time, Controllable Speech for Voice Agents - lmsys.org

    Higgs Audio v3 TTS on SGLang-Omni: Real-Time, Controllable Speech for Voice Agents lmsys.org

    最高第 1600:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  23. 23
    OPD Support in Miles - lmsys.org

    OPD Support in Miles lmsys.org

    最高第 400:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  24. 24
    SGLang and Miles Add Day-0 Support for NVIDIA Nemotron 3 Ultra for Long-Running Autonomous Agents - lmsys.org

    SGLang and Miles Add Day-0 Support for NVIDIA Nemotron 3 Ultra for Long-Running Autonomous Agents lmsys.org

    最高第 1700:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  25. 25
    Agent-Assisted SGLang Development: An Initial Exploration - lmsys.org

    Agent-Assisted SGLang Development: An Initial Exploration lmsys.org

    最高第 900:03 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  26. 26
    SGLang-Diffusion: Advanced Optimizations for Production-Ready Video Generation - lmsys.org

    SGLang-Diffusion: Advanced Optimizations for Production-Ready Video Generation lmsys.org

    最高第 2900:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  27. 27
    The next generation of speculative decoding: DFlash and Spec V2 - lmsys.org

    The next generation of speculative decoding: DFlash and Spec V2 lmsys.org

    最高第 1300:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时48分
  28. 28
    Unleashing Computational Power: Ultimate Latency Optimization of Qwen3 and Qwen3-VL on AMD MI300X Series - LMSYS Org

    Unleashing Computational Power: Ultimate Latency Optimization of Qwen3 and Qwen3-VL on AMD MI300X Series LMSYS Org

    最高第 3000:00 达到当日首次采集时已在榜21:23 观测离榜上榜 2 次(重入 1 次)累计约21小时8分
  29. 29
    Toward a Cleaner Quantization Stack in SGLang - lmsys.org

    Toward a Cleaner Quantization Stack in SGLang lmsys.org

    最高第 303:47 达到03:47 首次观测上榜当日结束时仍在榜上榜 3 次(重入 2 次)累计约17小时52分
  30. 30
    Bringing DeepSeek-V4 Flash RL Training to AMD Instinct MI355X GPUs with Miles - lmsys.org

    Bringing DeepSeek-V4 Flash RL Training to AMD Instinct MI355X GPUs with Miles lmsys.org

    最高第 700:00 达到当日首次采集时已在榜当日结束时仍在榜上榜 6 次(重入 5 次)累计约14小时44分
  31. 31
    Elastic EP in SGLang: Achieving Partial Failure Tolerance for DeepSeek MoE Deployments - lmsys.org

    Elastic EP in SGLang: Achieving Partial Failure Tolerance for DeepSeek MoE Deployments lmsys.org

    最高第 2500:00 达到当日首次采集时已在榜当日结束时仍在榜上榜 4 次(重入 3 次)累计约12小时52分
  32. 32
    Optimizing Ling-2.6-1T on TPU with SGLang-JAX: Hiding MoE Data Movement Behind Compute with One Pallas Kernel - LMSYS Org

    Optimizing Ling-2.6-1T on TPU with SGLang-JAX: Hiding MoE Data Movement Behind Compute with One Pallas Kernel LMSYS Org

    最高第 1200:03 达到00:03 首次观测上榜06:27 观测离榜上榜 5 次(重入 4 次)累计约4小时48分