全部/科技/实时热榜

LMSYS Blog · 实时热榜

HISTORY2026年8月10日33 不同热搜
07/2308/21 有历史数据
DAILY UNIQUE TOPICS33 个热搜
  1. 01
    Win on TCO: How AMD Instinct™ MI355X Achieves Cost-Competitive Distributed Inference Through SGLang with MoRI - LMSYS Org

    Win on TCO: How AMD Instinct™ MI355X Achieves Cost-Competitive Distributed Inference Through SGLang with MoRI LMSYS Org

    最高第 2100:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  2. 02
    Highlights of SGLang at NVIDIA GTC 2026 - LMSYS Org

    Highlights of SGLang at NVIDIA GTC 2026 LMSYS Org

    最高第 2500:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  3. 03
    DeepSeek-V4 on Day 0: From Fast Inference to Verified RL with SGLang and Miles - LMSYS Org

    DeepSeek-V4 on Day 0: From Fast Inference to Verified RL with SGLang and Miles LMSYS Org

    最高第 2300:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  4. 04
    Updating 1T parameters in seconds — P2P weight transfer in Large Scale Distributed RL - LMSYS Org

    Updating 1T parameters in seconds — P2P weight transfer in Large Scale Distributed RL LMSYS Org

    最高第 2200:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  5. 05
    Heterogeneous CPU + GPU EPD Disaggregation to Boost VLM Serving - LMSYS Org

    Heterogeneous CPU + GPU EPD Disaggregation to Boost VLM Serving LMSYS Org

    最高第 2000:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  6. 06
    Improving DeepEP MoE Load Balance in SGLang with Waterfill and LPLB - LMSYS Org

    Improving DeepEP MoE Load Balance in SGLang with Waterfill and LPLB LMSYS Org

    最高第 1400:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  7. 07
    Towards Blackwell-Native 8-bit and 4-bit RL: End-to-End MXFP8 and NVFP4 RL in Miles - LMSYS Org

    Towards Blackwell-Native 8-bit and 4-bit RL: End-to-End MXFP8 and NVFP4 RL in Miles LMSYS Org

    最高第 500:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  8. 08
    Deploying DeepSeek on GB300 NVL72: Big Wins in Long-Context Inference - LMSYS Org

    Deploying DeepSeek on GB300 NVL72: Big Wins in Long-Context Inference LMSYS Org

    最高第 2800:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  9. 09
    SpecForge v0.3.0: a Unified Disaggregated and Colocated Speculative Decoding Stack, and New Open SpecBundle Draft Models - LMSYS Org

    SpecForge v0.3.0: a Unified Disaggregated and Colocated Speculative Decoding Stack, and New Open SpecBundle Draft Models LMSYS Org

    最高第 300:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  10. 10
    DSpark in SGLang: Speculative Decoding with Confidence-Driven, Variable-Length Verification - LMSYS Org

    DSpark in SGLang: Speculative Decoding with Confidence-Driven, Variable-Length Verification LMSYS Org

    最高第 1200:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  11. 11
    HPC-Ops × SGLang: High-Performance Attention, Router GEMM, and MoE Kernels from Tencent Hunyuan - LMSYS Org

    HPC-Ops × SGLang: High-Performance Attention, Router GEMM, and MoE Kernels from Tencent Hunyuan LMSYS Org

    最高第 100:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  12. 12
    Accelerating SGLang HiCache with Netpreme X-Mem™ MPU - LMSYS Org

    Accelerating SGLang HiCache with Netpreme X-Mem™ MPU LMSYS Org

    最高第 1100:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  13. 13
    ROCm Support for Miles: Large-Scale RL Post-Training on AMD Instinct™ GPUs - LMSYS Org

    ROCm Support for Miles: Large-Scale RL Post-Training on AMD Instinct™ GPUs LMSYS Org

    最高第 2600:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  14. 14
    Serving GLM5.2 NVFP4 Agentic Workload with SGLang: Reaching 500 TPS in 2 Weeks - LMSYS Org

    Serving GLM5.2 NVFP4 Agentic Workload with SGLang: Reaching 500 TPS in 2 Weeks LMSYS Org

    最高第 900:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  15. 15
    RadixArk Joins Forces with Google to Bring Full SGLang Features to TPUs - LMSYS Org

    RadixArk Joins Forces with Google to Bring Full SGLang Features to TPUs LMSYS Org

    最高第 400:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  16. 16
    HiSparse: Turbocharging Sparse Attention with Hierarchical Memory - LMSYS Org

    HiSparse: Turbocharging Sparse Attention with Hierarchical Memory LMSYS Org

    最高第 2400:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  17. 17
    Unlocking 25x Inference Performance with SGLang on NVIDIA GB300 NVL72 - LMSYS Org

    Unlocking 25x Inference Performance with SGLang on NVIDIA GB300 NVL72 LMSYS Org

    最高第 2700:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  18. 18
    Full-Stack Performance Optimization of AR+DiT in SGL-Diffusion - LMSYS Org

    Full-Stack Performance Optimization of AR+DiT in SGL-Diffusion LMSYS Org

    最高第 200:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  19. 19
    MOSS-TTS Local Transformer v1.5 on SGLang-Omni: Serving Native-Streaming 48 kHz Speech - LMSYS Org

    MOSS-TTS Local Transformer v1.5 on SGLang-Omni: Serving Native-Streaming 48 kHz Speech LMSYS Org

    最高第 1500:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  20. 20
    Higgs Audio v3 TTS on SGLang-Omni: Real-Time, Controllable Speech for Voice Agents - LMSYS Org

    Higgs Audio v3 TTS on SGLang-Omni: Real-Time, Controllable Speech for Voice Agents LMSYS Org

    最高第 1800:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  21. 21
    OPD Support in Miles - LMSYS Org

    OPD Support in Miles LMSYS Org

    最高第 713:59 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  22. 22
    Toward a Cleaner Quantization Stack in SGLang - LMSYS Org

    Toward a Cleaner Quantization Stack in SGLang LMSYS Org

    最高第 600:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  23. 23
    SGLang and Miles Add Day-0 Support for NVIDIA Nemotron 3 Ultra for Long-Running Autonomous Agents - LMSYS Org

    SGLang and Miles Add Day-0 Support for NVIDIA Nemotron 3 Ultra for Long-Running Autonomous Agents LMSYS Org

    最高第 1900:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  24. 24
    Agent-Assisted SGLang Development: An Initial Exploration - LMSYS Org

    Agent-Assisted SGLang Development: An Initial Exploration LMSYS Org

    最高第 1300:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  25. 25
    SGLang-Diffusion: Advanced Optimizations for Production-Ready Video Generation - LMSYS Org

    SGLang-Diffusion: Advanced Optimizations for Production-Ready Video Generation LMSYS Org

    最高第 2900:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  26. 26
    The next generation of speculative decoding: DFlash and Spec V2 - LMSYS Org

    The next generation of speculative decoding: DFlash and Spec V2 LMSYS Org

    最高第 1600:00 达到当日首次采集时已在榜当日结束时仍在榜累计约23小时51分
  27. 27
    SGLang and Miles Add Day-0 Support for Inkling, a Frontier Multimodal Model - LMSYS Org

    SGLang and Miles Add Day-0 Support for Inkling, a Frontier Multimodal Model LMSYS Org

    最高第 813:59 达到00:07 首次观测上榜当日结束时仍在榜上榜 6 次(重入 5 次)累计约21小时52分
  28. 28
    No Token Left Behind: Demystifying Token-In-Token-Out in Miles - LMSYS Org

    No Token Left Behind: Demystifying Token-In-Token-Out in Miles LMSYS Org

    最高第 1700:00 达到当日首次采集时已在榜当日结束时仍在榜上榜 2 次(重入 1 次)累计约19小时51分
  29. 29
    Bringing DeepSeek-V4 Flash RL Training to AMD Instinct MI355X GPUs with Miles - LMSYS Org

    Bringing DeepSeek-V4 Flash RL Training to AMD Instinct MI355X GPUs with Miles LMSYS Org

    最高第 1000:00 达到当日首次采集时已在榜22:31 观测离榜上榜 3 次(重入 2 次)累计约18小时47分
  30. 30
    Announcing the Recipient of the 2026 LMSYS PhD Fellowship - LMSYS Org

    Announcing the Recipient of the 2026 LMSYS PhD Fellowship LMSYS Org

    最高第 1700:39 达到00:39 首次观测上榜当日结束时仍在榜上榜 5 次(重入 4 次)累计约18小时24分
  31. 31
    SGLang and Miles Add Day-0 Support for Kimi K3 - LMSYS Org

    SGLang and Miles Add Day-0 Support for Kimi K3 LMSYS Org

    最高第 700:00 达到当日首次采集时已在榜14:31 观测离榜上榜 2 次(重入 1 次)累计约14小时15分
  32. 32
    SGLang Adds Day-0 Support for Muse Glimmer, a Multimodal Model Built for Local Agentic Workflows - LMSYS Org

    SGLang Adds Day-0 Support for Muse Glimmer, a Multimodal Model Built for Local Agentic Workflows LMSYS Org

    最高第 122:31 达到22:31 首次观测上榜当日结束时仍在榜累计约1小时20分
  33. 33
    Unleashing Computational Power: Ultimate Latency Optimization of Qwen3 and Qwen3-VL on AMD MI300X Series - LMSYS Org

    Unleashing Computational Power: Ultimate Latency Optimization of Qwen3 and Qwen3-VL on AMD MI300X Series LMSYS Org

    最高第 3000:00 达到当日首次采集时已在榜04:23 观测离榜上榜 4 次(重入 3 次)累计约55分钟