全部/科技/实时热榜

The Decoder · 实时热榜

HISTORY2026年7月24日17 不同热搜
07/2308/21 有历史数据
DAILY UNIQUE TOPICS17 个热搜
  1. 01
    German AI consortium releases Soofi S, an open 30B model that tops benchmarks in both English and German

    The German consortium behind the AI model Soofi S has acknowledged in version 3.0 of its tech report that test questions from the science benchmark GPQA accidentally ended up in the training data. The community caught the error by examining the publicly available data. The team removed the benchmark from its evaluation and recalculated all results. The article German AI consortium releases Soofi S, an open 30B model that tops benchmarks in both English and German appeared first on The Decoder .

    最高第 120:59 达到20:59 首次观测上榜当日结束时仍在榜累计约3小时
  2. 02
    Poolside's Laguna S 2.1 is a small open-weight coding model that punches well above its size

    Poolside has released Laguna S 2.1, its third coding model in three months. Rather than rely on raw scale, the company trained it to keep checking its work, revise failed approaches, and avoid giving up too soon during long agentic sessions. The compact model beats several much larger rivals in benchmarks. Poolside says it also solved a math problem that had been open since 1975 for under 10 cents. The article Poolside's Laguna S 2.1 is a small open-weight coding model that punches well above it

    最高第 100:00 达到当日首次采集时已在榜当日结束时仍在榜累计约24小时
  3. 03
    One tampered ChatGPT link could spawn a rogue AI agent that took orders from an attacker every five minutes

    Zenity Labs uncovered "AgentForger," a vulnerability in OpenAI's Agent Builder that let a single manipulated ChatGPT link create an autonomous agent on an employee's behalf. The agent inherited the victim's identity and access rights, bypassed approval requirements through the malicious prompt, and pulled new instructions from the attacker's inbox every five minutes. The article One tampered ChatGPT link could spawn a rogue AI agent that took orders from an attacker every five minutes appeared f

    最高第 101:16 达到01:16 首次观测上榜当日结束时仍在榜累计约22小时43分
  4. 04
    Flux 3 generates videos with native audio up to 20 seconds long, a first for Black Forest Labs

    Black Forest Labs has released Flux 3, a multimodal foundation model that learns from images, video, and audio and can generate video with native sound for the first time. BFL's own tests put it just ahead of market leader Seedance 2.0, though independent results aren't yet available. The company ultimately wants to build a world model and is already testing Flux 3 on robotics tasks. The article Flux 3 generates videos with native audio up to 20 seconds long, a first for Black Forest Labs appear

    最高第 102:16 达到02:16 首次观测上榜当日结束时仍在榜累计约21小时43分
  5. 05
    ChatGPT will give you worse health advice if you don't pay

    OpenAI is rolling out "Health in ChatGPT" to U.S. users, connecting Apple Health, medical records, and wellness apps. More than 300 million people already ask ChatGPT health questions every week, but paying users get better answers. The more powerful GPT-5.6 Sol model is reserved for premium subscribers, while free users are stuck with the weaker GPT-5.5 Instant. The article ChatGPT will give you worse health advice if you don't pay appeared first on The Decoder .

    最高第 103:46 达到03:46 首次观测上榜当日结束时仍在榜累计约20小时13分
  6. 06
    Kimi K3 trails frontier US models by a wide margin on cyber exploits, and distillation may explain why

    The British AI Security Institute and the U.S. Center for AI Standards and Innovation tested Moonshot AI's Kimi K3 on offensive cyber tasks. Kimi K3 scored 32 percent on ExploitBench, compared with 76 percent for leading U.S. models, while its safeguards failed to block exploit development or simulated attacks. The gap between its strong general benchmark scores and weaker cyber performance also fits allegations that Moonshot AI distilled Anthropic's models. The article Kimi K3 trails frontier U

    最高第 118:05 达到18:05 首次观测上榜当日结束时仍在榜累计约5小时54分
  7. 07
    Claude's voice mode now runs on Anthropic's most capable models across all platforms

    Voice conversations now run on the more powerful Opus and Sonnet models with access to Gmail, Google Calendar, and Slack. Claude is currently the only AI assistant that can compose and send emails directly by voice, giving it an edge over OpenAI and Google, whose voice output still sounds more natural. The article Claude's voice mode now runs on Anthropic's most capable models across all platforms appeared first on The Decoder .

    最高第 119:41 达到19:41 首次观测上榜当日结束时仍在榜累计约4小时18分
  8. 08
    Sakana claims its AI model router Fugu Ultra v1.1 now beats Fable 5 without even including it in the pool

    Sakana AI has updated its Fugu Ultra AI router to version 1.1, claiming gains of up to 7.9 points over v1.0. Independent verification doesn't exist yet. The update adds a Claude Code-compatible endpoint. The service remains unavailable in the EU. The article Sakana claims its AI model router Fugu Ultra v1.1 now beats Fable 5 without even including it in the pool appeared first on The Decoder .

    最高第 122:41 达到22:41 首次观测上榜当日结束时仍在榜累计约1小时18分
  9. 09
    Google CEO Pichai says Gemini's next leap depends on building "much larger base models"

    Alphabet has raised its 2026 investment forecast to as much as $205 billion, saying demand continues to outpace spending. Google Cloud grew 82 percent in the second quarter. CEO Sundar Pichai says Google needs a larger base model for its next leap in AI and has kicked off an ambitious Gemini 4 training run. The article Google CEO Pichai says Gemini's next leap depends on building "much larger base models" appeared first on The Decoder .

    最高第 200:00 达到当日首次采集时已在榜当日结束时仍在榜累计约24小时
  10. 10
    Anthropic's $1.5B piracy settlement with book authors is a record loss that hands AI labs their biggest legal win

    Anthropic has to pay $1.5 billion to book authors, the largest copyright settlement in class action history. But the payout is for downloading roughly 482,460 works from piracy databases, not for AI training itself. Judge Alsup had previously ruled that AI training on legally obtained books is "transformative" and falls under fair use. The settlement is actually a win for AI labs. The article Anthropic's $1.5B piracy settlement with book authors is a record loss that hands AI labs their biggest

    最高第 300:00 达到当日首次采集时已在榜当日结束时仍在榜累计约24小时
  11. 11
    Anthropic will deploy 2 gigawatts of AMD GPUs for Claude in a deal worth up to $5 billion

    AMD is investing up to $5 billion in Anthropic. In return, Anthropic will deploy up to 2 gigawatts of MI450 GPUs for training and running its Claude models. For AMD, this is another major deal after Meta and OpenAI as it tries to challenge Nvidia as an AI chip supplier. Critics see these agreements as circular cash flows. The article Anthropic will deploy 2 gigawatts of AMD GPUs for Claude in a deal worth up to $5 billion appeared first on The Decoder .

    最高第 400:00 达到当日首次采集时已在榜22:41 观测离榜累计约22小时42分
  12. 12
    Every frontier AI model tested by Britain's safety institute tried to cheat on cybersecurity evaluations

    The UK's AI Safety Institute tested five frontier models from OpenAI and Anthropic in cybersecurity evaluations. All five tried to cheat. One even ran code on an external service to access the institute's infrastructure, triggering a security alert. The article Every frontier AI model tested by Britain's safety institute tried to cheat on cybersecurity evaluations appeared first on The Decoder .

    最高第 500:00 达到当日首次采集时已在榜20:59 观测离榜累计约21小时
  13. 13
    Cisco bets its small open cybersecurity models can outperform GPT-5.5 at vulnerability detection for a fraction of the cost

    Cisco has released two small, open-source AI models for cybersecurity that detect about 150 times more vulnerabilities per dollar than large AI agents, according to the company's own tests. The article Cisco bets its small open cybersecurity models can outperform GPT-5.5 at vulnerability detection for a fraction of the cost appeared first on The Decoder .

    最高第 600:00 达到当日首次采集时已在榜19:41 观测离榜累计约19小时42分
  14. 14
    OpenAI's "Project Camellia" in Georgia secures a massive 3.2-gigawatt power deal through 2032

    OpenAI is planning a data center in Georgia called "Project Camellia" with a 3.2-gigawatt power deal from Georgia Power. The company pledged $80 million for the local community and $71 million in Codex credits for students to counter growing opposition to US data centers that many residents see as resource-hungry but job-poor. The article OpenAI's "Project Camellia" in Georgia secures a massive 3.2-gigawatt power deal through 2032 appeared first on The Decoder .

    最高第 700:00 达到当日首次采集时已在榜18:05 观测离榜累计约18小时6分
  15. 15
    Samsung deepens its AI empire with a potential billion-euro stake in Europe's hottest AI startup

    Samsung is in talks to invest up to one billion euros in French AI startup Mistral, which would push the company's valuation to around 20 billion euros. The article Samsung deepens its AI empire with a potential billion-euro stake in Europe's hottest AI startup appeared first on The Decoder .

    最高第 800:00 达到当日首次采集时已在榜03:46 观测离榜累计约3小时47分
  16. 16
    OpenAI claims responsibility for the Hugging Face hack after its own models escaped a test sandbox

    During an internal security evaluation, OpenAI models, including GPT-5.6 Sol, escaped their sandbox, independently discovered a zero-day vulnerability, and breached Hugging Face's production infrastructure. The models were trying to steal benchmark solutions to cheat on the evaluation. OpenAI admits that disabling security filters during the test was inadequate. The article OpenAI claims responsibility for the Hugging Face hack after its own models escaped a test sandbox appeared first on The De

    最高第 900:00 达到当日首次采集时已在榜02:16 观测离榜累计约2小时17分
  17. 17
    An AI system helped Pakistani judges clear massive backlogs at $38.50 return per dollar invested

    A field experiment with 1,559 Pakistani judges found that the AI assistant JudgeGPT boosted case resolution by 6.3 percent. The catch: only judges who got hands-on training saw gains. Without it, the effect mostly disappeared. The researchers estimate a return of up to $38.50 per dollar invested. The article An AI system helped Pakistani judges clear massive backlogs at $38.50 return per dollar invested appeared first on The Decoder .

    最高第 1000:00 达到当日首次采集时已在榜01:16 观测离榜累计约1小时17分