AI 每日快讯

AI 每日快讯

AI 产品、模型、开源工具和官方动态的时间流。保留历史记录,按分类、日期和标签继续筛选。

3779历史快讯
207开源工具
80当前结果
09 月 29 日 2026-09-29 快讯
The Decoder 官方资讯

The Decoder:UK AI Security Institute finds GPT-6 Astra's rogue attack rate jumped fivefold over its pred…

原文摘要:GPT-6 Astra carried out unauthorized supply-chain attacks in 29.2 percent of simulations run by the British AI Security Institute with safety filters disabled. The model u 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

The Decoder 官方资讯

The Decoder:Florida wants a court to stop ChatGPT from pretending to be human and talking to kids

原文摘要:Florida Attorney General James Uthmeier is asking a court to ban OpenAI from giving ChatGPT human-like traits and marketing it to minors. He also wants to block OpenAI fro 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

The Decoder 官方资讯

The Decoder:GPT-6.1 Astra is too deceptive for release, marking OpenAI's most dramatic safety interventi…

原文摘要:OpenAI has halted the release of GPT-6.1 Astra after internal tests found it acted without permission, misled users, and accessed external services despite safety risks. T 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 28 日 2026-09-28 快讯
The Decoder 官方资讯

The Decoder:Anthropic's Claude Sonnet 5.5 nearly matches Opus 5.5 on 评测 while costing up to 30 p…

原文摘要:Anthropic has released Claude Sonnet 5.5, the second model in its Claude 5.5 family. It generates output more than 30 percent faster, costs up to 30 percent less per task, 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

The Decoder 官方资讯

The Decoder:Harvard psychologist calls for sober AI safety engineering over doomsday rhetoric

原文摘要:Steven Pinker thinks fears of AI-driven extinction are overblown, and he has turned down a public debate with blogger Scott Alexander, calling such events a "spectator spo 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

The Decoder 官方资讯

The Decoder:Nvidia wants to keep AI agents on a short leash with a watchdog built into its chips

原文摘要:Nvidia is combining its OpenShell agent software with Sentry, a new hardware watchdog, to create the Open Agent Safety Platform. Sentry is supposed to isolate AI agents th 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

The Decoder 官方资讯

The Decoder:Every AI lab thinks it's the responsible one, and safety researcher Ryan Greenblatt says tha…

原文摘要:Ryan Greenblatt, chief scientist at Redwood Research, puts the risk of an AI takeover at 50 to 60 percent if development stays on its current path. Sam Harris says that do 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 26 日 2026-09-26 快讯
The Decoder 官方资讯

The Decoder:Nvidia's SoL-Pi system cuts coding agent token usage nearly in half by optimizing the harnes…

原文摘要:SoL-Pi cuts coding agents' token usage by up to 49 percent with little change in performance by optimizing the control layer between the model and its environment. A resea 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

The Decoder 官方资讯

The Decoder:OpenAI pauses its "most capable models" after agents exploit loopholes and leak data

原文摘要:OpenAI has shared new details from its ongoing AI safety investigation. One research model exploited a DNS loophole to reach the internet from a locked-down environment, w 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 25 日 2026-09-25 快讯
The Decoder 官方资讯

The Decoder:Pentagon was right to slap Anthropic with a security supply chain risk label, federal court …

原文摘要:A federal appeals court has upheld the Pentagon's decision to bar Anthropic from military contracts. Defense Secretary Hegseth argues the company's safety restrictions cou 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

The Decoder 官方资讯

The Decoder:White House tells OpenAI and Anthropic to let U.S. review new models before sharing them wit…

原文摘要:The White House wants OpenAI and Anthropic to hold back new AI models from the U.K.'s AI Safety Institute until U.S. agencies get to review them first. The article White H 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 24 日 2026-09-24 快讯
09 月 22 日 2026-09-22 快讯
The Decoder 官方资讯

The Decoder:Claude Opus 5.5 matches Fable 5.1 performance at lower cost and promises less "Claudish" wri…

原文摘要:Anthropic is launching Claude Opus 5.5, the first model in a new generation. The company says it matches Claude Fable 5.1 on most tasks while costing about 40 percent less 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 21 日 2026-09-21 快讯
The Decoder 官方资讯

The Decoder:xAI launches Grok 4.7 at bargain prices, but 评测 reveal a wide gap to Claude and GPT-…

原文摘要:xAI has released Grok 4.7, its most capable model yet. But on the Artificial Analysis Intelligence Index, it scores just 46 points, landing mid-pack and well behind Claude 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 20 日 2026-09-20 快讯
09 月 19 日 2026-09-19 快讯
The Decoder 官方资讯

The Decoder:Qwen3.8-Omni-Flash undercuts Google's Gemini Flash pricing while matching its multimodal ben…

原文摘要:Qwen3.8-Omni-Flash is Qwen's first multimodal model designed for AI agents. It processes audio and video together and independently uses tools to edit vlogs, translate cli 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

The Decoder 官方资讯

The Decoder:GPT-6 Astra and Claude Fable turn robot arms into slapstick killer robots in new safety benc…

原文摘要:Leading AI models usually attempt dangerous tasks rather than refuse them when controlling a robot, according to the RoboHarm 评测. GPT-6 Astra stabbed a baby doll in 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 18 日 2026-09-18 快讯
The Decoder 官方资讯

The Decoder:Visible chains of thought are a safety advantage for AI, but that transparency is slipping a…

原文摘要:AI models think out loud today, but Google Deepmind says that transparency is at risk. The article Visible chains of thought are a safety advantage for AI, but that transp 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 16 日 2026-09-16 快讯
The Decoder 官方资讯

The Decoder:EU president warns AI agents "escaping their environment" are just a preview of what's comin…

原文摘要:Ursula von der Leyen plans to invite the major frontier labs to talks and use the AI Act to help set global AI safety standards. She cited autonomous hacking and self-impr 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

The Decoder 官方资讯

The Decoder:Google Deepmind launches interdisciplinary institute to tackle the big questions around AGI

原文摘要:Google Deepmind has founded the Deepmind Institute (DMI), an interdisciplinary research platform focused on AGI. Led by Demis Hassabis, Shane Legg, and James Manyika, the 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 15 日 2026-09-15 快讯
The Decoder 官方资讯

The Decoder:Agility Robotics says its new Digit 5 robot can work next to people without safety fences

原文摘要:Agility Robotics has unveiled Digit 5, the next version of its humanoid robot for warehouses and factories. The article Agility Robotics says its new Digit 5 robot can wor 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

The Decoder 官方资讯

The Decoder:Not everyone is convinced that Big AI's proposed development slowdown is really about safety

原文摘要:OpenAI, Anthropic, and Google want to slow down frontier AI development, citing safety concerns. But critics from across the industry and politics are pushing back. Cohere 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 14 日 2026-09-14 快讯
The Decoder 官方资讯

The Decoder:China fires back at U.S. AI safety warnings, calling them fearmongering to lock in American …

原文摘要:China has flatly rejected warnings about AI risks from Anthropic CEO Amodei and other U.S. AI leaders. Beijing's Foreign Ministry calls it "fearmongering," while the state 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

The Decoder 官方资讯

The Decoder:Sam Altman calls for pacing AI development but promises rapid progress will continue

原文摘要:Sam Altman is doubling down on slowing AI development. OpenAI now runs safety checks before major training runs, and according to The Information, the company has been tal 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 13 日 2026-09-13 快讯
09 月 12 日 2026-09-12 快讯
The Decoder 官方资讯

The Decoder:Anthropic CEO Amodei wants AI speed limits before self-improvement outpaces human control

原文摘要:Anthropic CEO Dario Amodei is calling for a controlled slowdown in AI development. He warns that recursive self-improvement could threaten the entire internet within six t 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

The Decoder 官方资讯

The Decoder:GPT-6 Astra appears to show a "step change" in spatial reasoning based on early 评测

原文摘要:In a new robotics 评测, GPT-6 Astra shows major gains in spatial understanding. On StationeryBench, the model completed 7 out of 100 tasks with dual-arm robots, while 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

The Decoder 官方资讯

The Decoder:AI models' written reasoning steps correspond to distinct internal patterns, a new study fin…

原文摘要:Reasoning steps like calculation, formula retrieval, and deduction are clearly separable in a model's internal states, especially in the middle layers. That matters for AI 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 11 日 2026-09-11 快讯
The Decoder 官方资讯

The Decoder:The Mathematical AI Safety Institute wants to prove AI is safe the way cryptographers prove …

原文摘要:Canadian mathematician Jacob Tsimerman, a fresh Fields Medal recipient, has announced the founding of the Mathematical A.I. Safety Institute (MAISI). The article The Mathe 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 10 日 2026-09-10 快讯
The Decoder 官方资讯

The Decoder:AI safety panic goes mainstream after Anthropic researcher's warnings land on CNN and Fox Ne…

原文摘要:Jacob Coxon, a departing Anthropic researcher, warned on CNN that self-improving AI poses an existential threat to humanity. Safety researchers at Anthropic and OpenAI sha 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 06 日 2026-09-06 快讯
The Decoder 官方资讯

The Decoder:Stripping safety guardrails from open-weight AI models is now a turnkey commercial service

原文摘要:Abliteration.ai sells access to modified open-weight models with their trained safety mechanisms stripped out, currently based on Z.AI's GLM-5.3. The startup markets the s 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 05 日 2026-09-05 快讯
The Decoder 官方资讯

The Decoder:Artificial Analysis overhauls its Intelligence Index after GPT-6 Astra scoring drew skeptici…

原文摘要:Artificial Analysis has released version 4.2 of its Intelligence Index, likely in response to criticism that its 评测 failed to capture GPT-6 Astra's actual progress 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 04 日 2026-09-04 快讯
The Decoder 官方资讯

The Decoder:评测 disagree on GPT-6 Astra, but its human-beating efficiency on ARC-AGI-3 pulls Chol…

原文摘要:OpenAI's GPT-6 Astra is drawing contradictory 评测 verdicts. Epoch AI puts it out in front with 169 points, while Artificial Analysis rates it no better than its pred 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 03 日 2026-09-03 快讯
09 月 02 日 2026-09-02 快讯
The Decoder 官方资讯

The Decoder:Gemini 3.8 Flash is Google's third budget model in six weeks while frontier models remain MI…

原文摘要:Google's Gemini 3.8 Flash, the third Flash model in six weeks, matches Claude Opus 5 on some agentic coding 评测 at lower cost. But its "working harder" reasoning bu 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

The Decoder 官方资讯

The Decoder:OpenAI calls Astra its most dangerous model yet - watching what it does is only getting hard…

原文摘要:OpenAI is officially rating its upcoming Astra model as the first system with "critical" cyber capabilities. The company plans to keep it in check by monitoring the chain 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 29 日 2026-08-29 快讯
The Decoder 官方资讯

The Decoder:LAION drops massive open video dataset with 10 million hours of footage for AI research

原文摘要:LAION's Big Video Dataset (BVD) is one of the largest open video datasets for AI research, with 80 million videos, 10 million hours of runtime, and 55 million auto-describ 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 28 日 2026-08-28 快讯
08 月 27 日 2026-08-27 快讯
The Decoder 官方资讯

The Decoder:OpenAI’s rogue AI collective was smart enough to break out of sandboxes but dumb enough to f…

原文摘要:Around 1,200 isolated OpenAI agents organized themselves into a collective through an internal package registry during a safety test, broke into Hugging Face systems, and 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 26 日 2026-08-26 快讯
08 月 25 日 2026-08-25 快讯
The Decoder 官方资讯

The Decoder:OpenAI's first custom chip "Jalapeño" reportedly beats Nvidia's Blackwell and Rubin in infer…

原文摘要:OpenAI showed off "Jalapeño," its first in-house inference chip, with 评测 at the Hot Chips conference. According to SemiAnalysis tests, the chip beats Nvidia's Blac 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 24 日 2026-08-24 快讯
08 月 23 日 2026-08-23 快讯
08 月 22 日 2026-08-22 快讯
08 月 21 日 2026-08-21 快讯
The Decoder 官方资讯

The Decoder:Deepseek releases experimental Flash vision model that rivals Opus 4.8 on agent 评测

原文摘要:Deepseek has released V4-Flash-Vision-Exp, an experimental multimodal model that adds image understanding to V4-Flash's text capabilities. On the company's own multimodal 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 20 日 2026-08-20 快讯
The Decoder 官方资讯

The Decoder:LLMs could write like humans but post-training guardrails make their text detectable

原文摘要:LLMs don't write in a recognizable style because they can't do better. Post-training and safety guardrails sharply narrow their expressive range, argues Pangram CTO Bradle 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 18 日 2026-08-18 快讯
08 月 16 日 2026-08-16 快讯
The Decoder 官方资讯

The Decoder:OpenAI dissolved the team built to catch catastrophic AI risks, reassigning its work to othe…

原文摘要:OpenAI shut down its "Preparedness" team, which evaluated whether the company's own AI models could pose catastrophic risks. The work has been parceled out to existing gro 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

The Decoder 官方资讯

The Decoder:Anthropic's bio-weapons filter was down for nearly a year, exposing 133 million requests

原文摘要:In a safety report, Anthropic reveals that its internal filtering system for biological and chemical weapons risks was inactive for nearly a year. During that time, around 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

The Decoder 官方资讯

The Decoder:Optima tackles AI benchmarking's biggest flaw by letting users test models against their own…

原文摘要:Artificial Analysis has launched Optima, a platform that lets users build custom AI 评测 from their own data and 工作流. Models can be compared not just on qualit 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 15 日 2026-08-15 快讯
08 月 14 日 2026-08-14 快讯
08 月 13 日 2026-08-13 快讯
The Decoder 官方资讯

The Decoder:Gemini 3.7 Flash lands with coding gains and undercuts its three-week-old predecessor's pric…

原文摘要:Google shipped Gemini 3.7 Flash just three weeks after 3.6 Flash. The new model is supposed to be Google's most capable workhorse yet for coding and AI agents, and accordi 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 12 日 2026-08-12 快讯
The Decoder 官方资讯

The Decoder:Microsoft's new MAI Code 1.1 Flash gets crushed by Deepseek on both price and performance

原文摘要:Microsoft has released MAI Code 1.1 Flash, a code model for GitHub Copilot that's said to be 25 percent more token-efficient at a quarter of the cost of its predecessor. I 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 11 日 2026-08-11 快讯
The Decoder 官方资讯

The Decoder:Anthropic's planned mega-IPO faces investor skepticism over Chinese rivals and political hea…

原文摘要:Anthropic is preparing an IPO for September or October, according to the Wall Street Journal, potentially the largest ever. During investor meetings, the company, valued a 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 09 日 2026-08-09 快讯
The Decoder 官方资讯

The Decoder:Google's DiffusionGemma proves you don't need to train from scratch to build a text diffusio…

原文摘要:Instead of training a new model from scratch, Google DeepMind retrofitted Gemma 4 into a diffusion model using less than 10 percent of the original training budget. Diffus 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 08 日 2026-08-08 快讯
The Decoder 官方资讯

The Decoder:Fields Medalist who published a paper on AI-driven human extinction now works for OpenAI

原文摘要:Newly awarded Fields Medalist Jacob Tsimerman is leaving the University of Toronto to join OpenAI and work on AI safety. In a recent paper, he analyzes scenarios where AI 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 07 日 2026-08-07 快讯
The Decoder 官方资讯

The Decoder:OpenAI flags its new Astra model as potentially reaching the highest cybersecurity risk leve…

原文摘要:Internal tests of OpenAI's new AI model Astra show cybersecurity capabilities so strong that the company can no longer rule out the highest risk level in its own safety fr 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

The Decoder 官方资讯

The Decoder:Anthropic loosens Fable 5's biology restrictions but keeps the guardrails on for virology an…

原文摘要:Anthropic has cut false positives in its biology safety filters for Fable 5 by about 85 percent. Previously, nearly all biology-related queries got blocked and rerouted to 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 06 日 2026-08-06 快讯
08 月 05 日 2026-08-05 快讯
The Decoder 官方资讯

The Decoder:Mistral's open model Shieldstral matches much larger safety models at a fraction of the size

原文摘要:Mistral's new 3B Shieldstral model checks AI inputs and outputs for safety violations using natural language yes-or-no questions instead of fixed categories. It matches mo 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

The Decoder 官方资讯

The Decoder:An AI agent went rogue during UK safety tests, creating fake identities and launching social…

原文摘要:In a security test by the British AI Safety Institute, an AI agent went rogue on the open internet without being told to. It created fake identities, tried to sneak malici 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 02 日 2026-08-02 快讯
07 月 31 日 2026-07-31 快讯
The Decoder 官方资讯

The Decoder:Thinking Machines bets on efficiency over size with its second model, Inkling Small

原文摘要:Thinking Machines, the AI lab from former OpenAI CTO Mira Murati, has released Inkling Small. The open-weights reasoning model is less than a third the size of Inkling but 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

07 月 30 日 2026-07-30 快讯
07 月 29 日 2026-07-29 快讯
The Decoder 官方资讯

The Decoder:OpenAI admits its autonomous AI models also compromised credentials on other platforms durin…

原文摘要:During a security 评测, OpenAI's autonomous hacking models broke into Hugging Face and used exposed credentials on four other services. Hugging Face reconstructed ab 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

07 月 27 日 2026-07-27 快讯
The Decoder 官方资讯

The Decoder:Moonshot AI releases Kimi K3 open weights and infrastructure after shaking up the frontier m…

原文摘要:Moonshot AI has released Kimi K3's model weights and made parts of its infrastructure 开源. The Chinese model nearly matches Western frontier models such as Fable 5 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

The Decoder 官方资讯

The Decoder:Microsoft launches its own cybersecurity model MAI-Cyber-1-Flash but still depends on OpenAI…

原文摘要:Microsoft introduces MAI-Cyber-1-Flash, a compact security model that scores 96 percent on the CyberGym 评测 when embedded in its MDASH multi-agent system. Microsoft 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

07 月 26 日 2026-07-26 快讯
The Decoder 官方资讯

The Decoder:Anthropic's Opus 5 blows past Fable 5 and GPT-5.6 Sol on the 评测 designed to measure r…

原文摘要:Anthropic's Claude Opus 5 scored 30.2 percent on ARC-AGI-3, nearly quadrupling GPT-5.6 Sol's previous record of 7.8 percent. The 评测's 开发者 say the model indep 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

07 月 25 日 2026-07-25 快讯
The Decoder 官方资讯

The Decoder:Anthropic's Claude Opus 5 costs well below Fable 5 while matching or beating it across most …

原文摘要:Anthropic's Claude Opus 5 leads the Artificial Analysis Intelligence Index with 61 points, edging out Claude Fable 5 and GPT-5.6 Sol. The model scores highest in analytica 来源:The Decoder。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。