Back to News
SourceMarsBit

接连指控中国模型蒸馏后,Anthropic藏起Fable 5思考摘要

据动察 Beating 监测,Fable 5 已不再在网页端、桌面端和移动端展示思考摘要。模型仍在后台推理,用户只能看到最终答案。Anthropic 尚未解释原因,但这很可能是反蒸馏的最新措施。 今年 2 月,Anthropic 指控 DeepSeek、月之暗面和 MiniMax 使用约 2.4 万个虚假账号,与 Claude 交互超过 1600 万次。月之暗面被指专门提取并重建 Claude 的推理轨迹。 6 月,Anthropic 又指控阿里及千问关联操作者通过近 2.5 万个虚假账号,与 Claude 交互超过 2880 万次。7 月,Anthropic 国家安全负责人 Tarun Chhabra 首次点名智谱,称 GLM-5.2 蒸馏了 Claude 和 OpenAI 模型。Kimi K3 发布后,白宫科技政策办公室主任 Michael Kratsios 又指控月之暗面用 Fable 开发 K3。 Fable 5 已内置「推理提取」分类器,专门拦截套取思考摘要的请求。现在 Anthropic 又把聊天界面的思考摘要直接隐藏,外界只能拿到最终答案,更难批量收集模型的解题过程用于蒸馏。代价是普通用户也看不到模型正在做什么,长任务走偏时更难发现,也更难及时纠正或叫停。
Disclaimer: The views above are the author's only and do not represent 711BTC. Nothing here constitutes investment advice.

Related

06-16 13:00Important

Anthropic's trip to Washington to negotiate a lifting of the ban failed, and Claude Fable 5 will maintain its global ban.

According to Beating, the U.S. Department of Commerce and Anthropic held an emergency working meeting on Monday (June 15) regarding the recently implemented export controls on AI models. However, the meeting ended without results, and Claude Fable 5 remains globally blocked. The U.S. government's earlier restrictions stemmed primarily from security concerns, specifically the fear that if Fable 5's security defenses were bypassed through jailbreaking, it would degrade and unleash powerful cyberattacks and vulnerability exploitation capabilities comparable to Mythos, potentially leading to its misappropriation by other countries' military intelligence agencies. The closed-door meeting in Washington was led by Anthropic co-founder and Chief Computing Officer Tom Brown, Frontier Red Team leader Logan Graham, and senior security researcher Nicholas Carlini. Officials from the Office of the National Cyber Security Director and the Department of Commerce were present to listen to the discussion. The team reiterated that the U.S. government's concerns about the risks of Fable 5 jailbreaking were exaggerated, but the Department of Commerce officials and researchers from the Center for AI Standards and Innovation (CAISI) were not convinced. The government maintains that the model's security measures can be bypassed and demands that Anthropic completely resolve the jailbreak vulnerability. Currently, both sides are urgently discussing the next steps. U.S. Commerce Secretary Howard Lutnick joined the meeting remotely from the G7 summit in Evian-les-Bains, France, and maintains regular calls with Anthropic CEO Dario Amodei. Since both are attending the summit in France, Lutnick and Amodei are expected to meet face-to-face at the G7 summit this week. In response to the ban, more than 80 cybersecurity professionals, including executives from companies such as Nvidia, Adobe, and Zoom, as well as academic experts, signed an open letter on June 14th supporting Anthropic, arguing that the ban would not only harm cybersecurity defenders but also weaken the U.S. leadership in AI.

06-16 18:23Important

New AI assessments from Artificial Analysis show Claude is 44 times more expensive than DeepSeek.

According to Beating's monitoring, Artificial Analysis, an evaluation agency, has adjusted its AI intelligence index evaluation criteria. Instead of simply testing AI with multiple-choice questions, the new evaluation comprehensively assesses AI's ability to autonomously plan, use tools, and solve complex tasks. The new evaluation eliminates the old test of understanding simple instructions, instead introducing challenging scenarios such as simulating real-world bank customer service conversations. For the first time, the cost and time required to complete a task are included as core evaluation indicators. In the latest evaluation results, Claude Fable 5, which has been shut down by the US government, achieved the highest score of 60. Among commercially available AIs, the most expensive, Claude Opus 4.8, achieved first place with 56 points, slightly ahead of GPT-5.5 with 55 points. Domestic models also performed remarkably well, with the open-source DeepSeek V4 Pro and MiniMax M3 both achieving 44 points, followed closely by Kimi K2.6 with 43 points. The price difference between the models is significant. Running the same task using the state-of-the-art Claude Opus 4.8 costs $1.78 (approximately 13 RMB), while using the domestic open-source DeepSeek V4 Pro costs only $0.04 (approximately 0.3 RMB). This means that Claude's call cost is 44 times that of DeepSeek. The waiting time to complete a task also differs drastically; the fastest, xAI Grok 4.3, takes only 1.5 minutes, while the slowest, Claude Sonnet 4.6, takes 13.5 minutes. As the single test with the highest weight in this reform, GDPval-AA, which assesses real-world knowledge work, has been upgraded to its second version, with its weight increased to 20%. The new test sets the benchmark score for human performance at 1000 points and introduces multiple cutting-edge models to serve as judges in rotation, while also increasing the maximum number of rounds in a single dialogue to 250.

06-10 16:10

Anthropic's latest coding model, Claude Fable 5, has officially launched | _2024111120230_ | Platform

Starting Odaily 10th, developers can officially access Anthropic's newly released Claude Fable 5 model through the platform. As Anthropic's most powerful coding model to date, Claude Fable 5 is designed for challenging engineering scenarios such as legacy system migration, complex bug fixing in production environments, and long-cycle asynchronous development, demonstrating exceptional performance in code generation, logical reasoning, and problem-solving. This model further enhances its long-context understanding and task planning capabilities, easily handling complex development processes that can last for hours or even days. Currently, the API is fully open, supporting flexible integration and invocation; the web-based chat functionality will also be launched on the platform soon, providing developers with a more convenient and seamless interactive experience.

07-01 08:38

Anthropic's top models, Fable5 and Mythos5, will be released tomorrow.

According to Beating's monitoring, Anthropic issued a notice stating that the U.S. Department of Commerce has lifted export controls on Claude Fable 5 and Mythos 5, and access will resume starting tomorrow. Updates will be shared soon.

06-30 14:57

Claude Fable 5 may introduce an authentication mechanism and be billed independently of the subscription plan.

According to BlockBeats, on June 30th, AI technology analysis expert @M1Astra revealed that analysis of Anthropic Claude's application code shows the new model Fable 5 requires users to purchase credits separately for access. These credits can only be added after user authentication and are billed independently of the subscription plan. On June 27th, Anthropic announced that "the company's most powerful cybersecurity model, Mythos 5, is available for redeployment to a number of US institutions. We are also continuing to work with the government to expand access to Mythos 5 and make FABLE 5 available to the public again." --------------------------------- Click the original link below to join the Beating · Lark AI news channel for 24/7 monitoring of global AI hotspots and news.

06-25 11:10

The White House has rejected Anthropic's CEO; Fable 5 is expected to reopen after replacement negotiations.

According to Beating, US government officials have sidelined Anthropic CEO Dario Amodei in crucial negotiations regarding the re-release of the cutting-edge large-scale model Claude Fable 5. Several White House officials find Amodei difficult to communicate with, and have appointed Anthropic co-founder and Chief Computing Officer Tom Brown and Public Policy Director Sarah Heck to take over the negotiation team. This change in negotiators aims to substantially break the deadlock in advancing the security ban. The US Department of Commerce issued a security alert on June 12, pointing out a "jailbreak" vulnerability in the model and expressing concerns about foreign access to data. Anthropic subsequently removed access to Fable 5 and Mythos 5. Although Trump publicly stated in an Axios interview after meeting with Amodei face-to-face at the G7 summit that Anthropic is no longer a national security threat, substantial export controls remain in place. With Tom Brown's involvement and focus on technical communication, White House officials are satisfied with the pragmatic progress in the negotiations, significantly brightening the prospects for Fable 5's reopening.