raw / articles
articles
116 个原始文档
51CTO MiMo V2.5 OpenClaw Agent Harness 2026
Token效率国内第一!MiMo V2.5 Pro登顶开源Agent王者;罗福莉:OpenClaw是巨大分水岭,模型与Harness需同步演进,MLA不符合Agent范式 51CTO技术栈 玉澄 编辑 玉澄 今天凌晨,小米突然开源了其最强模型 MiMo V2.5 系列!其中 MiMo V2.5 Pro 在 Agent 基准测试上登顶开源第一,超过了 Dee…
机器之心 Vibe Coding Anthropic Erik Schluntz 2026
如何正确Vibe Coding?这是来自Anthropic编程智能体负责人的大师课 原创 机器之心 北京 编辑|+0、Panda 如果摔断了手、打了两个月石膏,工作却不能停,程序员该怎么办?Anthropic 的研究员、《构建高效智能体》合著者 Erik Schluntz 的答案是:全权交给 Claude。 如今,随着 AI 强势重塑软件行业的规则,Vib…
如何在 90 天内打造每月 1 万美元的 Claude 自动化业务
Source tweet: https://x.com/adrianpunk115/status/2051494055692116286?s=46 X Article: http://x.com/i/article/2051355782780977152 Author: Adrian Punk (@AdrianPunk115) Published: 202…
特朗普:人在北京,刚下飞机
作者/账号:笔记侠 / 老贾 发布时间:2026年5月13日 20:28 北京 注:微信公众号适配器 bb browser 抓取失败,改用 live browser 提取正文;这里保存的是经去重整理后的原文要点版,保留文章主要结构与关键事实。 内容来源:汇编至公开资料整理。 5月12日,空军一号停在安克雷奇机场的停机坪上。黄仁勋从黑色SUV里出来,登上飞往…
新智元 Karpathy 锯齿状智能 Sequoia 2026
AI能改10万行代码,却让你走路去洗车!Karpathy戳破「锯齿状智能」 新智元报道 2026年4月29日 今天最先进的大模型,可以一口气重构一个10万行的代码库,也会在你想要洗车的时候,建议你走路去50米外的洗车店。 为什么同一个模型,一会儿它表现得像一位超级工程师,一会儿却又像一个刚毕业的实习生? 这是Karpathy近日在Sequoia AI As…
A framework for AI development transparency
Overview Anthropic proposes a targeted transparency framework for frontier AI development to ensure public safety and accountability during the period before comprehensive safety …
Accenture and Anthropic Multi-Year Partnership Summary
Overview Accenture and Anthropic have announced a major partnership expansion to help enterprises transition from AI pilots to full scale production deployment. The collaboration …
Activating AI Safety Level 3 Protections
Anthropic has proactively activated its AI Safety Level 3 (ASL 3) Deployment and Security Standards for the launch of Claude Opus 4 . This is a precautionary measure, as the compa…
AGI Hunt 黄仁勋炮轰Anthropic CEO 2026
老黄穿着西装上了个播客,然后开炮了。 这期播客叫「Memos to the President」,由华盛顿 AI 智库 SCSP(特别竞争研究项目)主办。主持人问:AI 行业是不是正在被「污名化」了? 01 别吓跑放射科医生 十年前 Hinton 预测 AI 将取代放射科医生,结果十年后放射科反而出现用工荒。黄仁勋:说服人不做放射科医生是有害的;说服人不学…
AI Agent 模型横测 GPT DeepSeek MiniMax Xiaomi 2026
千元横测GPT、DeepSeek、Xiaomi、MiniMax的最强模型,我找到了跟Agent们的绝配 卡尔的AI沃茨 2026年4月30日 事情是酱的 这天我在AA榜上看前28的模型感到有点陌生。上周太集中发的后果就是光在用GPT 5.5了,小米的Mimo V2.5 Pro,DeepSeek V4 Pro还没有放在Agent的场景上测。 所以我跟钱包一拍…
Aligning on child safety principles
Apr 23, 2024 Alongside other leading AI companies, we are committed to implementing robust child safety measures in the development, deployment, and maintenance of generative AI t…
Anthropic acquires Bun as Claude Code reaches $1B milestone
Dec 3, 2025 Claude is the world's smartest and most capable AI model for developers, startups, and enterprises. Claude Code represents a new era of agentic coding, fundamentally c…
Anthropic and Amazon expand collaboration for up to 5 gigawatts of new compute
Apr 20, 2026 We have signed a new agreement with Amazon that will deepen our existing partnership and secure up to 5 gigawatts (GW) of capacity for training and deploying Claude, …
Anthropic and NEC collaborate to build Japan's largest AI engineering workforce
Apr 24, 2026 NEC Corporation will use Claude as it builds one of Japan's largest AI native engineering organizations, making it available to approximately 30,000 NEC Group employe…
Anthropic Announcements: New Models & Computer Use
Overview On October 22, 2024, Anthropic announced three major developments: 1. An upgraded Claude 3.5 Sonnet model. 2. A new Claude 3.5 Haiku model. 3. A groundbreaking computer u…
anthropic claude opus 4 7 release 2026
Title: Introducing Claude Opus 4.7 URL Source: https://www.anthropic.com/news/claude opus 4 7 Markdown Content: Our latest model, Claude Opus 4.7, is now generally available. Opus…
Anthropic Economic Futures Program Launch
Date: June 27, 2025 Program Overview Anthropic has launched the Anthropic Economic Futures Program , a new initiative to support research and policy development focused on AI's ec…
Anthropic Economic Index: Insights from Claude 3.7 Sonnet
Source: Anthropic Date: Mar 27, 2025 Report: Second release from the Anthropic Economic Index, analyzing usage data from Claude.ai following the launch of Claude 3.7 Sonnet. Key F…
Anthropic Economic Index: Summary
Overview The Anthropic Economic Index is a new initiative launched on February 10, 2025, to understand AI's effects on labor markets and the economy over time. Its initial report …
Anthropic Election Safeguards Update (April 24, 2026)
Core Objective To ensure Claude provides accurate, impartial, and balanced information during elections, acting as a positive force for the democratic process. Key Safeguards and …
Anthropic expands partnership with Google and Broadcom for multiple gigawatts of next-generation compute
Apr 6, 2026 We have signed a new agreement with Google and Broadcom for multiple gigawatts of next generation TPU capacity that we expect to come online starting in 2027. This sig…
Anthropic Invests $100M in Claude Partner Network
Date: March 12, 2026 Source: Anthropic News Overview Anthropic has launched the Claude Partner Network , a program designed to support partner organizations that help enterprises …
Anthropic is endorsing SB 53
Date: September 8, 2025 Core Announcement Anthropic officially endorses SB 53 , a California bill governing powerful AI systems from frontier developers. This support follows less…
Anthropic Partners with Allen Institute and HHMI to Accelerate Scientific Discovery
Date: February 2, 2026 Core Challenge: Modern biological research generates data at an unprecedented scale, but transforming it into validated insights remains a bottleneck due to…
Anthropic Updates to Consumer Terms and Privacy Policy
Effective Date: August 28, 2025 Action Required by: October 8, 2025 (for existing users) Key Changes & User Choice Anthropic is updating its Consumer Terms and Privacy Policy to a…
Anthropic's Approach to Understanding and Addressing AI Harms
Overview Anthropic has shared its evolving framework for assessing and mitigating a broad spectrum of potential AI harms, from catastrophic risks to everyday concerns. This approa…
Anthropic's Collaboration with US CAISI & UK AISI
Source: Anthropic News, Sep 12, 2025 Core Focus: Strengthening AI model safeguards through government partnership. Overview Anthropic has established an ongoing partnership with t…
Anthropic's Responsible Scaling Policy (RSP) Version 3.0 - Summary
Source: Anthropic, February 24, 2026 Purpose: A voluntary framework to mitigate catastrophic risks from AI systems, updated to reinforce successes, address shortcomings, and incre…
Anthropic's Updated Responsible Scaling Policy (RSP) Summary
Core Purpose & Commitment Risk Governance Framework : Updated RSP mitigates potential catastrophic risks from frontier AI systems. Core Commitment : "We will not train or deploy m…
Apple's Xcode now supports the Claude Agent SDK
Apple's Xcode now supports the Claude Agent SDK \ Anthropic Product Apple's Xcode now supports the Claude Agent SDK Feb 3, 2026 Apple's Xcode is where developers build, test, and …
Best Practices for Claude Code - Claude Code Docs
Title: Best Practices for Claude Code Claude Code Docs URL Source: https://www.anthropic.com/engineering/claude code best practices Markdown Content: Best Practices for Claude Cod…
Building Safeguards for Claude
Claude is designed to amplify human potential while ensuring capabilities are channeled toward beneficial outcomes. The Safeguards team identifies misuse, responds to threats, and…
Challenges in Red Teaming AI Systems
June 12, 2024 Overview Red teaming is the adversarial testing of technological systems to identify vulnerabilities. Anthropic emphasizes that while red teaming is critical for AI …
Claude 2
Overview Announcement Date: July 11, 2023 Key Product: Claude 2, a new AI model from Anthropic Access Points: API for businesses and a new public beta website at claude.ai Availab…
Claude 2.1
Release Date: November 21, 2023 Availability: API (via Console) and powering the claude.ai chat experience (free and Pro tiers). Key Advancements 1. 200K Token Context Window Capa…
Claude 3 Haiku: our fastest model yet
Mar 13, 2024 Today we're releasing Claude 3 Haiku, the fastest and most affordable model in its intelligence class. With state of the art vision capabilities and strong performanc…
Claude 3.5 Sonnet on GitHub Copilot
Oct 29, 2024 Starting today, the new Claude 3.5 Sonnet begins rolling out on GitHub Copilot, enabling developers to choose Claude 3.5 Sonnet for coding—directly in Visual Studio C…
Claude 3.7 Sonnet & Claude Code Summary
Key Announcement Claude 3.7 Sonnet is Anthropic's most intelligent model to date and the first hybrid reasoning model on the market. It can produce near instant responses or exten…
Claude and Alexa+
Feb 26, 2025 Today, we're announcing that Claude models are helping power Alexa+. This collaboration is part of our ongoing partnership with Amazon to deliver advanced AI technolo…
Claude for Creative Work - Summary
Source: Anthropic (April 28, 2026, updated May 1, 2026) Key Announcement Anthropic has released a set of connectors that integrate Claude directly into the software used by creati…
Claude Haiku 4.5 Summary
Overview Model : Claude Haiku 4.5, Anthropic's latest small model. Release Date : October 15, 2025. Key Positioning : Delivers near frontier performance (comparable to the former …
Claude in Amazon Bedrock: Approved for use in FedRAMP High and DoD IL4/5 workloads
Jun 11, 2025 Today, we're announcing that Claude models are approved for use in FedRAMP High and DoD Impact Level 4 and 5 workloads through Amazon Bedrock in AWS GovCloud (US) reg…
Claude is now generally available in Xcode
Sep 15, 2025 Developers can now connect their Claude account to Xcode 26 to power coding intelligence features with Claude Sonnet 4. Xcode is Apple's integrated development enviro…
Claude now available in Microsoft Foundry and Microsoft 365 Copilot
Nov 18, 2025 Today we announced that Microsoft and Anthropic are expanding our partnership. As part of the partnership, Claude Sonnet 4.5, Haiku 4.5, and Opus 4.1 models are now a…
Claude Opus 4.1
Aug 5, 2025 Today we're releasing Claude Opus 4.1, an upgrade to Claude Opus 4 on agentic tasks, real world coding, and reasoning. We plan to release substantially larger improvem…
Claude Opus 4.6 Summary
Overview Anthropic has released Claude Opus 4.6 , its most advanced model to date, featuring significant improvements in coding, reasoning, and long context performance. It is ava…
Claude Sonnet 4.6: Summary
Release Date: February 17, 2026 Key Highlights Most capable Sonnet model yet , with full upgrades in coding, computer use, long context reasoning, agent planning, knowledge work, …
Claude's Constitution: Summary
Overview Anthropic's Constitutional AI (CAI) is a method for training AI models like Claude using an explicit set of principles (a "constitution") to guide behavior, rather than r…
Claude's New Constitution: Summary
Overview Anthropic has published a new, detailed constitution for its AI model, Claude. This document serves as the foundational guide for Claude's values, behavior, and identity,…
Codex goal 功能 2026
Codex 推出 /goal 功能,不达目标,不罢休 AGI Hunt J0hn 2026年5月1日 OpenAI 给 Codex CLI 加了个新命令,叫 /goal。你给它设一个目标,它就一直跑,跨多轮不丢上下文,不达目的不罢休。 这个功能随 Codex CLI 0.128.0 版本发布,目前还是实验性功能,需要手动开启。 Ralph Loop 要理解…
Collaborate with Claude on Projects
Jun 25, 2024 Our vision for Claude has always been to create AI systems that work alongside people and meaningfully enhance their workflows. As a step in this direction, Claude.ai…
Contextual Retrieval in AI Systems
Sep 19, 2024 Core Problem Traditional Retrieval Augmented Generation (RAG) systems often fail because they remove context when encoding information into chunks, leading to poor re…
Cursor 如何改进模型 Harness 2026 05 01
Cursor如何改进模型Harness 釉蓝 Yoryon 2026年5月1日 2026 年 4 月 30 日,Cursor 工程团队发了一篇博客 Continually improving our agent harness,介绍他们如何迭代 agent harness。 Harness 指模型和用户之间的中间层:系统提示词、工具描述、上下文管理、错误处…
DeepSeek V4最大的遗憾
henry 发自 凹非寺量子位 公众号 QbitAI DeepSeekV4的技术报告里有mHC,有CSA,有HCA,有Muon,有FP4…… 唯独没有Engram。 Engram去哪了? 这个话题一度成为网友们讨论的热点。 Engram在今年1月由DeepSeek和北大联合开源,主要研究大模型的记忆与效率问题。 自挂上arXiv的那一刻起,圈子里围绕它的探…
Deloitte will make Claude available to 470,000 people across its global network
Oct 6, 2025 Anthropic and Deloitte today announced an expanded alliance that will make Claude available to Deloitte people across its global network and develop new industry speci…
Detecting and Countering Malicious Uses of Claude: March 2025
Anthropic is committed to preventing misuse of Claude by adversarial actors while maintaining utility for legitimate users. Threat actors continuously explore methods to circumven…
Developing a computer use model
Oct 22, 2024 Claude 3.5 Sonnet can now interact with a computer like a human—moving a cursor, clicking, and typing via a virtual keyboard—by interpreting screenshots. This is a si…
Developing nuclear safeguards for AI through public-private partnership
Aug 21, 2025 Read the full post on red.anthropic.com Nuclear technology is inherently dual use: the same physics principles that power nuclear reactors can be misused for weapons …
Disrupting the first reported AI-orchestrated cyber espionage campaign
Overview Incident: In mid September 2025, Anthropic detected and disrupted a sophisticated, large scale cyber espionage campaign. Attribution: Assessed with high confidence to be …
Donating the Model Context Protocol and establishing the Agentic AI Foundation
Donating the Model Context Protocol and establishing the Agentic AI Foundation \ Anthropic Announcements Donating the Model Context Protocol and establishing the Agentic AI Founda…
Enabling Claude Code to work more autonomously
Sep 29, 2025 We're introducing several upgrades to Claude Code: a native VS Code extension, version 2.0 of our terminal interface, and checkpoints for autonomous operation. Powere…
Expanding our model safety bug bounty program
Aug 8, 2024 The rapid progression of AI model capabilities demands an equally swift advancement in safety protocols. As we work on developing the next generation of our AI safegua…
Frontier Model Security - Anthropic Summary
Overview Anthropic emphasizes that securing frontier AI models is a critical priority due to their rapidly increasing capabilities and strategic importance. The goal is to protect…
Frontier Threats Red Teaming for AI Safety - Summary
Overview Anthropic details its approach to "frontier threats red teaming," a specialized form of adversarial testing focused on national security risks (e.g., biosecurity, cyberse…
Golden Gate Claude
May 23, 2024 UPDATE: Golden Gate Claude was online for a 24 hour period as a research demo and is no longer available. If you'd like to find out more about our research on interpr…
GPT-5.5 与 Claude 4.7 提示词优化指南(2026)
范式转移:从「过程导向」到「结果导向」 2026 年的提示词工程正在经历一次根本性的范式转移。传统的提示词设计关注的是 过程 ——告诉模型"怎么做"(step by step 指令、链式思考、角色扮演)。而随着 GPT 5.5([[openai]])和 [[Claude Opus 4.7]] 等前沿模型推理能力的飞跃,最佳实践正在转向 结果导向 ——告诉模…
How Scientists Are Using Claude to Accelerate Research and Discovery
Overview Anthropic's Claude for Life Sciences suite and AI for Science program (providing free API credits) are enabling scientists to use Claude as a collaborative research partn…
Introducing 100K Context Windows
May 11, 2023 We've expanded Claude's context window from 9K to 100K tokens, corresponding to around 75,000 words! This means businesses can now submit hundreds of pages of materia…
Introducing Anthropic's AI for Science Program
May 5, 2025 Today, we're launching Anthropic's AI for Science program – a new initiative designed to accelerate scientific research and discovery through access to our API. This p…
Introducing Anthropic's first developer conference: Code with Claude
Apr 3, 2025 Today, we're announcing Code with Claude—our first developer conference—taking place on May 22, 2025 in San Francisco. Code with Claude is a hands on, one day event fo…
Introducing Claude 4: Summary
Overview Anthropic has announced the next generation of Claude models: Claude Opus 4 and Claude Sonnet 4 , released on May 22, 2025. These models set new standards for coding, adv…
Introducing Claude Design by Anthropic Labs
Announced: April 17, 2026 Product: Claude Design, a new Anthropic Labs product for collaborative visual creation. Core Model: Powered by Claude Opus 4.7 (Anthropic's most capable …
Introducing Claude Pro
Sep 7, 2023 Today, we're introducing a paid plan for our Claude.ai chat experience, currently available in the US and UK. Since launching in July, users tell us they've chosen Cla…
Introducing Claude Sonnet 4.5
Source: Anthropic Date: Sep 29, 2025 Core Announcement Claude Sonnet 4.5 is Anthropic's latest model, positioned as the best coding model in the world and the strongest for buildi…
Introducing Claude: Anthropic's Next-Generation AI Assistant
Overview Anthropic announced the broader availability of Claude , a next generation AI assistant based on research into training helpful, honest, and harmless AI systems . After a…
Introducing Labs
Jan 13, 2026 Our models are evolving at a rapid clip, and each new release brings another leap in capabilities. Building product experiences around these emerging capabilities req…
Introducing the Model Context Protocol
Nov 25, 2024 Today, we're open sourcing the Model Context Protocol(MCP), a new standard for connecting AI assistants to the systems where data lives, including content repositorie…
Introducing the next generation of Claude
Overview Anthropic announced the Claude 3 model family on March 4, 2024, setting new industry benchmarks. The family includes three models in ascending order of capability: Haiku …
lawrencew zen hero coding agent mvp 2026
教你跑通一个全自动 Coding Agent 的 MVP —— 大模型规划,小模型干活 我最近想验证一个反直觉的猜想: 在多 agent 编排里,非思考小模型反而比思考大模型更省钱。 为了验证这件事,我搭了个最小 MVP:hero coding。然后用同一个 harness 跑了三组模型:ChatGPT 5.4、Ling 2.6 flash、Ling 2.…
LLNL Expands Claude for Enterprise Deployment
Key Announcement Lawrence Livermore National Laboratory (LLNL) is expanding its deployment of Claude for Enterprise to its entire laboratory, making advanced AI capabilities avail…
Measuring political bias in Claude
Overview Anthropic details its approach to training and evaluating Claude for political even handedness —treating opposing viewpoints with equal depth, engagement, and quality. Th…
Notes from inside China's AI labs
Lessons from my trip to talk to most of the leading AI labs in China. Staring out the window on a new, high speed train from Hangzhou to Shanghai I’m gifted with views of dramatic…
Our framework for developing safe and trustworthy agents
Source: Anthropic (Aug 4, 2025) Core Idea: As AI evolves from assistants to autonomous agents, Anthropic shares a framework for responsible development, emphasizing safety, reliab…
Partnering with Mozilla to Improve Firefox's Security
Source: Anthropic (Mar 6, 2026) Core Finding: AI models (Claude Opus 4.6) can now independently identify high severity software vulnerabilities at unprecedented speed, demonstrati…
Progress from Anthropic's Frontier Red Team
Source: Anthropic Blog, March 19, 2025 Core Finding: Frontier AI models show "early warning" signs of rapid progress in dual use capabilities, approaching or exceeding undergradua…
Prompt Engineering for Business Performance - Summary
Executive Summary Core Value : Effective prompt engineering optimizes Claude's outputs, reduces deployment costs, and ensures on brand customer experiences. Proven Impact : A Fort…
Prompt Engineering for Claude's Long Context Window
Source: Anthropic (Sep 23, 2023) Core Focus: Techniques to maximize Claude's recall over its 100,000 token context window. Key Techniques Tested Two primary methods were evaluated…
Protecting the Wellbeing of Our Users: Anthropic's Safeguards
Overview Anthropic details its measures to ensure Claude handles sensitive conversations appropriately, focusing on suicide/self harm and sycophancy . The goal is to respond with …
Reflections on our Responsible Scaling Policy
Overview Anthropic published its first Responsible Scaling Policy (RSP) to address catastrophic safety failures and misuse of frontier models. The policy aims to turn high level s…
Releasing Claude Instant 1.2
Aug 9, 2023 Businesses working with Claude can now access our latest version of Claude Instant, version 1.2, available through our API. Claude Instant is our faster, lower priced …
ServiceNow Partners with Anthropic to Integrate Claude Across Platform and Workforce
Date: January 28, 2026 Source: https://www.anthropic.com/news/servicenow anthropic claude Partnership Overview ServiceNow has selected Anthropic's Claude as its default model for …
Shann Holmberg: AI Marketing Engineer Workflow
Main Tweet (2026 05 03) this is how you become an AI Marketing Engineer 1. prototype the workflow in hermes agent or openclaw. 2. run it two or three times against real work, corr…
sharbel hermes agent 15 features 2026
Title: Sharbel on X: "15 Hermes Agent features you've never touched" / X URL Source: https://x.com/sharbel/status/2049158152709382177 Published Time: Fri, 01 May 2026 15:01:51 GMT…
Sharing our compliance framework for California's Transparency in Frontier AI Act
Anthropic has published its Frontier Compliance Framework (FCF) to comply with California's Transparency in Frontier AI Act (SB 53) , which takes effect January 1, 2025 . This is …
Snowflake and Anthropic $200M Partnership Summary
Core Announcement Partnership: Multi year, $200 million strategic agreement between Snowflake and Anthropic. Goal: Bring agentic AI to global enterprises, enabling complex, multi …
Summary: Anthropic's Claude Code Security
Overview Anthropic has launched Claude Code Security , a new capability integrated into Claude Code on the web, now available in a limited research preview . It is designed to hel…
Summary: Anthropic's Responsible Scaling Policy (RSP)
Overview Anthropic has published its Responsible Scaling Policy (RSP) , a framework of technical and organizational protocols to manage catastrophic risks from increasingly capabl…
Summary: Claude 3.5 Sonnet Launch
Key Announcement Model : Claude 3.5 Sonnet is the first release in the Claude 3.5 model family . Performance : Outperforms competitor models and Claude 3 Opus on a wide range of e…
Summary: Claude Code & New Admin Controls for Business Plans
Key Features & Updates Premium Seats : Enterprise and Team plans can now upgrade to premium seats that bundle Claude (the app) and Claude Code (the coding agent) under one subscri…
Summary: Claude Opus 4.7 Release
Overview Model : Claude Opus 4.7, generally available as of April 16, 2026. Positioning : A significant upgrade to Opus 4.6, particularly for advanced software engineering and com…
Summary: Detecting and Preventing Distillation Attacks
Overview Anthropic has identified industrial scale, illicit distillation campaigns by three AI labs: DeepSeek, Moonshot AI, and MiniMax . These labs used over 24,000 fraudulent ac…
Summary: Introducing Claude Opus 4.5
Core Announcement Model : Claude Opus 4.5, released Nov 24, 2025 . Claim : "The best model in the world for coding, agents, and computer use," with significant improvements in eve…
Summary: The Case for Targeted Regulation - Anthropic
Core Argument Anthropic argues for urgently enacted, narrowly targeted AI regulation within the next 18 months to mitigate catastrophic risks (e.g., cyber, CBRN misuse) while pres…
Tailscale 显示 relay 不直连?90%是这个原因|附修复方法
Metadata Title: Tailscale 显示 relay 不直连?90%是这个原因|附修复方法 Author: 是万能的三叔 URL: https://www.bilibili.com/video/BV1cUdnBiEQ3/ Source: Bilibili video transcript (AI subtitle) Cleaned Read…
Teknium 推文:Hermes Agent Kanban 多智能体协调系统
作者: Teknium (@Teknium) — Nous Research 联合创始人 时间: 2026 05 04 02:09 (北京时间) 原文: Our first dive into Multi Agent Coordination and Cooperation is here, with Hermes Agent Kanban Orchest…
Testing and mitigating elections-related risks
1. The Core Testing Framework Anthropic utilizes a three stage iterative process to identify and remediate risks: A. Policy Vulnerability Testing (PVT) In depth, qualitative testi…
Testing Our Safety Defenses with a New Bug Bounty Program
Today, we're launching a new bug bounty program to stress test our latest safety measures. Similar to the program we announced last summer, we're challenging researchers to find u…
The 170-Line SOUL.md That Made My Hermes Agent Dangerous
Author: Tony Simons (@tonysimons ) Source: https://x.com/tonysimons /status/2051473178682118241?s=46 Article URL: http://x.com/i/article/2051460508943867904 Published: 2026 05 05T…
Third-Party Testing as a Key Ingredient of AI Policy
Source: Anthropic (March 25, 2024) Core Thesis: Effective third party testing for frontier AI systems is essential to prevent societal harm and is the best foundation for AI polic…
Threat Intelligence Report: Detecting and Countering Misuse of AI (August 2025)
Source: Anthropic Date: August 27, 2025 Key Finding: Threat actors have transitioned from using AI for advice to using Agentic AI for active operational execution, lowering techni…
U.S. Elections Readiness
Oct 8, 2024 2024 marks the first United States (U.S.) election cycle where generative AI tools are widely available. Since July 2023, we have taken concrete steps to help detect a…
Updating our Usage Policy
May 10, 2024 Today, we're updating the policies that protect our users and ensure our products and services are used responsibly. Our goal with these updates is to clarify which a…
Updating restrictions of sales to unsupported regions
Sep 4, 2025 Anthropic's Terms of Service prohibit use of our services in certain regions due to legal, regulatory, and security risks. However, companies from these restricted reg…
Usage policy update
Aug 15, 2025 Today, we're sharing some updates to our Usage Policy that reflect the growing capabilities and evolving usage of our products. Our Usage Policy serves as a framework…
Using Claude Code: The Unreasonable Effectiveness of HTML
Author: Thariq (@trq212) Tweet ID: 2052809885763747935 Tweet URL: https://x.com/trq212/status/2052809885763747935?s=46 Article URL: http://x.com/i/article/2052796100608974848 Publ…
Where the goblins came from | OpenAI
Title: Where the goblins came from URL Source: https://openai.com/index/where the goblins came from/ Markdown Content: Where the goblins came from OpenAI $1 [](https://openai.com/…