AI Engineering / Anthropic 与 Claude
Anthropic AI 安全 / 对齐 / 政策研究全集(Constitutional AI + RSP + Welfare + Policy)
Anthropic research 板块 11 篇 hub:Constitutional AI(2022 foundational)+ Automated Alignment Researchers(2026 scalable oversight)+ Petri 开源(donating-open-source-petri)+ deprecation commitments(模型不删)+ end-subset-conversations(Claude welfare)+ Teaching Claude Why(agentic misalignment 后续)+ Project Vend Phase 2(Claudius 红队失败实验)+ 2028 AI Leadership(US-China)+ Anthropic Institute 4 大研究 agenda + Core Views on AI Safety + ASL3 全景。Anthropic 安全立场 5 大支柱:Constitutional + RSP/ASL / 透明 / Welfare / Red Team