KNOWLEDGE NOTES

Knowledge Notes

Explore curated knowledge, course notes and methods, with their existing content and sources preserved.

These notes are in Chinese; English translations are not yet available.

287knowledge notes
AI Engineering / Anthropic 与 Claude

Anthropic AI 安全 / 对齐 / 政策研究全集(Constitutional AI + RSP + Welfare + Policy)

Anthropic research 板块 11 篇 hub:Constitutional AI(2022 foundational)+ Automated Alignment Researchers(2026 scalable oversight)+ Petri 开源(donating-open-source-petri)+ deprecation commitments(模型不删)+ end-subset-conversations(Claude welfare)+ Teaching Claude Why(agentic misalignment 后续)+ Project Vend Phase 2(Claudius 红队失败实验)+ 2028 AI Leadership(US-China)+ Anthropic Institute 4 大研究 agenda + Core Views on AI Safety + ASL3 全景。Anthropic 安全立场 5 大支柱:Constitutional + RSP/ASL / 透明 / Welfare / Red Team

AI安全模型对齐治理Chinese notes