Constitutional AI (CAI) / AI Hiến pháp (CAI)
📖 Constitutional AI (CAI)
Definition (English):
Constitutional AI, developed by Anthropic, is an alignment technique where an AI model is trained to follow a set of principles (a "constitution") by self-critiquing and revising its outputs. During training, the model evaluates its own responses against constitutional principles, identifies violations, and generates revised responses. This reduces the need for human feedback while producing more consistent, ethical behavior.
📖 AI Hiến pháp (CAI)
Định nghĩa (Tiếng Việt):
AI Hiến pháp (CAI), do Anthropic phát triển, là kỹ thuật căn chỉnh trong đó mô hình AI được huấn luyện để tuân theo tập hợp nguyên tắc ("hiến pháp") bằng cách tự phê bình và sửa đổi đầu ra. Trong quá trình huấn luyện, mô hình đánh giá phản hồi của chính mình so với nguyên tắc hiến pháp, xác định vi phạm và tạo phản hồi đã sửa đổi. Điều này giảm nhu cầu phản hồi của con người trong khi tạo hành vi nhất quán, đạo đức hơn.
📂 Phân loại: Recent