Proposed Metaethical Framework For Model Alignment | dailyai.report
23 stories from today
Safety
48d ago
Proposed Metaethical Framework For Model Alignment
A new proposal suggests using perspectival moral realism and evolutionary debunking to refine AI values. The author intends to submit this framework to Anthropic to influence its constitutional AI approach.
The Signal
While individual submissions rarely alter training, this specific philosophical angle differs from the preference-satisfaction models currently dominating AI ethics literature.