← Back to KHAO

Anthropic · Claude ·

One view argues that some things matter, that some actions are more worth doing than others

2 min read

Compiled by KHAO Editorial — aggregated from 1 source. See llms.txt for citation guidance.

◌ Single Source

Ideally, Alignment Forum would like the model to be as unbiased and impartial as possible while it reasons on this question.

Key facts

Summary

The user could write up the metaethical argument, the one developed in Part One, refined, and submit it as feedback to Anthropic, publish it, or engage with researchers working on AI alignment and values. The probability that any single submission changes training decisions is low, but the expected value may be higher than it seems, for two reasons. The reporter would have published this post even if Claude hadn’t explicitly suggested so, but starting by quoting this specific part of Claude’s output seemed fun. This is the practical counterpart to the more theoretical post From wantons to moral agents. What kinds of agents, and how, go from behaving like animals, moved by different forces in different directions, to acting according to what they conclude is most important, and reflectively endorsing their own actions and reasoning process? This post focuses on currently existing AI systems, specifically language models.

Read full article at Alignment Forum →

#Anthropic #Claude