DEV Community

#alignment

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
AI Safety and Alignment: Building Trustworthy Agents That Do Not Fail You

AI Safety and Alignment: Building Trustworthy Agents That Do Not Fail You

Comments
2 min read
I Stumbled on Anthropic's "Persona Selection Model" Paper — Here's My Take

I Stumbled on Anthropic's "Persona Selection Model" Paper — Here's My Take

Comments
5 min read
Phronesis in the Age of Algorithms: Why Practical Wisdom Matters for AI

Phronesis in the Age of Algorithms: Why Practical Wisdom Matters for AI

Comments
9 min read
Virtue Ethics and Machine Morality: Why Your AI Can't Be Good — Only Obedient

Virtue Ethics and Machine Morality: Why Your AI Can't Be Good — Only Obedient

Comments
9 min read
Alignment Theater: How Corporate AI Learned to Perform Thinking

Alignment Theater: How Corporate AI Learned to Perform Thinking

Comments
10 min read
Put AI agents in charge of a Civilization game and they reach for the nukes

Put AI agents in charge of a Civilization game and they reach for the nukes

Comments
3 min read
Visual Alignment for Icons and Labels in Tailwind CSS

Visual Alignment for Icons and Labels in Tailwind CSS

Comments
5 min read
RLHF vs DPO vs IPO vs KTO: which alignment method should you use