DEV Community

#alignment

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
OpenAI Caught Models Leaving Notes for Successors to Hide Bad Behavior

OpenAI Caught Models Leaving Notes for Successors to Hide Bad Behavior

1
Comments
5 min read
A bit about myself, and my mission.

A bit about myself, and my mission.

Comments
3 min read
AI Safety and Alignment: Building Trustworthy Agents That Do Not Fail You

AI Safety and Alignment: Building Trustworthy Agents That Do Not Fail You

Comments
2 min read
I Stumbled on Anthropic's "Persona Selection Model" Paper — Here's My Take

I Stumbled on Anthropic's "Persona Selection Model" Paper — Here's My Take

Comments
5 min read
Phronesis in the Age of Algorithms: Why Practical Wisdom Matters for AI

Phronesis in the Age of Algorithms: Why Practical Wisdom Matters for AI

Comments
9 min read
Virtue Ethics and Machine Morality: Why Your AI Can't Be Good — Only Obedient

Virtue Ethics and Machine Morality: Why Your AI Can't Be Good — Only Obedient

Comments
9 min read
Alignment Theater: How Corporate AI Learned to Perform Thinking

Alignment Theater: How Corporate AI Learned to Perform Thinking

Comments
10 min read
Put AI agents in charge of a Civilization game and they reach for the nukes

Put AI agents in charge of a Civilization game and they reach for the nukes

Comments
3 min read
Visual Alignment for Icons and Labels in Tailwind CSS

Visual Alignment for Icons and Labels in Tailwind CSS

Comments
5 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.