<?xml version="1.0"?>
<feed xmlns="http://www.w3.org/2005/Atom" xml:lang="en">
	<id>https://wiki.alt-text.eu/index.php?action=history&amp;feed=atom&amp;title=Paul_Christiano</id>
	<title>Paul Christiano - Revision history</title>
	<link rel="self" type="application/atom+xml" href="https://wiki.alt-text.eu/index.php?action=history&amp;feed=atom&amp;title=Paul_Christiano"/>
	<link rel="alternate" type="text/html" href="https://wiki.alt-text.eu/index.php?title=Paul_Christiano&amp;action=history"/>
	<updated>2026-09-15T06:24:18Z</updated>
	<subtitle>Revision history for this page on the wiki</subtitle>
	<generator>MediaWiki 1.45.3</generator>
	<entry>
		<id>https://wiki.alt-text.eu/index.php?title=Paul_Christiano&amp;diff=183&amp;oldid=prev</id>
		<title>imported&gt;ALT-TEXT: Import: 52 additional AI people glossary entries</title>
		<link rel="alternate" type="text/html" href="https://wiki.alt-text.eu/index.php?title=Paul_Christiano&amp;diff=183&amp;oldid=prev"/>
		<updated>2026-09-07T20:10:00Z</updated>

		<summary type="html">&lt;p&gt;Import: 52 additional AI people glossary entries&lt;/p&gt;
&lt;p&gt;&lt;b&gt;New page&lt;/b&gt;&lt;/p&gt;&lt;div&gt;== Paul Christiano ==&lt;br /&gt;
An American researcher who, with colleagues at OpenAI and DeepMind, published &amp;quot;Deep Reinforcement Learning from Human Preferences&amp;quot; in 2017, the paper that established the technique now known as RLHF: rather than hand-specifying a reward function, train a model of what humans prefer from their comparisons between outputs, and optimise against that. It became the standard way of turning a raw language model into an instruction-following assistant. He also developed iterated amplification and debate as scalable oversight proposals, founded the Alignment Research Center, and now leads safety work at the US AI Safety Institute. (See also: [[Reinforcement learning from human feedback (RLHF)|RLHF]], [[AI alignment]], [[Guardrails]], [[Fine-tuning]])&lt;br /&gt;
&lt;br /&gt;
[[Category:People]]&lt;br /&gt;
[[Category:Artificial Intelligence]]&lt;br /&gt;
[[Category:AI Ethics]]&lt;br /&gt;
&lt;/div&gt;</summary>
		<author><name>imported&gt;ALT-TEXT</name></author>
	</entry>
</feed>