Value Alignment

2026

Deven R. Desai and Mark O. Riedl
Using Agency Law to Tame AI Agents
Berkeley Technology Law Review forthcoming (2026).
SSRN Journal bibtexAgentsPolicy and LawValue Alignment

2025

Mark O. Riedl and Deven R. Desai
AI Agents and the Law
Proceedings of the 2025 AAAI Conference on AI, Ethics, and Society (2025).
arXiv Conference bibtexAgentsPolicy and LawValue Alignment

Md Sultan Al Nahian, Tasmia Tasrin, Spencer Frazier, Mark Riedl, and Brent Harrison
The Goofus & Gallant Story Corpus for Practical Value Alignment
Proceedings of the International Conference on Machine Learning and Applications (2025).
arXiv Conference bibtexStorytellingValue Alignment

Deven Desai and M. O. Riedl
Responsible AI Agents
arXiv:2502.18359 (2025).
arXiv bibtexAgentsPolicy and LawValue Alignment

2024

Ashutosh Baheti, Ximing Lu, Faeze Brahman, Ronan Le Bras, Maarten Sap, and Mark O. Riedl
Improving Language Models with Advantage-based Offline Policy Gradients
Proceedings of ICLR 2024 (2024).
arXiv Conference bibtexLarge Language ModelsReinforcement LearningValue AlignmentAI Safety

2023

Xiangyu Peng, Christopher Cui, Wei Zhou, Renee Jia, and Mark O. Riedl
Story Shaping: Teaching Agents Human-like Behavior with Stories
Proceedings of the 2023 AAAI Conference on Artificial Intelligence and Interactive Digital Entertainment (2023).
arXiv Conference bibtexStorytellingInteractive StoriesAgentsReinforcement LearningValue Alignment

2022

Louis Castricato, Alexander Havrilla, Shahbuland Matiana, Michael Pieler, Anbang Ye, Ian Yang, Spencer Frazier, and Mark Riedl
Robust Preference Learning for Storytelling via Contrastive Reinforcement Learning
arXiv preprint arXiv:2210.07792 (2022).
arXiv bibtexStorytellingLarge Language ModelsReinforcement LearningValue Alignment

Md Sultan Al Nahian, Spencer Frazier, Brent Harrison, and Mark Riedl
Machine Learning Approaches for Principle Prediction in Naturally Occurring Stories
arXiv preprint arXiv:2212.06048 (2022).
arXiv bibtexStorytellingValue Alignment

2021

Ashutosh Baheti, Maarten Sap, Alan Ritter, and Mark O. Riedl
Just Say No: Analyzing the Stance of Neural Dialogue Generation in Offensive Contexts
Proceedings of EMNLP 2021 (2021).
arXiv Conference bibtexLarge Language ModelsValue AlignmentAI Safety

Md Sultan Al Nahian, Spencer Frazier, Brent Harrison, and Mark O. Riedl
Training Value-Aligned Reinforcement Learning Agents Using a Normative Prior
arXiv:2104.09469 (2021).
arXiv bibtexAgentsReinforcement LearningValue AlignmentAI Safety

2020

Spencer Frazier, Md Sultan Al Nahian, Mark O. Riedl, and Brent Harrison
Learning Norms from Stories: A Prior for Value Aligned Agents
Proceedings of the AAAI/ACM Conference on AI, Ethics, and Society (2020).
arXiv Conference bibtexStorytellingValue Alignment

Xiangyu Peng, S. Li, Spencer Frazier, and Mark O. Riedl
Reducing Non-Normative Text Generation from Language Models
International Conference on Natural Language Generation (2020).
arXiv Conference bibtexLarge Language ModelsValue AlignmentAI Safety

2016

Mark Riedl and Brent Harrison
Using Stories to Teach Human Values to Artificial Agents
Proceedings of the 2nd International Workshop on AI, Ethics and Society (2016).
PDF Workshop bibtexAgentsValue AlignmentAI Safety

Brent Harrsion and Mark O Riedl
Learning From Stories: Using Crowdsourced Narratives to Train Virtual Agents
Proceedings of the 2016 AAAI Conference on Artificial Intelligence and Interactive Digital Entertainment (2016).
PDF Conference bibtexAgentsValue AlignmentAI SafetyHuman Computation

Brent Harrison and Mark Riedl
Towards Learning From Stories: An Approach for Interactive Machine Learning
Proceedings of the AAAI'16 Workshop on Symbiotic Cognitive Systems (2016).
PDF Workshop bibtexAgentsValue AlignmentAI Safety

Brent Harrison, Siddhartha Banerjee, and Mark O. Riedl
Learning from Stories: Using Natural Communication to Train Believable Agents
Proceedings of the 2016 IJCAI Workshop on Interactive Machine Learning (2016).
PDF Workshop bibtexAgentsValue AlignmentAI Safety

2012

Boyang Li, D. Scott Appling, Stephen Lee-Urban, and Mark O. Riedl
Learning Sociocultural Knowledge via Crowdsourced Examples
Proceedings of the 4th Workshop on Human Computation (2012).
PDF Workshop bibtexStorytellingValue AlignmentHuman Computation