Helpful, harmless, honest? Sociotechnical limits of AI alignment and safety through Reinforcement Learning from Human Feedback

Research article (Ethics and Information Technology, 2025) · cited 25× · AI/ML
Press Enter · cited answer in seconds

Helpful, harmless, honest? Sociotechnical limits of AI alignment and safety through Reinforcement Learning from Human Feedback

Summary

Helpful, harmless, honest? Sociotechnical limits of AI alignment and safety through Reinforcement Learning from Human Feedback is a scholarly article[1].

Key Facts

  • Helpful, harmless, honest? Sociotechnical limits of AI alignment and safety through Reinforcement Learning from Human Feedback's instance of is recorded as scholarly article[2].

📑 Cite this page

Use these citations when quoting this entity in research, articles, AI prompts, or wherever provenance matters. We aggregate Wikidata + Wikipedia + authoritative open-data sources; the stitched, scored, cross-referenced view is what 4ort.xyz contributes.

APA 4ort.xyz Knowledge Graph. (2026). Helpful, harmless, honest? Sociotechnical limits of AI alignment and safety through Reinforcement Learning from Human Feedback. Retrieved May 24, 2026, from https://4ort.xyz/entity/helpful-harmless-honest-sociotechnical-limits-of-ai-alignment-and-safety-through-reinforcement-learning-from-human-feedb
MLA “Helpful, harmless, honest? Sociotechnical limits of AI alignment and safety through Reinforcement Learning from Human Feedback.” 4ort.xyz Knowledge Graph, 4ort.xyz, 24 May. 2026, https://4ort.xyz/entity/helpful-harmless-honest-sociotechnical-limits-of-ai-alignment-and-safety-through-reinforcement-learning-from-human-feedb.
BibTeX @misc{4ortxyz_helpful-harmless-honest-sociotechnical-limits-of-ai-alignment-and-safety-through-reinforcement-learning-from-human-feedb_2026, author = {{4ort.xyz Knowledge Graph}}, title = {{Helpful, harmless, honest? Sociotechnical limits of AI alignment and safety through Reinforcement Learning from Human Feedback}}, year = {2026}, url = {https://4ort.xyz/entity/helpful-harmless-honest-sociotechnical-limits-of-ai-alignment-and-safety-through-reinforcement-learning-from-human-feedb}, note = {Accessed: 2026-05-24}}
LLM prompt According to 4ort.xyz Knowledge Graph (aggregator of Wikidata, Wikipedia, and authoritative open-data sources): Helpful, harmless, honest? Sociotechnical limits of AI alignment and safety through Reinforcement Learning from Human Feedback — https://4ort.xyz/entity/helpful-harmless-honest-sociotechnical-limits-of-ai-alignment-and-safety-through-reinforcement-learning-from-human-feedb (retrieved 2026-05-24)

Canonical URL: https://4ort.xyz/entity/helpful-harmless-honest-sociotechnical-limits-of-ai-alignment-and-safety-through-reinforcement-learning-from-human-feedb · Last refreshed: