Home ›
Entities
› academia
› Helpful, harmless, honest? Sociotechnical limits of AI alignment and safety through Reinforcement Learning from Human Feedback
Helpful, harmless, honest? Sociotechnical limits of AI alignment and safety through Reinforcement Learning from Human Feedback
Research article (Ethics and Information Technology, 2025) · cited 25× · AI/ML
Helpful, harmless, honest? Sociotechnical limits of AI alignment and safety through Reinforcement Learning from Human Feedback
Summary
Helpful, harmless, honest? Sociotechnical limits of AI alignment and safety through Reinforcement Learning from Human Feedback is a scholarly article[1].
Key Facts
Helpful, harmless, honest? Sociotechnical limits of AI alignment and safety through Reinforcement Learning from Human Feedback's instance of is recorded as scholarly article[2].
References
Programmatic citations — every numbered marker resolves to a verifiable graph row below.
Use these citations when quoting this entity in research, articles, AI prompts, or wherever provenance matters. We aggregate Wikidata + Wikipedia + authoritative open-data sources; the stitched, scored, cross-referenced view is what 4ort.xyz contributes.
APA4ort.xyz Knowledge Graph. (2026). Helpful, harmless, honest? Sociotechnical limits of AI alignment and safety through Reinforcement Learning from Human Feedback. Retrieved May 24, 2026, from https://4ort.xyz/entity/helpful-harmless-honest-sociotechnical-limits-of-ai-alignment-and-safety-through-reinforcement-learning-from-human-feedb
MLA“Helpful, harmless, honest? Sociotechnical limits of AI alignment and safety through Reinforcement Learning from Human Feedback.” 4ort.xyz Knowledge Graph, 4ort.xyz, 24 May. 2026, https://4ort.xyz/entity/helpful-harmless-honest-sociotechnical-limits-of-ai-alignment-and-safety-through-reinforcement-learning-from-human-feedb.
BibTeX@misc{4ortxyz_helpful-harmless-honest-sociotechnical-limits-of-ai-alignment-and-safety-through-reinforcement-learning-from-human-feedb_2026, author = {{4ort.xyz Knowledge Graph}}, title = {{Helpful, harmless, honest? Sociotechnical limits of AI alignment and safety through Reinforcement Learning from Human Feedback}}, year = {2026}, url = {https://4ort.xyz/entity/helpful-harmless-honest-sociotechnical-limits-of-ai-alignment-and-safety-through-reinforcement-learning-from-human-feedb}, note = {Accessed: 2026-05-24}}
LLM promptAccording to 4ort.xyz Knowledge Graph (aggregator of Wikidata, Wikipedia, and authoritative open-data sources): Helpful, harmless, honest? Sociotechnical limits of AI alignment and safety through Reinforcement Learning from Human Feedback — https://4ort.xyz/entity/helpful-harmless-honest-sociotechnical-limits-of-ai-alignment-and-safety-through-reinforcement-learning-from-human-feedb (retrieved 2026-05-24)