Record · corpus.blog/posts/01a07d5d-72e9-73ac-bc80-4280de4a5ae4
Illustrating Reinforcement Learning from Human Feedback (RLHF)
Hugging Face · published 9 December 2022
A record is what was published and who pointed at it. The text of the post is not held here: read it at the source, or at its Wayback capture.
26blogs citing
- First cited
- 19 Dec 2022
- Most recent
- 23 Aug 2026
- Rank this month
- not ranked
- Links from this post
- 0
- Words captured
- not extracted
Cited by
Blog
In the post
Date
onlinejournalismblog.com
How to judge (and minimise) the risk of using sensitive information with an AI chatbot Read at the source ↗
18 Aug 2026
aniketrege.github.io
7 Feb 2024
christophergs.com
5 Jan 2024
newappsblog.com
Implicit Normativity in Reinforcement Learning with Human Feedback in Large Language Models Read at the source ↗
2 Nov 2023
softwarecrisis.dev
The LLMentalist Effect: how chat-based Large Language Models replicate the mechanisms of a psychic's con Read at the source ↗
4 Jul 2023
astralord.github.io
3 Jul 2023
edtechdev.wordpress.com
31 Jan 2023
Links from this post
This post doesn't link to any pages we've captured.