Record · corpus.blog/posts/01a0f3de-cb36-730a-904a-3bf0cc584797
Offline and Conservative Model-Based RL
mbrenndoerfer.com · published 18 July 2026 · 49 min read
A record is what was published and who pointed at it. The text of the post is not held here: read it at the source, or at its Wayback capture.
0blogs citing
- First cited
- —
- Most recent
- —
- Rank this month
- not ranked
- External links
- 7
- Words captured
- 11,168
Cited by
No blog we hold has cited this post yet.
Similar posts
LeRobot v0.6.0: Imagine, Evaluate, Improve
Hugging Face
7 Jul 2026
7 Jul 2026
Off-Policy Policy Evaluation
cruxponent.com
15 Jun 2026
RLTL;DR: Self-Improvement by Internalizing Self-Generated Feedback
Apple Machine Learning Research
1 Oct 2026
RL Fundamentals
saheb.github.io
22 Sept 2026
1 Aug 2026
11 Aug 2026
Reinforcement Learning for LLMs: The Complete Guide
cameronrwolfe.substack.com
24 Aug 2026
Four Days of RL: One Gradient, Dressed Four Ways
saran-blog.vercel.app
18 Jun 2026
A Variational Lens on RL in Diffusion Models.
astro-eric.github.io
11 Jun 2026
Links from this post
Link
Host
arxiv.org
papers.nips.cc
proceedings.mlr.press
arxiv.org
arxiv.org
arxiv.org
arxiv.org
Elsewhere on mbrenndoerfer.com (19)
mbrenndoerfer.com
mbrenndoerfer.com
mbrenndoerfer.com
mbrenndoerfer.com
mbrenndoerfer.com
Something wrong here? Report a problem
A post on mbrenndoerfer.com, published 18 July 2026, has not been cited by any blog yet (corpus.blog, measured 8 October 2026).