Blog · corpus.blog/blogs/queirozf.com/posts
queirozf.com
queirozf.com
2026
2025
Paper Summary: A General Theoretical Paradigm to Understand Learning from Human Preferencesoriginal ↗
21 Jul 2025
15 Jun 2025
27 Apr 2025
Paper Summary: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learningoriginal ↗
19 Apr 2025
Paper Summary: DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Modelsoriginal ↗
6 Apr 2025
2024
18 Oct 2024
Paper Summary: Few-shot Fine-Tuning vs In-context Learning: a Fair Comparison and Evaluationoriginal ↗
22 Jul 2024
31 Mar 2024
2023
16 Nov 2023