Record · corpus.blog/posts/01a07df3-316b-710b-935b-48311734f56e
Your AI Product Needs Evals
Hamel Husain · published 29 March 2024
A record is what was published and who pointed at it. The text of the post is not held here: read it at the source, or at its Wayback capture.
27blogs citing
- First cited
- 3 Apr 2024
- Most recent
- 5 Sept 2026
- Rank this month
- not ranked
- Links from this post
- 23
- Words captured
- 4,775
Cited by
Blog
In the post
Date
druce.ai
Agent Reliability Engineering: Field Notes on Building Predictable, Steerable Agents Read at the source ↗
5 Sept 2026
Elastic Blog
19 Aug 2026
leocavalcante.dev
Evaluating AI Agents Beyond the "Vibes Check": How to Measure What Actually Matters Read at the source ↗
29 Jun 2026
kig.re
23 Jun 2026
ivanturkovic.com
3 May 2026
blog.mariusvach.com
22 Apr 2025
Eugene Yan
27 Oct 2024
Links from this post
Link
Host
github.com
rechat.com
rechat.com
docs.pytest.org
metabase.com
langchain.com
langchain.com
lilacml.com
en.wikipedia.org
supervisely.com
https://developers.google.com/machine-learning/crash-course/classification/accuracy-precision-recall
developers.google.com
geteppo.com
youtube.com
vanishinggradients.fireside.fm
arize.com
humanloop.com
github.com
honeyhive.ai