Record · corpus.blog/posts/01a087d6-35eb-72e8-92d1-5292f73987aa
Super-fast deduplication of large datasets using Splink and DuckDB
Robin Linacre's blog: Post list · published 18 January 2024
A record is what was published and who pointed at it. The text of the post is not held here: read it at the source, or at its Wayback capture.
0blogs citing
- First cited
- —
- Most recent
- —
- Rank this month
- not ranked
- External links
- 16
- Words captured
- 2,267
Cited by
No blog we hold has cited this post yet.
Links from this post
Link
Host
duckdb.org
query.wikidata.org
en.wikipedia.org
wikidata.org
imai.fas.harvard.edu
moj-analytical-services.github.io
scholar.googleusercontent.com
moj-analytical-services.github.io
moj-analytical-services.github.io
moj-analytical-services.github.io
moj-analytical-services.github.io
Elsewhere on robinlinacre.com (4)
robinlinacre.com