Blog · 19 September 2026

Introducing corpus.blog

A citation index for the independent developer blogosphere, and why it counts blogs rather than posts.

We are corpus.blog: a citation index for the independent developer blogosphere. We read the feeds of a large number of developer blogs, and we record, for every post we find, its title, its date, and who links to it.

What is live today:

  • Most cited posts and blogs. A leaderboard of what the corpus is pointing at right now, and the blogs it is pointing at most.
  • Talked about. The terms whose use has spread across blogs this week, compared against what was normal before it.
  • A record page for every post and every blog we hold — who cites it, what it links to, and when it was published.
  • Reports that write down exactly how every figure on the site is computed, so a number here is checkable rather than trusted on faith.

The rule under all of it: every figure counts distinct blogs, never articles. A small number of prolific sources produce a large share of everything published, so ranking by volume would just surface whoever posts most. Counting each blog once per figure asks a different question — how many different people found this worth pointing at — and that is the question a citation was always supposed to answer.

One promise we are making in public, because it is easy to break by accident: we do not republish anyone's text. The pages on this site show titles, links, dates and counts. To read a post, you go to the blog that wrote it.

Claiming a blog — proving you control your own domain, and getting the record of your blog back, along with webmentions sent on your behalf — is coming, but is not open yet.

And this blog has a feed: the same kind of feed we read from every other blog in the corpus, at /blog/feed. We are trying to be, in every way we can measure, the sort of source we would want to hold.

Read more about what the crawler does and how to stop it.