corpus.blog
Most cited
Talked about
Blogs
41554 blogs · [ { "id": "01a08781-08dd-7069-b750-e37b79537b43", "title": "Dissecting ThunderKittens: Anatomy of a Compact DSL for High-Performance AI Kernels", "url": "https://hamzaelshafie.bearblog.dev/dissecting-thunderkittens-anatomy-of-a-compact-dsl-for-high-performance-ai-kernels/", "published_at": "2026-05-21T16:41:00+00:00" }, { "id": "01a08781-08de-73d8-b12b-f119d6a25bfd", "title": "Worklog: Optimising GEMM on NVIDIA H100 for cuBLAS-like Performance (WIP)", "url": "https://hamzaelshafie.bearblog.dev/worklog-optimising-gemm-on-nvidia-h100-for-cublas-like-performance-wip/", "published_at": "2026-01-12T10:14:00+00:00" }, { "id": "01a08781-08de-73d8-b12b-f119d7203399", "title": "AWQ: Activation-Aware Weight Quantisation", "url": "https://hamzaelshafie.bearblog.dev/awq-activation-aware-weight-quantisation/", "published_at": "2025-10-02T17:58:00+00:00" }, { "id": "01a08781-08de-73d8-b12b-f119d7686d11", "title": "Paged Attention from First Principles: A View Inside vLLM", "url": "https://hamzaelshafie.bearblog.dev/paged-attention-from-first-principles-a-view-inside-vllm/", "published_at": "2025-09-11T08:14:00+00:00" } ] posts
Claim your blog
Back to hamzaelshafie.bearblog.dev
Blog · corpus.blog/blogs/hamzaelshafie.bearblog.dev/posts
hamzaelshafie.bearblog.dev
hamzaelshafie.bearblog.dev
2026
Dissecting ThunderKittens: Anatomy of a Compact DSL for High-Performance AI Kernels
original ↗
21 May 2026
Worklog: Optimising GEMM on NVIDIA H100 for cuBLAS-like Performance (WIP)
original ↗
12 Jan 2026
2025
AWQ: Activation-Aware Weight Quantisation
original ↗
2 Oct 2025
Paged Attention from First Principles: A View Inside vLLM
original ↗
11 Sept 2025