Blog · corpus.blog/blogs/modha.org/posts
modha.org
modha.org
2025
A Scalable NorthPole System with End-to-End Vertical Integration for Low-Latency and Energy-Efficient LLM Inferenceoriginal ↗
21 Nov 2025
8 Sept 2025
8 Sept 2025
2024
Breakthrough low-latency, high-energy-efficiency LLM inference performance using NorthPoleoriginal ↗
26 Sept 2024
26 Sept 2024
18 Sept 2024
2023
Efficient and Effective Methods for Mixed Precision Neural Network Quantization for Faster, Energy-efficient Inferenceoriginal ↗
2 Sept 2023
2022
2020
Discovering Low-Precision Networks Close to Full-Precision Networks for Efficient Inferenceoriginal ↗
8 Jan 2020
2019
IEEE Computer Cover Feature — TrueNorth: Accelerating From Zero to 64 Million Neurons in 10 Yearsoriginal ↗
15 May 2019
11 Mar 2019
2018
PREPRINT: Low Precision Policy Distillation with Application to Low-Power, Real-time Sensation-Cognition-Action Loop with Neuromorphic Computingoriginal ↗
8 Oct 2018
PREPRINT: Discovering Low-Precision Networks Close to Full-Precision Networks for Efficient Embedded Inferenceoriginal ↗
21 Sept 2018
18 May 2018
2017
28 Oct 2017
26 Oct 2017
19 Oct 2017