Show HN: Running PrismML's Bonsai inside DRAM by breaking DDR4 timing rules

Hardware Hacking & Reverse Engineering, AI & Machine Learning, DevOps & Infrastructure(news.ycombinator.com)view on HackerNews
CaSAternary LLM inferenceedge AIDRAMmemory bushardware architecture

Author: pcdeni

Date: 7/23/2026

Article Summary:
A researcher proposes a new hardware architecture, CaSA, that performs ternary LLM inference directly inside COTS DRAM, bypassing the memory bus, to improve AI model execution on edge devices.