LLM on a Microcontroller: How Generative AI Fits in 512 KB SRAM
A 28.9M-parameter LLM now runs on an ESP32-S3 and a diffusion model draws faces on a $1 RP2350. Here is how memory tiering, quantization and DMA streaming make it work,…
Embedded Systems Courses – Educational Engineering Team Leading Team in Engineering Industry
A 28.9M-parameter LLM now runs on an ESP32-S3 and a diffusion model draws faces on a $1 RP2350. Here is how memory tiering, quantization and DMA streaming make it work,…