“Small Language Models for Edge AI: Trade-Offs and Quantization in Practice,” a Presentation from AMD
Dwith Chenna, MTS Product Engineer, AI Inference at AMD presents “Small Language Models for Edge AI: Trade-Offs and Quantization in Practice” at the May 2026 Embedded Vision Summit. Large language models are powerful but often impractical for embedded and on-prem systems due to latency, cost, privacy and memory constraints. Small… “Small Language Models for Edge […]










