Tag: edge ai constraints

AI

Runtime and memory constraints in Llama 3.2 edge deployment

Deploying Llama 3.2 3B on smart glasses via ExecuTorch presents challenges like static context length and lack of KV cache. While Meta Ray-Bans hold 80% market share, privacy concerns persist due to a 2.4% PII memorization rate during model inversion attacks.