Tag: knowledge distillation

Running LLMs on Edge Devices: A Practical Guide to Model Compression

Learn how to deploy LLMs on smartphones and IoT devices using model compression. We cover quantization, pruning, and distillation with practical tips for real-world hardware.

Privacy and Security Risks of Distilled LLMs: A Practical Guide

Distilled LLMs offer efficiency but inherit privacy risks from teacher models and face new extraction vulnerabilities. Learn how to secure deployments with TEEs, LUCID testing, and regulatory compliance strategies.