Keynote 14 - AI, Rewired: Building Edge Intelligence That Adapts, Distributes and Acts
This session explores how to overcome static model limits by deploying adaptive, high-performance edge intelligence on resource-constrained devices. We present a full-stack co-design framework combining conditional fine-tuning (IGU-LoRA), advanced quantization (L4Q), and RL-driven orchestration across heterogeneous silicon (CPU, GPU, and NPU). Learn how deployment-aware architectures like CoA-LoRA solve critical battery, thermal, and memory bandwidth constraints in real time. Discover how this technology powers the next wave of AI PCs to deliver seamless, low-latency agentic experiences directly on the device.


