Prepare, run, and develop generative AI applications on Qualcomm Dragonwing IoT platforms using Genie, Qualcomm AI Hub, and the QAIRT SDK.
Qualcomm® Dragonwing™ products bring generative AI (GenAI) to edge devices without relying solely on cloud infrastructure, lowering latency, enhancing privacy, and reducing costs for diverse IoT applications.
‘GenAI use-cases’ are not supported on Dragonwing™ IQ-2390 and Dragonwing™ IQ-615.
The following image illustrates the GenAI architecture, including supported use cases and applications, GenAI frameworks, and the lower-level backends and libraries.Dragonwing IQ products integrate a CPU, GPU, and Hexagon NPU for heterogeneous computing, optimized for large language models (LLMs), vision, and multimodal AI tasks. Leading GenAI models — including Llama, Whisper, Stable Diffusion, and LLaVA — are supported for IoT scenarios such as robotics, retail, and industrial automation.For responsive LLM performance, Dragonwing IQ supports multi-billion parameter models with techniques such as prefill acceleration and self-speculative decoding. To enable bring-your-own-model (BYOM) workflows and streamline deployment, use Qualcomm AI Hub, Qualcomm AI Runtime SDK execution providers, and Qualcomm Generative AI Inference Extensions (Genie).
Generative AI on Dragonwing delivers context-aware, adaptive, and privacy-preserving solutions at the edge, enabling automation, personalization, and operational efficiency for use cases such as autonomous drones, predictive maintenance, multimodal agents, and on-device retrieval-augmented generation (RAG).