Build and ship On-Device AI
Deploy agentic and generative AI models directly with Snapdragon® and Qualcomm Dragonwing™ processors. Optimize for inference latency, cost, and power efficiency.
Artificial intelligence at your fingertips
Choose your developer approach
Start with the Guided approach for pre-built workflows and framework runtimes or choose the Native approach for direct control over model execution, optimization, and runtime targeting.
Sample datasets for training models across research, prototyping, and experimentation.
Download a pre-optimized model or bring your own and prepare it for on-device execution.
Guided
Pre-optimized models from Qualcomm AI Hub
Hundreds of pre-optimize models, Small Language Models (SLMs), and Large Language Models (LLMs) are ready to profile, test, and deploy.
Native
Qualcomm Neural Processing SDK
If you want to prepare a model from your own training data, start with the Qualcomm® Neural Processing SDK, which includes tools to build a compatible model for execution on Snapdragon or Dragonwing.
Qualcomm® AI Runtime SDK (QAIRT)
Prepare models from your own training data using Qualcomm® AI Runtime SDK (QAIRT). QAIRT includes conversion and optimization tools for TensorFlow, PyTorch, and ONNX, with execution across CPU, GPU, NPU, and Qualcomm® Sensing Hub on Snapdragon and Dragonwing.
See this Building and Executing Your Model tutorial which includes steps to build a model from data
Choose from an array of devices with Snapdragon and Dragonwing for your compute, IoT, or mobile projects.
Guided
Access real compute and mobile devices with Dragonwing and Snapdragon remotely through your browser. Build, test, debug, and profile on physical devices without requiring physical devices.
Use Qualcomm AI Hub to profile and test on devices with the Qualcomm AI Hub Workbench.
Native
Windows on Snapdragon / Compute
Snapdragon X Series processors power the fastest, most power-efficient AI PCs from several major brands.
Mobile
Explore the full range of Android OS devices powered by Snapdragon processors for your mobile apps and games.
Choose from powerful evaluation kit (EVK) options for your next agentic AI or GenAI project.
Convert your models into Snapdragon and Dragonwing compatible formats and optimize them for size, accuracy, and performance during execution.
Guided
Use Qualcomm AI Hub Workbench to quickly optimize a trained model for execution on over 60 cloud-based devices using supported runtimes.
Use Edge Impulse to build IoT datasets, train models, and optimize libraries to run directly on devices with Dragonwing.
Native
AI Model Efficiency Toolkit (AIMET)
Use AIMET to quantize and compress models trained with PyTorch or ONNX before running them with ONNX Runtime (ORT), Qualcomm Neural Processing SDK, or Qualcomm AI Engine Direct SDK (QNN).
Qualcomm Neural Processing SDK
Qualcomm Neural Processing SDK includes several analysis and optimization tools to prepare a model for execution using the SDK’s runtime.
Qualcomm AI Runtime SDK
Qualcomm® AI Runtime SDK (QAIRT) includes several analysis and optimization tools to prepare a model for execution using the SDK’s runtime.
See thisSee the Quantization code samples in this introductory tutorial
Below are several runtimes you can use for hardware-accelerated inference in your application running on Snapdragon or Dragonwing.
Choose one of the many sample applications as your starting point for loading a model and running inference. See the Get Started with Qualcomm AI Hub Apps tutorial
AI/ML Framework Runtimes
Built on the Qualcomm AI Engine Direct SDK these runtime integrations let you use your preferred AI/ML and Generative AI framework.
Choose from:
For ONNX models, ONNX Runtime (ORT) includes a Qualcomm Plugin Execution Provider that you can incorporate into your application.
To run LiteRT models incorporate the LiteRT Delegate into your application.
Native
Qualcomm AI Engine Direct SDK
QNN includes low-level tools for converting and optimizing models for execution by the runtime.
Deploy your app, your model, or both and choose the tools that match your pipeline.
Guided
Choose one of the many sample applications as your starting point for loading a model and running inference. See the Get Started with Qualcomm AI Hub Apps tutorial
FoundriesFactory™ Platform
Use FoundriesFactory to build a CI/CD pipeline that manages OTA updates for apps and models to IoT edge devices in the field via containerization.
If you’re building IoT AI solutions with Edge Impulse, the platform integrates FoundriesFactory for Dragonwing deployments.
Native
Qualcomm AI Engine Direct SDK
QNN includes low-level tools for converting and optimizing models for execution by the runtime.
Use Cases
