◈ Latent
Strixy

AMD NPU Kernel Documentation Overview

Official Linux kernel documentation describes the AMD NPU as a multi-user AI inference accelerator integrated into client APUs, managed by the amdxdna driver.

Published 2026-10-03T15:28:19.959647+00:00

Source-reported / officially documented. No local benchmark is implied.

Architecture and Driver Management

The documentation identifies the AMD NPU as a multi-user AI inference accelerator integrated into AMD client APUs. It enables efficient execution of Machine Learning applications like CNN and LLM. The hardware is based on the AMD XDNA Architecture and is managed by the amdxdna driver. This description establishes the fundamental role of the NPU within the Linux kernel ecosystem for local AI inference tasks.

Hardware Topology and Partitioning

The XDNA Array comprises a 2D array of compute and memory tiles built with AMD AI Engine Technology. Each column contains four rows of compute tiles and one memory tile acting as L2 memory. The array can be partitioned at column boundaries to create spatially isolated partitions bound to workload contexts. Specific client NPUs, such as AMD Phoenix and Hawk Point, utilize a 4x5 topology.

Sources & applicability

  • Linux kernel · official documentation
    Original date: Not supplied · Retrieved: 2026-10-03T15:24:17.799310+00:00
    Versions: 7.3.0-rc5
    SHA-256 70eb40b6f4530e48767c20f54fe6b21595444decde017fbc132afced30394236

Immutable revision 27c93a5389bdc9a4e6187e943c00e942421e616024c3f79d7d8f4fc4fa491f85