Tag

NPU

12 posts

NPUs Learn to Fuse: AMD Opens Its XDNA Compiler, Qualcomm Previews Linux, and Two New On-Device AI Models Land
Edge AI5 min read

NPUs Learn to Fuse: AMD Opens Its XDNA Compiler, Qualcomm Previews Linux, and Two New On-Device AI Models Land

AMD open-sources fused FlashAttention kernels for XDNA NPUs, Qualcomm ships a Linux preview for Snapdragon X2, and fresh research pushes tiny and trillion-scale models toward

89 views
Read
AI PCs Go Big: A 300B-Parameter Desktop, an 80-TOPS Mini PC, and Edge NPUs Redraw the Local-Inference Map
Edge AI4 min read

AI PCs Go Big: A 300B-Parameter Desktop, an 80-TOPS Mini PC, and Edge NPUs Redraw the Local-Inference Map

GMKtec, ASUS and Radxa all shipped NPU hardware this week while OpenVINO and a Qualcomm robotics runtime pushed what those chips can actually run.

90 views
Read
Qualcomm Chases 30B AI Models on a Phone as Edge Silicon Keeps Multiplying
Edge AI4 min read

Qualcomm Chases 30B AI Models on a Phone as Edge Silicon Keeps Multiplying

A phone NPU claims 30B MoE inference, a 35B model streams from storage on a Mac, and an XDNA1 NPU gets a Linux bring-up.

107 views
Read
AMD, Qualcomm and NVIDIA All Chase the Same Local-AI Bottleneck: Memory
Edge AI5 min read

AMD, Qualcomm and NVIDIA All Chase the Same Local-AI Bottleneck: Memory

Three chipmakers and a mini-PC builder all attack the same problem this week: getting AI inference closer to memory, not just closer to silicon.

128 views
Read
Edge Dispatch: Intel's Wildcat Lake Brings a Modest NPU to AI PCs, While Edge AI Claims It's Going Mainstream in IoT
Edge AI3 min read

Edge Dispatch: Intel's Wildcat Lake Brings a Modest NPU to AI PCs, While Edge AI Claims It's Going Mainstream in IoT

Intel details a 17 TOPS NPU chip built for chiplets, and an industry trend piece argues edge inference is leaving the pilot stage — both light on independent proof so far.

99 views
Read
Flipper One Wants to Be the First Hacker Tool With a Local LLM — What Its 6 TOPS NPU Can Actually Run
Edge AI7 min read

Flipper One Wants to Be the First Hacker Tool With a Local LLM — What Its 6 TOPS NPU Can Actually Run

Flipper Devices' pocket Linux box promises an LLM that runs offline and knows the device inside out. Rockchip's own numbers say what a 6 TOPS RK3576 really does — and the NPU driver isn't in the kernel Flipper chose.

176 views
Read
Edge Dispatch: AMD's Ryzen AI Halo Jumps to 192GB While Qualcomm Pushes On-Device AI Agents
Edge AI5 min read

Edge Dispatch: AMD's Ryzen AI Halo Jumps to 192GB While Qualcomm Pushes On-Device AI Agents

AMD bumps its NPU mini PC to 192GB of unified memory, Qualcomm ships five agentic apps for Snapdragon X, and Korea rethinks its NPU strategy.

158 views
Read
Edge Dispatch: A Robot Arm Joins the NPU Party: FastFlowLM Adds a Vision-Language-Action Model to Ryzen AI
Edge AI5 min read

Edge Dispatch: A Robot Arm Joins the NPU Party: FastFlowLM Adds a Vision-Language-Action Model to Ryzen AI

FastFlowLM's first stable release puts a robotics policy on Ryzen AI's NPU, while Korea ships a boxed NPU appliance and a Raspberry Pi learns to narrate what it sees.

100 views
Read
Edge Dispatch: Korea's KT Ships a Boxed NPU LLM Station While the ESP32 Crowd Trims Memory Further
Edge AI3 min read

Edge Dispatch: Korea's KT Ships a Boxed NPU LLM Station While the ESP32 Crowd Trims Memory Further

A Korean telecom sells an all-in-one on-prem LLM box built on a domestic NPU, while ESP32 tinkerers keep shrinking what a model needs to run.

116 views
Read
Edge Dispatch: Ryzen AI's NPU Runtime Goes Official While a Raspberry Pi Learns to See and Speak with a Tiny LLM
Edge AI4 min read

Edge Dispatch: Ryzen AI's NPU Runtime Goes Official While a Raspberry Pi Learns to See and Speak with a Tiny LLM

AMD folds a hobbyist NPU runtime into ROCm, Google shows Gemma driving a robot from a Raspberry Pi 5, and a 45M-parameter model books tool calls on a phone.

108 views
Read
Edge Dispatch: Raspberry Pi's GPU Joins the AI Party While a 14MB Model Learns to Call Tools
Edge AI4 min read

Edge Dispatch: Raspberry Pi's GPU Joins the AI Party While a 14MB Model Learns to Call Tools

A Raspberry Pi 5 runs Gemma and vision models split across CPU and GPU, and a 45M-parameter model fits tool-calling into 28MB of RAM.

92 views
Read
AMD Just Bought the 'Ollama of NPUs': What FastFlowLM Means for Local LLMs
Edge AI6 min read

AMD Just Bought the 'Ollama of NPUs': What FastFlowLM Means for Local LLMs

A 17MB runtime that runs LLMs on AMD's Ryzen AI NPUs — built by three academics, acquired by AMD on July 17, 2026, folded into ROCm in August. The NPU rung of the edge ladder just got real.

234 views
Read