Local Deployment of DeepSeek-R1 Large Model on RK3588 Platform: NPU Edition

Local Deployment of DeepSeek-R1 Large Model on RK3588 Platform: NPU Edition

Using the NPU to run inference on large models can more effectively plan resource allocation, achieving more efficient applications. 1 Preparation Before Deployment To deploy large language models using the NPU on RK3588, the following files need to be prepared in advance, which include the corresponding files for Rockchip’s NPU and the DeepSeek inference model … Read more

Specifications and Parameters of the Ziguang Zhanrui T710/UDS710 Core Board – 4G Android Core Board Custom Solutions

Specifications and Parameters of the Ziguang Zhanrui T710/UDS710 Core Board - 4G Android Core Board Custom Solutions

The T710 uses the Zhanrui 4G UDS710 chip, which is a powerful module with rich interfaces in the WiFi version. It is a highly integrated embedded application processor that includes a quad-core ARM Cortex-A75 00P core and a quad-core ARM Cortex-A55MP as the application processor. It runs on the Android 10.0 system, supports 2.4G and … Read more

Comparative Analysis of Computing Acceleration Technologies: Technical Characteristics, Application Scenarios, and Industrial Ecosystems of GPUs, FPGAs, ASICs, TPUs, and NPUs

Comparative Analysis of Computing Acceleration Technologies: Technical Characteristics, Application Scenarios, and Industrial Ecosystems of GPUs, FPGAs, ASICs, TPUs, and NPUs

Source: DeepHub IMBA This article has a total of 4500 words, and it is recommended to read in 5 minutes. This article will delve into five main types of computing accelerators. In today’s rapidly evolving computing technology, traditional general-purpose processors (CPUs) are gradually being supplemented or replaced by dedicated hardware accelerators, especially in specific computing … Read more

Edge AI Chips: The Core Engine for Intelligent Applications

Edge AI Chips: The Core Engine for Intelligent Applications

The Edge AI chip is a processor specifically designed to efficiently run artificial intelligence algorithms on terminal devices such as smartphones, IoT devices, and autonomous vehicles. Through hardware-level optimization, they can achieve low power consumption and high real-time AI computing, forming the core hardware foundation for edge AI applications. Why do we need edge AI … Read more

Using NPU on RK Platform

Using NPU on RK Platform

With the development of AI intelligence, many chips more suitable for AI learning have been introduced following traditional CPUs and GPUs. This article introduces how to develop NPU chips based on the SDK provided by the RK platform. 1. Introduction to NPU Chips NPU stands for Neural Network Processing Unit. 2. Using RKNN 1. SDK … Read more

NXP Leads the Way by Integrating NPU into MCU

NXP Leads the Way by Integrating NPU into MCU

According to a report by Electronic Enthusiasts Network (by Cheng Wenzhi), IC Insights released the MCU sales data for 2021 a few days ago, revealing that NXP’s MCU sales reached $3.795 billion, ranking first. In fact, NXP is not only experiencing rapid growth in MCU sales but is also continuously innovating in product development. In … Read more

Limitations and Future Prospects of NPU in Large Model Applications

Limitations and Future Prospects of NPU in Large Model Applications

👇 Follow our official account and selectStar, to receive the latest insights daily Abstract The NPU (Neural Processing Unit) is a chip designed specifically for neural network computations, excelling in matrix operations and convolution tasks, characterized by low power consumption and high efficiency. It is mainly used for inference tasks on edge devices, such as … Read more

The Era of DeepSeek: ASIC Chips Crowned as Kings

The Era of DeepSeek: ASIC Chips Crowned as Kings

Since the emergence of ChatGPT at the end of 2022, followed by the hundred-model battle in 2023, and the recent releases of GPT-4.5 by OpenAI, Grok3 by xAI, Claude 3.7 Sonnet by Anthropic, and Llama4 by Meta, the iteration speed of large models has been accelerating. In China, there has been a surge in open-source … Read more

RKLLama: LLM Server and Client for Rockchip 3588/3576 Chips

RKLLama: LLM Server and Client for Rockchip 3588/3576 Chips

Address:https://github.com/NotPunchnox/rkllama RKLLama is an open-source server and client solution designed to run large language models (LLMs) optimized for the Rockchip RK3588 (S) and RK3576 platforms, and to interact with them. Unlike solutions such as Ollama or Llama.cpp, RKLLama fully utilizes the Neural Processing Units (NPU) on these devices, providing an efficient and high-performance solution for … Read more

Live Broadcast: Detailed Explanation and Practical Demonstration of Efficient AI Application Deployment Based on ‘Zhou Yi’ NPU

Live Broadcast: Detailed Explanation and Practical Demonstration of Efficient AI Application Deployment Based on 'Zhou Yi' NPU

Course Introduction The emergence of large models like DeepSeek has triggered explosive growth in AI applications, with a continuous rise in edge inference demand. The NPU (Neural Processing Unit), with its excellent AI acceleration capabilities and high energy efficiency, has become a key solution to meet terminal inference needs. Arm Technology’s self-developed ‘Zhou Yi’ NPU … Read more