chiprook
← Hardware
HardwareSeptember 29, 2026, 11:07

Forlinx launches 20-TOPS M.2 AI accelerator for local LLM inference

Forlinx has listed an M.2 AI accelerator card built on Rockchip's RK1820 and RK1828 coprocessors, delivering 20 TOPS of INT8 performance and up to 5GB of stacked DRAM. The M.2 2280 module targets local LLM, VLM and computer vision inference on Linux and Android, and a four-card PCIe cascade on the OK3588-C board can run 27B-31B parameter models at roughly 40W.

Forlinx launches 20-TOPS M.2 AI accelerator for local LLM inference
#Forlinx#Rockchip#RK3588
Read next
Hardware

Georgia Tech, Nvidia and Stanford's BOOST speeds LLM inference by 31%

Hardware

MLPerf Inference v6.1: 5.7x per-accelerator gains, 512-GPU run, Vera Rubin results

Software

Run local LLM in ChatGPT Desktop app with opencodex alone

Hardware

Apple reportedly developing enterprise AI inference server with M8 Ultra chips