Forlinx 20-TOPS M.2 AI accelerator supports PCIe cascading for local LLM inference

Forlinx Embedded has listed an M.2 AI accelerator card based on Rockchip’s RK1820 and RK1828 processors, providing 20 TOPS of INT8 computing performance and up to 5GB of integrated DRAM. The module uses an M.2 2280 interface and is designed to handle local AI inference, including large language models, vision-language models, and computer vision workloads on embedded Linux and Android systems. linuxgizmos.com/forlinx-20-tops-m-2-ai-...pcie-cascading-for-local-llm-inference

添加评论
点赞收藏
点踩分享查看原文
评论
?
参与讨论