Bringing BitNet to ExecuTorch via Vulkan

BitNet-style ternary brings LLM inference to ExecuTorch via its Vulkan backend, enabling much smaller, bandwidth-efficient models with portable GPU execution on edge devices. Presented at PyTorch Conference Europe 2026.

添加评论
点赞收藏
点踩分享查看原文
评论
?
参与讨论