curious if Jev-like models might accelerate us towards on-device AI native functionality on mobile devices

strong LLMs are a long way from running on phones:

  • slow memory bandwidth
  • only highly quantized MoE models will fit
  • lots of issues with power/heat/etc

Folks are building NPUs for the next gen of mobile but targeting pretty modestly sized LLMs

There’s lots of mobile UX that would benefit from fast/cheap AI decision models - from notifications, typing/texting, to in-app experiences like inboxes/calendars/etc

I always figured the other solution would be a model-on-a-chip that would encode an older-yet-useful model into the actual phone hardware itself, but maybe this will move faster?

添加评论
点赞收藏
点踩分享查看原文
评论
?
参与讨论