What's your biggest pain point when choosing between cloud GPU providers for LLM inference?[R]

Trying to understand how other people make this decision. Do you compare $/hr, $/token, throughput, reliability? Is there a tool or resource you rely on, or are you just doing the math manually? Asking because I'm an ML engineer who's been doing this in spreadsheets and wondering if I'm missing something obvious. submitted by /u/Technomadlyf [link] [comments]

添加评论
点赞收藏
点踩分享查看原文
评论
?
参与讨论