1 min readfrom Machine Learning

What's your biggest pain point when choosing between cloud GPU providers for LLM inference?[R]

Trying to understand how other people make this decision. Do you compare $/hr, $/token, throughput, reliability? Is there a tool or resource you rely on, or are you just doing the math manually?

Asking because I'm an ML engineer who's been doing this in spreadsheets and wondering if I'm missing something obvious.

submitted by /u/Technomadlyf
[link] [comments]

Want to read more?

Check out the full article on the original site

View original article