OpenRouter did not expect that a separate ecosystem of specialized inference providers would emerge to host open weight models, rather than hyperscalers dominating that layer.
OpenRouter's founding thesis assumed hyperscalers like Google, Amazon, and Azure would monopolize open weight model hosting, but specialized inference providers like Fireworks and Together proved far faster and more capable at serving these models.
transcript
Alex Atallah: one thing that we did not expect was that a an ecosystem of companies would emerge to host and serve the open weight models. Um like early on, it wasn't clear that that that market wasn't going to be a monopoly where like just, you know, the three hyperscalers serve all the open weight models and uh and start and startups don't, you know, they're they're really far behind. In reality, like you know, how often do you hear people running, you know, GLM on a hyperscaler? Never. Like they're using the the inference providers like Fireworks and Together.
explains mechanism · 1gives example · 1provides context · 2