Excited to see this go live with the @togethercompute team.
As more production workloads move to open models, reliable infrastructure becomes just as important as the models themselves.
Proud to see M3 launch with Provisioned Throughput. 🤝
We're introducing Provisioned Throughput: reserved inference capacity for frontier open models, with token-based pricing and a 99% uptime SLA. Serverless simpli...