Together AI and Y Combinator have announced a partnership to provide the first dedicated GPU cluster exclusively for Y Combinator's portfolio of AI-native startups. This initiative aims to solve the critical bottleneck of compute access by offering flexible, cost-effective infrastructure for training and inference without requiring long-term commitments.
- The cluster allows startups to reserve and provision GPUs directly through Together’s self-service portal with individual billing, keeping management separate from YC.
- Founders can spin up GPUs in minutes and benefit from long-term rates while maintaining the flexibility to scale for short-term sprints.
- The infrastructure supports a full range of needs, from single-node compute for early-stage teams to larger scaling requirements.
- Together AI leverages its research on attention mechanisms and Mamba architecture to improve inference speed and unit economics for workloads on the cluster.
The partnership addresses the growing difficulty and cost of securing compute capacity, enabling founders to access resources comparable to those of larger companies while retaining control over their usage.