Conifer
Local-first routing that cuts AI token costs by keeping most requests on your hardware.
- Category
- AI & Machine Learning
- Launched
- July 15, 2026
- Stage
- Unknown
- Pricing
- free
- Website
- conifer.build



The cloud has a pricing meter that never stops ticking, and for AI teams, every token is a coin dropped into that meter. Conifer, a new entrant from the Y Combinator launch roster, proposes a simple but radical fix: don't send all your requests to the cloud. Instead, route most of them to your own hardware, where the marginal cost is effectively zero. The company's tagline—'Stop paying cloud prices for every token'—is a direct appeal to every team that has watched its inference bill balloon as usage scales. Conifer's pitch is quantitative: 'We route 80% of requests to your hardware at no cost.' That number is the crux of its value proposition. It suggests that the vast majority of AI workloads don't need the cloud's elasticity or scale—they can run perfectly well on commodity GPUs or even CPUs sitting in a server closet. The cloud becomes an overflow valve, not the default destination. This is a contrarian stance in an industry where 'cloud first' is dogma, and it's worth unpacking.
Discovered via
- Y Combinator Launches · · Conifer