OneTriangle

The cheapest, fastest lightweight inference.

onetriangle.ai

Community-submitted · content unverified. This organization has not verified ownership or the information in this profile.

About

Inference has surpassed training as AI’s largest compute cost, accelerated by increasing demand for AI and agents. By using proprietary KV cache transfer tech, OneTriangle has cut inference costs by 20% and time by 40% compared to present standards. OneTriangle is a Y Combinator company (Summer 2026). YC-listed status: Active.

Location
San Francisco, CA, USA
Added via
web
Ownership
unclaimed

Where to go

Products

  • OneTriangle — Inference engine that cuts LLM inference costs and latency by 20% using KV cache transfer between a small prefill model and a large decode model. · Api

Links