Hyperbolic Labs is democratizing AI by breaking down barriers to computing power via its Open-Access AI Cloud, a GPU marketplace and AI inference service built on idle compute across the globe. This role builds inference capabilities on top of "Forge", the company's unified control plane, so customers can consume model tokens without managing GPUs.
Based in San Francisco, CA. Preferred experience includes inference optimization (quantization, batching, kernel tuning) and prior GPU cloud/inference-provider work.

Hyperbolic operates a decentralized, aggregated GPU marketplace that gives AI teams affordable, on-demand access to compute. Having recently closed a Series A, the company is scaling its supply operations and infrastructure to meet growing demand for AI compute.
Apply To This Job<<>>
Support us by letting the company know you found them on our website.
Go To the Offer