Custom
The GPUs you specify, in the count and location you need.
Tell us the GPUs and the count. We deliver a dedicated cluster and run it for you.
GPU capacity brought online
B300 servers delivered
sites for 20 enterprise customers
inference efficiency gain at Meta
Track record of Larch Labs' founders in prior roles.
The GPUs you specify, in the count and location you need.
A deployment fee well below any cloud's rental margin. See the math below.
24/7 operations and performance tuning included.
The math, side by side
| Typical neocloud 3-year lease | Larch | |
|---|---|---|
| What you pay in total | Hardware, site and operations, plus about 50% of the hardware cost as the cloud's return | Hardware, site and operations, plus a one-time 5% of the hardware cost as initial deployment fee |
| Who owns the GPUs | The cloud | You |
GPU model, count and timeline.
A fixed price and a delivery date.
We deploy, test and hand over, then keep it running.
Already own GPUs? We also operate and optimize existing clusters.
Dedicated training and inference capacity.
Lower cost per GPU-hour, so more margin per token.
Build-out by people who've done it at hyperscale.
Private capacity, delivered and run end to end.
Co-founder & CEO · Optimization
Mark spent more than 10 years at Meta working on AI infrastructure at production scale. He led optimization projects that more than doubled the inference GPU efficiency of a flagship AI system, and those techniques were adopted across Meta's core production fleet, saving billions of dollars in GPU spend.
At Larch Labs he leads the optimization practice: profiling customers' training and inference stacks, finding where the bottlenecks are, and lifting GPU efficiency on hardware they already own.
LinkedIn ↗Co-founder & CTO · Deployment and Operations
Alex has 15 years of experience building infrastructure, most recently deploying GPU capacity at hyperscale. He led delivery of ~30MW and ~2,000 B300 servers across 17 sites for 20 enterprise customers, representing billions of dollars in total infrastructure value.
At Larch Labs he runs deployment and operations end to end: site selection, cluster architecture, procurement, installation and burn-in, then 24/7 operations once the cluster is live, backed by long-standing supplier relationships.
Co-founder & CMO · Go-to-market and BD
Rulan spent years at ByteDance building enterprise products for BytePlus, the company's B2B cloud and data business. She has also invested in early-stage AI infrastructure companies with Eastlink Capital, so she knows the market from both the builder's and the investor's side.
At Larch Labs she leads go-to-market and business development: positioning, partnerships with OEMs, colocation providers and GPU financiers, and the path from first call to signed deployment.
LinkedIn ↗We'll get back to you shortly.
Thanks. We'll reply by email.