ProductHardwareLibraryCareersDocs
Book a call

The research lab focused on inference.

TrainingAutomatic fine-tuning from your traffic.InferenceOpen-model inference, pay based on your latency requirement.
NotesFrontier notes on inference, systems and hardware.MethodsHow we operate, and why.

Latest writing

  • Announcing Carat: An inference engine designed for Gemma 4Sep 2, 2026
  • ThesisJul 13, 2026
  • Affordable and open intelligenceSep 9, 2026
See all writing

Library

Notes & Methods

Announcing Carat: An inference engine designed for Gemma 4Designing Carat, an inference engine built specifically around the model architecture of Gemma 4.Sep 2, 2026
ThesisMost of the car is missing.Jul 13, 2026
Affordable and open intelligenceMore intelligence for less energy, open to everyone.Sep 9, 2026
What we’re buildingInference for long-running work.Sep 9, 2026
Sequencing and depthStart in software. Go down the stack.Sep 9, 2026
In personHard problems benefit from shared context.Sep 9, 2026
What we look forAgency, optimism, and exceptional ability.Sep 9, 2026
Getting out the gateWin the first users by being meaningfully better.Sep 9, 2026
ValuesFigure it out. Do the hard thing. Primitive first.Sep 9, 2026
BrandEverything is considered.Sep 9, 2026

Product

  • Training
  • Inference
  • Service tiers
  • Models and pricing

Developers

  • Quickstart
  • API reference
  • Errors
  • Status

Company

  • Thesis
  • Careers
  • Library

Legal

  • Privacy
  • Terms
  • Acceptable use
  • DPA
  • Security

© 2026 Gradiated Ltd

Cambridge, --:--:-- GMT