ProductHardwareLibraryCareersDocs
Book a call

The research lab focused on inference.

TrainingAutomatic fine-tuning from your traffic.InferenceOpen-model inference, pay based on your latency requirement.
NotesFrontier notes on inference, systems and hardware.MethodsHow we operate, and why.

Latest writing

  • Announcing Carat: An inference engine designed for Gemma 4Sep 2, 2026
  • ThesisJul 13, 2026
  • Affordable and open intelligenceSep 9, 2026
See all writing

Fast and cheap inference post-trained on your workload.

We turn your traces into an RL environment, build the evals, and post-train an open model automatically. You just change your base URL.

Get API KeyView Docs

Latest News

Read More
Announcing Carat: An inference engine designed for Gemma 4Designing Carat, an inference engine built specifically around the model architecture of Gemma 4.Sep 2, 2026
ThesisMost of the car is missing.Jul 13, 2026
Affordable and open intelligenceMore intelligence for less energy, open to everyone.Sep 9, 2026

Product

  • Training
  • Inference
  • Service tiers
  • Models and pricing

Developers

  • Quickstart
  • API reference
  • Errors
  • Status

Company

  • Thesis
  • Careers
  • Library

Legal

  • Privacy
  • Terms
  • Acceptable use
  • DPA
  • Security

© 2026 Gradiated Ltd

Cambridge, --:--:-- GMT