Skip to main content
Models power your AI applications. Add a model to a project to deploy it, then make inference requests. A model in your project comes from one of four sources: The catalog is the recommended default. New models appear there server-side without requiring a platform update.

Add and deploy a model

Once a model is in your project’s registry, deploy it for inference:
The model becomes available within a few minutes.

Import from Hugging Face

For models not in the Adaptive catalog, import from Hugging Face directly:
Import runs as an asynchronous job — the model appears in the registry once conversion finishes.
Catalog import (the recommended path) is currently UI-only. Open the Add Model dialog in your project and pick Import from Adaptive ML’s catalog.
A training run saves snapshots at configurable intervals (see checkpoint_frequency). Any saved checkpoint can be promoted to a standalone model in the registry, then evaluated or deployed like any imported model.Promotion is currently UI-only — open a run and pick Promote checkpoint on a saved checkpoint. SDK support uses the underlying GraphQL mutation directly:
Promotion runs as an asynchronous copy job. Promoted models retain full provenance back to the source run.
LoRA backbone footgun: when you promote a LoRA checkpoint, the platform records the backbone reference in the model’s metadata but does not automatically attach the backbone to your project. If the backbone isn’t already in the project, inference against the promoted LoRA fails with a “model not found” error. Add the backbone with add_to_project before deploying.
See SDK Reference for all model methods.