OffRail

Overview

How OffRail exposes NSFW-capable generation models through a managed multimodal API.

OffRail fronts a focused catalogue of NSFW-capable generation models with one managed API. Your application addresses an OffRail model ID; the gateway handles eligible upstream access, request metering, and the delivery lifecycle for that modality.

Public release scope

The first public release is intentionally focused:

  • image, video, chat, and speech generation routes;
  • usage-based billing without a subscription requirement;
  • managed upstream routing behind each public model ID;
  • a live model catalogue as the source of truth for availability.

Self-service console access and API key provisioning are still in private preview. There is no separate demo route or public dashboard workflow in this release.

What stays consistent

Every public route uses the same core contract:

  1. Authenticate with an OffRail API key.
  2. Send the selected public model ID in the request.
  3. Let OffRail resolve an eligible upstream mapping.
  4. Receive a response or track a task using the documented delivery shape.
  5. Pay for measured usage in the route's native billing unit.

What changes by modality

ModalityPrimary inputResultLifecycle
ChatMessages and controlsText or structured responseSync or streaming
ImagePrompt and optional referenceImage URL, base64, or taskSync or task
VideoPrompt and optional referencesPollable video taskAsync
AudioText and voice selectionBinary audio payloadSync

Model-specific sizes, durations, voices, and prices are exposed on the public catalogue and the relevant API reference page.

Managed routing

OffRail separates the model your application requests from the upstream provider that serves it. This lets the public model ID remain stable while eligible provider selection and supported fallback happen behind the gateway.

Fallback never changes the route contract. A video request remains an asynchronous video task; an audio request remains an audio response.

Usage-based billing

Billing follows the work performed by each route:

  • chat is metered by input and output tokens;
  • images are metered per generated image;
  • video is metered per generated second;
  • speech is metered by input characters.

See Pricing for the public billing model and each model detail page for its current starting rate.

Next step

Continue to the Quickstart when you have a private preview API key.

On this page