Skip to main content
Choose a model based on the capability, speed, modality, and efficiency your application needs. You can change models without changing the rest of your API integration.
Not sure where to start? Use Lume 3 for most applications. Choose Lume 3 Max for the most demanding reasoning and coding tasks, Lume 2 for lightweight high-volume workloads, or Lume 3 Omni when you need multimodal capabilities.

Choose a model

Lume 3

Recommended defaultStrong general-purpose performance with an efficient cost profile. Start here for most text, coding, analysis, and application workloads.

Lume 3 Max

Highest capabilityThe flagship Lume model for advanced reasoning, complex coding, mathematics, physics, and demanding multi-step tasks.

Lume 2

Highest efficiencyA cost-efficient text model for lightweight, well-defined, and high-volume workloads.

Lume 3 Omni

MultimodalA multimodal model with tool use for workflows that need more than text-only processing.

Model comparison


Context window

Lume 3-generation models support context windows of up to 400K tokens. The context window includes the information available to the model during a request, such as instructions, messages, documents, tool context, and generated output. Large context windows are useful for long documents, codebases, extended conversations, and workflows that need substantial supporting information.
More context does not automatically produce better results. Send only the information that is relevant to the task whenever possible.

Model IDs

Use the model ID exactly as documented when making an API request.
For example:
To use another model, change the model value. Your API key and Base URL remain the same.

Next steps

Quickstart

Make your first API request with a Lume model.

Pricing

Compare pricing across the Lume model family.

Integrations

Connect OverControl to supported tools and OpenAI-compatible clients.