Nucleus builds and runsyour enterprise LLM

We fine-tune models on your data and deploy them as private API endpoints, with dashboards, automation, scheduling and monitoring built in. Your secure LLM, at ChatGPT-level intelligence, with full data privacy.

National Neuroscience Institute logo
National University Hospital logo
Mandai Wildlife Group logo
SingHealth logo
Infocomm Media Development Authority logo
Institute for Adult Learning logo
Ministry of Education logo
Nanyang Polytechnic logo
Institute of Singapore Chartered Accountants logo
Ministry of Health logo
National Neuroscience Institute logo
National University Hospital logo
Mandai Wildlife Group logo
SingHealth logo
Infocomm Media Development Authority logo
Institute for Adult Learning logo
Ministry of Education logo
Nanyang Polytechnic logo
Institute of Singapore Chartered Accountants logo
Ministry of Health logo
National Neuroscience Institute logo
National University Hospital logo
Mandai Wildlife Group logo
SingHealth logo
Infocomm Media Development Authority logo
Institute for Adult Learning logo
Ministry of Education logo
Nanyang Polytechnic logo
Institute of Singapore Chartered Accountants logo
Ministry of Health logo
National Neuroscience Institute logo
National University Hospital logo
Mandai Wildlife Group logo
SingHealth logo
Infocomm Media Development Authority logo
Institute for Adult Learning logo
Ministry of Education logo
Nanyang Polytechnic logo
Institute of Singapore Chartered Accountants logo
Ministry of Health logo
Trusted by leading institutions
Product

Fine-tune. Serve. Keep the weights.

Train open models on your data across five modalities, serve them behind private endpoints, and export the weights whenever you want. Three meters run the bill: training, inference and storage.

49 open models, five modalities

Fine-tuning from managed jobs
to raw training steps


Language, vision, image, video and audio bases. Start with a managed job and go deeper when the task asks for it, all the way down to individual forward-backward passes.

  • Managed jobs. Upload examples, pick a base model, start a run. Scheduling, GPUs and the production handoff are handled.
  • Custom recipes. Your base, your data, your method. LoRA ranks up to 64 and evaluations billed on the same meters.
  • Loop-level access. Forward-backward passes and optimiser steps as first-class API calls. RL, DPO and distillation are yours to write.
Training tokens from $0.30 / 1M
Datasets feeding a fine-tuning coredataset.jsonlLoRA r=32step 1,240 / 2,000
Models

Choose your base model

Open models across language, vision, image, video and audio. Fine-tune on your data, serve the result as a private endpoint, and keep the weights.

zai-org · Language

GLM 5.3

zai-org/GLM-5.3

MoE · 780B params · 40B active

Context
1M
LoRA rank
16
Fine-tune
$4.60
Serve
$2.95
Qwen · Vision

Qwen3-VL 235B Instruct

Qwen/Qwen3-VL-235B-A22B-Instruct

MoE · 235B params · 22B active

Context
128K
LoRA rank
32
Fine-tune
$2.40
Serve
$1.60
black-forest-labs · Image generation

FLUX.2 dev

black-forest-labs/FLUX.2-dev

Flow transformer · 32B params

Context
Up to 4MP
LoRA rank
32
Fine-tune
$1.40
Serve
$8.00
Lightricks · Video generation

LTX-2

Lightricks/LTX-2

DiT · native audio + video

Context
Up to 4K · 10s
LoRA rank
32
Fine-tune
$1.60
Serve
$5.50
Qwen · Embedding & reranking

Qwen3 Embedding 8B

Qwen/Qwen3-Embedding-8B

Bi-encoder · 8B params · 4096 dims

Context
32K
LoRA rank
64
Fine-tune
$0.08
Serve
$0.05
openai · Language

GPT-OSS 120B

openai/gpt-oss-120b

MoE · 117B params · 5.1B active

Context
128K
LoRA rank
32
Fine-tune
$0.68
Serve
$0.45
Pricing

Near-frontier intelligence at 2% the cost

Three meters run the bill: training tokens, served tokens and storage. Reserve dedicated GPUs when you need guaranteed capacity. Idle time costs nothing.

Training tokens
from$0.03/ 1MFloor rate across 47 trainable bases.
Served tokens
from$0.02/ 1MInput and output tokens cost the same.
Storage
flat$0.10per GB / monthDatasets, checkpoints and exported weights.
Dedicated GPUs
from$1.99/ GPU-hourL40S to B200, billed by the second.
CompanyAlibabaNVIDIADeepSeekOpenAIZ AIGoogleKimiMistralMiniMaxMetaOtherreads more than one modalityunscored, shown in the lanewithdrawal scheduled
Use cases

Built with Nucleus

Agents that connect to the tools a business already runs, and deliver the result a team used to spend a day on. Every workflow runs on the Nucleus inference stack, ready to run on your data.

47 trucks rerouted while they were still moving GPS traces, traffic and weather become live route changes, fuel efficiency per vehicle, and an SMS to every customer whose window slipped.

Quickstart

From dataset to a model you own

A handful of API calls take a JSONL file to a fine-tuned model served behind an OpenAI-compatible endpoint. Prefer to drive the training loop yourself? Forward-backward passes and optimiser steps are first-class API calls too.

nucleus quickstart
01_upload_dataset.sh

Bring a JSONL file of examples. That's the only prerequisite.

Teams

Private models for every team

Fine-tune, evaluate, serve, and export models your organisation owns, with a workflow for every team that touches them.

ML Engineers

Fine-tune and deploy custom models on your enterprise data with full control over training pipelines and inference endpoints.

Research Teams

Drive the training loop call by call. RL, DPO, distillation, and custom losses, with the GPUs handled for you.

Product Teams

Ship features on models that speak your domain. Keep your existing OpenAI or Anthropic client code and swap the base URL.

Data Teams

Turn proprietary datasets into evaluated, versioned models with managed fine-tuning jobs. Upload a JSONL file, get back a private endpoint.

Platform Teams

Serve models behind OpenAI- and Anthropic-compatible endpoints, with usage tracking and rate limits built in.

Enterprise Leaders

Own your AI strategy. Training data stays in your account, weights are exportable, and models can run wherever policy requires.

Scale

Built to scale. Built to be owned.

Nucleus handles training and inference at any scale, from the first fine-tuning job to full production traffic, so your team ships models instead of managing GPUs.

  • Your model weights

    100% yours

  • Compatible APIs

    OpenAI + Anthropic

Tokens served
10.9M
+13.4%

Your tuned and private LLM. At a fraction of the cost.