Batch Processing

Large jobs, at half the price.

Submit a million requests. Get results within 24 hours. Same models, same APIs, half the cost. Perfect for evals, bulk classification, dataset processing.

Async by default

Submit a job, get a webhook, download results. No need to hold a connection.

24-hour SLA

Most jobs complete in 2-4 hours. Hard guarantee: 24 hours or the job is free.

50% of sync pricing

Same quality, same reliability, half the cost. The trade is time, not quality.

Up to 100K per job

Submit up to 100,000 requests per batch. Larger jobs queued automatically.

Standard format

JSONL input, JSONL output. Drop-in compatible with the OpenAI Batch API.

Persistent storage

Results stored for 30 days. Download via signed URL.

Pricing

Token-efficient, with volume discounts and burn incentives.

Chat batch
0.0002QUBIC / 1K tokens

50% of sync rate

Embedding batch
0.00006QUBIC / 1K tokens

60% of sync rate

Image batch
0.0040QUBIC / image

50% of sync rate

Reasoning batch
0.0012QUBIC / 1K tokens

50% of sync rate

Pricing is illustrative. Final rates are governed by on-chain parameters and may vary based on network state.

Staking requirements

Tier-based access. Higher stakes unlock better economics and more capacity.

TierRequired stakeAccess
Builder50M QUBICBatch API enabled
Startup150M QUBICLarger jobs, priority queue
Business500M QUBICCustom windows, dedicated capacity
EnterpriseCustomOn-prem, custom SLA

Example

Drop-in compatible with the OpenAI SDK.

batch.py
from aigarth import Aigarth

client = Aigarth(api_key="sk-aigarth-...")

# Submit a batch of classification requests
with open("requests.jsonl") as f:
    job = client.batches.create(
        input_file="file-abc123",
        endpoint="/v1/chat/completions",
        completion_window="24h",
    )

# Wait for completion (or use webhook)
result = job.wait()
print(f"Processed {result.request_counts.completed} requests")

Enterprise benefits

Everything in the standard tier, plus the things enterprises need.

Ready to get started?

Open the console, generate an API key, and run your first call in minutes.