Use cases
Workload-first guidance for software development, documents, data-heavy jobs, support, research, local inference, and GPU hosting.
Use case
Software development
Use subscriptions for humans and open-model APIs for routine loops; reserve frontier models for planning, architecture, reviews, and complex bugs.
Use case
Data-heavy processing
Batch non-urgent work, validate cache rates, and compare open-model APIs before committing to dedicated GPUs.
Use case
Research workflows
Use cheap collection and extraction lanes, then spend frontier tokens only on synthesis, review, and difficult reasoning.
Use case
Document processing
Process large document volumes with open-model extraction, batch routes, cache discipline, and a small frontier evaluator sample.
Use case
Support automation
Price routine support drafts separately from policy-sensitive escalations, human review, and frontier-model evaluation.
Use case
Local inference
Benchmark open models locally before renting GPUs, then promote only models that pass quality, latency, and utilization tests.
Use case
GPU hosting strategy
Compare GPU rental, storage, egress, utilization, interruption risk, and operations overhead before self-hosting a workload.