New: SLMs outperform frontier LLMs on narrow enterprise tasks by up to 34% — Read the ScaleDown AI benchmark →
Products
⚡ Platform 🧠 Model Catalog 📖 Documentation
Solutions
🏥 Healthcare ⚖️ Legal 📈 Finance 🔧 Manufacturing
More
📊 Benchmarks 💳 Pricing 🔑 Sign In Deploy a model →
Model Catalog

The right model
for every task.

Browse task-specific SLMs built for precision, or frontier LLMs for open-ended reasoning. Every model is one click from production.

Browse all models → Compare benchmarks
Task-Specific SLMs

Precision models for narrow enterprise tasks

Fine-tuned on domain-specific corpora. Faster, cheaper, and more accurate than frontier models on the tasks they're built for.

Healthcare

MedCode-SLM-3B

Medical ICD-10/11 coding, clinical note classification, and diagnosis extraction. 94.2% accuracy — 18 points above GPT-4 on MIMIC-III coding benchmark.

3.2B params12ms p50$0.00004/1k tok
Legal

LexSLM-7B

Contract clause extraction, obligation identification, and risk flagging. Fine-tuned on 4M contract clauses. 91% F1 on CUAD benchmark.

7B params28ms p50$0.00012/1k tok
Finance

FinSLM-1.3B

Real-time earnings sentiment, SEC filing classification, and risk signal extraction. 97% accuracy on FinBERT benchmark. Fast enough for live trading pipelines.

1.3B params5ms p50$0.000018/1k tok
Manufacturing

IndustrySLM-2B

SCADA log parsing, predictive maintenance signals, and anomaly detection. 88% recall on critical equipment failure events, 72 hours in advance.

2B params8ms p50$0.000028/1k tok
Genomics

GenomeSLM-6B

Variant interpretation, gene expression classification, and pathogenicity prediction. Pre-trained on 500M genomic sequences from NCBI and Ensembl.

6B params22ms p50$0.00009/1k tok
Customer Support

SupportSLM-1B

Intent classification, ticket routing, and resolution suggestion for enterprise support queues. 95% routing accuracy. Deploy in your support platform via webhook.

1B params4ms p50$0.000014/1k tok
Frontier LLMs

Open-ended reasoning at enterprise scale

For agentic workflows, complex reasoning, and tasks where domain breadth matters more than depth.

Meta

Llama 3.1 405B

Meta's flagship open-weight model. Best-in-class for reasoning, coding, and agentic workflows. Runs on 8×H100 with tensor parallelism.

Mistral

Mistral Large 2

Efficient frontier model with 128k context. Excellent multilingual support and function calling. Ideal for RAG pipelines and document analysis.

Microsoft

Phi-3 Medium

14B parameter model punching above its weight. Competitive with models 5× larger on reasoning benchmarks. Cost-effective frontier option.

Ready to deploy your exact-fit model?

Start free in under 60 seconds. No GPU reservations. No minimum commits.

Deploy now — it's free → Read the docs