Everything you need to deploy, fine-tune, route, and observe SLMs and LLMs on Quantum Junction.
Deploy your first model in under 60 seconds. Covers installation, authentication, and your first inference call.
Full OpenAI-compatible REST API documentation. Deployments, inference, fine-tuning, and routing endpoints.
The official Quantum Junction Python SDK. Async-first, type-annotated, fully compatible with LangChain and LlamaIndex.
Three steps: install the SDK, authenticate, deploy.
All endpoints follow the OpenAI API specification. Drop in your Quantum Junction API key and base URL — your existing code works without modification.
Official SDKs for Python, TypeScript/Node, and Go. All async-first, fully typed. Compatible with LangChain, LlamaIndex, and Haystack.
Subscribe to deployment events, fine-tuning job completion, and inference anomalies via webhook. HMAC-signed payloads.
Change two lines — your base_url and api_key — and your existing OpenAI-based code runs on Quantum Junction with no other changes.
Start free in under 60 seconds. No GPU reservations. No minimum commits.