Serverless inference at scale.
OpenAI and Anthropic SDK compatible. Deploy chat and coding model inference instantly. Pay per 1M tokens, scale automatically. No infrastructure to manage.
Live model catalog.
Access our full roster of chat and coding models. Real pricing updated as rates change. OpenAI and Anthropic SDK compatible.
Base URL: https://api.runars.ca/v1
| Model | Input ($/M) | Output ($/M) | Context |
|---|---|---|---|
| Aion Labs Aion 2.0🇨🇦 | $2.2784/M | $4.5568/M | 128K |
| Deepseek V4 Flash🇨🇦 | $0.2057/M | $0.4114/M | 1M |
| Deepseek V4 Flash 0731🇨🇦 | $0.1314/M | $0.2645/M | 1M |
| Deepseek V4 Pro🇨🇦 | $1.3313/M | $3.9938/M | 1M |
| Deepseek V4 Pro 0813🇨🇦 | $1.3313/M | $3.9938/M | 1M |
| GLM 4.5🇨🇦 | $0.9509/M | $3.4866/M | 128K |
| GLM 4.5 Air🇨🇦 | $0.3018/M | $1.7433/M | 128K |
| GLM 4.6🇨🇦 | $0.9509/M | $3.4866/M | 198K |
| GLM 4.7🇨🇦 | $0.9286/M | $3.4866/M | 128K |
| GLM 5🇨🇦 | $1.3930/M | $4.4575/M | 198K |
| GLM 5.1🇨🇦 | $2.2188/M | $6.9733/M | 200K |
| GLM 5.2🇨🇦 | $0.9509/M | $3.0310/M | 1M |
| GLM 5.3🇨🇦 | $2.2188/M | $6.9733/M | 1M |
| GLM 5.3 Flash🇨🇦 | $0.2089/M | $0.6965/M | 1M |
| Kimi K3🇨🇦 | $6.0512/M | $30.2560/M | 1M |
| Minimax M2.7🇨🇦 | $0.6051/M | $2.4205/M | 198K |
| Muse Spark 1.2🇨🇦 | $3.0616/M | $10.4095/M | 1M |
| Qwen 3.8 Max🇨🇦 | $5.0427/M | $15.1280/M | 1M |
| Qwen 3.5 35B A3B🇨🇦 | $0.6303/M | $2.5213/M | 256K |
| Qwen 3.6 35B A3B🇨🇦 | $0.2017/M | $2.0171/M | 262K |
| Qwen 3.6 27B🇨🇦 | $0.6965/M | $4.6432/M | 256K |
Pricing shown in CAD. Billing by token. All prices subject to change — check docs for current rates.
Built for any workload.
Inference, training, batch jobs, agents. If it runs in a container, it runs on Runars.
3.0 Use Cases →
Inference
Run LLM inference endpoints that scale with demand.
Fine-tuning
Train and fine-tune models on your data, your schedule.
AI Agents
Long-running autonomous tasks without server management.
Batch Processing
Compute-heavy workloads that run when needed, scale to zero.
Ready to run AI in Canada?
Start with the inference API for instant compatibility. Scale to GPU rental when you're ready. No lock-in, no setup fees.
npm install @anthropic-ai/sdk