Model Reference

plugsky-max — heavy reasoning, zero drama

plugsky-max is Plugsky's heavy reasoning model: 256K context, ~0.71s median latency, multimodal input, and strong tool calling.

plugsky-max: heavy reasoning with 256K context, ~0.71s latency, multimodal support. Plugsky's largest mainstream reasoning model.

Key points

  • OpenAI-compatible API: change the base URL and API key to switch.
  • Flat-rate plans with unlimited fair-use — see the pricing section.
  • Private deployment options: VPC, on-prem and air-gapped.

Explore more: all articles · free tools · docs · what is Plugsky.

Get started in minutes

OpenAI-compatible API with 30+ models, free trial, and a 99.9% uptime SLA. No code changes required.

Start Free → Read the docs