Model Reference

plugsky-max — heavy reasoning, zero drama

plugsky-max is Plugsky's heavy reasoning model: 256K context, ~0.71s median latency, multimodal input, and strong tool calling.

Summary

ModelIdentityContextLatency*Best for
plugsky-maxQwen 3.5-class (MoE)256K tokens~0.71s medianDeep reasoning, multimodal, long documents, agents

*Median measured latency on the live API at https://api.plugsky.com/v1.

Specs

PropertyValue
Model familyQwen 3.5-class (MoE, 397B)
Context window256K tokens
Measured latency~0.71s median on the live API
CapabilitiesReasoning, multimodal, tool calling
Available onBuilder plan and above

When to use plugsky-max

  • Deep reasoning — the hardest prompts your product sees.
  • Multimodal — images plus text in one call.
  • Long-document analysis — 256K context for contracts, papers, logs.

FAQ

Does plugsky-max accept images?

Yes — multimodal input is supported. Dedicated vision models are also available (e.g. plugsky-vision-fast).

Which plan do I need?

Builder and above include plugsky-max access.

Get started in minutes

OpenAI-compatible API with 30+ models, free trial, and a 99.9% uptime SLA. No code changes required.

Start free trial → Read the docs