Languages

How do you build Polish AI apps with Plugsky?

Polish is Latin-script with diacritics and heavy inflection: nouns, adjectives and verbs change form across seven cases and multiple genders. Plugsky serves Polish on the standard OpenAI-compatible endpoint — set base_url to https://api.plugsky.com/v1 and send UTF-8 text to /v1/chat/completions. Pair embeddings with stemming for search, pin Pan or Pani for formality, and use plugsky-embed-multilingual for Polish and English retrieval.

Key facts

Polish scriptLatin script with diacritics ą, ć, ę, ł, ń, ó, ś, ź and ż
InflectionSeven grammatical cases and gender produce many surface forms per word
RegisterFormal Pan and Pani versus informal ty — keep one per surface
TokenisationInflected endings and consonant clusters split into subwords — measure with the token calculator
API compatibilityOpenAI-compatible POST https://api.plugsky.com/v1/chat/completions — chat, streaming, JSON mode and function calling
Models30+ models behind one endpoint, from free tiers to frontier reasoning
Free tierFree plan with plugsky-micro and plugsky-lite, no card required
Product statusChat, streaming, JSON mode, function calling, embeddings, RAG and agents live; audio, images, moderation, files, batch, fine-tuning, assistants and responses coming soon

TL;DR

  • Polish runs on the standard endpoint with one base_url change.
  • Inflection multiplies word forms — pair embeddings with stemming for search.
  • Add diacritic-insensitive keys for mobile users.
  • Pin Pan or Pani formality per surface and keep it consistent.
  • plugsky-embed-multilingual covers Polish and English retrieval.

How it works, step by step

  1. Create a free Plugsky key and set base_url to https://api.plugsky.com/v1.
  2. Send real Polish text to /v1/chat/completions and compare two or three models.
  3. Measure tokens with the token calculator on inflected, diacritic-heavy text.
  4. Add stemming for keyword search so inflected forms match, alongside embeddings.
  5. Choose Pan or Pani versus ty per surface, and state it in the system prompt.
  6. For RAG, embed with plugsky-embed-multilingual and review a Polish gold set with native speakers.
1Create a freePlugsky key and setbase_url to2Send real Polishtext to/v1/chat/completions3Measure tokens withthe tokencalculator on4Add stemming forkeyword search soinflected forms5Choose Pan or Paniversus ty persurface, and state6For RAG, embed withplugsky-embed-multilingualand review a Polish

Try it yourself

Open the RAG sandbox →

How Plugsky handles Polish text

Polish needs no special endpoint: UTF-8 text with diacritics goes in and comes back. The language-specific work is inflection and formality. Polish marks seven grammatical cases across nouns, adjectives and pronouns, and gender affects endings too, so a single word appears in many forms depending on its role in the sentence.

Formality is explicit rather than implied: Pan and Pani with third-person verb forms for formal address, ty for informal. Mixing them inside one conversation is the failure Polish readers notice first, so fix the choice per surface.

Tokenisation and cost in Polish

Polish text carries diacritics and long consonant clusters, and inflected endings can split into subwords, so counts differ from English in both directions. Formal and legal writing compounds these effects with longer sentence structures.

  • Measure on inflected text, not on dictionary forms.
  • Keep diacritics for display and embedding; fold them for search keys only.
  • Watch hyphenated and abbreviated forms in official writing.
  • Re-measure whenever models change.

Polish retrieval and RAG

Inflection is the classic Polish retrieval problem: the same concept appears in several case forms, and keyword search treats them as different strings. One multilingual collection built with plugsky-embed-multilingual handles that better, and stemming covers the keyword path.

  • Add stemming or lemmatisation for keyword search on top of embeddings.
  • Add diacritic-insensitive keys for mobile users who skip accents.
  • Test inflected query forms against base-form documents.
  • Keep one embedding model per collection.

Code example: a Polish request

Point your OpenAI client at https://api.plugsky.com/v1; only the content and system prompt change.

client = OpenAI(base_url="https://api.plugsky.com/v1", api_key=os.environ["PLUGSKY_API_KEY"])

client.chat.completions.create(model="plugsky-pro", messages=[{"role": "system", "content": "Odpowiadaj po polsku, w formalnym stylu z formą Pan lub Pani"}, {"role": "user", "content": "Streść tę umowę w trzech punktach"}])

Start on the free plan with plugsky-micro and plugsky-lite, then compare paid models on a Polish gold set. See the docs for the API reference.

Honest comparison

CapabilityPlugskyPolish apps todayBuilding in-house
API compatibilityOpenAI-compatible — change base_url and model nameVaries by provider and SDKFull rewrite
Polish text handlingUTF-8 with diacritics, inflection-aware search and Pan/Pani promptsDepends on provider tokeniser and preprocessingYou build normalisation, stemming and evals
Token budgetFixed tokeniser per model; measure inflected text and chunk to fitVaries by provider and modelYou host and tune each tokeniser
Multilingual retrievalplugsky-embed-multilingual for Polish and English RAGOften needs a separate embedding vendorYou serve and maintain embeddings
SovereigntyCloud, VPC, on-prem and air-gapped with residency optionsUsually US/EU public endpointsYou own the full stack

Frequently asked questions

Can Plugsky handle Polish text?

Yes. The API accepts UTF-8 Polish input with diacritics on the OpenAI-compatible chat endpoint. Compare two or three models on your own prompts before choosing.

How does Polish inflection affect search?

The same word appears in many case and gender forms, so keyword search misses inflected queries. Use embeddings and add stemming or lemmatisation for keyword keys.

Should prompts use Pan or Pani?

Formal service, B2B and public-sector flows use Pan or Pani with third-person verb forms; consumer and community products often use ty. State the choice in the system prompt.

Do I need diacritic-insensitive search?

Many users type without Polish diacritics on mobile. Add an accent-insensitive path while keeping diacritics for display and embeddings.

Is there a multilingual embedding model?

Yes — plugsky-embed-multilingual is part of the 30+ model catalogue and supports Polish and English retrieval from one collection.

Can I keep data in the EU?

Plugsky supports cloud, VPC, on-prem and air-gapped deployment with residency options; confirm specifics with the docs and the enterprise team.

Is there a free plan?

Yes — plugsky-micro and plugsky-lite are free with no card, and a 14-day full-access trial covers the paid catalogue.