Languages

How do you build Turkish AI apps with Plugsky?

Turkish runs on the same OpenAI-compatible endpoint: set base_url to https://api.plugsky.com/v1 and send UTF-8 text to /v1/chat/completions. Agglutinative suffix chains fragment into many tokens, and Turkish i casing breaks naive search code. Measure real text, apply Turkish case rules, and use plugsky-embed-multilingual for mixed Turkish and English retrieval.

Key facts

Turkish textLatin script, left-to-right; dotted and dotless i affect casing
TokenisationAgglutinated suffix chains split into many subwords — measure with the Plugsky token calculator
Registersiz and sen change verb forms throughout — keep one per surface
API compatibilityOpenAI-compatible POST https://api.plugsky.com/v1/chat/completions — chat, streaming, JSON mode and function calling
Models30+ models behind one endpoint, from free tiers to frontier reasoning
Free tierFree plan with plugsky-micro and plugsky-lite, no card required
DeploymentPlugsky cloud, your VPC, on-prem and air-gapped options
Product statusChat, streaming, JSON mode, function calling, embeddings, RAG and agents live; audio, images, batch and fine-tuning coming soon

TL;DR

  • Turkish needs only a base_url change on the standard endpoint.
  • Suffix chains fragment heavily — measure tokens on real text.
  • Apply Turkish-aware casing (İ/i and I/ı) in search.
  • Pin siz or sen formality per surface.
  • plugsky-embed-multilingual covers Turkish and English retrieval.

How it works, step by step

  1. Create a Plugsky API key on the free plan (no card) and set base_url to https://api.plugsky.com/v1.
  2. Send a small set of real Turkish prompts to /v1/chat/completions and compare output across two or three models.
  3. Count tokens for those prompts with the Plugsky token calculator and set chunk sizes that fit your model context.
  4. Normalise text before indexing or prompting: apply Turkish case-folding rules and standardise loanword spelling.
  5. For RAG, embed with plugsky-embed-multilingual and test cross-language queries alongside Turkish-only queries.
  6. Score candidate models on a Turkish gold set with native-speaker review, then cut production traffic over.
1Create a PlugskyAPI key on the freeplan (no card) and2Send a small set ofreal Turkishprompts to3Count tokens forthose prompts withthe Plugsky token4Normalise textbefore indexing orprompting: apply5For RAG, embed withplugsky-embed-multilingualand test6Score candidatemodels on a Turkishgold set with

Try it yourself

Open the token calculator →

How Plugsky handles Turkish text

Turkish uses the Latin alphabet and reads left to right, with dotted and dotless i (İ/i and I/ı) that change case differently from English.

Formal and informal address (siz versus sen) and the spelling of modern loanwords are the main register choices.

Evaluate on formal and informal Turkish samples; check suffix chains, vowel harmony and correct dotted or dotless i usage.

Tokenisation and cost in Turkish

Turkish agglutination stacks suffixes into long words such as evlerinizdekiler that tokenisers split heavily, so token counts run high relative to word counts, and case folding with Turkish i rules matters for search.

  • Apply Turkish-aware case folding (İ to i, I to ı) in search code.
  • Keep suffix chains intact in display text.
  • Test formal and informal address separately.
  • Count tokens on real user text with long verbs and suffixes.

Turkish retrieval and RAG

Turkish retrieval improves with stemming and Turkish-aware casing; keep original forms for display and embeddings, and add stems for keyword search.

  • Use plugsky-embed-multilingual for Turkish and English corpora.
  • Apply Turkish case-folding rules in keyword search.
  • Consider stemming for agglutinated forms.
  • Evaluate formal and informal query sets.

Code example: a Turkish request

Point your existing OpenAI client at https://api.plugsky.com/v1 and pass Turkish text in the content field — no language flag and no separate endpoint. Streaming, JSON mode and function calling keep the same request shapes.

client = OpenAI(base_url="https://api.plugsky.com/v1", api_key=os.environ["PLUGSKY_API_KEY"])

client.chat.completions.create(model="plugsky-pro", messages=[{"role": "user", "content": "Bu sözleşmeyi Türkçe olarak üç maddede özetle."}])

Start on the free plan with plugsky-micro and plugsky-lite, then compare paid models on a Turkish gold set before cutover. See the docs for request details.

For production, log the model name and your normalisation settings with each request, and re-run the Turkish gold set whenever either changes — language quality regressions usually come from prompt or preprocessing drift, not from the model alone.

Honest comparison

CapabilityPlugskyTurkish workflow todayBuilding in-house
API compatibilityOpenAI-compatible — change base_url and model nameVaries by provider and SDKFull rewrite
Turkish text handlingTurkish-aware casing with suffix-aware retrievalDepends on provider tokeniser and prompt hygieneYou build normalisation, segmentation and evals
Token budgetFixed tokeniser per model; measure with the Plugsky token calculator and chunk to fitVaries by provider and modelYou host and tune each tokeniser
Multilingual retrievalplugsky-embed-multilingual available for cross-language RAGOften needs a separate embedding vendorYou serve and maintain embeddings
SovereigntyCloud, VPC, on-prem and air-gapped with residency optionsUsually US/EU public endpointsYou own the full stack

Frequently asked questions

Can Plugsky handle Turkish text?

Yes. The API accepts UTF-8 Turkish input on the OpenAI-compatible chat endpoint; output quality depends on the model, so compare two or three on your own prompts before choosing.

How do I estimate token usage for Turkish?

Turkish agglutination stacks suffixes into long words such as evlerinizdekiler that tokenisers split heavily, so token counts run high relative to word counts, and case folding with Turkish i rules matters for search. Use the token calculator at /tools/llm-token-calculator before sizing context windows or chunk lengths.

Siz or sen?

Use siz for business and formal products, sen for consumer apps. Keep the choice consistent across prompts and examples.

Why does Turkish casing break search?

The dotted and dotless i follow Turkish rules: İ lowercases to i and I lowercases to ı. Standard case folding gets this wrong, so apply Turkish rules.

Is there a multilingual embedding model?

Yes — plugsky-embed-multilingual is part of the 30+ model catalogue and is built for cross-language retrieval. Keep one embedding model per vector collection.

Can I keep data in my region?

Plugsky supports cloud, VPC, on-prem and air-gapped deployment with data-residency options; confirm your requirements with the docs and the enterprise team.

How do I migrate an existing app?

Change base_url to https://api.plugsky.com/v1 and map the model name. Streaming, JSON mode, function calling and embeddings keep the same request shapes.

Is there a free plan?

Yes — the free plan includes two free models, plugsky-micro and plugsky-lite, with no card. A 14-day full-access trial unlocks the paid catalogue.