epithre-prme

Premium tier for long-context workloads. Use when prompts exceed 32K tokens or when you need the strongest reasoning available.

Capabilities

Capability Notes
Tier Premium
Context window 180,000 tokens (~135K English words / ~140K Indonesian)
Max output 16,384 tokens
Modalities Text only (no vision)
Tool use Yes
Extended thinking Yes; recommended for hard problems
Structured output Yes; json_object + json_schema
Prompt caching Yes; especially valuable here given long stable contexts

When to use

When NOT to use

Pricing

Performance characteristics

Strongest in our lineup at:

Latency

Plan for higher latency vs omni. Use streaming UX or background-job pattern.

Caveats

See also