Playground
Each run uses one credit first; if credits run out, it uses the actual input tokens instead. This is the same for all five models.
FIVE MODELS · ONE ACCOUNT
Laya AI started with two compact Laya models for short text. It now also runs Jev, Mercury Decide and Jev-Omni, so longer documents and media get typed answers from the same account, wallet and API key.
START FROM YOUR INPUT
The question types are the same everywhere. What changes is how much each model can read and what kind of input it accepts.
One message, review or comment in English
A short message in Chinese, Spanish, Arabic or another language
A full ticket thread, contract clause or JSON record
Mid-length text where you want a second, independent model
A photo, voice note or short clip
SIDE BY SIDE
| Model | Made by | Input | Size limit | Questions | API | Now |
|---|---|---|---|---|---|---|
Laya Englishlaya-english | Convai Innovations | Short English text | 512 tokens per question | Yes/no, choice, score | Yes | … |
Laya Multilinguallaya-multilingual | Convai Innovations | Short text, 100+ languages | 1,024 tokens per question | Yes/no, choice, score | Yes | … |
Jevjev-latest | TypeSafe | Long text and JSON | 64k tokens per request | Yes/no, choice, score | Yes, also jev-1.13.0 and jev-preview | … |
Mercury Decidemercury-decide | Inception | Text and JSON | 32,768 tokens per request | Yes/no, choice, score | Yes | … |
Jev-Omnijev-omni | Independent, open weights on Gemma 4 12B | Image, audio, video or text | 8,192 tokens; 6 MiB file | One choice per request | Playground only | … |
Oversized input is rejected before any charge on every model; nothing is silently cut. If a model is offline, requests to it fail before billing and are never answered by another model.
SWITCHING
Every text model uses the same request: a state, a model and typed questions. Start on Laya English, and when a case turns out to need more context, change model to jev-latest or mercury-decide without touching your questions or parsing code. The Decisions-compatible endpoint also accepts ~typesafe/jev-latest and inception/mercury-decide.
{
"model": "jev-latest",
"state": "…the full support thread…",
"questions": {
"escalate": { "type": "noul", "instructions": "Escalate to engineering now?" }
}
}BILLING
Each run uses one credit first; if credits run out, it uses the actual input tokens instead. This is the same for all five models.
Requests use paid input tokens first, then one credit for the whole request when tokens are insufficient. Output tokens are free. Failed requests are refunded.
See plans →