No Python, no model on your server
Running Laya means Python, PyTorch and around 800 MB of model files. Most web hosting can't do that. With Laya API it's one HTTPS request from PHP, JavaScript or anything else.
Free open beta, Jev-compatible request format
Ask your software a question in plain words. Laya API sends back JSON your code can branch on: a yes/no probability, a picked option, or a score. Every answer comes with a confidence value. No free text to parse.
Three answer types, one request format.
Send a text and a question with the answers you allow. You get back the pick, how likely every option was, and how sure the model is.
POST /api/v1/systemone?wait=10
{
"state": "I was charged twice for March.",
"model": "laya-latest",
"questions": {
"team": {
"type": "choice",
"instructions": "Which team handles this?",
"criteria": {
"billing": "Payments and refunds",
"support": "Bugs and errors",
"sales": "Plans and pricing"
}
}
}
}
{
"status": "completed",
"result": {
"answers": {
"team": {
"type": "choice",
"choice": "billing",
"probabilities": {
"billing": 0.91,
"support": 0.06,
"sales": 0.03
},
"confidence": 0.93
}
},
"usage": { "input_tokens": 41 }
}
}
Laya is an open model. You could run it yourself, if your server can handle it.
Running Laya means Python, PyTorch and around 800 MB of model files. Most web hosting can't do that. With Laya API it's one HTTPS request from PHP, JavaScript or anything else.
Jev is still in early access. Laya API is open to everyone in the free beta, with the same request format. Moving between the two is a change of URL and model name.
We run Laya API in Germany, under the GDPR. Need a data processing agreement (Art. 28 GDPR)? We'll sign one with you.
Requests run as jobs. That keeps your app responsive: wait for the answer in the same call
with ?wait=10, or fire and forget
and fetch the job later. Every job also shows up in your dashboard.
Same request body, same answers object. The only
difference: the answer arrives inside a job object, and
?wait=10 returns it in the same call.
Create an account, generate a key and start sending requests. No credit card. Beta keys are limited to 60 requests per minute.
Output tokens are free. An answer is a handful of numbers, not paragraphs of text, so you’ll only ever be billed for what you send. Pricing will be published before the beta ends.