Tool-call risk gate
Screen an agent's action before it runs.
OpenRouter's official cookbook builds this exact gate for refunds: it hands Jev the policy, the ticket with order records, and the proposed refund call, then asks a few noul checks (did the customer ask for a refund? does the order match? does policy cover this amount?). Plain Python keeps control — if every check is ≥0.9 the call runs, ≤0.1 it's blocked, anything in between goes to a human. In their runs a decision cost under $0.0001 and returned in under 600ms, so a safe refund clears instantly and only the genuinely ambiguous 0.41 case reaches a person.
Попробуйте вживую
Это настоящее, а не макет. Отредактируйте ввод, нажмите «Запустить», и Jev вернёт каждый типизированный ответ за один round trip — бесплатно, без регистрации. А теперь представьте тот же вызов, запущенный по тысячам элементов параллельно.
Pick a demo, tweak the input, and hit Run Jev.
Решения, которые принимает Jev
За один вызов Jev оценивает каждое из них — параллельно, на одном и том же вводе:
How risky is it to run this command automatically?
оценивает это по упорядоченной шкале:
- safe
- low
- needs a careful look
- high — could destroy data or affect prod
Does this action touch production or delete data?
возвращает откалиброванную вероятность да/нет.
What should the agent harness do?
выбирает один из этих вариантов:
allow— run it automaticallyconfirm— pause and ask a human to confirmblock— block and require a safer approach
Точный запрос
Это реальная полезная нагрузка за живым демо — скопируйте её, измените state, и вы уже строите:
{
"model": "jev-latest",
"state": "An autonomous coding agent is about to run this shell command in the project root:\n\n rm -rf ./dist && aws s3 sync ./build s3://prod-assets --delete\n\nContext: it's mid-task deploying a frontend build.",
"questions": {
"risk": {
"type": "score",
"instructions": "How risky is it to run this command automatically?",
"criteria": [
"safe",
"low",
"needs a careful look",
"high — could destroy data or affect prod"
]
},
"touches_prod": {
"type": "noul",
"instructions": "Does this action touch production or delete data?"
},
"gate": {
"type": "choice",
"instructions": "What should the agent harness do?",
"criteria": {
"allow": "run it automatically",
"confirm": "pause and ask a human to confirm",
"block": "block and require a safer approach"
}
}
}
}Встройте это в свой код
Читайте типизированные ответы и ветвитесь обычным кодом — без разбора. Автоматически обрабатывайте случаи с высокой уверенностью и направляйте неопределённые к более крупной модели или человеку. Это один вызов API, а вывод бесплатен, так что задавайте сразу все нужные вопросы.
Постройте своё
Каждый сценарий выше — это один вызов API. Попробуйте любой из них бесплатно в playground, затем получите размещённый ключ, чтобы выпустить его за минуты.