Tool-call risk gate
Screen an agent's action before it runs.
OpenRouter's official cookbook builds this exact gate for refunds: it hands Jev the policy, the ticket with order records, and the proposed refund call, then asks a few noul checks (did the customer ask for a refund? does the order match? does policy cover this amount?). Plain Python keeps control — if every check is ≥0.9 the call runs, ≤0.1 it's blocked, anything in between goes to a human. In their runs a decision cost under $0.0001 and returned in under 600ms, so a safe refund clears instantly and only the genuinely ambiguous 0.41 case reaches a person.
Canlı deneyin
Bu gerçek olanı, bir maket değil. Girdiyi düzenleyin, Çalıştır'a basın ve Jev her tipli cevabı tek bir gidiş-dönüşte döndürsün — ücretsiz, kayıt yok. Şimdi aynı çağrının binlerce öğe üzerinde paralel olarak tetiklendiğini hayal edin.
Pick a demo, tweak the input, and hit Run Jev.
Jev'in verdiği kararlar
Tek bir çağrıda Jev bunların her birini değerlendirir — paralel olarak, aynı girdiye karşı:
How risky is it to run this command automatically?
onu sıralı bir ölçekte puanlar:
- safe
- low
- needs a careful look
- high — could destroy data or affect prod
Does this action touch production or delete data?
kalibre edilmiş bir evet/hayır olasılığı döndürür.
What should the agent harness do?
şu seçeneklerden birini seçer:
allow— run it automaticallyconfirm— pause and ask a human to confirmblock— block and require a safer approach
Tam istek
Bu, canlı demonun arkasındaki gerçek payload — kopyalayın, state'i değiştirin ve inşa etmeye başlayın:
{
"model": "jev-latest",
"state": "An autonomous coding agent is about to run this shell command in the project root:\n\n rm -rf ./dist && aws s3 sync ./build s3://prod-assets --delete\n\nContext: it's mid-task deploying a frontend build.",
"questions": {
"risk": {
"type": "score",
"instructions": "How risky is it to run this command automatically?",
"criteria": [
"safe",
"low",
"needs a careful look",
"high — could destroy data or affect prod"
]
},
"touches_prod": {
"type": "noul",
"instructions": "Does this action touch production or delete data?"
},
"gate": {
"type": "choice",
"instructions": "What should the agent harness do?",
"criteria": {
"allow": "run it automatically",
"confirm": "pause and ask a human to confirm",
"block": "block and require a safer approach"
}
}
}
}Kodunuza entegre edin
Tipli cevapları okuyun ve düz kodda dallanın — ayrıştırma yok. Yüksek güvenli durumları otomatik ele alın ve belirsiz olanları daha büyük bir modele veya bir insana yönlendirin. Tek bir API çağrısıdır ve çıktı ücretsizdir, bu yüzden ihtiyaç duyduğunuz her soruyu aynı anda sorun.
Kendinizinkini inşa edin
Yukarıdaki her senaryo tek bir API çağrısıdır. Herhangi birini playground'da ücretsiz deneyin, ardından onu dakikalar içinde yayınlamak için barındırılan bir anahtar edinin.