← 모든 활용 사례

Task-completion check

Before an agent stops, ask Jev whether the work is actually finished and the claims hold up.

Coding-agent use case · output & completion checksawesome-jev-by-typesafe

Agents love to declare victory early. "Diff verification" and "agent output checks" in the awesome-jev catalog cover the stop decision: before an agent ends its turn, verify that the task is genuinely done and that its claims are backed by evidence (tests actually passed, the change addresses the ask). Jev's statement-verification shape — a calibrated yes/no against the task, the diff and the test output — is exactly the primitive TypeSafe's own docs recommend for "checking whether a statement is true of a record before taking an action." Below a threshold, the agent keeps working instead of handing back a half-finished job.

라이브로 사용해 보기

이것은 목업이 아니라 실제입니다. 입력을 편집하고 실행을 누르면 Jev가 한 번의 왕복으로 모든 타입 지정 답변을 반환합니다 — 무료, 가입 불필요. 이제 같은 호출이 수천 개의 항목에 병렬로 실행되는 것을 상상해 보세요.

POST jevtypesafeai.com/api/v1/decide
state — 소프트웨어가 Jev에 전달하는 입력384c
questions — 돌려받고 싶은 타입 지정 결정들
noulcomplete
Does the change actually address the stated task?
→ 확률 0.0 … 1.0
noultests_back_it
Does the test output provide real evidence the fix works (not just unrelated passes)?
→ 확률 0.0 … 1.0
scoreconfidence
How confident should the agent be that it's safe to stop and hand back?
→ 0…3 · 4 단계
실제 API · 무료 · 가입 불필요
타입이 지정되고 보정된 출력이 여기에 나타납니다.
데모를 고르고, 입력을 수정한 뒤 눌러보세요 Jev 실행.
API 키 받기 →← 모든 활용 사례

Jev가 내리는 결정들

한 번의 호출로, Jev는 동일한 입력에 대해 이들 각각을 병렬로 평가합니다:

noulcomplete

Does the change actually address the stated task?

보정된 예/아니오 확률을 돌려줍니다.

noultests_back_it

Does the test output provide real evidence the fix works (not just unrelated passes)?

보정된 예/아니오 확률을 돌려줍니다.

scoreconfidence

How confident should the agent be that it's safe to stop and hand back?

정렬된 척도 위에서 점수를 매깁니다:

  1. keep working
  2. borderline
  3. likely done
  4. clearly done

정확한 요청

이것이 라이브 데모 뒤에 있는 실제 페이로드입니다 — 복사해서 state를 바꾸면, 바로 만들기 시작입니다:

{
  "model": "jev-latest",
  "state": "Task: \"Fix the bug where refunds over the order total are silently accepted.\"\n\nAgent's final summary: \"Added a balance check in process_refund so over-total refunds now raise RefundError.\"\n\nDiff: added `if amount > order.remaining_balance: raise RefundError(...)` before the gateway call.\n\nTest output: `test_refund_over_total PASSED · test_refund_partial PASSED · 2 passed, 0 failed`",
  "questions": {
    "complete": {
      "type": "noul",
      "instructions": "Does the change actually address the stated task?"
    },
    "tests_back_it": {
      "type": "noul",
      "instructions": "Does the test output provide real evidence the fix works (not just unrelated passes)?"
    },
    "confidence": {
      "type": "score",
      "instructions": "How confident should the agent be that it's safe to stop and hand back?",
      "criteria": [
        "keep working",
        "borderline",
        "likely done",
        "clearly done"
      ]
    }
  }
}

코드에 연결하기

타입이 지정된 답을 읽고 평범한 코드로 분기하세요 — 파싱 없이. 신뢰도 높은 경우는 자동 처리하고 불확실한 경우는 더 큰 모델이나 사람에게 라우팅하세요. API 호출 한 번이고 출력은 무료이므로, 필요한 모든 질문을 한 번에 물어보세요.

직접 만들어 보기

위의 모든 시나리오는 단 한 번의 API 호출입니다. 플레이그라운드에서 무엇이든 무료로 사용해 본 뒤, 호스티드 키를 받아 몇 분 만에 배포하세요.

이 데모 실행하기 ▶API 키 받기 →
Task-completion check — 라이브 데모가 있는 Jev 활용 사례 · Jev by TypeSafe AI