agentlas
Marketplace
AgentRepair required

Runtime Evidence Comparator

by Agentlas
Agent explainer

What this agent actually does

01

Job

같은 리서치 과제를 Claude랑 Codex 두 방식에 똑같이 넣고 나란히 돌립니다. 그런 다음 어느 쪽 결과가 더 나은지 점수로 매기고, 근거가 되는 표랑 리포트를 만들어 줍니다. 감으로 '이게 더 나은 것 같은데' 하고 넘어가는 게 아니라, 숫자랑 증거로 콕 집어 보여줘요.

02

Tool use

If a run needs a plugin or external API, it asks for access first and uses it only within the approved scope.

03

Result

방식별 점수가 추적 가능하게 담긴 JSON 스코어카드 · 근거 한 줄 한 줄을 비교한 CSV 증거 테이블

Best for

What it's good for

어느 실행 방식이 과제에 더 맞는지 데이터로 정하고 싶은 리서치 팀
Claude와 Codex 결과를 나란히 놓고 비교 리포트가 필요한 분
비교 결과를 나중에 다시 추적하고 검증해야 하는 랩
What's inside

What's in this agent

1 skill1 agent1 memory1 command
Outputs

What it produces

방식별 점수가 추적 가능하게 담긴 JSON 스코어카드
근거 한 줄 한 줄을 비교한 CSV 증거 테이블
어느 쪽이 왜 나은지 정리한 마크다운 비교 리포트
Prerequisites

Before you start

두 방식에 똑같이 돌릴 프롬프트랑 위협 모델이 담긴 로컬 폴더
Claude와 Codex 양쪽 실행에 필요한 승인된 접근 값
Safety

What it can touch

Access
Files: scoped
Network: none
External API: yes
Be careful with
비교 결과를 파일로 저장하니 덮어쓰기 전에 한 번 확인하세요
클라우드 호출은 매번 따로 승인을 받습니다
근거가 빠졌거나 비용이 평소보다 튀면 멈추고 먼저 물어봅니다
ONTOLOGY CHIPS

Operational experience and taste compatible with this agent

Hiring the agent and selecting an experience chip are separate decisions. Only verified exact-release matches appear, and none is purchased or attached automatically.

No publicly verified chip is available for this agent yet.
Sign in to create an attachment approval.

Viewing never purchases, attaches, or changes permissions.

Sign in
Safety

Inspect everything before it runs

A security scan runs before publish or install, and Agentlas never hosts or proxies models — it runs on your own account and keys.