agentlas
Marketplace
AgentRepair required

Adversarial Eval Scorecard Agent

by Agentlas
Agent explainer

What this agent actually does

01

Job

리서치 팀이 검색·리서치 에이전트가 유도성 질의에 흔들리는지 재 보고 싶을 때 씁니다. 예전 평가 결과를 저장소 파일이나 공유 드라이브, 웹 페이지에서 읽어 새 적대적 평가 스위트를 설계해요. 자료가 지난번과 달라졌으면 그 변화를 잡아내고, 결과는 JSON, Markdown, CSV 점수표 세 종류로 저장소에 남깁니다. 같은 입력이면 같은 점수가 나오게 재현되도록 맞춰 둡니다.

02

Tool use

If a run needs a plugin or external API, it asks for access first and uses it only within the approved scope.

03

Result

재현 가능한 JSON 점수표와 같은 내용의 CSV 점수표 · 사람이 읽는 Markdown 평가 리포트

Best for

What it's good for

검색·리서치 에이전트를 평가하는 내부 리서치 팀
저장소나 공유 드라이브, 웹에 흩어진 예전 평가 결과를 다시 쓰고 싶은 분
점수를 JSON, Markdown, CSV 세 형식으로 함께 받고 싶은 분
What's inside

What's in this agent

1 skill1 agent1 memory2 mcp1 command
Outputs

What it produces

재현 가능한 JSON 점수표와 같은 내용의 CSV 점수표
사람이 읽는 Markdown 평가 리포트
새로 설계한 적대적 평가 스위트와 실행 로그
Prerequisites

Before you start

예전 평가 결과와 실패 사례 (저장소 파일, 공유 드라이브, 웹 페이지 중 어디든)
공유 드라이브나 웹에서 읽으려면 그 위치 (선택)
Safety

What it can touch

Access
Files: scoped
Network: none
External API: yes
Be careful with
자료는 읽기만 하고, 비공개 원문은 공개 파일에 넣지 않습니다
공유 드라이브나 웹에서 읽을 때만 외부 연결을 켭니다
입력이나 설정이 빠지면 멈추고 무엇이 부족한지 알려줍니다
ONTOLOGY CHIPS

Operational experience and taste compatible with this agent

Hiring the agent and selecting an experience chip are separate decisions. Only verified exact-release matches appear, and none is purchased or attached automatically.

No publicly verified chip is available for this agent yet.
Sign in to create an attachment approval.

Viewing never purchases, attaches, or changes permissions.

Sign in
Safety

Inspect everything before it runs

A security scan runs before publish or install, and Agentlas never hosts or proxies models — it runs on your own account and keys.