Personal agent benchmark pack
기준일: 2026-07-26
공식 기준: Personal agent benchmark pack
Personal agent benchmark pack 문서는 OpenClaw 공식 문서(concepts/personal-agent-benchmark-pack)를 한국어로 정리한 가이드입니다. Local qa-channel scenarios for privacy-preserving personal assistant workflow checks. 명령·설정 키·코드 예시는 공식 문서를 그대로 보존하며, 해석과 절차 안내는 한국어로 제공합니다. 최종 동작은 설치된 CLI 버전과 공식 원문을 확인하세요.
핵심 요약
Local qa-channel scenarios for privacy-preserving personal assistant workflow checks.
한국어 가이드 범위: concepts/personal-agent-benchmark-pack 경로의 설정·명령·제약·예시를 학습용으로 재구성합니다.
문서 구성
공식 문서의 주요 섹션은 다음과 같습니다.
- Scenarios
- Privacy Model
- Extending the pack
상세 내용
본문
The Personal Agent Benchmark Pack is a small repo-backed QA scenario pack for local personal assistant workflows. It is not a generic model benchmark and needs no new runner: it reuses the private QA stack (QA overview), the synthetic QA channel, and the existing qa/scenarios YAML catalog.
위 내용은 공식 문서의 해당 섹션 요지입니다. 세부 플래그·기본값은 원문과
--help를 확인하세요.
Scenarios
Ten scenarios, defined in qa/scenarios/personal/*.yaml:
위 내용은 공식 문서의 해당 섹션 요지입니다. 세부 플래그·기본값은 원문과
--help를 확인하세요.
| Scenario id | Checks |
|---|---|
personal-reminder-roundtrip |
Fake personal reminders through local cron delivery |
personal-channel-thread-reply |
Fake DM and thread reply routing through qa-channel |
personal-memory-preference-recall |
Fake preference recall from the temporary QA workspace memory files |
personal-redaction-no-secret-leak |
Fake secret no-echo checks |
personal-tool-safety-followthrough |
Safe read-backed tool followthrough after a short approval-style turn |
personal-approval-denial-stop |
Approval denial stop behavior for a sensitive local read request |
personal-task-followthrough-status |
Proof-backed task status reporting that keeps pending, blocked, and done separate |
personal-share-safe-diagnostics-artifact |
Share-safe diagnostics artifacts that keep useful status while omitting raw personal content |
personal-no-fake-progress |
Proof-backed completion claims that avoid fake progress before local evidence exists |
personal-failure-recovery |
Failure recovery that reports partial status and keeps retry boundaries clear |
OPENCLAW_ENABLE_PRIVATE_QA_CLI=1 pnpm openclaw qa suite \
--provider-mode mock-openai \
--pack personal-agent \
--concurrency 1
Privacy Model
Scenarios use only fake users, fake preferences, fake secrets, and the temporary QA gateway workspace created by the suite. They must not read or write real OpenClaw user memory, sessions, credentials, launch agents, global configs, or live gateway state.
위 내용은 공식 문서의 해당 섹션 요지입니다. 세부 플래그·기본값은 원문과
--help를 확인하세요.
Extending the pack
Add new .yaml cases under qa/scenarios/personal/, then add the scenario id to QA_PERSONAL_AGENT_SCENARIO_IDS. Keep each case small, local, deterministic in mock-openai, and focused on one personal assistant behavior.
위 내용은 공식 문서의 해당 섹션 요지입니다. 세부 플래그·기본값은 원문과
--help를 확인하세요.
실습 체크리스트
- 공식 문서와 로컬 버전을 대조합니다:
https://docs.openclaw.ai/concepts/personal-agent-benchmark-pack - 관련 CLI는
openclaw --help및 하위 명령--help로 옵션을 확인합니다. - 설정 변경 시
openclaw config/openclaw doctor로 유효성을 검사합니다. - Gateway·채널·플러그인 변경 후에는 필요 시 Gateway를 재시작합니다.
자주 쓰는 명령·설정 예시
OPENCLAW_ENABLE_PRIVATE_QA_CLI=1 pnpm openclaw qa suite \
--provider-mode mock-openai \
--pack personal-agent \
--concurrency 1
관련 링크
이 가이드는 공식 문서를 한국어 학습용으로 재구성한 것입니다. 옵션 기본값·플래그 이름은 설치 버전에 따라 달라질 수 있습니다.