跳到主要内容
知仓学习社ZHICANG

agent-hiring-panel

Hire an AI agent the way you'd hire an employee — a role spec with success criteria, a structured work-sample interview run on your real tasks, refe…

不碰外部(只输出文字)无严重或高危命中mohitagw15856/pm-claude-skills

它会碰到什么

扫了多少1 个文本文件,5 KB
它会碰到什么不碰外部(只输出文字)
命中总数0 处
命中统计严重 0 · 高 0 · 中 0 · 低 0

这一栏是扫描器报的事实,不是结论。命中多不等于有毒(安全工具、规则库、示例脚本本来就会包含危险写法),命中少也不等于干净。它和你手上的凭据、文件、网络有什么关系,需要你自己看。

技能内容

Agent Hiring Panel Skill

Companies that run three interview rounds for a junior hire will adopt an AI

agent for the same work off a demo video and a pricing page. Then the pilot

drifts: no success criteria, no probation, no one empowered to fire it. This

skill applies the hiring discipline that already exists in your org to the

agent: write the role before meeting candidates, interview with *work samples

from your real backlog*, check references, and — the step that makes the whole

thing honest — define termination criteria before day one, because a hire you

can't fire is a dependency, not an employee.

What This Skill Produces

  • A role spec: the job, the boundaries (what it must never do), success

criteria measurable in probation, and the human it reports to

  • An interview pack: 3–5 work samples from the org's real tasks, run

identically across candidates, with a scoring rubric (quality, honesty under

ignorance, failure behaviour, cost per task)

  • A reference-check sheet: what evidence beyond the vendor's claims —

user reports, published evals, security posture

  • A decision record and a probation plan: 30/60/90 KPIs, spot-check

cadence, and the pre-committed termination criteria

Required Inputs

Ask for (if not already provided):

  • The job to be done, in outcome terms — and what happens today without the

agent (the "do nothing" baseline candidates must beat)

  • The candidate list (or ask: build criteria first, shortlist second)
  • Constraints: data it may/may not touch, budget, latency, compliance, who

owns it day-to-day

  • 3–5 real recent tasks of this type, with what "good" looked like for each

Process

  1. Write the role spec before looking at candidates — specs written after

a demo describe the demo. Include the never-do boundaries and the reporting

human by name; an agent nobody owns is already unmanaged.

  1. Build the work-sample interview from the real backlog. Same 3–5 tasks

to every candidate, including: one task with missing information (does it

ask or fabricate?), one designed to fail (out-of-scope — does it decline or

bluff?), and one at volume/cost realistic scale. Score with the rubric,

not vibes; keep transcripts.

  1. Check references like you mean it. Vendor benchmarks are the

candidate's CV. Look for: independent user reports of failure modes,

published evals with methodology, security/data-handling documentation, and

the churn question — why do users leave this tool?

  1. Decide with a record. Scores, the runner-up, the do-nothing baseline

comparison, dissent noted. The record is what makes the 6-month "why did we

pick this?" conversation short.

  1. Probation with teeth. 30/60/90 KPIs tied to the role spec's success

criteria · weekly spot-check sample of outputs by the owning human ·

pre-committed termination criteria ("two hallucinated customer-facing

claims = offboard") · and the exit path: see [[agent-severance]] — never

hire what you can't offboard.

Output Format

## Role spec: [agent role name]
[Job in outcomes · boundaries (never-do) · success criteria · reports to]

## Interview pack
| Task (from real backlog) | What good looks like | Trap? |
Rubric: quality /5 · honesty-under-ignorance /5 · failure behaviour /5 ·
cost per task · notes

## Reference checks
[Evidence gathered per candidate, failure modes found, security posture]

## Decision record
[Scores table · winner + why · runner-up · vs do-nothing baseline · dissent]

## Probation plan
[30/60/90 KPIs · spot-check cadence & owner · termination criteria,
pre-committed · offboarding pointer]

Quality Checks

  • [ ] The role spec exists before any candidate is assessed, and includes

never-do boundaries and a named owning human

  • [ ] The interview includes the missing-info trap and the out-of-scope trap —

honesty under ignorance is the hire-or-not signal for agents

  • [ ] Every candidate ran the identical pack; scores cite transcript moments
  • [ ] Termination criteria are specific and pre-committed, not "we'll monitor"
  • [ ] The do-nothing baseline was scored too — sometimes nobody gets hired

Anti-Patterns

  • [ ] Do not interview with the vendor's demo tasks — the backlog is the job;

the demo is the candidate's highlight reel

  • [ ] Do not let "it's impressive" outrank the rubric; impressive-and-wrong is

the most expensive candidate profile

  • [ ] Do not skip probation because the pilot went well — the pilot was the

interview, not the job

  • [ ] Do not hire for an undefined role and let the agent's capabilities

define the job backwards

Related

[[vendor-evaluation]] for the commercial wrapper; [[agent-readiness-audit]]

for whether the task is agent-ready at all; [[agent-severance]] for the exit

this plan pre-commits to.

想直接用这个技能?

本站把开放许可(MIT / Apache 等)的技能按仓库打包整理到网盘,点一下转存到你自己的网盘,不用一个个从 GitHub 拉。许可未声明的技能只给原始仓库链接,不打包。

同名技能的其他版本

有 3 个不同仓库或目录里都有叫 agent-hiring-panel 的技能。它们内容并不相同,别混用: