Every candidate uses an agent now. Kismet sends a real assessment in a private repo, captures how they drive the agent (prompts, edits, commits), and returns evals where every score cites the lines it rests on.
The Kismet extension reports what they do by hand: saves, terminal commands, and how they review each diff. Keystrokes never leave the editor.
savesterminal commandsdiff review
Claude Code in the terminal
Capture settings ship inside the repo. Any Claude Code session started there reports automatically, including every edit the agent proposed and whether they kept it.
committed hooksper-edit diffssecrets redacted
Company-paid AI keys. Give each candidate one key with a budget. It stops at the deadline, on submit, or once the budget is spent.
Timed & recorded
Timed when it matters. Recorded when you need it.
Each is a per-assessment setting, and candidates see which ones are on before they accept.
1:24left
Timed assessments
The clock starts when they press Start, not when they accept. When time's up, the work is collected automatically.
Recording screen and mic42:10
promptcommit
Screen recording
Whole screen and microphone from Start to submit, with a pre-flight check and alerts if the recording drops.
Q1Q2Q3
Recorded follow-up
After they submit, they reopen their own code and walk you through it, one question at a time.
Evaluation
Every score cites its evidence.
Three evals run on every submission. Each claim links to the lines, prompt or commit it came from. A score with no evidence doesn't ship.
41export async function listOrders(cursor?: string) {
42 if (cursor && !isCursor(cursor)) throw new BadCursor();
One copilot across your whole workspace. It reads your roles, candidates, sessions and company context, and cites every fact it states. When you ask it to act, it proposes the change and waits for you.
@mention candidates, roles and assessments
Nothing changes until you confirm
Read-only MCP access from Claude Code
Priya ran the tests before 5 of 6 commits and rejected 3 agent editsprompt #7orders.ts:42. Sam committed agent output without running it twicecommit a1f3c09.
Proposed actionSet Priya → Shortlisted?
ConfirmCancel
Trust
Transparent by design.
Before they start, candidates read exactly what is captured and what isn't. Nothing is recorded until they accept.
What you see
Prompts, agent replies and the tools the agent ran
Every edit the agent proposed, and whether they kept it
Terminal command lines, never their output, with secrets removed
Commits and pushes to the assessment repo
Screen and microphone from Start to submit, only if the assessment records
What you never see
Keystrokes, or anything outside the assessment repo
The model's private reasoning
Their Claude account, billing or credentials
Anything after they submit
disclosure accepted before startingversioned, so the terms never shiftsecrets redactedrecordings expire automatically
Got an assessment invite?
Sign in with the email it was sent to. Work in your own editor and Claude Code, and see exactly what's shared before you start.