engineering screens
you can defend.
Codesolara opens a real editor, terminal and file tree in the candidate's browser, with an AI assistant — Solara — in the panel beside them. What they ask it, what they keep, and what they throw out becomes a scorecard that shows how they actually think.
// what you get back
A scorecard where every claim cites the moment it came from — and any claim whose evidence doesn't hold is withdrawn, not guessed.
Reverted the assistant's timezone change after noticing it dropped the DST offset, then asked for a failing test before re-attempting.
Early prompts named the symptom but not the surface; the assistant explored three files before reaching the scheduler. Later prompts were specific.
The scheduler was corrected; the feed still computes its own window and disagrees with it after a DST transition.
Withdrawn by the verifier: the cited probe exercised the scheduler, not the feed, so it cannot support this claim. Left out of the score rather than marked down.
Illustrative — not a real candidate.
// what the candidate sees
A real workspace. An AI pair. An engineer who reviews, corrects, and verifies the work. An evidence-backed scorecard.
// how it works
Send one link
Pick a task, invite a candidate. No account for them, nothing to install — the workspace opens in their browser.
They build, with Solara
A realistic task in a real editor and terminal, with an assistant that can read, write and run code beside them — and it asks before it runs a command.
You get the scorecard
The session becomes a structured read on the candidate, every claim citing the moment it came from. Automatically — no panel review.
Real work, not whiteboards
A realistic task in a real editor, terminal and file tree — with Solara in the panel beside it.
You see the judgment calls
Solara changes the code; the candidate reads it, runs it, and fixes or undoes what is wrong. What they do next is the signal a diff alone can't give you.
The scorecard is the product
Every attempt becomes an evidence-backed read on the candidate, on the dimensions you chose to weigh.
Checked, then double-checked
A second pass verifies that the evidence behind each score actually holds. When it can't, the scorecard says so instead of guessing.
Explainable & defensible
Every score cites its evidence. Humans stay in the loop and can override anything — on the record.
Free for the whole team
Unlimited seats. Invite every reviewer. You pay when a candidate actually interviews.
The AI-free test is not the only alternative — here is how the other platforms handle AI, and what "fluency with AI" actually means if you are going to measure it.
You pay per attempt.
Every seat is free.
No subscription, no per-recruiter tax, nothing that expires. Three depths of interview — they differ by how long the candidate gets and how much they can lean on the AI assistant before it runs out.
A first read
- 60 minutes
- AI assistant for targeted help
- A focused, single-surface task
- Full scorecard
The working interview
- 60 minutes
- AI assistant for a full working session
- A realistic feature in an existing codebase
- Full scorecard
The onsite replacement
- 90 minutes
- AI assistant to lean on throughout
- An ambiguous brief with real trade-offs
- Full scorecard
// we're setting per-attempt pricing with our first teams — ask and we'll quote you plainly
just changed.