The Take-Home Invitation, Scored to Do, Cap, or Decline
You can take a specific take-home invitation, score it on five dimensions, and pick do, cap, or decline with the exact reply already written.
A take-home invitation lands in your inbox and the clock starts before you have decided anything. Most published guides assume you have already said yes and coach the build; this one sits one step upstream, at the fork where you choose to do it in full, cap and negotiate its scope, or decline without killing the process. It is for candidates mid-loop who need a defensible answer today, not a philosophy of unpaid work. What follows is a scored rubric: five dimensions, three actions, and the exact reply for each.
What this framework decides
This framework turns one take-home invitation into one of three actions with a matching reply already written. You score the invite on five dimensions, map the scores to Do, Cap, or Decline, and send the reply that fits.
The three actions are not "yes" and "no." They are:
- Do - complete the assignment as scoped, confirm and ask one clarifying question.
- Cap - agree in principle, state a time box, offer a live walkthrough for anything beyond it.
- Decline - thank them, offer a substitute, keep the door open.
The reason a bare yes/no fails is that the invite itself is often ambiguous. A task close to the live product can be genuine signal or free work. A "two-hour" exercise can cost five once you count environment setup. The score exists to resolve that ambiguity into an action you can defend to yourself and to the recruiter.
That gap is the whole reason to score rather than reflexively comply. Companies fan assignments out cheaply, and each candidate pays full price. Capping effort beats maximizing it because the expected return on a marginal hour is low.
The five dimensions you score
Score every invite on the same five dimensions, in this order: realistic hours, free-work risk, stage and defense, evaluation criteria, and your market leverage. Each dimension proves something specific, and each has a way it lies to you.
Realistic hours, not stated hours
What it proves: whether the exercise is inside the norms practitioners accept. What it looks like when it lies: the stated number is almost always low. Assignment authors consistently underestimate their own exercises, so double the estimate before you judge it. One candidate spent five or more hours just setting up an unfamiliar environment for a task billed at one to two hours, then canceled.
The practitioner consensus is a three-hour ceiling for an unpaid exercise, with most of the signal already present by hour two. A HireVue engineering director put the outer bound plainly: assignments that take two days are too much to ask.
Free-work risk
What it proves: whether the output is a check or a deliverable. What it looks like when it lies: product relevance masquerades as legitimacy. A task can be close to the live product because that is what the team does all day, or because they want free labor. You resolve it by asking whether the output will be used and whether it is a toy slice or a real thing.
Named red flags from the field: the task is tied too closely to the company's actual product, there are multiple revision rounds, the output is publishable or deliverable such as real marketing campaigns or client-specific problems, and the role is constantly open. A product leader's counterpoint keeps this honest: for product roles, unless the assignment is truly badly designed, candidates should not spend more than a few hours, and the output will necessarily be lacking and not usable by the company.
Stage and defense
What it proves: whether a human is invested and whether your judgment will be examined. What it looks like when it lies: a polished-looking process with no debrief scheduled. The single cleanest free-work tell is a missing defense session. If no one plans to discuss your work, the work itself, not your reasoning, is the deliverable they wanted.
Evaluation criteria
What it proves: how graders will read the work, which tells you what to optimize. What it looks like when it lies: purposefully vague instructions, which some teams use deliberately to see how you think. Vagueness is not automatically bad, but paired with long hours it is a negotiate trigger.
Structured rubrics dominate real grading. Slack grades its exercise against more than 30 predetermined criteria to limit bias, looking for clean, readable, performant, maintainable output. Reviewers also read commit history: breaking work into logical commits helps the reviewer follow your line of thinking. This is why a scoped, explainable three-hour slice can outscore an over-engineered build.
Market leverage
What it proves: how hard you can push without losing the process. What it looks like when it lies: you overestimate it. Leverage is real only with competing options. Competing offers or scarce skills raise it; nothing lined up lowers it. With nothing lined up, negotiating aggressively can cost you the process.
The five dimensions, outermost to innermost
- Realistic hoursDoubled estimate against the 3-hour ceiling
- Free-work riskProduct proximity, revisions, reusable output, perpetual role
- Stage and defenseFollows a human, has a scheduled debrief
- Evaluation criteriaStated rubric or purposeful vagueness
- Market leverageCompeting offers or scarce skills, honestly read
Hours mapped to action
Hours are the fastest cut. Map the doubled estimate against the thresholds below before you weigh anything softer. This table is the spine of the Do/Cap/Decline call.
| Hours (author estimate) | Recommended action | Source |
|---|---|---|
| 3 or less, post-conversation | Do | rankid.dev |
| 4 to 8, vague or no criteria | Cap and negotiate | rankid.dev |
| 2 to 4 early, 4 to 8 senior | Reasonable range | signalroster.com |
| Over 6 | Consider decline | hiredkit.ai |
Note what doubling does here. A task billed at four hours becomes eight once you correct for author optimism, which pushes it from the reasonable band toward the decline line. The estimate you score is never the estimate you were quoted.
The adoption numbers that set your expectations
Take-homes are common and their pass rates are low, so treat each one as a bet with a known bad payout, not a formality you owe the company. The benchmarks below explain why capping beats maximizing.
| Metric | Value | Source |
|---|---|---|
| Companies using take-homes | 68% | hiredkit.ai (citing CoderPad 2025) |
| Candidates passing after completion | ~20% | hiredkit.ai |
| Hybrid take-home plus live review | 41% | hiredkit.ai |
| Slack grading criteria | 30+ | indeed.com/hire |
The 41% hybrid figure matters for your defense-session check. Pairing the take-home with a live review is now common enough that its absence is a real signal, not just bad luck. When a review is scheduled, your explainable slice is what carries the round.
Companies fan assignments out cheaply and each candidate pays full price, which is why capping beats maximizing.
The procedure, invite to reply
Run these steps in order the moment the invite lands. The whole pass takes about ninety minutes of thinking and produces one action plus one sent reply. Sources disagree on whether to assess free-work risk or stage first; either order works as long as you complete both before you decide.
From invite to sent reply
- Log the invite's hard factsRecord estimated hours, stage in the loop, whether a human has spoken to you, the stated criteria, and whether a defense session is scheduled. Done means every dimension has a value.
- Double the time estimateAssume true cost is roughly double the author's number, because authors consistently underestimate their own exercises. Done means a realistic hour figure.
- Score the free-work riskCheck task-to-product proximity, revision rounds, deliverable reusability, any NDA or Confidential stamp, and whether the role is perpetually open. Done means a high, medium, or low label.
- Check the stage and defenseConfirm whether the task follows a human conversation and whether a debrief is scheduled. No defense session raises risk. Done means the stage is confirmed.
- Weigh market leverageCompeting offers or scarce skills raise leverage to negotiate or decline; nothing lined up lowers it. Done means an honest leverage read.
- Land on Do, Cap, or DeclineMap the scores to one action. Low risk, capped hours, scheduled defense means Do; vague, long, or defense-free means Cap; high free-work risk plus leverage means Decline.
- Send the matching replyDo means confirm and ask one clarifying question. Cap means state a time box and offer a live walkthrough. Decline means thank, offer a substitute, keep the door open.
- Time-box the build if cappingStop at the cap and document your tradeoffs. Done means a working slice plus a note of what you would do with more time.
- Prepare the defenseBe ready to explain approach, tradeoffs, and what you would do with more time. Done means you can walk a reviewer through the work in ten minutes.
Mapping scores to the three actions
Once every dimension has a label, the mapping is mechanical. The matrix below reads free-work risk against leverage, which are the two variables that most often flip a borderline case.
Free-work risk against leverage
The clean cases resolve on hours and defense alone. Doubled hours at or under three, after a human conversation, with a scheduled debrief, is a Do. Doubled hours between four and eight, or a vague brief, or no prior conversation, is a Cap. High free-work risk with a missing defense session and any real leverage is a Decline.
The messy cases are where leverage does the work. High risk with nothing else lined up means you cap and insist on a review call rather than refuse outright, because a flat refusal can end a process you cannot yet replace.
The reply scripts
Each action has a reply that keeps the process alive. Send one of these three, adapted to your specifics. The point of scripting them is that the hard thinking is already done by the time you write, so you do not soften a Cap into a Do out of nerves.
Thanks for sending this over - happy to take it on. Before I start, one quick clarification: [the single most scope-relevant question, e.g. "should I prioritize a working slice over full coverage if I hit the time box?"]. I'll aim to keep it within the estimated time and will note any tradeoffs I make so we can talk them through on the review call.
Send when hours and risk are low and a defense call exists. Ask exactly one question so you look engaged, not high-maintenance.
I'd like to move forward on this. To keep it fair to both sides, I'll time-box the build to [N] hours and hand over a working slice with a short note on the tradeoffs and what I'd do with more time. I'd also be glad to walk through it live on a call so you can probe my reasoning directly - any work beyond the time box I'd want to scope together first. Does that work?
Send when the brief is vague, the doubled hours run long, or no conversation has happened yet. The 30-minute line is a real reported phrasing that advanced processes.
Thanks for the invitation. Rather than a new build, I'd like to offer a faster way to see the same signal: I can walk you through [a comparable past project / a portfolio piece / public work at LINK] on a live call and answer any questions about my approach and tradeoffs. That should give you a strong read on how I work. Would that be a workable substitute here?
Send when free-work risk is high and you have leverage. Offer public work you already own so nothing new is built for free. Never confront the ethics of the practice.
Two phrasings from the field are worth keeping in these scripts. Offering to go over the work on a live call, with anything past thirty minutes compensated, advanced processes. Getting the recruiter on the phone to talk through the project, rather than emailing back and forth, also advanced them. Naming a genuine competing offer to ask to expedite or skip a step is a legitimate move when you actually hold one.
How this goes wrong
The framework fails in predictable ways, and the failures are more costly than the occasional lost opportunity. Each mode below has a check that catches it before it costs you.
- Trusting the stated hours. A "2-hour" task can need unfamiliar tooling setup; one candidate lost five or more hours before canceling. Check: scope the setup and environment before starting the clock.
- Mistaking product-relevance for legitimacy. A task close to the live product can be genuine signal or free work. Check: ask whether the output will be used and whether it is a toy slice or a real deliverable.
- Over-delivering to win. You pass the screen but signal that you ignore scope, and rejections happen for "not enough effort" even at a cap, because graders weight judgment, not volume. Check: cap openly and defend the cap, do not hide it.
- Confronting the recruiter as a decline tactic. Calling the practice discriminatory can end the process cleanly rather than negotiate it. One candidate did exactly this and the process ended. Check: decline with a substitute, never with an accusation.
- Declining with no substitute. A bare no reads as disinterest. Sources stress offering an alternative and not ghosting. Check: every decline carries a portfolio, past-project, or public-work offer.
- Assuming a defense session exists. No debrief scheduled means padded or AI-assisted work goes unchecked, and that a genuine skills-check may not be happening at all. Check: ask if there is a review call.
- Ignoring leverage. With nothing lined up, negotiating aggressively can cost the process; leverage is real only with competing options. Check: read your leverage honestly before you push.
Leverage is a local number
Your ability to decline or negotiate depends on how thin the market is around you, and that varies enormously by geography. In Refolk's index of professional profiles, the screening-stage recruiter population differs by nearly two orders of magnitude across markets.
| Market | Recruiter / TA profiles | Ratio vs Germany |
|---|---|---|
| United States | 22,950 | 82.6x |
| United Kingdom | 940 | 3.4x |
| Germany | 278 | 1.0x |
The mechanism is simple. In a thin market a single decline is more visible and harder to replace, so a substitution offer preserves optionality where a flat refusal would burn it. In a deep market you have more room to walk. In Refolk's index, the top US employers of these screening-stage recruiters include Snowflake, Blue Origin, and Google; in the German sample, Google and AWS lead. Knowing who owns the stage at your target helps you judge how the take-home will actually be graded, since Greenhouse recommends the hiring manager for the role create the take-home test, and the same manager often runs the defense.
If you want to pressure-test your own leverage, look at who actually issues and grades these assignments in your field before you reply. Refolk writes your resume from your history, tailors it to each posting, and scores your fit, so you walk into the take-home fork already knowing which processes are worth the hours.
Before you hit send
Run this checklist against your reply before it leaves your outbox. It catches the failures above and confirms your action matches your score.
Reply-ready check
- The time estimate is doubled and scored against the 3-hour ceiling and 6-hour decline line.
- Free-work risk is labeled high, medium, or low with a stated reason.
- You have confirmed whether a defense or review call is scheduled.
- Your leverage read is honest, based on options you actually hold.
- The action, Do, Cap, or Decline, matches the risk-versus-leverage matrix.
- A Cap names a specific time box and offers a live walkthrough.
- A Decline offers a concrete substitute and keeps the door open, with no ethics confrontation.
- If you are building, you can defend approach and tradeoffs in ten minutes.
Keeping this current
The thresholds move, so re-check the mechanism rather than memorizing a number. Take-home adoption was reported at 68% of companies with a hybrid live-review pairing at 41%, and both figures trend upward year over year; when you re-run this, look for the current adoption and hybrid rates from a fresh coding-platform report rather than trusting last cycle's numbers. The pass-after-completion rate near 20% and the completion rate near 30% are the two figures that justify capping, so watch whether they shift. Everything else in this framework is structural: double the estimate, find the defense session, read your leverage locally, and let the score, not the invite's tone, choose your action.
Questions job seekers ask
Should I do this take-home assignment or not?
Score it first. If the doubled time estimate is three hours or less, it arrives after you have spoken to a human, the criteria are stated, and a debrief call is scheduled, do it. If it is vague, long, or lands before any conversation, cap and negotiate the scope. If the task is a close clone of the live product with no defense session and you have leverage, decline with a substitute. The decision is the five-dimension score, not your mood on the day.
How do I decline a take-home without burning the bridge?
Never send a bare no, which reads as disinterest. Thank them, offer a concrete substitute such as walking through prior work or a portfolio or public work, and keep the door open. One hiring-side commenter says they accept public work in place of the assignment. Avoid confronting the recruiter about the ethics of the practice; candidates who called it discriminatory reported the process ended cleanly rather than opening a negotiation.
How many hours is a take-home reasonably worth?
Practitioner sources converge on a low ceiling of three hours for an unpaid exercise, with most of the signal an employer can gather already present by hour two. Typical ranges run 2 to 4 hours for early screens and 4 to 8 hours for senior roles. Consider declining above six hours. Double the author's estimate first, since authors systematically undercount, and one candidate lost 5 or more hours just setting up a supposedly 1 to 2 hour task.
Can I offer a work sample instead of a take-home?
Yes, and it is one of the strongest moves. Reported substitutes are walking the interviewer through past work projects, presenting a portfolio more thoroughly, or a live session. Offer public work you already own so nothing new is built for free. Be prepared to be dropped for some roles if you refuse to participate at all, which is why a substitute offer, not a flat refusal, keeps you in the running.
What are the red flags that a take-home is free work?
The task sits too close to the live product, there are multiple revision rounds, the output is publishable or deliverable such as real marketing campaigns or client-specific problems, the brief is un-timeboxed, and the role is constantly open. The cleanest single tell is no scheduled defense session. If no one plans to discuss your work, the work itself is the deliverable they wanted, not your judgment.
Does capping my effort hurt my chances of passing?
Less than you fear. Structured rubrics reward scoped, explainable work over polish. Slack grades against more than 30 predetermined criteria and reviewers read commit history to follow your thinking, so a defensible three-hour slice with documented tradeoffs can outscore an over-engineered build. Rejections do happen for not enough effort, so the fix is to state your time box openly and defend your judgment, not to hide that you capped.
Put this to work
Paste your career in once. Every application after that is written for you.
Drop a resume or a LinkedIn URL. I rank the live openings against it, rewrite the resume and write a cover letter for the best of them, and fill in the employer's form when you press the button. You read, you decide what goes out.
01Drop your resume
A PDF or a LinkedIn URL. About a minute, once.
02I rank the openings
Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.
03Each one is written up
Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.
- New matches ranked and written before you are up.
- Every bullet stays inside what your history supports. Nothing invented.
- Queued, submitted, interviewing, offer: one screen, not a spreadsheet.
500 free credits on sign-up. No card. Nothing is sent until you say so.