# The Take-Home Invitation, Scored to Do, Cap, or Decline

*You can take a specific take-home invitation, score it on five dimensions, and pick do, cap, or decline with the exact reply already written.*

- Canonical URL: https://www.refolk.ai/candidates/guides/take-home-scored-do-cap-decline
- Pillar: Interviewing
- Format: Framework
- Published: 2026-09-23
- Last reviewed: 2026-09-23
- Reading time: 16 min

A take-home invitation lands in your inbox and the clock starts before you have decided anything. Most published guides assume you have already said yes and coach the build; this one sits one step upstream, at the fork where you choose to do it in full, cap and negotiate its scope, or decline without killing the process. It is for candidates mid-loop who need a defensible answer today, not a philosophy of unpaid work. What follows is a scored rubric: five dimensions, three actions, and the exact reply for each.

## What this framework decides

This framework turns one take-home invitation into one of three actions with a matching reply already written. You score the invite on five dimensions, map the scores to Do, Cap, or Decline, and send the reply that fits.

The three actions are not "yes" and "no." They are:

- **Do** - complete the assignment as scoped, confirm and ask one clarifying question.
- **Cap** - agree in principle, state a time box, offer a live walkthrough for anything beyond it.
- **Decline** - thank them, offer a substitute, keep the door open.

The reason a bare yes/no fails is that the invite itself is often ambiguous. A task close to the live product can be genuine signal or free work. A "two-hour" exercise can cost five once you count environment setup. The score exists to resolve that ambiguity into an action you can defend to yourself and to the recruiter.

**~20% - Candidates who pass after completing a take-home**

Cited by hiredkit.ai. About 30% even complete the assignments they are sent, so most poured-in hours are sunk.

That gap is the whole reason to score rather than reflexively comply. Companies fan assignments out cheaply, and each candidate pays full price. Capping effort beats maximizing it because the expected return on a marginal hour is low.

## The five dimensions you score

Score every invite on the same five dimensions, in this order: realistic hours, free-work risk, stage and defense, evaluation criteria, and your market leverage. Each dimension proves something specific, and each has a way it lies to you.

### Realistic hours, not stated hours

What it proves: whether the exercise is inside the norms practitioners accept. What it looks like when it lies: the stated number is almost always low. Assignment authors consistently underestimate their own exercises, so double the estimate before you judge it. One candidate spent five or more hours just setting up an unfamiliar environment for a task billed at one to two hours, then canceled.

The practitioner consensus is a three-hour ceiling for an unpaid exercise, with most of the signal already present by hour two. A HireVue engineering director put the outer bound plainly: assignments that take two days are too much to ask.

### Free-work risk

What it proves: whether the output is a check or a deliverable. What it looks like when it lies: product relevance masquerades as legitimacy. A task can be close to the live product because that is what the team does all day, or because they want free labor. You resolve it by asking whether the output will be used and whether it is a toy slice or a real thing.

Named red flags from the field: the task is tied too closely to the company's actual product, there are multiple revision rounds, the output is publishable or deliverable such as real marketing campaigns or client-specific problems, and the role is constantly open. A product leader's counterpoint keeps this honest: for product roles, unless the assignment is truly badly designed, candidates should not spend more than a few hours, and the output will necessarily be lacking and not usable by the company.

### Stage and defense

What it proves: whether a human is invested and whether your judgment will be examined. What it looks like when it lies: a polished-looking process with no debrief scheduled. The single cleanest free-work tell is a missing defense session. If no one plans to discuss your work, the work itself, not your reasoning, is the deliverable they wanted.

> **Rule:** No scheduled defense, no free build
>
> If there is no review call planned to discuss your approach and tradeoffs, treat the free-work risk as high until proven otherwise. Ask directly whether a debrief is scheduled before you start the clock.

### Evaluation criteria

What it proves: how graders will read the work, which tells you what to optimize. What it looks like when it lies: purposefully vague instructions, which some teams use deliberately to see how you think. Vagueness is not automatically bad, but paired with long hours it is a negotiate trigger.

Structured rubrics dominate real grading. Slack grades its exercise against more than 30 predetermined criteria to limit bias, looking for clean, readable, performant, maintainable output. Reviewers also read commit history: breaking work into logical commits helps the reviewer follow your line of thinking. This is why a scoped, explainable three-hour slice can outscore an over-engineered build.

### Market leverage

What it proves: how hard you can push without losing the process. What it looks like when it lies: you overestimate it. Leverage is real only with competing options. Competing offers or scarce skills raise it; nothing lined up lowers it. With nothing lined up, negotiating aggressively can cost you the process.

#### The five dimensions, outermost to innermost

1. **Realistic hours** - Doubled estimate against the 3-hour ceiling
2. **Free-work risk** - Product proximity, revisions, reusable output, perpetual role
3. **Stage and defense** - Follows a human, has a scheduled debrief
4. **Evaluation criteria** - Stated rubric or purposeful vagueness
5. **Market leverage** - Competing offers or scarce skills, honestly read

*Score from the outside in; the stated hours are the surface, leverage is the core that sets how hard you can push.*

## Hours mapped to action

Hours are the fastest cut. Map the doubled estimate against the thresholds below before you weigh anything softer. This table is the spine of the Do/Cap/Decline call.

| Hours (author estimate) | Recommended action | Source |
|---|---|---|
| 3 or less, post-conversation | Do | rankid.dev |
| 4 to 8, vague or no criteria | Cap and negotiate | rankid.dev |
| 2 to 4 early, 4 to 8 senior | Reasonable range | signalroster.com |
| Over 6 | Consider decline | hiredkit.ai |

Note what doubling does here. A task billed at four hours becomes eight once you correct for author optimism, which pushes it from the reasonable band toward the decline line. The estimate you score is never the estimate you were quoted.

> **Watch out:** Setup time hides inside the estimate
>
> The stated hours count the task, not the environment. Scope the tooling and setup before you start the clock. A candidate lost 5+ hours on setup for a "1-2 hour" task and canceled with nothing to show.

## The adoption numbers that set your expectations

Take-homes are common and their pass rates are low, so treat each one as a bet with a known bad payout, not a formality you owe the company. The benchmarks below explain why capping beats maximizing.

| Metric | Value | Source |
|---|---|---|
| Companies using take-homes | 68% | hiredkit.ai (citing CoderPad 2025) |
| Candidates passing after completion | ~20% | hiredkit.ai |
| Hybrid take-home plus live review | 41% | hiredkit.ai |
| Slack grading criteria | 30+ | indeed.com/hire |

The 41% hybrid figure matters for your defense-session check. Pairing the take-home with a live review is now common enough that its absence is a real signal, not just bad luck. When a review is scheduled, your explainable slice is what carries the round.

> Companies fan assignments out cheaply and each candidate pays full price, which is why capping beats maximizing.

## The procedure, invite to reply

Run these steps in order the moment the invite lands. The whole pass takes about ninety minutes of thinking and produces one action plus one sent reply. Sources disagree on whether to assess free-work risk or stage first; either order works as long as you complete both before you decide.

#### From invite to sent reply

1. **Log the invite's hard facts** - Record estimated hours, stage in the loop, whether a human has spoken to you, the stated criteria, and whether a defense session is scheduled. Done means every dimension has a value.
2. **Double the time estimate** - Assume true cost is roughly double the author's number, because authors consistently underestimate their own exercises. Done means a realistic hour figure.
3. **Score the free-work risk** - Check task-to-product proximity, revision rounds, deliverable reusability, any NDA or Confidential stamp, and whether the role is perpetually open. Done means a high, medium, or low label.
4. **Check the stage and defense** - Confirm whether the task follows a human conversation and whether a debrief is scheduled. No defense session raises risk. Done means the stage is confirmed.
5. **Weigh market leverage** - Competing offers or scarce skills raise leverage to negotiate or decline; nothing lined up lowers it. Done means an honest leverage read.
6. **Land on Do, Cap, or Decline** - Map the scores to one action. Low risk, capped hours, scheduled defense means Do; vague, long, or defense-free means Cap; high free-work risk plus leverage means Decline.
7. **Send the matching reply** - Do means confirm and ask one clarifying question. Cap means state a time box and offer a live walkthrough. Decline means thank, offer a substitute, keep the door open.
8. **Time-box the build if capping** - Stop at the cap and document your tradeoffs. Done means a working slice plus a note of what you would do with more time.
9. **Prepare the defense** - Be ready to explain approach, tradeoffs, and what you would do with more time. Done means you can walk a reviewer through the work in ten minutes.

## Mapping scores to the three actions

Once every dimension has a label, the mapping is mechanical. The matrix below reads free-work risk against leverage, which are the two variables that most often flip a borderline case.

#### Free-work risk against leverage

Horizontal axis runs from Low leverage to High leverage. Vertical axis runs from Low free-work risk to High free-work risk.

| Quadrant | What it means |
| --- | --- |
| Low risk, low leverage | Do it, cap only if the doubled hours run long |
| Low risk, high leverage | Do it, or cap and offer a walkthrough to save time |
| High risk, low leverage | Cap hard and demand a defense call, or decline with a substitute |
| High risk, high leverage | Decline and offer public work, you can afford to |

*Read your risk label against your honest leverage to break a tie between Cap and Decline.*

The clean cases resolve on hours and defense alone. Doubled hours at or under three, after a human conversation, with a scheduled debrief, is a Do. Doubled hours between four and eight, or a vague brief, or no prior conversation, is a Cap. High free-work risk with a missing defense session and any real leverage is a Decline.

The messy cases are where leverage does the work. High risk with nothing else lined up means you cap and insist on a review call rather than refuse outright, because a flat refusal can end a process you cannot yet replace.

## The reply scripts

Each action has a reply that keeps the process alive. Send one of these three, adapted to your specifics. The point of scripting them is that the hard thinking is already done by the time you write, so you do not soften a Cap into a Do out of nerves.

**Do - confirm and clarify**

```
Thanks for sending this over - happy to take it on. Before I start, one quick clarification: [the single most scope-relevant question, e.g. "should I prioritize a working slice over full coverage if I hit the time box?"]. I'll aim to keep it within the estimated time and will note any tradeoffs I make so we can talk them through on the review call.
```

*Send when hours and risk are low and a defense call exists. Ask exactly one question so you look engaged, not high-maintenance.*

**Cap - time-box and offer a walkthrough**

```
I'd like to move forward on this. To keep it fair to both sides, I'll time-box the build to [N] hours and hand over a working slice with a short note on the tradeoffs and what I'd do with more time. I'd also be glad to walk through it live on a call so you can probe my reasoning directly - any work beyond the time box I'd want to scope together first. Does that work?
```

*Send when the brief is vague, the doubled hours run long, or no conversation has happened yet. The 30-minute line is a real reported phrasing that advanced processes.*

**Decline - substitute and keep the door open**

```
Thanks for the invitation. Rather than a new build, I'd like to offer a faster way to see the same signal: I can walk you through [a comparable past project / a portfolio piece / public work at LINK] on a live call and answer any questions about my approach and tradeoffs. That should give you a strong read on how I work. Would that be a workable substitute here?
```

*Send when free-work risk is high and you have leverage. Offer public work you already own so nothing new is built for free. Never confront the ethics of the practice.*

Two phrasings from the field are worth keeping in these scripts. Offering to go over the work on a live call, with anything past thirty minutes compensated, advanced processes. Getting the recruiter on the phone to talk through the project, rather than emailing back and forth, also advanced them. Naming a genuine competing offer to ask to expedite or skip a step is a legitimate move when you actually hold one.

## How this goes wrong

The framework fails in predictable ways, and the failures are more costly than the occasional lost opportunity. Each mode below has a check that catches it before it costs you.

- **Trusting the stated hours.** A "2-hour" task can need unfamiliar tooling setup; one candidate lost five or more hours before canceling. Check: scope the setup and environment before starting the clock.
- **Mistaking product-relevance for legitimacy.** A task close to the live product can be genuine signal or free work. Check: ask whether the output will be used and whether it is a toy slice or a real deliverable.
- **Over-delivering to win.** You pass the screen but signal that you ignore scope, and rejections happen for "not enough effort" even at a cap, because graders weight judgment, not volume. Check: cap openly and defend the cap, do not hide it.
- **Confronting the recruiter as a decline tactic.** Calling the practice discriminatory can end the process cleanly rather than negotiate it. One candidate did exactly this and the process ended. Check: decline with a substitute, never with an accusation.
- **Declining with no substitute.** A bare no reads as disinterest. Sources stress offering an alternative and not ghosting. Check: every decline carries a portfolio, past-project, or public-work offer.
- **Assuming a defense session exists.** No debrief scheduled means padded or AI-assisted work goes unchecked, and that a genuine skills-check may not be happening at all. Check: ask if there is a review call.
- **Ignoring leverage.** With nothing lined up, negotiating aggressively can cost the process; leverage is real only with competing options. Check: read your leverage honestly before you push.

> **Tip:** Commits are part of the defense
>
> Reviewers read git history to follow your thinking. If the task is code, break it into logical commits so the reviewer can see your development process, not just the endpoint. This is cheap and it directly feeds the rubric.

## Leverage is a local number

Your ability to decline or negotiate depends on how thin the market is around you, and that varies enormously by geography. In Refolk's index of professional profiles, the screening-stage recruiter population differs by nearly two orders of magnitude across markets.

| Market | Recruiter / TA profiles | Ratio vs Germany |
|---|---|---|
| United States | 22,950 | 82.6x |
| United Kingdom | 940 | 3.4x |
| Germany | 278 | 1.0x |

The mechanism is simple. In a thin market a single decline is more visible and harder to replace, so a substitution offer preserves optionality where a flat refusal would burn it. In a deep market you have more room to walk. In Refolk's index, the top US employers of these screening-stage recruiters include Snowflake, Blue Origin, and Google; in the German sample, Google and AWS lead. Knowing who owns the stage at your target helps you judge how the take-home will actually be graded, since Greenhouse recommends the hiring manager for the role create the take-home test, and the same manager often runs the defense.

If you want to pressure-test your own leverage, look at who actually issues and grades these assignments in your field before you reply. [Refolk](/candidates) writes your resume from your history, tailors it to each posting, and scores your fit, so you walk into the take-home fork already knowing which processes are worth the hours.

Ask me this: `Technical recruiters at US software companies who own the screening and assessment stage of hiring.` - [run the search](https://www.refolk.ai/start?q=Technical%20recruiters%20at%20US%20software%20companies%20who%20own%20the%20screening%20and%20assessment%20stage%20of%20hiring.).

*Returns named screening-stage recruiters so you can see who issues and grades the take-home before you decide how hard to push.*

## Before you hit send

Run this checklist against your reply before it leaves your outbox. It catches the failures above and confirms your action matches your score.

#### Reply-ready check

- [ ] The time estimate is doubled and scored against the 3-hour ceiling and 6-hour decline line.
- [ ] Free-work risk is labeled high, medium, or low with a stated reason.
- [ ] You have confirmed whether a defense or review call is scheduled.
- [ ] Your leverage read is honest, based on options you actually hold.
- [ ] The action, Do, Cap, or Decline, matches the risk-versus-leverage matrix.
- [ ] A Cap names a specific time box and offers a live walkthrough.
- [ ] A Decline offers a concrete substitute and keeps the door open, with no ethics confrontation.
- [ ] If you are building, you can defend approach and tradeoffs in ten minutes.

## Keeping this current

The thresholds move, so re-check the mechanism rather than memorizing a number. Take-home adoption was reported at 68% of companies with a hybrid live-review pairing at 41%, and both figures trend upward year over year; when you re-run this, look for the current adoption and hybrid rates from a fresh coding-platform report rather than trusting last cycle's numbers. The pass-after-completion rate near 20% and the completion rate near 30% are the two figures that justify capping, so watch whether they shift. Everything else in this framework is structural: double the estimate, find the defense session, read your leverage locally, and let the score, not the invite's tone, choose your action.

## Frequently asked questions

### Should I do this take-home assignment or not?

Score it first. If the doubled time estimate is three hours or less, it arrives after you have spoken to a human, the criteria are stated, and a debrief call is scheduled, do it. If it is vague, long, or lands before any conversation, cap and negotiate the scope. If the task is a close clone of the live product with no defense session and you have leverage, decline with a substitute. The decision is the five-dimension score, not your mood on the day.

### How do I decline a take-home without burning the bridge?

Never send a bare no, which reads as disinterest. Thank them, offer a concrete substitute such as walking through prior work or a portfolio or public work, and keep the door open. One hiring-side commenter says they accept public work in place of the assignment. Avoid confronting the recruiter about the ethics of the practice; candidates who called it discriminatory reported the process ended cleanly rather than opening a negotiation.

### How many hours is a take-home reasonably worth?

Practitioner sources converge on a low ceiling of three hours for an unpaid exercise, with most of the signal an employer can gather already present by hour two. Typical ranges run 2 to 4 hours for early screens and 4 to 8 hours for senior roles. Consider declining above six hours. Double the author's estimate first, since authors systematically undercount, and one candidate lost 5 or more hours just setting up a supposedly 1 to 2 hour task.

### Can I offer a work sample instead of a take-home?

Yes, and it is one of the strongest moves. Reported substitutes are walking the interviewer through past work projects, presenting a portfolio more thoroughly, or a live session. Offer public work you already own so nothing new is built for free. Be prepared to be dropped for some roles if you refuse to participate at all, which is why a substitute offer, not a flat refusal, keeps you in the running.

### What are the red flags that a take-home is free work?

The task sits too close to the live product, there are multiple revision rounds, the output is publishable or deliverable such as real marketing campaigns or client-specific problems, the brief is un-timeboxed, and the role is constantly open. The cleanest single tell is no scheduled defense session. If no one plans to discuss your work, the work itself is the deliverable they wanted, not your judgment.

### Does capping my effort hurt my chances of passing?

Less than you fear. Structured rubrics reward scoped, explainable work over polish. Slack grades against more than 30 predetermined criteria and reviewers read commit history to follow your thinking, so a defensible three-hour slice with documented tradeoffs can outscore an over-engineered build. Rejections do happen for not enough effort, so the fix is to state your time box openly and defend your judgment, not to hide that you capped.

---

*From the Refolk guide library. I revise these guides rather than replacing them, so the current version is always at https://www.refolk.ai/candidates/guides/take-home-scored-do-cap-decline*
