Methodology · The Execution Paradox · Q3 2026
N=603 responses were collected; one was a junk response, so all results are computed over the 602 analysable ones, which were collected between 2026-01-18 and 2026-03-07. Recruitment was targeted rather than open: 517 respondents (86%) came through the Prolific participant panel against a screen described in full below, 33 through launch partners, 14 through an organization and 3 through other channels, with 35 carrying no recorded source. The sample is side-builders, founders, freelancers, business owners and employees (30.6% / 20.3% / 17.6% / 15.1% / 12.1% respectively, plus 4.2% on early startup teams and one respondent who described themselves differently). Data-quality checks flagged 4 of 602 possible straightliners (0.7%, within-respondent SD below 0.5 across 11 numeric items); they were retained, and removing them moves no published mean by more than 0.004. The burnout composite (8 items, each normalized to 0 to 1 and the average rescaled to 0 to 5) showed acceptable internal consistency, Cronbach's α = 0.797. Bootstrap 95% percentile intervals (2,000 resamples) were used for means and Wilson score intervals for proportions. Analysis is complete-case per metric with no imputation. Results describe this specific screened sample, not the general working population.
Sample
| Work situation | % | Respondents |
|---|---|---|
| Side builder | 30.6% | 184 |
| Founder | 20.3% | 122 |
| Freelancer | 17.6% | 106 |
| Business owner | 15.1% | 91 |
| Employee | 12.1% | 73 |
| Early startup team | 4.2% | 25 |
| Described themselves otherwise | 0.2% | 1 |
| Channel | Respondents | % of sample |
|---|---|---|
| Prolific participant panel | 517 | 85.9% |
| Launch partners | 33 | 5.5% |
| No recorded source | 35 | 5.8% |
| Through an organization | 14 | 2.3% |
| Reddit, a Tally form, other | 3 | 0.5% |
The panel recruitment was screened, not open. It applies to the 517 respondents recruited through the participant panel, who had to meet every one of these criteria:
- fluent in English
- in full-time employment
- within an age band (23 to 45, and 25 to 43 on some batches)
- holding a job position, with a stated number of years of work experience
- using AI at work weekly or more often
- doing at least a quarter of their work on a computer
- a stated level of education
- a stated employer or organization type
- currently running something entrepreneurial
- a panel approval rate of 99 to 100
- not a respondent to any earlier batch, so nobody answered twice
The remaining 85 respondents came through Flowdealer's own channels and were not screened on these criteria.
Cohort cross-tabs
| Say-do-gap flags | Respondents | % of sample | Mean burnout | Scored |
|---|---|---|---|---|
| 0 | 320 | 53.2% | 1.47 | 316 |
| 1 | 164 | 27.2% | 2.12 | 162 |
| 2 | 66 | 11.0% | 2.43 | 66 |
| 3 | 38 | 6.3% | 2.96 | 37 |
| 4 | 10 | 1.7% | 3.17 | 10 |
| 5 | 3 | 0.5% | 3.49 | 3 |
| 6 | 1 | 0.2% | 3.75 | 1 |
| Hours per week | Respondents | Productive best | Completion | Quality of life | Burnout |
|---|---|---|---|---|---|
| <40 | 94 | 2.83 | 3.51 | 3.00 | 1.89 |
| 40-60 | 348 | 3.11 | 3.71 | 3.12 | 1.79 |
| 60-80 | 134 | 3.07 | 3.51 | 2.63 | 2.11 |
| 80-100 | 17 | 3.00 | 3.53 | 2.76 | 2.16 |
| 100+ | 9 | 3.89 | 3.44 | 2.78 | 2.16 |
Instruments
| Column | Question as asked | Response format | Direction |
|---|---|---|---|
| q1_roles | What roles do you perform? | Select any of eleven named functions, plus free entry | Neither |
| q2_work_situation | What's your work situation? | Side builder · Founder · Freelancer · Business owner · Employee · Early startup team, plus free entry | Neither |
| q3_productive_best | How close to your "productive best" are you right now? | Slider, 0 to 5 | Higher is better |
| q4_quality_of_life | How would you rate your current quality of life? | Slider, 0 to 5 | Higher is better |
| q5_hours_per_week | How many hours per week do you spend working? | Under 40 · 40 to 60 · 60 to 80 · 80 to 100 · 100 or more | Neither |
| q6_plans_schedule | Do you actively plan your schedule? | Yes, regularly · Sometimes, inconsistently · No, I don't plan | Neither |
| q7_replan_frequency | How often do you have to re-plan your day due to unexpected changes? | Almost never · Few times a month · Few times a week · Almost daily · Daily | Higher is worse |
| q8_text | What helps you follow through on your commitments and responsibilities? What doesn't work? | Free text | Neither |
| q9_external_pressure_needed | How often do you find it difficult to follow through without external pressure? | Almost never · Few times a month · Few times a week · Almost daily · Daily | Higher is worse |
| q10_priority_difficulty | How often do you find it difficult to decide what your next priority should be? | Almost never · Few times a month · Few times a week · Almost daily · Daily | Higher is worse |
| q11_overcommit_frequency | How often do you commit to more than you can actually deliver? | Almost never · Few times a month · Few times a week · Almost daily · Daily | Higher is worse |
| q12_plan_completion | By the end of a typical week, how much of what you planned actually gets done? | Slider, 0 to 5 | Higher is better |
| q13_text | When executing your plan is harder than it should be, what's usually the reason? | Free text | Neither |
| q14_primary_drain | Which drains you most on a typical workday? | Cognitive load · Social interactions · Emotional demands · Physical fatigue · All drain me roughly equally | Neither |
| q15_text | If you could fix one thing about your work ethic or execution, what would it be? | Free text | Neither |
| q16_tool_count | How many tools do you actively juggle to manage work and life? | 1 to 3 · 4 to 6 · 7 to 9 · 10 or more | Neither |
| q17_tools_used | Select your 3 to 5 most frequently used tools for managing work | Select from a named tool list, plus free entry | Neither |
| q18_text | What do you dislike the most about your current tools and systems? | Free text | Neither |
| q19_drive | How hungry are you to achieve more? | Slider, 0 to 5 | Higher is better |
| q20_wellbeing_priority | How important is it that success doesn't cost you your health or sanity? | Slider, 0 to 5 | Higher is better |
| q21_overwhelm_frequency | How often do you feel overwhelmed by life, commitments and/or responsibilities? | Almost never · Few times a month · Few times a week · Almost daily · Daily | Higher is worse |
| q22_introversion_extraversion | Where would you place yourself on the introversion to extraversion scale? | Slider, 0 to 5 | Neither |
| q23_neurodivergence_impact | How much does neurodivergence (ADHD, autism, dyslexia, or similar) impact your daily work? | Not applicable to me · Minimally · Moderately · Significantly | Neither |
| q24_age | What is your age? | Age in years; bucketed into cohorts only afterwards | Neither |
| q25_text | Anything else you want to share? | Free text, optional | Neither |
Direction is the scale's direction, not its coding. It says which end of the answer is the better outcome, and nothing about reverse-scoring: the burnout composite reverses productive best, quality of life and plan completion when it combines them, and all three are listed here as higher is better, because they are.
| Item | α without this item | Verdict |
|---|---|---|
| Overwhelm frequency | 0.761 | Keep |
| Priority difficulty | 0.759 | Keep |
| Overcommit frequency | 0.778 | Keep |
| External pressure needed | 0.772 | Keep |
| Replan frequency | 0.792 | Keep |
| Plan completion | 0.770 | Keep |
| Quality of life | 0.781 | Keep |
| Productive best | 0.783 | Keep |
| Segment | Definition | % | Respondents | 95% CI (Wilson) |
|---|---|---|---|---|
| Screener: high (pain score ≥ 4) | composite pain ≥ 4 | 57% | 343 | [53.0%, 60.9%] |
| Screener: mid (2–3) | composite pain 2–3 | 38% | 229 | [34.2%, 42.0%] |
| Screener: low (< 2) | composite pain < 2 | 5% | 30 | [3.5%, 7.0%] |
The pain screener above was collected at intake to size an audience, and no figure in this edition uses it. The analysis runs on the burnout composite and the profiles it assigns.
| Band | Profiles it collects | Respondents |
|---|---|---|
| Severe | Full Burnout | 75 |
| Struggling | Planning Chaos, Overwhelmed, Demand Overload | 272 |
| At Risk | Diffuse Strain, Productive but Paying, Quiet Struggle | 88 |
| Healthy | Coping, Thriving | 160 |
These bands are failure types read off the profile assignment, not rungs on a severity ladder: they name how execution breaks, not how badly. There is no score threshold behind them. The severity ladder is in section 2 of the report.
Analysis
| Method | How it was used |
|---|---|
| Correlations | Pearson throughout, except drive against completion, which is reported as a Spearman rank correlation |
| Confidence intervals | bootstrap percentile, 2,000 resamples, for means |
| Wilson score intervals | for proportions |
| Cohen's d | effect size for differences between two group means |
| Multiple comparisons | not corrected; this edition is exploratory and every figure should be read that way |
| Missing data | complete-case per metric, no imputation, so a table's N can be lower than the sample |
| Check | Pass | Value |
|---|---|---|
| completeness | ✓ | Lowest analysis column 89% (age); the optional closing free-text question is 52% and carries no figure |
| straightliners | ✓ | 4 of 602 (0.7%) below SD 0.5 across 11 numeric items |
| reliability | ✓ | Cronbach's α = 0.797, acceptable |
| sample-size | ✓ | N=602 |
| Metric | All respondents | Flagged removed | Shift |
|---|---|---|---|
| Productive best | 3.065 | 3.064 | -0.001 |
| Quality of life | 2.977 | 2.973 | -0.004 |
| Overwhelm frequency | 2.836 | 2.831 | -0.004 |
| Drive | 3.947 | 3.946 | 0.000 |
| Burnout composite | 1.888 | 1.884 | -0.004 |
The flagged respondents were kept. The table above is why the count is worth stating and not worth arguing about: on the pipeline's own convention it is four people, on the other common convention it is one, and dropping all four moves no published mean by more than 0.004.
| Metric | Scale | Mean | 95% CI (bootstrap percentile) | Respondents |
|---|---|---|---|---|
| Productive best | 0 to 5, higher is better | 3.06 | [2.99, 3.14] | 602 |
| Quality of life | 0 to 5, higher is better | 2.98 | [2.89, 3.06] | 602 |
| Overwhelm frequency | 1 to 5, higher is worse | 2.84 | [2.74, 2.93] | 602 |
| Drive | 0 to 5, higher is better | 3.95 | [3.86, 4.03] | 602 |
Specific LLM extraction prompts, confidence thresholds, and profile-assignment cutoffs are proprietary methodology and not publicly released. Profiles are assigned by rule, not clustering. The aggregate data behind every figure in this report is published under CC-BY-NC 4.0, and the files are linked above.
Limitations
| Contrast | Split at | Below the split | At or above | Difference (95% CI) | Cohen’s d | Separates? |
|---|---|---|---|---|---|---|
| Role count and the burnout composite | 4 or more roles | 1.87 (392 people) | 1.92 (203 people) | 0.05 [-0.09, 0.19] | 0.06 (negligible) | No |
This is the one known-groups check we ran, and it did not discriminate. If the burnout composite measured strain from juggling roles, people carrying four or more roles should score higher than people carrying fewer; they score 0.05 higher on a 0-to-5 scale, and the interval on that difference straddles zero. Read it as a null result, not as validation. It is also a weak test of the composite: section 10 of the report publishes role count as a non-predictor of almost everything, so a null here is what that finding predicts. A contrast that should separate is the better test, and this edition does not publish one.
This edition has limits, and they are worth stating plainly. The sample is a convenience sample of high-agency working people, N=602, and it does not represent the general working population. Most of it was recruited through a paid participant panel against the screen set out above, so the people here are, by construction, in full-time work, inside an age band, doing most of their work on a computer, already using AI at work weekly, and running something of their own. Every figure should be read as describing people like that. The design is cross-sectional: every figure in the report is an association measured at a single point in time, and nothing here establishes cause. The instrument collected no physiological, sleep, time-tracking or buffer-time measures, and it followed nobody over time, so any claim that would need those is attributed external research or it is not made. Neurodivergence is one self-report item asking how much ADHD, autism, dyslexia or similar impacts daily work, answered on a four-point scale from not applicable to significantly. It measures impact, not diagnosis.
Three groups are small enough to name. The burnout taxonomy defines nine profiles; the ninth, Quiet Struggle, matched a single respondent and is excluded from every published figure. Productive but Paying holds 15 respondents, so its numbers are directional rather than precise. So is the planning inversion: only 41 respondents never plan at all, and the gap between them and the inconsistent planners does not reach statistical significance.
Two definitions carry more weight than their wording suggests. The severity groups in section 2 are cumulative rather than a partition, and each one is a threshold rather than a label. In crisis means a burnout composite of 3.0 or above. In distress means severe on any one of three measures: a burnout composite of 3.0 or above, a quality of life of 1 or below, or daily overwhelm. Under strain means a quality of life of 2 or less on the 0 to 5 scale, or overwhelm almost daily. Each group contains the one before it, so the shares do not sum. Thriving is the apex profile, outside crisis and distress, though two of its 37 respondents fall under strain. The pain screener in the segment-definitions table above was collected at intake and is not used in this edition’s analysis, which runs on the burnout composite and the profiles it assigns.
Profile detail
| Profile | Overwhelm | Overcommit | Priority difficulty | External pressure | Replan | Wellbeing priority |
|---|---|---|---|---|---|---|
| The Replan Loop | 2.05 | 1.74 | 2.12 | 1.85 | 3.30 | 4.17 |
| The Holding Pattern | 2.46 | 2.07 | 2.12 | 2.07 | 2.62 | 4.11 |
| The Red Zone | 4.16 | 3.77 | 3.75 | 3.48 | 3.96 | 3.55 |
| The Slow Burn | 3.43 | 2.76 | 3.00 | 2.83 | 3.21 | 4.12 |
| The Volume Trap | 4.25 | 2.01 | 2.24 | 2.10 | 2.74 | 3.90 |
| The Pressure Cooker | 2.26 | 2.87 | 2.11 | 2.89 | 2.42 | 4.15 |
| The Benchmark | 1.62 | 1.24 | 1.27 | 1.27 | 1.62 | 4.35 |
| The Hidden Cost | 3.13 | 2.20 | 2.27 | 2.33 | 3.13 | 3.33 |