3 cumulative citations
View corpus contextTask-level analysis of 193,497 UK Civil Service vacancies finds wide variation in AI exposure even within identical job titles; LLM-driven redesigns point to augmentation and productivity gains—strategic leadership, complex problem-solving and stakeholder management remain human strengths rather than roles being broadly automated.
Citation observations
Cumulative provider counts captured on specific dates; providers are never combined.
The adoption of generative artificial intelligence (AI) is predicted to lead to fundamental shifts in the labour market, resulting in displacement or augmentation of AI-exposed roles. To investigate the impact of AI across a large organisation, we assessed AI exposure at the task level within roles at the UK Civil Service (UKCS). Using a novel dataset of UKCS job adverts, covering 193,497 vacancies over 6 years, our large language model (LLM)-driven analysis estimated AI exposure scores of 1,542,411 tasks. By aggregating AI exposure scores for tasks within each role, we calculated the mean and variance of job-level exposure to AI, highlighting the heterogeneous impacts of AI, even for seemingly identical jobs. We then use an LLM to redesign jobs, focusing on task automation, task optimisation, and task reallocation. We find that the redesign process leads to tasks where humans have comparative advantage over AI, including strategic leadership, complex problem resolution, and stakeholder management. Overall, automation and augmentation are expected to have nuanced effects across all levels of the organisational hierarchy. Most economic value of AI is expected to arise from productivity gains rather than role displacement. We contribute to the automation, augmentation and productivity debates as well as advance our understanding of job redesign in the age of AI.
Summary
Main Finding
Using a large LLM-driven task-level analysis of 193,497 UK Civil Service job adverts (1,542,411 extracted tasks), the authors show that AI’s labour-market impact is heterogeneous: most jobs are only partially exposed to generative AI, and the largest economic value will likely come from productivity gains via job redesign (reallocating freed-up time to higher‑value human tasks) rather than mass role displacement.
Key Points
- Dataset and scope
- Nearly all UK Civil Service (UKCS) job adverts from 16 Jan 2019 to 3 Dec 2024: 193,497 vacancies, 1,542,411 tasks.
- Covers 37 departments, 28 professions, and all grades; sample over‑represents mid/high grades (HEO, SEO, G6/G7).
- Task-level AI exposure
- Tasks were extracted from job adverts and each scored for AI exposure with an LLM.
- Aggregated to job-level by mean and variance of task exposures, exposing substantial heterogeneity even among similar job titles.
- Distribution of job exposure: >60% medium exposure, 20% low exposure (concentrated among senior roles — ~47% of SCS), ≈18% high exposure (could be fully automated).
- Job redesign via LLM
- For jobs not declared fully automatable, an LLM removed automatable tasks and proposed redesigned task mixes that reinvest saved time into activities where humans have comparative advantage.
- New task composition after redesign: 26% strategic leadership, 18% complex problem resolution, 17% stakeholder management/communication; declines in administrative support, records management, and some data‑processing tasks.
- Higher-grade roles skew toward leadership/complex problem solving; lower grades see more stakeholder/communication tasks.
- Thresholding and economic magnitudes
- Introducing a conservative automation threshold θ ≥ 80: ~145,864 jobs (≈75% of dataset) qualify for some degree of redesign rather than full automation.
- Estimated potential benefits for UKCS (under their model and assumptions): ~£5.2bn in productivity gains from redesigned roles and ~£1.1bn in cost reductions from roles dominated by highly automatable tasks.
- Robustness and context
- Authors vary the redesign process and thresholds; results are robust in that new tasks generally emphasize human‑comparative strengths.
- Results are context‑specific (public sector, advertised tasks) and depend on LLM judgments and assumptions about reinvestment of time.
Data & Methods
- Data sources
- GRID (Government Recruitment Information Database): job adverts (193,497 vacancies).
- UK Civil Service Statistics (UKCSS): departmental and workforce context; used to map grades, estimate median salaries by grade/department.
- Pipeline
- Task segmentation: extract discrete tasks from each job advert (LLM-assisted).
- AI exposure scoring: use an LLM to assign an exposure score to each task (1,542,411 tasks scored).
- Job aggregation: compute job-level mean and variance of task exposure; identify tasks above exposure thresholds.
- Thresholding: define a critical threshold θ above which tasks/jds are considered automatable; jobs with sufficient mass above θ are treated as fully automatable, otherwise eligible for redesign.
- Job redesign: apply an LLM to remove automatable tasks and propose replacements or reallocated tasks focused on human comparative advantage (leadership, complex problem solving, stakeholder engagement, risk/quality management).
- Economic simulation: combine redesigned task time reallocations with salary bands and staffing counts to estimate productivity gains and cost reductions.
- Adjustments and checks
- Iterative proportional fitting to adjust sample representation to the overall UKCS distribution by department/grade/profession.
- Robustness: alternative redesign constraints and variations in θ to test sensitivity.
- Limitations (methodological caveats noted by authors)
- Reliance on vacancy text (may differ from on‑the‑job tasks).
- Exposure and redesign depend on LLM judgments and prompt design.
- Salary/cost estimates approximate (using grade/department median salaries).
- Focused on one large public-sector employer; generalisability requires applying the pipeline to other datasets.
Implications for AI Economics
- Granularity matters: task-level measurement reveals large within-occupation heterogeneity. Aggregate/occupation-level exposure indices risk overstating or misstating displacement risk.
- Productivity vs displacement trade-off:
- Many exposed jobs are only partially automatable; firms and public organisations may capture most value by redesigning jobs and reallocating human effort toward complementary tasks, producing productivity gains rather than mass layoffs.
- Policy and firm strategies should prioritise redesign and reallocation mechanisms (training, workflow redesign, incentives) to capture gains.
- Labour reallocation and skill demand:
- Demand will shift toward human‑centric activities (leadership, complex problem solving, stakeholder management, risk/quality oversight). Upskilling and career‑path design should anticipate these shifts.
- Senior/expert roles appear less automatable on average; complementarities may amplify returns to expertise, affecting wage and inequality dynamics.
- Measurement and policy evaluation:
- LLMs can be used both as measurement tools and prescriptive designers for job redesign at scale, but their outputs need validation (field trials, task-time studies) before large policy or staffing decisions.
- Threshold choice (θ) for deciding between automation versus redesign is economically consequential; organisations should consider conservative thresholds and pilot trials to assess net welfare effects.
- Research agenda:
- Extend within‑firm, task-level analyses across sectors and countries to map heterogeneity in exposure and redesign potential.
- Empirically evaluate realized productivity gains from redesign (randomised or quasi-experimental deployments) to validate monetary estimates and measure spillovers (task creation, new occupations).
Assessment
Claims (8)
| Claim | Direction | Outcome | Confidence & Evidence | Details |
|---|---|---|---|---|
| We assessed AI exposure at the task level within roles at the UK Civil Service (UKCS). Automation Exposure | null_result | AI exposure scores at the task level |
Reading fidelity
high
Study strength
medium
|
not reported
|
| We used a novel dataset of UKCS job adverts covering 193,497 vacancies over 6 years. Other | null_result | number of job adverts / dataset coverage |
Reading fidelity
high
Study strength
high
|
n=193497
|
| Our LLM-driven analysis estimated AI exposure scores for 1,542,411 tasks. Automation Exposure | null_result | count of tasks with estimated AI exposure scores |
Reading fidelity
high
Study strength
medium
|
n=1542411
|
| Aggregating task-level AI exposure scores to the job-level reveals heterogeneity in AI exposure (mean and variance) even for seemingly identical jobs. Automation Exposure | mixed | mean and variance of job-level AI exposure |
Reading fidelity
high
Study strength
medium
|
not reported
|
| We used an LLM to redesign jobs, focusing on task automation, task optimisation, and task reallocation. Task Allocation | null_result | job redesign outputs (tasks reclassified into automation, optimisation, reallocation) |
Reading fidelity
high
Study strength
medium
|
not reported
|
| The redesign process produces tasks where humans have comparative advantage over AI, including strategic leadership, complex problem resolution, and stakeholder management. Task Allocation | positive | types/categories of tasks where humans retain comparative advantage |
Reading fidelity
medium
Study strength
medium
|
not reported
|
| Automation and augmentation are expected to have nuanced effects across all levels of the organisational hierarchy. Organizational Efficiency | mixed | heterogeneity of AI impacts across organisational levels |
Reading fidelity
high
Study strength
speculative
|
not reported
|
| Most economic value of AI is expected to arise from productivity gains rather than role displacement. Firm Productivity | positive | source of economic value from AI (productivity gains versus role displacement) |
Reading fidelity
medium
Study strength
speculative
|
not reported
|