VOLOVLOAutomation risk & transition, task by task
AskOccupationsMajorsBusinessFoundersChangesNotesMethod
Search occupations, majors…
EN
  • English
  • 简体中文
  • 日本語
  • Español
  • Português
  • Français
VLO
VOLO

Understanding how automation changes work — task by task, with the evidence shown and the uncertainty admitted.

AskOccupationsMajorsBusinessFoundersChangesNotesMethodWill AI replace my job?AboutRole diagnosisOpen dataOntologyVOLO ProPrivacyTerms
© 2026 VOLO
Occupations
All occupations
AI / software
Translator / InterpreterBank tellerCopywriterContent moderatorCustomer service representativeAdministrative assistantSoftware tester / QA engineerData entry clerkGraphic designerParalegalVideo editorAccountant / BookkeeperMarketing specialistFrontend developerTax preparer / tax agentVoice actorMedical coderData analystInsurance claims handlerTechnical writer / documentation engineerJunior software developerInsurance underwriterHR / recruiterLoan officer / credit officerFinancial analystProcurement / supply chain specialistJournalistSales / account managerReal estate agentNetwork engineerIllustratorIT support specialist / helpdeskAuditorSupply chain plannerManagement consultantPhotographerBackend developerAI researcherProduct / UX designerBusiness systems ownerE-commerce operations specialistActuaryAnimator and VFX artistWriter and authorRadiologistData engineerProject managerData scientistCourt reporterEditor and proofreaderLawyerMedical assistant / clinic assistantQuantity surveyorContent creatorQuantitative analystMachine learning engineerExperienced software engineerDevOps / platform / SRE engineerBusiness analystMusician and composerProduct managerPharmacistPartnerships / channel managerSecurity analyst (SOC)Financial adviser / financial plannerInterpreterCompliance officerLibrarianArchitectFirst-line manager / team supervisorCounsellor / therapistRetail salesperson / shop assistantCivil / structural engineerElectrical engineerIndustrial engineerSecurity guardInterior designerUniversity lecturerBus driverInsurance agentMechanical engineerSchool teacherGeneral practitioner / primary care doctorAirline pilotWaiter / restaurant serverRadiographer / radiologic technologistAuto mechanic / vehicle technicianPolice officerSonographerAircraft maintenance technicianAir traffic controllerVeterinarianPhysiotherapist / rehabilitation therapistDentistSurgeonConstruction workerSocial workerRegistered nurseCare worker / nursing assistantPlumberAI implementation lead
RPA / self-service
Government service clerkOperations coordinatorMetro train driverReceptionist / front deskPharmacy technician
Robotics
Retail cashier / shop assistantContainer port workerWarehouse workerAssembly line workerMedical laboratory technicianActor and modelWelderChef / cookCleaner / janitorFarmerFirefighterCabin crewElectrician
Autonomous driving
Ride-hail / taxi driverTruck driverDelivery rider / courier
Majors
All majorsEnglish / Foreign languagesComputer scienceAccountingPsychologyJournalism / CommunicationFinanceLawVisual communication designMarketingNursingBusiness administrationEducation and teacher trainingArchitecturePublic administrationMedicineHospitality and tourism managementEconomicsInformation systems
Guides
Ask VOLOFor businessFor foundersRecent changesNotesRole diagnosisFirst AI experimentVOLO ProMethod & evidenceAboutFollow an occupationSearch
You are reading as:I have a jobI am studyingI run a companyI am building something
On this pageWhat is happeningTask breakdownRecent changesFor youWhat it means for youWhat you can doHow we knowWhat these rest on
Occupations›Data scientist

Get told when a verified record lands on this occupation → · Mark which of these tasks are yours (VOLO Pro, free during the launch) →

Data scientist

Turns a business or research question into something data can answer: frames it, cleans and explores the data, builds and validates statistical or predictive models, designs and reads experiments, and explains to decision-makers what the result does and does not show. AI agents can now write much of the analysis code, but on public benchmarks the best still solve about a third of realistic data-analysis tasks, and in causal questions they struggle most. Banking supervisors still expect models to be challenged by independent people, and US projections expect data science jobs to grow.

Technology, finance & researchAssessed 2026-09-30
Tasks automating
0of 6
2 being augmented
Still human-led
3of 6
1 new task
Evidence-backed judgements
2of 6
7 verified records
Test this week · first of 3 directions

Take one recent analysis you did and write down whether it showed cause or only association, and what experiment would have settled it.

See all 3 ↓
50/100
Automation impact indexLow confidence

This is not a probability of losing your job. It combines how much of the role's task load is exposed to automation with how far adoption has actually gone — useful for comparing occupations on one consistent basis, and for nothing else.

Where this applies

Written for data scientists who build models and run experiments to inform decisions — in technology companies, banks and insurers, research and the public sector. Analysts working mainly in SQL and dashboards have their own page, as do engineers who build and run learned systems in production. The evidence is benchmarks of AI agents on data-science tasks, banking supervisors' model-risk guidance, rules on auditing automated decisions, and US labour projections; it establishes what agents can do on test tasks and what rules expect of people, not how data scientists' time has changed.

What is happening

What is actually changing#

The unit of analysis is the task, not the job title. A role is not replaced — its task mix shifts.

AutomatingBeing augmentedStill human-ledNew taskStriped: our inference, not yet backed by a verified record

Each tile is one task. Its size is how much of the job it is; its colour is where the task is heading. Click a tile to see what the judgement does not establish.

Core task
Framing the question and the metric
Still human-led≈ Platform inference
What this does NOT mean

This rests on the nature of the work and on O*NET's candidate task list; no record here measures how framing is done or who does it.

Read this task in full →
Significant task
Cleaning and exploring the data
Being augmented✓ Evidence-backed
What this does NOT mean

Benchmark tasks are not a company's data, and the EU duties for high-risk systems apply later; nothing here measures time spent on cleaning.

Read this task in full →Make this your first AI experiment at work →
Core task
Building and validating the model
Being augmented≈ Platform inference
What this does NOT mean

Benchmarks date quickly and test short tasks; they do not show how models are built on real projects or how much of the work agents now do.

Read this task in full →Make this your first AI experiment at work →
Significant task
Designing experiments and reading causes
Still human-led≈ Platform inference
What this does NOT mean

The benchmark tested 2024-era models on textbook questions; newer models may do better, and it does not measure experiments on real products.

Read this task in full →
Core task
Explaining what it means, and what it does not
Still human-led≈ Platform inference
What this does NOT mean

A projection is not a result, and it counts jobs, not which tasks those jobs involve.

Read this task in full →
Significant task
Challenging and auditing models
New task✓ Evidence-backed
What this does NOT mean

The banking guidance is not enforceable and covers large banks; the New York rule covers hiring tools in one city. Neither measures how many people do this work.

Read this task in full →

Is this your job? Say so and this page narrows to your share of it.

A job title is a bundle of tasks bought together, and no two people hold the same bundle. Nothing is sent anywhere — it stays in this browser.

Framing the question and the metricStill human-led≈ Platform inferenceCleaning and exploring the dataBeing augmented✓ Evidence-backedBuilding and validating the modelBeing augmented≈ Platform inferenceDesigning experiments and reading causesStill human-led≈ Platform inferenceExplaining what it means, and what it does notStill human-led≈ Platform inferenceChallenging and auditing modelsNew task✓ Evidence-backed

Read all 6 tasks in full — direction, reasoning and limits →

Recent changes#

2023202420252026today2023-04-06 · Policy mandateEmployers may not use an automated hiring tool unless it had a bias audit within the past year by an auditor independent of the tool, under New York City's rules2024-02-27 · CapabilityThe strongest model reached 58% accuracy on statistical and causal questions from textbooks and papers, and models struggled to use causal knowledge and data together2024-06-13 · ConstraintData for high-risk AI must follow governance practices covering cleaning and bias checks, and such systems must be effectively overseeable by people, under the EU's AI Act2024-09-12 · CapabilityThe best AI agent solved only 34.12% of realistic data-analysis tasks on DSBench, a benchmark of 466 analysis and 74 modelling tasks from real competitions2024-10-09 · CapabilityThe best language models reached only 30.5% accuracy on DA-Code, a benchmark of agent-based data-science coding tasks2026-04-17 · ConstraintModels should face effective challenge from people with the expertise, independence and standing to change them, US bank supervisors say, while leaving generative and agentic AI out of scope2026-08-27 · ForecastData scientist employment will grow 35 percent from 2025 to 2035, much faster than average, with about 24,800 openings a year, the US Bureau of Labor Statistics estimated
Can move a judgementCannot move one (forecast, capability demo…)
Forecast2026-08-27Verified 2026-09-29
Data scientist employment will grow 35 percent from 2025 to 2035, much faster than average, with about 24,800 openings a year, the US Bureau of Labor Statistics estimated

United States. The statistics bureau's occupational outlook projects that employment of data scientists will grow 35 percent from 2025 to 2035, much faster than the average for all occupations, with about 24,800 openings a year on average over the decade. It is a projection for one country, made while firms integrate AI systems into their work; it counts jobs, not which tasks those jobs involve.

A named person with standing publicly predicted something, on a date, in an attributable statement. It is recorded so that who said what, and when, stays checkable — and it never moves a task's assessment, because a prediction is not an observation. Its value arrives later: the record sits on the same page as the evidence about that occupation, so anyone reading the forecast reads the record of what happened next beside it. That is the reckoning; this site publishes no verdict on whether a forecast came true.

U.S. Bureau of Labor Statistics — Occupational Outlook Handbook: Data Scientists (Last modified date: August 27, 2026) ↗Full impact card →
Constraint2026-04-17Verified 2026-09-29
Models should face effective challenge from people with the expertise, independence and standing to change them, US bank supervisors say, while leaving generative and agentic AI out of scope

United States. The banking agencies' revised model-risk guidance replaces SR 11-7 of 2011. It says effective challenge is performed by individuals with the appropriate expertise to conduct a critical and objective challenge, sufficient independence to maintain objectivity, and the organisational standing and influence to effect change; it states that generative AI and agentic AI models are not within its scope; and it says the guidance does not set enforceable standards, so non-compliance will not result in supervisory criticism. It is most relevant to banking organisations with over $30 billion in total assets.

Failure, rollback, regulation or cost is suppressing adoption. Can lower an assessment or widen its uncertainty.

Board of Governors of the Federal Reserve System, OCC and FDIC — SR 26-2, Revised Guidance on Model Risk Management (April 17, 2026), superseding SR 11-7 ↗Full impact card →
Capability2024-10-09Verified 2026-09-29
The best language models reached only 30.5% accuracy on DA-Code, a benchmark of agent-based data-science coding tasks

An academic benchmark of data-science coding tasks that require an agent to work with real data and write code, including wrangling and analysis. It reports that, with its baseline agent, the current best language models achieve only 30.5% accuracy, leaving ample room for improvement. It tests 2024-era models on designed tasks.

A demo, benchmark or paper shows the task can be done. Updates what the technology can do — not what employers will do.

Huang et al. — DA-Code: Agent Data Science Code Generation Benchmark for Large Language Models, arXiv 2410.07331 (submitted 9 Oct 2024; EMNLP 2024) ↗Full impact card →
Capability2024-09-12Verified 2026-09-29
The best AI agent solved only 34.12% of realistic data-analysis tasks on DSBench, a benchmark of 466 analysis and 74 modelling tasks from real competitions

A benchmark of 466 data-analysis tasks and 74 data-modelling tasks drawn from Eloquence and Kaggle competitions, used to test AI agents built on models such as GPT-4o, Claude and Gemini. It reports that the best agent solved only 34.12% of the data-analysis tasks and reached a 34.74% relative performance gap on modelling tasks. It tests models available in 2024–2025 on competition tasks, not work on a company's data; one author's employer develops AI models.

A demo, benchmark or paper shows the task can be done. Updates what the technology can do — not what employers will do.

Jing et al. (UT Dallas, Tencent AI Lab, USC) — DSBench: How Far Are Data Science Agents from Becoming Data Science Experts?, arXiv 2409.07703 (submitted 12 Sep 2024; revised 11 Apr 2025) ↗Full impact card →
Constraint2024-06-13Verified 2026-09-29
Data for high-risk AI must follow governance practices covering cleaning and bias checks, and such systems must be effectively overseeable by people, under the EU's AI Act

European Union. Article 10 of the AI Act says training, validation and testing data sets for high-risk AI systems shall be subject to data governance and management practices, including data preparation such as annotation, labelling and cleaning, and examination in view of possible biases; Article 14 says high-risk systems shall be designed so that they can be effectively overseen by natural persons. The obligations bind providers of high-risk systems; under the 2026 amending regulation, the high-risk obligations for systems listed in Annex III apply from 2 December 2027.

Failure, rollback, regulation or cost is suppressing adoption. Can lower an assessment or widen its uncertainty.

European Parliament and Council — Regulation (EU) 2024/1689 (Artificial Intelligence Act), Articles 10 and 14, consolidated text of 27 July 2026 ↗Full impact card →
Capability2024-02-27Verified 2026-09-29
The strongest model reached 58% accuracy on statistical and causal questions from textbooks and papers, and models struggled to use causal knowledge and data together

An academic benchmark of 411 questions with data sheets from textbooks, online learning materials and academic papers, testing statistical and causal reasoning. It reports that the strongest model, GPT-4, achieves 58% accuracy, and that models have difficulty with data analysis and causal reasoning and struggle to use causal knowledge and provided data simultaneously. It tests 2024-era models on textbook-style questions, not experiments on real products.

A demo, benchmark or paper shows the task can be done. Updates what the technology can do — not what employers will do.

Liu et al. — Are LLMs Capable of Data-based Statistical and Causal Reasoning? Benchmarking Advanced Quantitative Reasoning with Data (QRData), arXiv 2402.17644 (submitted 27 Feb 2024; Findings of ACL 2024) ↗Full impact card →
Policy mandate2023-04-06Verified 2026-09-29
Employers may not use an automated hiring tool unless it had a bias audit within the past year by an auditor independent of the tool, under New York City's rules

New York City. The final rule implementing the city's law on automated employment decision tools says an employer or employment agency may not use or continue to use such a tool if more than one year has passed since its most recent bias audit; that a bias audit must at a minimum calculate the selection rate and impact ratio for each category; and that an auditor is not independent if it is or was involved in using, developing or distributing the tool. The tools it covers include machine learning, statistical modelling and data analytics. It applies to hiring and promotion in one city.

Regulation, subsidy or public procurement is requiring or funding adoption — the mirror of a constraint. It shows adoption is being required, not that it has happened, so one mandate is never enough on its own; two independent ones are.

New York City Department of Consumer and Worker Protection — Notice of Adoption of Final Rule, Use of Automated Employment Decision Tools (implementing Local Law 144 of 2021; effective July 5, 2023) ↗Full impact card →
What it means for you

What this means for you#

If you are starting out

If you are starting out, the parts of the job an AI agent does best — writing routine analysis and modelling code — are the parts junior roles used to be built on. Build your value where the evidence says agents are weakest: framing the question, designing experiments and reasoning about causes, and explaining results to people who must decide.

If you are experienced

Expect to review and direct generated code more than write it, and expect more of your time to go to validation, documentation and audits as rules on automated decisions spread. Your standing to challenge a model — and to be believed when you say where it stops — is the part the rules are written around.

Your options#

Four directions, each with its real constraints and one thing you can test this week. Continuing as you are is a legitimate choice — it just has to be a chosen one.

Stay and strengthen

Stay a data scientist, and move towards experiments and causal work

Causal reasoning and experiment design are where the benchmarks show models struggling most, and decisions depend on them.

Real constraints

Experimentation roles sit mainly in large product companies and research; smaller firms may not run enough tests to need one.

Test this week

Take one recent analysis you did and write down whether it showed cause or only association, and what experiment would have settled it.

Reshape the role

Take on model validation and audit

Banking supervisors expect independent challenge of models and New York City requires independent bias audits of hiring tools; the work needs people who understand models and are not the ones who built them.

Real constraints

Independence means not auditing your own team's models, which may mean moving teams or employers; the banking guidance is not binding.

Test this week

Find out whether your organisation has a model validation or model risk function, and what it checks before a model goes live.

Adjacent move

Move towards decision support: product or policy analytics

Explaining what a result means to the people deciding is the part of the job least touched by automation in the evidence here.

Real constraints

These roles reward domain knowledge and communication over technical depth, and pay can differ.

Test this week

Ask one decision-maker you work with which of your recent results changed a decision, and why.

Common questions#

Will AI replace data scientists?

Not on present evidence. AI agents can write much of the analysis code, but on public benchmarks the best still solve about a third of realistic data-analysis tasks, and they are weakest at causal reasoning. Rules on models expect independent people to challenge and audit them, and US projections expect data scientist jobs to grow. What changes is the mix: less hand-written code, more framing, checking and explaining.

How long do I have before this job disappears?

We do not answer that with a number of years. There is a signal you can watch instead: whether AI agents start matching people on causal and experimental questions in benchmarks, and whether rules on models stop requiring an independent person to challenge them. The first is not true today; the second is the opposite of the current direction.

How well can AI agents do data science today?

On DSBench, a benchmark of 466 data-analysis and 74 data-modelling tasks from real competitions, the best agent solved about 34% of the analysis tasks; on DA-Code the best models reached about 30.5% accuracy; and on QRData, a benchmark of statistical and causal questions, the strongest model reached 58%. Benchmarks date quickly, but they show where agents are strongest and where they struggle.

Is data science still a good career with AI?

US projections expect data scientist employment to grow 35 percent from 2025 to 2035, far faster than average. The work is shifting towards framing questions, designing experiments, validating and auditing models and explaining results — the parts where agents are weakest and rules expect a person.

How we know

What these judgements rest on#

2 of 6 task judgements on this page are backed by a verified event and 4 are platform inference, each labelled where it appears. Behind them sit 2 technology dimensions, a reconstructed trajectory since language models reached the public, and 7 verified events.

See which technologies, how it got here, and the method →

Where it sits in the official classification: skills, knowledge, related jobs →

Other roles in the same function#

A company divides its work into functions before it divides it into jobs. These sit in Technology & data alongside this one — a fact about org charts, not a judgement that they are similar or that they are changing in the same direction.

Junior software developer · Experienced software engineer · Frontend developer · Backend developer · Data engineer · Machine learning engineer · AI researcher · Software tester / QA engineer · Data analyst · Business analyst · Business systems owner · IT support specialist / helpdesk · Security analyst (SOC) · DevOps / platform / SRE engineer · Network engineer · Technical writer / documentation engineer