VOLOVLOAutomation risk & transition, task by task
AskOccupationsMajorsBusinessFoundersChangesNotesMethod
Search occupations, majors…
EN
  • English
  • 简体中文
  • 日本語
  • Español
  • Português
  • Français
VLO
VOLO

Understanding how automation changes work — task by task, with the evidence shown and the uncertainty admitted.

AskOccupationsMajorsBusinessFoundersChangesNotesMethodWill AI replace my job?AboutRole diagnosisOpen dataOntologyVOLO ProPrivacyTerms
© 2026 VOLO
Guides
Ask VOLOFor businessFor foundersRecent changesNotesRole diagnosisFirst AI experimentVOLO ProMethod & evidenceAboutFollow an occupationSearch
Occupations
All occupations
AI / software
Translator / InterpreterBank tellerCopywriterContent moderatorCustomer service representativeAdministrative assistantSoftware tester / QA engineerData entry clerkGraphic designerParalegalVideo editorAccountant / BookkeeperMarketing specialistFrontend developerTax preparer / tax agentVoice actorMedical coderData analystInsurance claims handlerTechnical writer / documentation engineerJunior software developerInsurance underwriterHR / recruiterLoan officer / credit officerFinancial analystProcurement / supply chain specialistJournalistSales / account managerReal estate agentNetwork engineerIllustratorIT support specialist / helpdeskAuditorSupply chain plannerManagement consultantPhotographerBackend developerAI researcherProduct / UX designerBusiness systems ownerE-commerce operations specialistActuaryAnimator and VFX artistWriter and authorRadiologistData engineerProject managerData scientistCourt reporterEditor and proofreaderLawyerMedical assistant / clinic assistantQuantity surveyorContent creatorQuantitative analystMachine learning engineerExperienced software engineerDevOps / platform / SRE engineerBusiness analystMusician and composerProduct managerPharmacistPartnerships / channel managerSecurity analyst (SOC)Financial adviser / financial plannerInterpreterCompliance officerLibrarianArchitectFirst-line manager / team supervisorCounsellor / therapistRetail salesperson / shop assistantCivil / structural engineerElectrical engineerIndustrial engineerSecurity guardInterior designerUniversity lecturerBus driverInsurance agentMechanical engineerSchool teacherGeneral practitioner / primary care doctorAirline pilotWaiter / restaurant serverRadiographer / radiologic technologistAuto mechanic / vehicle technicianPolice officerSonographerAircraft maintenance technicianAir traffic controllerVeterinarianPhysiotherapist / rehabilitation therapistDentistSurgeonConstruction workerSocial workerRegistered nurseCare worker / nursing assistantPlumberAI implementation lead
RPA / self-service
Government service clerkOperations coordinatorMetro train driverReceptionist / front deskPharmacy technician
Robotics
Retail cashier / shop assistantContainer port workerWarehouse workerAssembly line workerMedical laboratory technicianActor and modelWelderChef / cookCleaner / janitorFarmerFirefighterCabin crewElectrician
Autonomous driving
Ride-hail / taxi driverTruck driverDelivery rider / courier
Majors
All majorsEnglish / Foreign languagesComputer scienceAccountingPsychologyJournalism / CommunicationFinanceLawVisual communication designMarketingNursingBusiness administrationEducation and teacher trainingArchitecturePublic administrationMedicineHospitality and tourism managementEconomicsInformation systems
Enter as:I have a jobI am studyingI run a companyI am building something
Recent changes›Data scientist›2024-02-27
CapabilityCognitive automation2024-02-27

The strongest model reached 58% accuracy on statistical and causal questions from textbooks and papers, and models struggled to use causal knowledge and data together

Data scientistoccupation page →
Event date / reported
2024-02-27 · reported 2024-06-09
Evidence stage
CapabilityA demo, benchmark or paper shows the task can be done. Updates what the technology can do — not what employers will do.
Tasks this bears on
Designing experiments and reading causes
Setting up the A/B test or the study so the answer can be trusted, and saying whether a change caused the result or merely came with it.
Still human-led≈ Platform inference
Where this applies
An academic benchmark of 411 questions with data sheets from textbooks, online learning materials and academic papers, testing statistical and causal reasoning. It reports that the strongest model, GPT-4, achieves 58% accuracy, and that models have difficulty with data analysis and causal reasoning and struggle to use causal knowledge and provided data simultaneously. It tests 2024-era models on textbook-style questions, not experiments on real products.
What this means
Telling cause from coincidence is where the models tested were weakest — and it is the question most decisions rest on. That part of the data scientist's work is the least assisted by the tools so far.
What it does not yet show
A 2024 benchmark on textbook questions; newer models may do better, and it does not measure real experiments.
What you can check
Open arXiv 2402.17644 (QRData) and find "The strongest model GPT-4 achieves an accuracy of 58%".
Does it change the assessment?
No — and this stage does not move it either. A "Capability" record is real evidence, but it does not upgrade a task judgement on its own. The 1 linked judgement above stand where they were.
Source
Liu et al. — Are LLMs Capable of Data-based Statistical and Causal Reasoning? Benchmarking Advanced Quantitative Reasoning with Data (QRData), arXiv 2402.17644 (submitted 27 Feb 2024; Findings of ACL 2024) · verified 2026-09-29 · Claude (VOLO agent) · interpreted 2026-09-29 · Claude (VOLO agent)
Primary source — published by the party that did this, or the authority of record. No co-signature needed.
All changes for Data scientist →All recent changes →How events become evidence →