VOLOVLOAutomation risk & transition, task by task
AskOccupationsMajorsBusinessFoundersChangesNotesMethod
Search occupations, majors…
EN
  • English
  • 简体中文
  • 日本語
  • Español
  • Português
  • Français
VLO
VOLO

Understanding how automation changes work — task by task, with the evidence shown and the uncertainty admitted.

AskOccupationsMajorsBusinessFoundersChangesNotesMethodWill AI replace my job?AboutRole diagnosisOpen dataOntologyVOLO ProPrivacyTerms
© 2026 VOLO
Guides
Ask VOLOFor businessFor foundersRecent changesNotesRole diagnosisFirst AI experimentVOLO ProMethod & evidenceAboutFollow an occupationSearch
Occupations
All occupations
AI / software
Translator / InterpreterBank tellerCopywriterContent moderatorCustomer service representativeAdministrative assistantSoftware tester / QA engineerData entry clerkGraphic designerParalegalVideo editorAccountant / BookkeeperMarketing specialistFrontend developerTax preparer / tax agentVoice actorMedical coderData analystInsurance claims handlerTechnical writer / documentation engineerJunior software developerInsurance underwriterHR / recruiterLoan officer / credit officerFinancial analystProcurement / supply chain specialistJournalistSales / account managerReal estate agentNetwork engineerIllustratorTravel agent / advisorIT support specialist / helpdeskAuditorSupply chain plannerManagement consultantPhotographerBackend developerAI researcherProduct / UX designerBusiness systems ownerE-commerce operations specialistActuaryAnimator and VFX artistWriter and authorRadiologistData engineerProject managerData scientistCourt reporterEditor and proofreaderLawyerMedical assistant / clinic assistantQuantity surveyorContent creatorQuantitative analystMachine learning engineerExperienced software engineerDevOps / platform / SRE engineerBusiness analystMusician and composerGIS analyst / cartographerProduct managerDietitian / nutritionistPharmacistPartnerships / channel managerSecurity analyst (SOC)Financial adviser / financial plannerInterpreterPathologistChip design engineerCompliance officerLibrarianArchitectFirst-line manager / team supervisorCounsellor / therapistRetail salesperson / shop assistantCivil / structural engineerElectrical engineerIndustrial engineerSecurity guardInterior designerUniversity lecturerPsychologistBus driverInsurance agentMechanical engineerFashion designerSchool teacherGeneral practitioner / primary care doctorAirline pilotDriving instructorWaiter / restaurant serverRadiographer / radiologic technologistAuto mechanic / vehicle technicianPolice officerSonographerAircraft maintenance technicianDental hygienistAnaesthesiologist / anaesthetistAir traffic controllerVeterinarianPhysiotherapist / rehabilitation therapistDentistSurgeonConstruction workerSocial workerJudge / magistrateRegistered nurseCare worker / nursing assistantPlumberAI implementation lead
RPA / self-service
Government service clerkOperations coordinatorMetro train driverReceptionist / front deskPharmacy technician
Robotics
Retail cashier / shop assistantContainer port workerWarehouse workerAssembly line workerMedical laboratory technicianActor and modelWelderChef / cookCleaner / janitorFarmerFirefighterHVAC technicianCabin crewElectrician
Autonomous driving
Ride-hail / taxi driverTruck driverDelivery rider / courier
Majors
All majorsEnglish / Foreign languagesComputer scienceAccountingPsychologyJournalism / CommunicationFinanceLawVisual communication designMarketingNursingBusiness administrationEducation and teacher trainingArchitecturePublic administrationMedicineHospitality and tourism managementEconomicsInformation systems
Enter as:I have a jobI am studyingI run a companyI am building something
Recent changes›Judge / magistrate›2024-01-02
CapabilityCognitive automation2024-01-02

General-purpose language models were wrong about random federal court cases between 58% and 88% of the time, researchers found

Judge / magistrateoccupation page →
Event date / reported
2024-01-02
Evidence stage
CapabilityA demo, benchmark or paper shows the task can be done. Updates what the technology can do — not what employers will do.
Tasks this bears on
Researching the law
Finding and reading the statutes, precedents and authorities a case turns on.
Being augmented✓ Evidence-backed
Where this applies
A study of general-purpose language models asked specific, verifiable questions about random federal court cases, finding legal hallucinations between 58% of the time with ChatGPT 4 and 88% with Llama 2, and that models often fail to correct a user's incorrect legal assumptions. Older models from 2023, from the same research group as the study of legal research tools on this page.
What this means
General chatbots were worse than specialised tools at the facts of cases — a reason judiciaries warn against public chatbots.
What it does not yet show
Older models; the same lab as the other study here, so not independent of it.
What you can check
Open arXiv 2401.01301 and find "between 58% of the time with ChatGPT 4 and 88% with Llama 2".
Does it change the assessment?
No — and this stage does not move it either. A "Capability" record is real evidence, but it does not upgrade a task judgement on its own. The 1 linked judgement above stand where they were.
Source
Dahl et al. — Large Legal Fictions: Profiling Legal Hallucinations in Large Language Models (arXiv 2401.01301, submitted 2 Jan 2024) · verified 2026-09-30 · Claude (VOLO agent) · interpreted 2026-09-30 · Claude (VOLO agent)
Primary source — published by the party that did this, or the authority of record. No co-signature needed.
All changes for Judge / magistrate →All recent changes →How events become evidence →