Skip to content

An Evaluation of Epic EHR AI Outpatient Chart Summarization

Epic Generative Artificial Intelligence Chart Summarization Tool to Reduce Ambulatory Provider Cognitive Task Load: A Randomized Controlled Trial

Status
Active, not recruiting
Phases
Unknown
Study type
Interventional
Source
ClinicalTrials.gov
Registry ID
NCT07438743
Enrollment
284
Registered
2026-02-27
Start date
2026-02-23
Completion date
2026-05-01
Last updated
2026-03-06

For informational purposes only — not medical advice. Sourced from public registries and may not reflect the latest updates. Terms

Conditions

Quality Improvement

Keywords

Physician Task Load, Professional Fulfillment, Artificial Intelligence, Chart Summarization, System Usability

Brief summary

This is a RCT of 284 outpatient physicians at a large academic health system, randomized 1:1 to an electronic health record (EHR) produced generative AI outpatient chart summarization tool or a usual-care control group. The 90 day study will observe the effects of the tool prior to system-wide roll out of the tool.

Detailed description

The primary aim of this study is to evaluate the impact of an EHR developed generative AI outpatient chart summarization tool on self-reported physician-task load score (PTL), comparing the tool to a control group. Exploratory outcomes include EHR-derived time metrics (Caboodle and Signal), Professional fulfillment Index (PFI), usability (SUS), provider satisfaction and productivity, and patient experience item results from CG-CAHPS. We will also evaluate whether AI literacy modifies adoption and effect of the tool using the short-form Meta AI Literacy Scale (MAILS). On an exploratory basis, we will also perform adjustments based on provider specialty, access to an ambient-listening AI scribe, panel complexity, provider age group, provider sex, and time-varying effects by month over the study period. Enrolled participants are randomized to one of two groups. Randomization will be stratified by whether the participant has an active AI scribe license, and covariate-constrained randomization will be performed within strata to improve balance on baseline PTL (NASA-TLX-adapted score) and a modified baseline chart review time (Caboodle-derived). Due to the nature of the intervention, participants cannot be blinded to group assignment. The primary purpose of the initiative is to improve quality, efficiency, and business operations at University of California, Los Angeles (UCLA) Health and will inform the operational implementation of the tool across all providers within the UCLA Health System. Nevertheless, the UCLA study team plans to rigorously examine and publish the impact of this intervention across the health system, which is why the study team pre-registered the initiative.

Interventions

OTHERGenAI Chart Summarization

Epic's generative AI chart summarization tool summarizes a subset of a patient's notes. Use of the tool is optional and intended solely to provide a summary for providers and does not provide clinical decision support. The system automatically selects recent notes or a provider can manually select specific notes of interest. The number of notes summarized is limited by the character constraints of the EHR, 24,000 English characters or 30 notes. The system uses AI to generate a short summary of relevant information. The summaries are meant to be used as a tool to aid providers and are not intended to be placed in clinical notes. The summaries created are currently not stored in the patient's chart.

Sponsors

University of California, Los Angeles
Lead SponsorOTHER

Study design

Allocation
RANDOMIZED
Intervention model
PARALLEL
Primary purpose
HEALTH_SERVICES_RESEARCH
Masking
SINGLE (Investigator)

Eligibility

Sex/Gender
ALL
Healthy volunteers
No

Inclusion criteria

* Ambulatory care providers within the UCLA Health system including physicians and advanced practice providers (APPs), such as nurse practitioners and physician assistants with at least one half-day clinic session per week. * Providers complete baseline pre-survey

Exclusion criteria

• Trainee providers (e.g., residents, medical students), and psychologists

Design outcomes

Primary

MeasureTime frameDescription
Change from Baseline Physician Task LoadBaseline and 90 days after initial exposure to the interventionPhysician task load adapted from the NASA Task Load Index (TLX), a validated tool for assessing EHR-related cognitive task load in four sub-scales (mental demand, temporal demand, physical demand, and effort). This outcome is adapted to capture the task of pre-charting, defined for this study as the practice of reviewing patient information in the EHR before a patient visit to prepare for the encounter. Each sub-scale is rated from 0 (low) to 100 (high) and is aggregated to a 0-400 point scale. No patient level information will be collected for this outcome measure.

Secondary

MeasureTime frameDescription
Change in Modified Total Chart Time Per EncounterBaseline, after 60 days of exposure to the intervention, and after 90 days of exposure to the intervention.Using Caboodle, Epic's enterprise data warehouse, we will use a customized metric to measure clinician time spent reviewing the patient's chart. Based on internal validation where clinicians could access the chart summarization tool, the metric includes Caboodle Tier 1 activities corresponding to Clinical Review (all activities), Documentation limited to pre-charting activity only, and Other limited to Navigator-related activity, as well as the activity dedicated to the chart summarization tool. All note-writing, order entry, and encounter-signing activities are excluded. Time post checkout is also excluded. Change will be assessed relative to a 6-month retrospective baseline. No patient level information will be collected for this outcome measure.
Change from Baseline Professional Fulfillment Index ScoreBaseline and 90 days after initial exposure to the interventionThe Professional Fulfillment Index (PFI) is a validated 16-item instrument that uses a 5-point Likert scale (0-4) to measure professional fulfillment, work exhaustion, and interpersonal disengagement. Burnout is reported based on combined results of the work exhaustion and interpersonal disengagement subscales. A higher score indicates greater level of exhaustion. No patient level information will be collected for this outcome measure.
Change from Baseline Self-Reported Pre-Charting EffectivenessBaseline and 90 days from initial exposure to the interventionProviders will answer questions regarding self-reported effectiveness and efficiency in pre-charting. The questions use a 5-point Likert scale to measure these metrics, with a higher score indicating greater self-reported efficiency and effectiveness.
Provider Satisfaction Scores90 days after initial exposure to the interventionSelf-reported satisfaction survey that includes physician reported effect of chart summarization tool on pre-charting efficiency and effectiveness, and other potential unintended consequences. No patient level information will be collected for this outcome.
System Usability Scale90 days from initial exposure to the interventionThe system usability scale (SUS) is a ten-item questionnaire that uses a 5-point Likert scale to measure different aspects of system usability. The total scale ranges from 0-100. A higher SUS score indicates higher usability.
Safety EventsOver the 90 day period following after initial exposure to the interventionClinician-reported safety of AI-generated chart summaries, including perceived frequency of clinically significant errors and occurrence of major safety events during the intervention period. No patient level information will be collected for this outcome.
Qualitative Tool-Specific EHR FeedbackOver the 90 days following initial exposure to the interventionAggregated real-time feedback submitted by clinicians through the EHR- native summarization interface during the intervention period, including ratings of accuracy, completeness, and hallucinations. No patient level information will be collected for this outcome.
Change from Baseline Consumer Assessment of Healthcare Providers and Systems Clinician & Group Survey (CG-CAHPS) MetricBaseline and 90 days after initial exposure to the interventionThe Consumer Assessment of Healthcare Providers and Systems Clinician \& Group Survey (CG-CAHPS) is a standardized patient feedback survey measuring experiences with providers. We will examine changes in the CG-CAHPS mean scores compared to 6 months prior to the intervention for the question "In the last 6 months, how often did this doctor seem to know the important information about your medical history".
Change from Baseline Epic Signal (Activity) Metric: Time Outside Scheduled HoursBaseline and 90 days after the initial exposure to the interventionWe will examine change from a retrospective baseline 6 months prior to enrollment in Signal metrics including time outside scheduled hours per scheduled day. Using this data will determine how a provider's time is utilized in the EHR. No patient level information will be collected for this outcome measure.
Clinician Relative Value Units (RVUs) per Week60 days after exposure to the intervention, and 90 days after exposure to the interventionAverage relative value units (RVUs) per week during intervention months 2 and 3, adjusted for baseline (6-month pre-intervention period)

Countries

United States

Contacts

PRINCIPAL_INVESTIGATORJohn N Mafi, MD, MPH

Division of General Internal Medicine & Health Services Research, David Geffen School of Medicine at the University of California, Los Angeles

PRINCIPAL_INVESTIGATORPaul J Lukac, MD, MBA, MS

UCLA Health Information Technology, UCLA Health

Outcome results

None listed

Source: ClinicalTrials.gov · Data processed: Mar 7, 2026