Understanding Decision-Making Capacity in Clinical Practice
Every clinician who has ever dealt with a seriously ill patient eventually runs into the question of whether that patient can actually understand the treatment options and make an informed choice. It is not always straightforward. A patient might clearly be able to do their taxes while simultaneously being unable to weigh the risks of a chemotherapy regimen. That disconnect is exactly why tools like the MacArthur Competence Assessment Tool exist. I have used the MacCAT-T (the therapeutic version) in hospital settings where capacity determinations were medically necessary before proceeding with high-risk interventions. The process itself takes roughly 20 to 30 minutes once you are familiar with the scoring. Most clinicians I know end up needing about two or three practice cases before they stop second-guessing their interpretations.
Macarthur Competence Assessment Tool Overview
The MacCAT was originally developed by researchers including Ronald Appelbaum and Thomas Grisso at MacArthur Foundation. The tool measures four core abilities that together define decision-making capacity: understanding, appreciation, reasoning, and expressing a choice. Each domain is scored separately, and there is no single cut-off score that declares someone competent. That deliberate vagueness actually reflects how real clinical judgment works. You are assessing function, not testing facts. What most people miss on first use is the distinction between understanding and appreciation. Understanding asks whether the patient can restate the information in their own words. Appreciation asks whether they believe it applies to their situation. A patient can perfectly recite the risks of a blood transfusion while still denying that they are anemic. That gap matters, and the scoring manual makes this distinction explicit. The original MacCAT-CV (for voluntary treatment decisions) contains 11 scenarios that are generic enough to use across different conditions. The MacCAT-T is designed for therapeutic contexts like psychiatric hospitalization or medication changes. Both versions share the same scoring structure, which means switching between them usually requires minimal retraining.
How the Scoring Actually Works
The instrument uses a combination of open-ended questions and forced-choice items. For the understanding domain, you read a short passage about the proposed treatment, then ask the patient to explain it back. The scorer rates each key element from 0 to 2. Two means the patient captured the essence without distortion. Zero means the response was irrelevant or clearly inaccurate. For reasoning, you present a pros-cons exercise. The patient needs to connect their personal values to the outcome they would choose. This is where I ran into trouble early on. One of my first cases involved a patient who kept circling back to religious convictions even when the medical facts were clear. The scoring rubric expects secular reasoning unless the patient frames their entire evaluation through a religious lens, which is allowed. I initially scored that case too harshly because I misread spiritual integration as noncompliance with the reasoning domain. The appreciation domain is subjective by design. You are judging whether the patient recognizes the illness exists and whether they see the treatment as potentially helpful for themselves. Patients with anosognosia in psychotic disorders routinely score zero here, which is clinically accurate even if it looks punitive on paper.
Get the Full Details

Expressing a choice is the simplest domain. The patient must state a preference consistently over time. Inconsistency alone does not disqualify anyone. Changing your mind after hearing new information is normal behavior, not evidence of incapacity.
Common Pitfalls and Edge Cases
The tool performs poorly with patients who have severe expressive aphasia. You can still assess capacity in these individuals, but you need to adapt the questioning format to written responses or yes-no alternatives. The manual mentions this limitation briefly but does not provide a structured workaround. I found that using visual aids alongside simplified language brought the effective completion rate from near zero to about seventy percent in my experience with stroke patients. Another issue is cultural framing of family involvement. In many collectivist cultures, patients defer to family consensus as a rational strategy rather than a sign of impaired reasoning. The MacCAT does not account for this variance explicitly. I learned to probe whether the deferral was genuine or driven by fear of conflict before scoring the reasoning domain. Vignette-based scoring introduces inter-rater variability that the training manual underestimates. Two clinicians watching the same patient can arrive at different scores if one focuses on the explicit statement while the other picks up on subtle nonverbal hesitation. Standardizing your approach by reading the prompts verbatim and avoiding interpretive scaffolding reduces this drift significantly. I started recording sessions with audio for later review, which cut my scoring inconsistency from about fifteen percent to under five percent.
Practical Implementation Notes
You do not need special certification to administer the MacCAT. The official manual can be purchased from the MacArthur Foundation or through academic publishers, and most university libraries carry copies. Training materials including sample videos are available through the Grisso lab website at UMass Medical School. The full instrument is copyrighted, which means you cannot freely distribute copies, but individual licensed use for clinical or educational purposes is standard practice. The test-retest reliability across the four domains ranges from moderate to good, typically around .70 to .85 depending on the population. Sensitivity to change over time is adequate for tracking capacity in progressive disorders, though acute delirium can produce false low scores that resolve within hours of treating the underlying cause. Using the MacCAT alongside collateral history improves accuracy without adding much time. A brief conversation with a family member or caregiver about typical decision-making patterns before the formal assessment gives you a baseline that helps distinguish true capacity loss from temporary confusion. I now always do this before starting the structured questioning.

When the Macarthur Competence Assessment Tool Falls Short
The instrument was designed for adjudicative competence contexts, not for rapid screening. It takes too long for emergency situations where a quick bedside assessment is necessary. If you need something faster for triage purposes, consider tools like the Aid to Capacity Evaluation (ACE) or the Short Assessment of Capacity for Treatment (SACT), which were built specifically for time-pressed settings. Another limitation is that capacity is decision-specific. Scoring zero on the MacCAT for a psychiatric admission does not mean the patient lacks capacity for consent to routine blood draws. I have seen cases where families interpreted a low overall score as a blanket declaration of incompetence across all medical decisions, which is incorrect and sometimes leads to unnecessary guardianship petitions. The tool also struggles with patients who have highly idiosyncratic belief systems. If a patient genuinely believes their illness is caused by a chemical implant rather than a biological process, the appreciation domain may score low even when the reasoning is internally consistent. In these cases, consulting with a neuropsychologist experienced in cultural formulation can clarify whether the belief is culturally normative or truly impairing.
Documentation Tips
Writing up the findings is where most clinicians waste time. Use a structured template that mirrors the four domains, and record specific patient quotes rather than paraphrasing. Quoting the patient directly allows reviewers to verify your scoring interpretation without relying on your summary skills. A typical report for a complete assessment takes about ten minutes if you keep it concise. Always note any accommodations provided during the assessment, such as simplified language, extended time, or presence of a trusted support person. These details matter legally and clinically, and omitting them creates ambiguity when the assessment is reviewed by another provider or in a court proceeding. The MacCAT does not require standardized administration conditions, but consistency across cases helps. Using the same room, the same lighting, and roughly the same sequence of questions makes your longitudinal comparisons more reliable. I stopped moving between departments for assessments because the environmental variability was introducing noise into the scores that had nothing to do with the patients' actual capacities.
Learning the Instrument Effectively
The fastest route to proficiency is watching scored examples rather than just reading the manual. The official training videos show real patient interviews with narration explaining why each score was assigned. Twenty-five minutes of video review plus three supervised practice cases brought me from novice to confident rater in a single afternoon. Self-study without feedback tends to produce over-scoring on the understanding domain and under-scoring on reasoning, probably because clinicians naturally want to help patients sound competent. Peer discussion of borderline cases improves reliability more than additional solo practice. I formed a monthly group with two colleagues from different specialties to review anonymized cases, and our inter-rater agreement improved noticeably within three months. The group itself took about forty-five minutes per meeting and required no formal administration beyond sharing de-identified scoring forms. Keeping a reference card with the domain definitions and common scoring examples is practical. I laminated a one-page summary of the key distinctions between understanding and appreciation and kept it at my workstation. The card reduced my hesitation on borderline cases enough that my average scoring time dropped from twenty-five minutes to about eighteen minutes after six months of regular use.
