Getting Real With the BRP-2
The Behavior Rating Profile Second Edition Brp 2 is a teacher- and parent-report instrument designed to assess behavioral and emotional functioning in children and adolescents ages 5 through 18. It produces standard scores across several clinical and adaptive scales, giving you a profile rather than a single number. That sounds useful until you actually have to interpret the thing with real data in front of you. I picked up a copy about four years ago for a school-based assessment project. The manual is dense, the normative sample is dated enough that you should be cautious about applying it to contemporary populations, and the scoring itself is straightforward but the clinical interpretation is where people routinely go off the rails.
What the Instrument Actually Measures
BRP-2 gives you raw scores on scales like Internalizing Problems, Externalizing Problems, Attention Problems, Social Problems, Learning Problems, and a handful of subscales under each umbrella. From those raw scores you compute T-scores or standard scores using the appropriate norm table based on the respondent type and child's age band. The manual provides tables for teacher ratings and parent ratings separately because the distributions differ between those groups. One thing beginners miss: the instrument does not give you a diagnosis. It gives you indicators. A high score on the Anxiety scale, for example, might suggest elevated anxious behavior but it does not rule out ADHD-related restlessness mimicking anxiety or vice versa. Comorbidity is the norm in school-age populations, not the exception, and BRP-2 scores reflect that complexity rather than resolving it.
Scoring Workflow
Here is the practical sequence I use and recommend: Complete all items on the rating form. Check for missing responses before moving forward. Any skipped item needs to be followed up with the rater because the norm-referenced scores assume complete data. Incomplete protocols are generally scored only if fewer than five items are missing on any scale, but the manual specifies exact thresholds and I follow them precisely rather than eyeballing it. Convert raw scores using the appropriate norm table. If you are working with a child whose age falls between two norm bands, interpolate. Do not round the age down or up arbitrarily. I have seen assessors round an 8-year-11-month-old to the 9-year-old norms because they were pressed for time. That is not acceptable. The difference between those norm bands can shift a T-score by three to five points, which is the difference between "average" and "borderline clinical."
Get the Full Details

After computing all standard scores, draw the profile. The visual comparison between scales is where patterns actually emerge. A child might have a clinically elevated score on one scale but the relative height matters more than the absolute cutoff. An elevation of 65 on the Social Problems scale is far more meaningful when the next highest scale is 52 than when it is 63.
A Real Problem I Encountered
Last year I was working with a second-grade referral where the teacher ratings and parent ratings showed dramatically different profiles. The teacher reported significant Externalizing Problems with a T-score of 71 on the Hyperactivity scale. The parent report showed a nearly flat profile, everything between 42 and 49. The school psychologist wanted to proceed with a full evaluation based on the teacher data alone. Something felt wrong about that approach. I pulled the child's work samples, reviewed classroom observations from two different periods, and cross-checked with the special education teacher who worked with the child for part of the day. The hyperactivity symptoms were significantly more pronounced in unstructured settings and during transitions, but the child was compliant and engaged during direct instruction. The discrepancy between teacher and parent ratings was real, but the interpretation needed to account for setting effects rather than treating the higher rating as the definitive picture. The workaround was to supplement the BRP-2 data with structured classroom observations and a behavior intervention plan trial over three weeks before moving forward with any diagnostic conclusions. The BRP-2 scores stayed elevated on the teacher form, but the observational data reframed how I understood them. That process added about two weeks to the evaluation timeline but prevented a misattribution that could have led to an inappropriate placement recommendation.
Common Pitfalls
Raters tend to compress their ratings toward the middle of the scale when they are filling out these forms quickly. I see it constantly. A teacher who rates every scale between 45 and 55 is probably not describing a child with no behavioral concerns. They are probably rushing through the form. I flag those profiles and request a rater conference before accepting the results. The manual acknowledges this tendency and recommends rater training, but most schools do not provide that training. Another frequent error is using teacher norms when the rating was completed by a parent, or mixing norm groups across age bands. The scoring software built into the official manual does not allow this, but the pencil-and-paper versions rely on the administrator matching the correct table. I learned this the hard way when a colleague accidentally scored a parent report using the teacher norm table. The T-scores came out approximately five points lower across the board, which shifted two clinical scales below cutoff and changed the entire clinical impression. That error went unnoticed for six months. A third pitfall is treating the composite scores as more reliable than the subscale scores. The Internalizing and Externalizing composites are derived from multiple subscales and have better reliability estimates, but they also lose specificity. A high composite can result from many different underlying profiles. I always report subscale scores alongside composites and never rely on the composite alone for decision-making.

Limitations Worth Stating Clearly
The normative sample for BRP-2 was collected in the early 1990s. Cultural norms, diagnostic labeling practices, and classroom expectations have shifted substantially since then. A T-score of 65 today may represent a different behavioral threshold than it did when the norms were established. I factor this into my interpretations and note it in any written report. This is not a dealbreaker for the instrument, but it is a limitation that requires conscious adjustment rather than blind reliance on the published cutoffs. The instrument also lacks strong sensitivity to subtle or internalizing presentations in younger children. A seven-year-old with significant anxiety may not show enough behavioral disturbance on the BRP-2 to generate elevated scores, especially if the rating form emphasizes observable external behaviors. I supplement the BRP-2 with a separate anxiety-specific measure in those cases rather than assuming a non-elevated score means no anxiety is present. If you are looking for a more contemporary instrument with updated norms and stronger psychometric properties for specific populations, the Behavior Assessment System for Children, Third Edition (BASC-3) or the Achenbach System of Empirically Based Assessment (ASEBA) are reasonable alternatives. BRP-2 remains a valid tool when used correctly and when its limitations are acknowledged, but it is not the most robust option available for every situation.
Practical Takeaways
Use both teacher and parent ratings whenever possible. Discrepancies between raters are data, not noise. Invest time in rater training or at minimum a brief rater conference before accepting completed forms. Never score using the wrong norm table. Always draw the profile graph because the visual pattern reveals what the numbers alone obscure. Report subscale scores, not just composites. Document norm sample age limitations in your interpretive notes. And when the scores do not match what you observe in the child, trust the observation and note the discrepancy rather than forcing the data to fit.