Article

How to Measure Career Readiness in College Students

July 27, 2026 ยท 5 min read

How to Measure Career Readiness in College Students
Photograph by RDNE Stock project on Pexels.

There are three practical ways to measure career readiness in college students: ask the student, ask the people who supervised the student's work, and review what the student actually produced. Each one evidences something different. Self-assessment evidences confidence and self-awareness. Supervisor observation evidences behavior in a real work setting. Artifact review evidences skill at a single point in time under known conditions. If you only have budget and attention for one, none of the three is sufficient on its own, and self-report is the weakest of the three by a clear margin.

Student self-assessment

Self-assessment is cheap, scalable, and the only method that captures the student's internal state. You learn whether a student believes she can lead a meeting, whether she recognizes gaps in her own technical skills, and whether an internship changed how she sees herself. Those are real outcomes. Career development is partly a story students tell about themselves, and if that story does not change, the experience probably did not land.

What self-assessment cannot do is establish that the competency exists. The size of the gap is well documented. NACE's Job Outlook research has repeatedly found students rating themselves proficient at far higher rates than employers rate them: on professionalism and work ethic, roughly 90 percent of students versus roughly 43 percent of employers; on communication, roughly 80 percent versus roughly 42 percent. Teamwork is the one competency where the two groups come close. That is not a measurement error to be corrected. It is the finding: students and employers are using different standards, and only one of those standards matters at the point of hire.

Self-assessment also has a specific failure mode that catches assessment coordinators off guard. Students who learn the most sometimes rate themselves lower after an experience than before it, because they now understand what competence actually requires. Pre-post self-report can show a decline that represents growth. If you run pre-post designs, plan for retrospective pretest items ("looking back, how would you rate your skill at the start?") alongside the standard pre-post pair, and expect the two to disagree.

Supervisor and employer observation

The supervisor who watched a student work for twelve weeks has evidence no one else has: whether the student showed up prepared, took feedback without defensiveness, escalated a problem before it became a crisis, and behaved professionally when no one was checking. This is the closest thing to a direct measure of career readiness as employers define it, because employers are defining it.

The constraints are real. Response rates are the first one. Site supervisors are busy, they are not your employees, and a twenty-minute survey at the end of a summer will be ignored. Instruments that take five minutes and use behaviorally anchored items get returned; long Likert batteries do not. Second, supervisor ratings compress toward the top. A supervisor who likes a student, or who does not want to jeopardize a placement, will mark "exceeds expectations" across the board. Halo effects are strong enough that item-level variation often disappears. You can reduce this with forced-choice items, with rubrics that describe observable behavior rather than traits, and by asking for one specific example in an open text field, which also gives you quotable qualitative evidence for accreditation narratives.

Third, supervisor quality varies. A student placed with a hands-off supervisor at a thin site gets rated on very little actual observation. Multi-rater designs help here, which is the argument for collecting faculty supervisor input alongside employer input rather than treating either as definitive. This is the structure the Career Readiness Report uses: the student's self-assessment sits next to independent ratings from the employer and the faculty supervisor on the same eight NACE competencies, so the gap between the three is visible rather than averaged away.

Artifact review

Artifact review means faculty or trained raters scoring student work products against a rubric: a written report, a presentation recording, a portfolio, a capstone deliverable, a code repository. The AAC&U VALUE rubrics were built for exactly this and are already mapped closely enough to several NACE competencies (written communication, critical thinking, teamwork, intercultural knowledge) that you do not need to build from scratch.

Artifact review is the strongest evidence of demonstrated skill and the only one that produces a durable record a reviewer can re-examine. It is also the most expensive. Norming raters takes a half day minimum, and inter-rater reliability below about 0.70 makes your results hard to defend. Sampling is usually the answer: score 100 artifacts well rather than 800 badly.

The honest limit of artifact review is scope. It cannot evidence professionalism, initiative, or how a student handles ambiguity. A polished report says nothing about whether the student produced it the night before after three missed deadlines. Several NACE competencies are behavioral and simply do not leave artifacts.

A defensible combination

Map each competency to the method that can actually evidence it, then stop trying to measure everything with everything.

  • Communication, critical thinking, and technology: artifact review on a sample, with a VALUE-based rubric.
  • Professionalism, teamwork, equity and inclusion, leadership: supervisor observation with behaviorally anchored items.
  • Career and self-development: self-assessment, which is the appropriate instrument here because the construct is partly internal.

Collect self-assessment on all eight anyway, because the gap between self-rating and supervisor rating is itself a useful metric and a strong coaching tool. A student who rates herself 4.5 on communication while her supervisor rates her 2.5 has a specific, actionable problem, and that conversation is more valuable than either number alone.

For accreditation and program review, the pattern that holds up is a small number of competencies measured with observation or artifacts and reported with sample sizes and response rates stated plainly, plus self-report data labeled as self-report. Reviewers are not fooled by a 92 percent proficiency figure that turns out to be students grading themselves. Labeling it honestly costs nothing and protects everything else in the report.

The Career Readiness Report is free for every college and university. Open now, in beta.

Create your institution