C921 Assessment and Evaluation Strategies for Measuring Student Learning, catalog number NURS 6004, is the three-CU measurement course in the MSN Nursing Education specialty, covering the design, development, implementation and evaluation of student achievement outcomes in nursing education programmes. It is the most technical course in the track. Writing well is not enough here, because the deliverables usually require you to build an instrument and then defend its properties.
What NURS 6004 is actually testing
Assessment in education has a vocabulary that students often use loosely and this course does not. Assessment measures an individual learner. Evaluation judges a course or a programme. Formative assessment informs learning while it is happening and usually carries no weight. Summative assessment certifies achievement at a boundary. Getting those four terms right throughout a paper is a small thing that signals a great deal.
Underneath the vocabulary sits the real content: validity, reliability and fairness. Validity asks whether the instrument measures what it claims to measure, which in nursing means asking whether a multiple choice question about a medication error really tells you whether this student would prevent one. Reliability asks whether the same performance would receive the same score twice, which is why rubrics with vague descriptors fail. Fairness asks whether the instrument disadvantages students for reasons unrelated to the competency, which in nursing frequently means reading load, cultural assumptions inside a scenario, or a technology requirement.
The third element is item and instrument craft. Writing a defensible multiple choice item, building an analytic rubric whose levels are actually distinguishable, deciding what a clinical evaluation tool should observe and what it should ignore. This is a craft with rules, and deliverables in this course usually test whether you know them.
Turning scored aspects into a section plan
The rubric lives in your Course of Study rather than the WGU catalog. Count the aspects before you build anything. Each scores independently on a three-point scale and each needs a 2, which in a measurement course means an elegant instrument with no validity argument attached will still come back.
The word budget, worked. Suppose seven scored aspects and directions asking for about 2,000 words of narrative alongside whatever instrument the task requires. Reserve 130 for an introduction naming the outcome being measured and 100 for the close, leaving 1,770 across seven aspects, or about 253 each. Then weight it. The validity aspect deserves 360 because it has to make an argument rather than a claim. The fairness aspect deserves 300, since it needs specific threats and specific mitigations. That leaves 1,110 for five aspects at about 222 each, which is enough when the instrument itself carries part of the evidence.
Build the blueprint before the items. A table with content areas down the side, cognitive levels across the top, and a target number of items in each cell tells you what to write and simultaneously constitutes your content validity evidence. Items written first and blueprinted afterwards always cluster in the easiest content and the lowest cognitive level.
A structure that fits an assessment design deliverable
Task directions win where they specify a format. Where they leave it open, this arrangement puts the measurement argument where evaluators look for it.
| Section | What belongs in it | What earns the aspect |
|---|---|---|
| Outcome measured | The course or programme outcome the instrument targets, quoted | A single clear outcome; a vague target makes validity unarguable |
| Purpose and stakes | Formative or summative, what the result is used for, who sees it | Stakes stated, because they set the standard of rigour required |
| Blueprint | Content areas by cognitive level with item or criterion counts | Weighting justified against the outcome and against practice importance |
| The instrument | Items, rubric criteria or observation checklist as required | Craft rules followed; flawed items undercut every later argument |
| Validity argument | Content, construct and consequential evidence for this use | Evidence assembled as an argument rather than an assertion of validity |
| Reliability plan | How consistency is achieved, including rater training where relevant | A concrete mechanism such as anchor examples or double marking |
| Fairness analysis | Threats to fair measurement and what you do about each | Specific threats such as reading load or scenario assumptions |
| Scoring and standard | How scores are produced and how the pass point was set | A defensible standard-setting rationale rather than a round number |
| Use of results | What the educator and the programme do with the data | Feedback to students and improvement of teaching, both named |
| References | APA list of measurement and nursing education assessment literature | Measurement claims cited to measurement sources |
Where a rubric is the deliverable, the test of quality is whether two markers who have never met would place the same work at the same level. That means descriptors with observable differences rather than adverbs. Consistently, mostly and sometimes are not a scale.
Evidence craft in a measurement course
This is the course where sloppy sourcing shows fastest, because the claims are technical and checkable.
- Cite measurement concepts to measurement literature. Validity as an argument, reliability estimates and standard setting all have their own scholarship, and general teaching sources will not carry those claims.
- Use nursing education assessment research for anything about nursing students, including clinical evaluation instruments, which have their own body of work.
- Name any published instrument you adapt and cite it. Adapting a validated tool is legitimate and changes its properties, which you should say.
- Report reliability with the statistic and the context. An estimate quoted with no sample and no setting is decoration.
- Give the standard-setting method a source. Choosing a pass mark is a documented procedure, not an opinion, and evaluators in this course know the methods by name.
- Quote sparingly. Item-writing rules and definitions are heavily reproduced, and WGU runs submissions through a similarity check.
The habit that most improves a deliverable here is writing the consequences of being wrong. If this instrument passes a student who is not safe, what happens. If it fails a student who is competent, what happens. Consequential validity is the part of measurement that matters most in a professional programme and the part students most often leave out.
What separates Competent from a submission sent back
Aspects score on their own, and validity and fairness are the two that come back most.
- The blueprint exists and the instrument matches it.
- Validity is argued with evidence rather than asserted as a property.
- Reliability has a mechanism attached, not a hope.
- Fairness names specific threats and specific mitigations.
- The pass standard has a stated method behind it.
Performance assessment work at WGU can be revised and resubmitted with no grade penalty, so a return costs time rather than standing. Time is the scarce thing in a flat-rate six-month term, and C921 is a three-CU course that usually sits directly in front of the field experience and capstone in the education specialty. A delay here has a way of moving two courses rather than one.
Six mistakes that cost time in C921
- Using assessment and evaluation interchangeably. The course distinguishes them and so should every sentence.
- Writing items before the blueprint. Unblueprinted items cluster in easy content and low cognitive levels.
- Rubric levels separated by adverbs. If two markers cannot agree, the rubric is unreliable no matter how it looks.
- Claiming validity as a property. Validity belongs to a use, not to an instrument, and the aspect wants an argument.
- Setting a pass mark by convention. Seventy five percent because it is standard is not a rationale. Name a method.
- Stopping at the score. The use-of-results aspect asks what changes for the student and for the teaching, and it is frequently the thinnest section.
How support works on this course
C921 is the course where the education specialty gets technical, and the help that works is technical too. Send the rubric from your Course of Study with the task directions and the first pass is the blueprint, because it doubles as your content validity evidence and it governs everything you build afterwards. From there you get item or rubric craft checked against the rules, a validity argument assembled from real evidence types, a reliability mechanism, a fairness analysis with named threats, and a standard-setting rationale with a method behind it.
The boundaries hold. Objective assessments at WGU are proctored, so we prepare only, never sit them, and never ask for portal credentials. On field-based courses in this track we never complete practice hours, contact mentors or sites, sign placement paperwork or fill in hour logs.
Questions students ask about C921
Is C921 the same course as NURS 6004?
How much statistics does C921 require?
Can I build a clinical evaluation tool instead of a written test?
Building an assessment instrument for C921?
Send your Course of Study rubric and the task directions. We start with the blueprint, then check item or rubric craft and build the validity argument.
Where C921 sits in WGU's programs
The July 2026 catalog places this code in 3 current WGU programs. Open a program page for the complete standard path and term positions. The live Degree Plan remains authoritative after transfer credit, substitutions, and mentor planning.
The assessments, one by one
The public catalog does not publish this course's PA/OA identity or task count. WGU Tutors publishes at most one PA manual per course and only from a WGU-controlled public rubric. Until that source exists, PA help begins from the student's real Course of Study and OA support remains preparation only.