Sunday, November 25, 2012

Journal Norm-Referenced and Criterion-Referenced Tests (


Main Title: Norm-Referenced and Criterion-Referenced Test

Stephanie Grant

Ashford University

Learning & Assessment for the 21st Century

Prof. Darrell Rice

November 22, 2012


 

Abstract

    The purpose of this journal is to compare the difference and similarities between a norm-referenced and criterion-referenced test as well as how it will summarize how they each will be used in the assessment performance (place or rank) compared in other students.  You may have had a test in a class and your teacher showed you the class results which mapped out into a curve.  Whereas the majority of the student score was in the middle, others were very low and a couple of students score was on the high end.  During my time in school (1980s’) when this occurred the teacher would curve the grades so the ones with lower scores/grades would receive a grade/score of a C verses a D.  Even though by doing this you may help the student to receive a C verses D it doesn’t help them if they don’t fully understand the concept of the test or if they are actually learning you are just passing them on.  Which I found out by doing research that this is a normal distribution found in norm-referenced testing.

 

 

 

 

 

 

 

 

 

 

 

 

 

Norm-Referenced tests – refers to where a student stands compared to other students.  Certain kind of test data helps determine a student’s “place or rank”.  This is accomplished by comparing the student’s performance to a norm or average of performances by other, similar students.  (Kubiszy, Tom & Borich, Gary, Educational Testing & Measurement Classroom Application and Practice 9th edition (2010))

Standardized examinations such as the SAT are norm-referenced tests.  The goal is to rank the set of examinees so that decisions about their opportunity for success (e.g. college entrance) can be made.

 

Criterion-Reference test – tells us about a student’s level of proficiency in or mastery of some skill or set of skills.  It compares a student’s performance to a standard of mastery criterion.  CRT- information conveys to a comparison with a criterion or absolute standard, it also helps the teacher decide whether a student needs more or less work on some skill or set of skills.  The norm-referenced tests helps determine a student’s place or rank whereas the criterion-referenced tests says nothing about the student’s place or rank compared to other students but is useful only for certain types of decisions.  (Kubiszy, Tom & Borich, Gary, Educational Testing & Measurement Classroom Application and Practice 9th edition (2010))

The results of a Criterion-Reference tests is usually “pass or fail” and are used in making decisions about a job entry, certification, or licensure.  A national board medical exam is an example of a CRT.  Either the examinee has the skills to practice the profession, in which case he or she is licensed, or does not.

The norm-referenced test, in contrast, tends to be general; it measures a variety of specific and general skills at once, but fails to measure them thoroughly.  Also with norm-referenced test you get an estimate of ability in a variety of skills in a much shorter time than you could through a battery of criterion-referenced tests.  Criterion-referenced tests must be very specific if they are to yield information about individual skills.  This is both an advantage and disadvantage. Using a specific test enables you to be relatively certain that your students have mastered or failed to master the skill in question.  The major disadvantage of criterion-referenced tests is that many such tests would be necessary to make decisions about the multitude of skills typically taught in the average classroom. (Kubiszy, Tom & Borich, Gary, Educational Testing & Measurement Classroom Application and Practice 9th edition)

TABLE 5.1 Comparing Norm-Referenced and Criterion-Referenced Tests

Dimension                                          NRT                                                    CRT

Average of number of students who get an item right
50%
80%
Compares a student’s performance to
The performance of other students.
Standards indicative of mastery.
Breadth of content sampled
Broad, covers many objectives.
Narrow, covers a few objectives
Comprehensiveness of content sampled
Shallow, usually one or two items per objective.
Comprehensive, usually three or more items per objective.
Variability
Since the meaningfulness of a norm-referenced score basically depends on the relative position of the score in comparison with other scores, the more variability or spread of scores, the better.
The meaning of the score does not depend on comparison with other scores: It flows directly from the connection between the items and the criterion.  Variability may be minimal.
Item Construction
Items are chosen to promote variance or spread items that are “too easy” or “too hard” are avoided. One aim is to produce good “distractor options.”
Items are chosen to reflect the criterion behavior. Emphasis is placed on identifying the domain of relevant responses.
Reporting and interpreting consideration
Percentile rank and standard scores used {relative rankings}.
Number succeeding or failing or range of acceptable performance used {e.g., 90% proficiency achieved, or 80% of class reached 90% proficiency}
 
 
 

 

(Kubiszy, Tom & Borich, Gary, Educational Testing & Measurement Classroom Application and Practice 9th edition)(Chart form chapter 5 page 94)

Criterion-Referenced Assessment may be used in all three types of evaluation:

Assessment For Learning
Assessment As Learning
Assessment Of Learning
Assessment for learning is ongoing, diagnostic, and formative. It is for ongoing planning. It is not used for grading report cards.
Assessment as learning actively involves students.  It is ongoing, and involves self and peer assessment.  It provides students with the opportunity to use the feedback to improve learning. Allows time for self-edit.
Assessment of learning occurs at end of year or at key stages. It is summative. It is for grading and report cards.
Assessment for learning involves criterion-referenced criteria based Prescribed Learning Outcomes identified in the provincial curriculum, reflecting performance in relation to a specific learning task
Assessment as learning includes teacher-supported self-assessment or peer-assessment based on criteria, with a subsequent opportunity to bring work up to standard.
Summative teacher evaluation in the classroom is essentially criterion-referenced (based on Prescribed Learning Outcomes); norm-referenced assessment (comparing student achievement to that of others) appears on standardized tests such as the FSA and the Government finals

 

 

So within it all the Norm-Referenced test (standardized test) basically shows the students growth over a period of time, and the Criterion-Referenced test you can see what the student accomplished; overall they both are put into learning / developing perspective for the student.

 

References

 

(Kubiszy, Tom & Borich, Gary, Educational Testing & Measurement Classroom Application and Practice 9th edition)



(Kubiszy, Tom & Borich, Gary, Educational Testing & Measurement Classroom Application and Practice 9th edition)(Chart form chapter 5 page 94)


 
http://www.stephanie4095.blogspot.com