A
B
C
D
E
100
Which of the numerical scales is used to determine your semester GPA?
What is Ratio Scale. GPA’s are used for direct comparisons and can be mathematically manipulated. The numbers on the scale are equidistant from each other and there is an absolute zero value (0.0 GPA).
100
In a university, all students are given a new student identification number upon registration. These numbers are on what scale?
What is Nominal Scale
100
What term describes a set of data in which the mean is less than the most frequently occurring scores?
What is positively skewed.
100
In a normal distribution, most students’ scores fall at the edges of the bell curve.
What is false.
100
True or false: Large sets of data are organized and understood through methods known as Measures of Central Tendency.
What is false. Large sets of data are organized and understood through methods known as descriptive statistics.
200
True or False: A multimodal distribution means that there are three or more modes.
True, a multimodal distribution has three or more modes.
200
True or false: In order for a score to be considered significant it must be at least two standard deviations from the mean.
What is false: Any test score that is 1 standard deviation above or below the mean score is considered significant.
200
Which number is representative of a raw score? A. Total test items = 25 B. Correct answers = 20 C. Incorrect answers = 5 D. Ratio = 20/25
What is B
200
This term is used when a national sample of students of the same age/grade take the same assessment and establish a mean and standard deviation. A. Norm-referenced B. Standardized C. Derived scores D. Measures of central tendency
What is A Norm-referenced is the term used to define a national sample of students the same age/grade who attempted the same items.
200
One would expect that the number of classes a college student attends in a specific course and the final exam grade in that course would have a ___.
What is Positive correlation
300
A test instrument may have good reliability; however, that does not guarantee that the test has ___.
What is validity
300
Organizing data to see how the data cluster is called A. Measures of Dispersion B. Measures of Central Tendency C. Frequency D. Normal Distribution
What is B Explanation: Measures of Central Tendency describes a way to organize data to see how the data cluster, or are distributed around a numerical representation of the average score.
300
The number of times a student moves during elementary school may likely have a ___ _____ to the student’s achievement scores in elementary school.
What is negative correlation.
300
The smaller the ______ ___ __ ________, the more reliable the test.
What is standard error of measurement
300
What are the four steps required to calculating variance.
Step 1: To calculate the amount of distance of each score from the mean, subtract the mean for the set of data from each score.  Step 2: Find the square of each of the difference scores found in Step 1 (multiply each difference score by itself).  Step 3: Find the total of all of the squared score differences. This is called the sum of squares.  Step 4: Calculate the average
400
Comparing the scores of a new assessment instrument with the scores of an older test may help to establish which of the following? A. content validity B. internal reliability C. alternative forms of validity D. coefficient alpha
What is A Explanation: One of the ways to establish if a new instrument assesses the same traits as an older version of the assessment is to establish content validity—the consistency of items on an instrument to measure a skill, trait or domain.
400
Reliability coefficients are expressed: A. between +.5 and -.5 B. between + 1 and 0 C. between - 1 and 0. D. between -1 and +1.
What is D Explanation: Reliability coefficients are expressed between -1 and +1.
400
When an instrument can predict performance on some other variable, it is considered to have A. content validity B. construct validity C. predicative validity D. criterion-related validity
What is C Explanation: Predictive Validity measures how well an instrument can predict performance on some other variable.
400
What is reliability?
Reliability—the dependability or consistency of an instrument across time or items.
400
What is validity?
The degree to which an instrument measures what it was designed to measure.
500
How would the use of standard error of measurement help in making educational decisions for students?
Standard Error of Measurement (SEM) is founded on the principle that error exists in testing and a student’s obtained score may not be representative of their true score. This error could be related to a variety of variables including hunger or environmental issues, to name a few. Determining the SEM allows the teacher to determine if the student’s scores are a reasonable estimate of their abilities.
500
What is test-retest reliability?
Test-retest reliability is when the same test is given over a period of time to determine if the trait being measured is one that is stable over time.
500
What is Interrater Reliability?
Interrater Reliability establishes the consistency of a test across examiners. It is not a measure to determine if the instrument is internally consistent.
500
What is one of the disadvantages of a repeated measures test?
One of the disadvantages of repeated measures (test-retest) design is that students may remember test items (practice effects) and score higher on the assessment the second time.
500
What type of validity does the SAT purport to have?
The SAT is a measure of predictive validity. Predictive validity measures how well an instrument can predict performance on another variable. Basically, the SAT is purported to predict how well a student will perform in course work at the collegiate level.