Skip to main content
Study NotesLET Secondary · Assessment of LearningReal content

LET Secondary Assessment of LearningStatistics, Grading and Interpreting Assessment ResultsStudy Notes

Study notes for Statistics, Grading and Interpreting Assessment Results that match the LET Secondary 2026 syllabus. Built to mirror how Professional Regulation Commission (PRC) structures LET Secondary Assessment of Learning questions, these notes walk through each concept with examples, formulas, and practice questions designed for time-pressured exam conditions.

Exam context

Professional Regulation Commission (PRC) runs the Licensure Examination for Professional Teachers — Secondary on Bi-annual. Its Assessment of Learning section sits under a "Core" weighting, and Statistics, Grading and Interpreting Assessment Results is the 4th chapter in the 5-chapter LET Secondary Assessment of Learning rotation. The LET Secondary passing mark is Weighted average of 75% with no grade below 50%, and the most recent 2026 paper drew about a meaningful share of questions from Assessment of Learning.

Statistics, Grading and Interpreting Assessment Results - Study Notes

As a Filipino elementary teacher, you assess student learning daily—through quizzes, performance tasks, and summative exams. But raw scores alone tell you nothing. You must understand the statistical foundations that transform scores into meaningful feedback, interpret patterns in class performance, and assign grades fairly under DepEd Order No. 8, s. 2015. This chapter equips you with the statistical toolkit and grading knowledge required by the Licensure Examination for Teachers and demanded by your professional responsibility under the Code of Ethics for Professional Teachers (RA 7836). You will master measures of central tendency and variability, the normal curve, standard scores, and the K-12 grading system. The LET loves computation items here, especially skewness and standard scores—concepts that separate proficient teachers from those who guess.

Summary

This chapter equipped you with the statistical and practical knowledge required to transform raw assessment scores into meaningful feedback and fair grades. You master measures of central tendency (mean, median, mode) and variability (range, SD, IQR) to describe class performance and identify individual learners. You understand percentiles and quartiles to communicate relative standing. You recognize the properties of the normal curve and the critical concept of skewness—named after the tail, not the hump—to infer test difficulty and class readiness. You compute and interpret standard scores (z and T) to compare performance across subjects and assessments, a cornerstone of fair grading. You grasp correlation as a tool for identifying relationships between variables (attendance and achievement, study time and grades) while guarding against causal claims. You know DepEd Order No. 8, s. 2015 cold: the three components (WW, PT, QA), the subject-specific weights, the step-by-step grade computation, the transmutation table, and the descriptors that appear on report cards. Finally, you synthesize all this knowledge to interpret real classroom data, identify pupils needing enrichment and intervention, adjust instruction, and communicate results to families—all while ensuring fairness and non-discrimination under the Code of Ethics for Professional Teachers (RA 7836) and child-protection law (RA 7610). Assessment is not an end in itself but a means to understand your learners deeply and serve them justly. The LET tests computation items (z-scores, grade calculation) and scenario interpretation (skewness, standard scores, grading decisions). Master the worked examples, practice with real classroom data, and you will be well-prepared.

Sections

A measure of central tendency is a single number that represents an entire set of scores. It answers the question: 'What is the typical score in this class?' Three measures serve this purpose, each with strengths and weaknesses. **The Mean (Arithmetic Average)** The mean is the sum of all scores divided by the number of scores: Mean = ΣX / N. It uses every score in the distribution, making it mathematically elegant and the most commonly reported measure. However, this very strength is also its weakness: the mean is highly sensitive to extreme values (outliers). If one student scores 20 while the rest score 80–95, that single outlier drags the mean down significantly, painting a false picture of class performance. In a class where 25 students average 88 but one scored 20, the mean might be 87—hiding the fact that nearly all students performed well. **The Median (Middle Score)** The median is the middle point when all scores are arranged in order from lowest to highest. If the number of scores (n) is odd, the median is the middle score itself. If n is even, the median is the average of the two middle scores. The median is the 50th percentile (P50) and is completely unaffected by outliers. In that same class where one student scored 20, the median would remain around 88 because that outlier sits at one end and does not shift the middle. The median is the best representative measure for skewed distributions (distributions with a long tail on one side). **The Mode (Most Frequent Score)** The mode is the score that appears most often in a distribution. A distribution can be unimodal (one mode), bimodal (two modes), multimodal (many modes), or have no mode if all scores appear equally often. The mode is the only measure usable with purely nominal (categorical) data, such as favorite subjects (English, Math, Science). It is less refined mathematically but useful for identifying the 'most popular' score or category. **Worked Example** Seven Grade 3 students score on a reading fluency assessment: 82, 85, 85, 88, 90, 92, 94 words per minute. - **Mean:** (82 + 85 + 85 + 88 + 90 + 92 + 94) ÷ 7 = 616 ÷ 7 = **88 words per minute** - **Median:** Arranged in order, the 4th (middle) score is **88 words per minute** - **Mode:** The score 85 appears twice; all others appear once, so **mode = 85 words per minute** **When to Use Which Measure** Use the **mean** when data is roughly symmetric and contains no extreme outliers—it is the most powerful statistically. Report the **median** when a distribution is skewed or contains outliers; it gives a more honest picture of the typical student. Use the **mode** when describing categorical data or when you need the most common or popular response. In your elementary classroom, if a Grade 5 Math test produces scores clustered around 80–85 with no outliers, report the mean. If one student scored 30 while the others scored 75–95, report the median alongside the mean to tell the full story.

Heading

Measures of Central Tendency: Where Scores Center

Examples

  • Ten Grade 2 pupils' Mathematics scores: 75, 78, 80, 80, 82, 85, 85, 85, 90, 92. Mean = 842 ÷ 10 = 84.2. Median = (82 + 85) ÷ 2 = 83.5 (average of 5th and 6th scores). Mode = 85 (appears three times).
  • In a class where most pupils scored 85–90 but one scored 30 (due to illness and makeup), the mean might be 80, but the median remains 87. The median better represents how the healthy class performed.
  • A survey asks Grade 1 pupils: 'What is your favorite subject?' Responses: English (12 votes), Mathematics (18 votes), Science (10 votes), Araling Panlipunan (5 votes). The mode is Mathematics. Mean and median are not meaningful here.

Key Points

  • Mean is the arithmetic average; most sensitive to outliers; best for symmetric data.
  • Median is the middle score; unaffected by outliers; the 50th percentile; best for skewed data.
  • Mode is the most frequent score; the only measure for nominal/categorical data.
  • All three measures coincide (are equal) in a perfectly symmetrical distribution.
  • Always consider the context and shape of your distribution when choosing a measure.

Central tendency tells you where scores cluster; variability (or dispersion) tells you how widely they scatter. Two classes might both have a mean of 80, but one class's scores might range from 78 to 82 (tight, homogeneous group), while another's range from 40 to 100 (scattered, heterogeneous group). Variability reveals this hidden picture. **The Range** The range is the simplest measure of variability: Range = Highest score − Lowest score. In the reading fluency example (82, 85, 85, 88, 90, 92, 94), the range is 94 − 82 = 12 words per minute. The range is quick to compute but crude because it depends on only two scores (the extremes). One outlier can inflate the range dramatically. If a class of 25 has one student score 100 and another score 20, the range is 80, which might not reflect the spread of the other 23 students. **The Standard Deviation (SD)** The standard deviation is the most reliable and widely used measure of variability. Conceptually, it is the average distance of scores from the mean. A small SD means scores cluster tightly around the mean (a homogeneous, tightly-knit class); a large SD means scores are spread far from the mean (a heterogeneous, diverse class). The variance is the standard deviation squared; both communicate spread, but SD is in the same units as the original scores, making it easier to interpret. SD formula (population): SD = √[Σ(X − Mean)² / N] SD formula (sample): SD = √[Σ(X − Mean)² / (N − 1)] **Worked Example: Computing Standard Deviation Conceptually** Five Grade 4 pupils score on a Math quiz: 4, 6, 8, 10, 12 points. - **Step 1: Find the mean.** (4 + 6 + 8 + 10 + 12) ÷ 5 = 40 ÷ 5 = 8 - **Step 2: Find each score's deviation from the mean (X − Mean).** - 4 − 8 = −4 - 6 − 8 = −2 - 8 − 8 = 0 - 10 − 8 = +2 - 12 − 8 = +4 - **Step 3: Square each deviation.** 16, 4, 0, 4, 16 - **Step 4: Sum the squared deviations.** 16 + 4 + 0 + 4 + 16 = 40 - **Step 5: Divide by N (for a population) or N − 1 (for a sample).** 40 ÷ 5 = 8 (variance) - **Step 6: Take the square root.** √8 ≈ 2.83 **Interpretation:** The standard deviation is approximately 2.83 points. On average, pupils' scores deviate about 2.83 points from the mean of 8. The spread is moderate. **Homogeneous vs. Heterogeneous Classes** Imagine two Grade 5 English classes, both with a mean reading comprehension score of 75: - **Class A (homogeneous):** Scores are 73, 74, 74, 75, 75, 76, 76, 77. SD ≈ 1.2. Scores cluster tightly; the class learned together at a similar pace. - **Class B (heterogeneous):** Scores are 50, 60, 70, 75, 80, 85, 90, 100. SD ≈ 15.7. Scores sprawl widely; the class has advanced readers, struggling readers, and everyone in between. Class A's small SD signals that a whole-class reteaching approach might work. Class B's large SD signals the need for differentiated instruction, with small groups at different levels. SD informs your pedagogical decisions as much as it describes data.

Heading

Measures of Variability: How Spread Out Scores Are

Examples

  • Two Grade 3 Math classes both average 80. Class X: scores 78, 79, 80, 81, 82 (SD ≈ 1.4). Class Y: scores 50, 70, 80, 90, 100 (SD ≈ 20.6). Both have the same mean, but Class X is homogeneous and ready for grade-level content; Class Y needs differentiated instruction.
  • A Science performance task (out of 100 points) yields scores: 88, 89, 90, 91, 92. Mean = 90, SD ≈ 1.4 (very tight). Most pupils demonstrated mastery; this class is ready to advance.
  • A vocabulary assessment (out of 50 words) yields scores: 10, 20, 30, 40, 50. Mean = 30, SD ≈ 14.1 (very spread out). The class spans from non-readers to strong readers; small-group reading instruction is essential.

Key Points

  • Variability (dispersion) describes how spread out scores are around the mean.
  • Range = highest − lowest; quick but crude, affected by extreme values only.
  • Standard deviation is the average distance of scores from the mean; most reliable measure of spread.
  • Small SD = homogeneous class (tight clustering); large SD = heterogeneous class (wide scatter).
  • Variance = SD²; both convey spread, but SD is in original units and more interpretable.
  • Use variability alongside central tendency: same mean does not mean same class composition.

While central tendency and variability describe the distribution as a whole, percentiles, quartiles, and deciles allow you to locate an individual student's position within the group. These are rank-order measures that answer: 'How did this student compare to peers?' **Percentiles and Percentile Ranks** A percentile divides a distribution into 100 equal parts. A percentile rank (PR) of 75 means the student scored as well as or better than 75% of the comparison group. This is a norm-referenced interpretation: the student's standing is defined relative to classmates, not by percent-correct or absolute mastery. A pupil who scores at the 75th percentile in Grade 4 Math did not necessarily answer 75% of the items correctly; rather, that pupil's score is higher than or equal to 75% of classmates' scores. **Quartiles (Dividing into Fourths)** Quartiles divide an ordered distribution into four equal parts: | Quartile | Equivalent Percentile | Meaning | |---|---|---| | **Q1 (First Quartile)** | P25 (25th percentile) | 25% of scores fall at or below this point | | **Q2 (Second Quartile)** | P50 (50th percentile) = **Median** | 50% of scores fall at or below; middle of distribution | | **Q3 (Third Quartile)** | P75 (75th percentile) | 75% of scores fall at or below this point | **Important:** Q2 is always equal to the median and the 50th percentile. These three names describe the exact same point in the distribution. **Deciles (Dividing into Tenths)** Deciles divide an ordered distribution into ten equal parts, labeled D1 through D9: - **D5 = P50 = Q2 = Median** (this chain of equalities is a classic LET test item) - D1 = P10, D2 = P20, D3 = P30, and so on up to D9 = P90. Deciles are common on standardized tests like the National Achievement Test (NAT) and Metrobank-MTAP-DepEd Math Challenge, which often report student standing by decile. **The Interquartile Range (IQR)** IQR = Q3 − Q1. The IQR captures the spread of the middle 50% of scores, excluding the bottom 25% and top 25%. Like the median, the IQR is resistant to outliers; if one student scored extremely low or high, the IQR would remain unchanged. For this reason, researchers often pair the median and IQR to describe skewed distributions. **Worked Example: Quartiles and Percentiles** Twenty Grade 5 pupils' Language Arts scores (arranged lowest to highest): 60, 65, 68, 70, 72, 74, 76, 78, 78, 80, 81, 82, 84, 85, 86, 87, 88, 90, 92, 95 - **Q1 (25th percentile):** At position 0.25 × 20 = 5, approximately 72. One quarter of pupils scored 72 or below. - **Q2 (50th percentile = Median):** At position 0.50 × 20 = 10, approximately (80 + 81) ÷ 2 = 80.5. Half the class scored 80.5 or below. - **Q3 (75th percentile):** At position 0.75 × 20 = 15, approximately 86. Three quarters scored 86 or below. - **IQR = Q3 − Q1 = 86 − 72 = 14 points.** The middle 50% of pupils span 14 points. If you report that a pupil scored at the 75th percentile, you mean that pupil is in the top quarter of the class. A pupil at the 25th percentile is in the bottom quarter and likely needs remediation.

Heading

Percentiles, Quartiles, and Deciles: Dividing the Distribution

Examples

  • A Grade 4 pupil scores at the 60th percentile on a Reading Comprehension test. This means the pupil did as well as or better than 60% of classmates (not 60% correct). If the class is 30, the pupil likely outscored about 18 classmates.
  • A Science performance task: Q1 = 70 points, Q2 = 80 points, Q3 = 88 points. The bottom 25% scored 70 or below; the middle 50% scored between 70 and 88; the top 25% scored 88 or above. A pupil who scored 85 is in the top half but below the top quarter.
  • A standardized test reports a pupil at D7 (7th decile). This is equivalent to P70, meaning the pupil is in the top 30% of the country or region—a strong performance warranting recognition and enrichment.

Key Points

  • Percentile rank (PR) = relative position in the group (percent-correct is different).
  • PR 75 means scored as well as or better than 75% of the group.
  • Q1 = P25, Q2 = P50 = Median, Q3 = P75; these describe the same distribution points.
  • D5 = P50 = Q2 = Median (classic LET chain of equalities to memorize).
  • IQR = Q3 − Q1; captures the middle 50% of scores; resistant to outliers.
  • Use percentiles to communicate student standing to parents and to identify learners needing support.

The normal curve (or normal distribution) is a bell-shaped, symmetrical curve that describes the distribution of many natural and human phenomena: heights, weights, intelligence, and standardized test scores. Understanding the normal curve is essential because it underpins standard scores and helps you interpret class performance. **Properties of the Normal Curve** 1. **Symmetrical:** The curve is mirror-image symmetrical about the vertical line at the mean. The left and right halves are identical. 2. **Central Tendency Coincides:** In a perfectly normal distribution, the mean, median, and mode all fall at the exact center of the curve and are equal. 3. **Asymptotic Tails:** The tails of the curve approach but never touch the horizontal baseline. Theoretically, extreme scores are always possible, no matter how unlikely. 4. **Single Peak:** The curve has one mode (the highest point) at the center. It is unimodal. **The 68-95-99.7 Rule (Empirical Rule)** Under a normal curve, fixed proportions of scores fall within specified standard deviation units from the mean: - **Within 1 SD of the mean:** approximately **68%** of scores (34% on each side of the mean) - **Within 2 SD of the mean:** approximately **95%** of scores - **Within 3 SD of the mean:** approximately **99.7%** of scores (nearly all) **Practical Application in Your Classroom** Suppose a Grade 5 Mathematics assessment has a mean of 80 and a standard deviation of 5. - About 68% of pupils scored between 75 (80 − 5) and 85 (80 + 5). - About 95% scored between 70 (80 − 10) and 90 (80 + 10). - About 99.7% scored between 65 (80 − 15) and 95 (80 + 15). If a pupil scores 90 (2 SDs above the mean), that pupil is in the top ~2.5% of the class—a notably high score. If a pupil scores 70 (2 SDs below the mean), that pupil is in the bottom ~2.5%—a notably low score warranting intervention. **When Data Approximates the Normal Curve** Large, unselected groups (national standardized tests, thousands of pupils) typically approximate normality. Small classroom samples often deviate. A Grade 2 classroom of 25 pupils rarely produces a perfect bell curve. Nevertheless, understanding normality helps you recognize when your class deviates from expectations and why. **Limitations of Assuming Normality** Not all distributions are normal. Data can be skewed (asymmetrical), bimodal (two peaks), or uniform (flat). Before assuming a normal distribution, examine your data visually (via histogram or dot plot) or test for normality. Assuming normality when data is skewed can lead to misinterpretation.

Heading

The Normal Curve: The Bell-Shaped Blueprint

Examples

  • A Grade 6 National Achievement Test (NAT) Reading score has a national mean of 50 and SD of 10. A pupil from your school scores 70 (2 SDs above mean, top ~2.5% nationally). This pupil is an exceptionally strong reader and should receive enrichment.
  • Your Grade 3 class of 22 pupils takes a spelling quiz (out of 20 words). The mean is 15 and SD is 2. A pupil scores 19 (2 SDs above mean, top ~2.5% of your class). Another scores 11 (2 SDs below mean, bottom ~2.5% of your class). The second pupil needs spelling intervention.
  • A Mathematics performance task (out of 100 points) yields a class mean of 78 and SD of 8. About 68% of your class scored between 70 and 86 points. Those below 70 (bottom 16%) need reteaching; those above 86 (top 16%) are ready for acceleration.

Key Points

  • Normal curve is bell-shaped, symmetrical, with tails approaching but not touching the baseline.
  • In a normal distribution, mean = median = mode at the center.
  • The 68-95-99.7 rule: ~68% within 1 SD, ~95% within 2 SD, ~99.7% within 3 SD.
  • A score 2 or more SDs from the mean is notably extreme (top or bottom ~2.5%).
  • Large, unselected groups approximate normality; small classroom samples often deviate.
  • Always check if your data actually appears normal before interpreting using the normal curve.

Skewness measures the asymmetry of a distribution—whether it leans toward one end or is balanced in the middle. Skewness is named after the tail (the direction the tail points), not the hump (where most scores pile up). This single fact trips up countless test-takers. Master it cold, and you will outperform peers on the LET. **Positively Skewed (Right-Skewed) Distributions** In a positively skewed distribution: - **The tail points toward the high end (right).** - **Most scores pile up at the low end (left).** - **The long tail pulls the mean toward the high end, so Mean > Median > Mode.** - **Interpretation:** Most pupils scored low; only a few scored high. The test was difficult for the class. **Visual:** Imagine a histogram where the hump is on the left and a long thin tail stretches toward the right. The mean is dragged rightward by that tail and exceeds the median. **Why does this happen?** A test is too hard, so most pupils fail (low scores). A few advanced pupils squeeze out high scores, creating a tail at the top end. The mean, sensitive to outliers, is pulled upward. **Negatively Skewed (Left-Skewed) Distributions** In a negatively skewed distribution: - **The tail points toward the low end (left).** - **Most scores pile up at the high end (right).** - **The long tail pulls the mean toward the low end, so Mean < Median < Mode.** - **Interpretation:** Most pupils scored high; only a few scored low. The test was easy for the class. **Visual:** The hump is on the right and a thin tail stretches left. The mean is dragged leftward and falls below the median. **Why does this happen?** A test is too easy, so most pupils get high scores. A few struggling pupils score low, creating a tail at the bottom end. The mean is pulled downward. **Symmetrical (Normal) Distribution** When a distribution is symmetrical: - **No tail; the curve is balanced.** - **Mean = Median = Mode.** - **Interpretation:** Balanced test difficulty; the class spread is normal. **Summary Table: The Skewness Decoder** | Distribution | Tail Points to | Scores Pile Up | Test Difficulty | Measure Order | |---|---|---|---|---| | **Positively Skewed** | HIGH (right) | LOW (left) | DIFFICULT | Mean > Median > Mode | | **Negatively Skewed** | LOW (left) | HIGH (right) | EASY | Mean < Median < Mode | | **Normal (Symmetrical)** | Neither | Center | Balanced | Mean = Median = Mode | **Classic LET Scenario Items** **Scenario 1:** "After a Grade 4 Mathematics quiz, Teacher Rosa finds that most of her pupils obtained very high scores, with only a few scoring below 50. What is the shape of the score distribution?" - Most scores are **HIGH**, tail is at the **LOW** end. - The distribution is **negatively skewed** (left-skewed). - The quiz was **easy** for the class. - **Mean < Median < Mode.** **Scenario 2:** "After a Grade 5 English test, Teacher Juan observes that most pupils scored between 50 and 65, with only three pupils scoring above 90. What can be inferred?" - Most scores are **LOW**, tail is at the **HIGH** end. - The distribution is **positively skewed** (right-skewed). - The test was **difficult** for the class. - **Mean > Median > Mode.** **Worked Example: Computing Skewness Indicators** Grade 3 pupils' Reading Fluency Assessment (words per minute): 45, 50, 52, 55, 58, 60, 65, 70, 75, 145 (one advanced reader). - **Mean:** (45 + 50 + 52 + 55 + 58 + 60 + 65 + 70 + 75 + 145) ÷ 10 = 715 ÷ 10 = 71.5 - **Median:** (58 + 60) ÷ 2 = 59 (the average of the 5th and 6th scores when arranged in order) - **Mode:** No mode (all scores appear once) - **Order:** Mean (71.5) > Median (59) > Mode (none), so the distribution is **positively skewed**. - **Interpretation:** Most pupils scored between 45 and 75 wpm (modest reading speed), but one pupil scored 145 wpm, pulling the mean upward. The class needs reading support; this single advanced reader is an outlier. **Teaching Implication:** In a positively skewed class distribution, the median and mode better represent the 'typical' pupil than the mean. You might report: "The typical pupil reads at 59 wpm (median), but the class average is 71.5 wpm due to one advanced reader." This distinction guides your instruction: small-group reading intervention for the bulk of the class while providing enrichment to the advanced reader.

Heading

Skewness: The LET's Favorite Trap (And How to Master It)

Examples

  • A Grade 2 phonics quiz scores: 80, 82, 85, 88, 90, 92, 95, 98, 100, 100 (eight pupils at 90–100, two below). Mean ≈ 91, Median = 92. Mean < Median suggests slight negative skew (easy test, most pupils mastered phonics).
  • A Grade 6 Advanced Mathematics assessment scores: 25, 30, 35, 40, 45, 50, 88 (one honors student). Mean = 45.3, Median = 40. Mean > Median indicates positive skew (hard test; most pupils struggled, one advanced pupil excelled).
  • A perfectly balanced Grade 4 Science test: 70, 75, 80, 80, 85, 90, 95. Mean = 82.1, Median = 80, Mode = 80. Mean ≈ Median ≈ Mode indicates normal distribution (balanced difficulty).

Key Points

  • Skewness is named after the TAIL, not the hump where scores pile up.
  • Positively skewed: tail RIGHT, scores LOW, test DIFFICULT, Mean > Median > Mode.
  • Negatively skewed: tail LEFT, scores HIGH, test EASY, Mean < Median < Mode.
  • Normal (symmetrical): Mean = Median = Mode, balanced test difficulty.
  • The mean is dragged toward the tail; use median for skewed distributions.
  • LET loves scenario items on skewness—reason through the logic rather than memorizing blindly.

Raw scores from different tests cannot be compared directly. A score of 40 on a 50-item quiz means something entirely different from 40 on an 80-item exam. A score of 80 in a Grade 4 Mathematics test is not comparable to 80 in a Grade 5 Mathematics test, which has different items and a different norm group. Standard scores solve this problem by converting raw scores to a common scale expressed in standard deviation units. They allow apples-to-apples comparison of performance across tests, subjects, and even years. **The z-Score: Standard Deviation Units** The z-score expresses a raw score in terms of standard deviations from the mean: **z = (X − Mean) / SD** where: - X = the raw score - Mean = the class or group mean - SD = the standard deviation The z-score tells you **how many standard deviations a score lies above (+) or below (−) the mean**. A z-score of +2.0 means the score is 2 SDs above the mean; a z-score of −1.5 means 1.5 SDs below. The z distribution always has a mean of 0 and an SD of 1, making it a universal scale. **Interpreting z-Scores** - **z = 0:** The raw score equals the mean. - **z > 0 (positive):** The raw score is above the mean. - **z < 0 (negative):** The raw score is below the mean. - **z ≥ ±2.0:** The score is notably extreme (top or bottom ~2.5% under normality). - **z ≥ ±3.0:** The score is very rare (top or bottom ~0.15% under normality). **Worked Example 1: Computing and Interpreting z-Scores** A Grade 5 class takes a Mathematics test with a class mean of 80 and SD of 4. Three pupils' scores: - **Ana scored 88:** z = (88 − 80) / 4 = 8 / 4 = **+2.0** (2 SDs above the mean, top ~2.5% of the class; exceptional performance) - **Ben scored 76:** z = (76 − 80) / 4 = −4 / 4 = **−1.0** (1 SD below the mean, bottom 16%; below grade-level expectation) - **Cara scored 80:** z = (80 − 80) / 4 = 0 / 4 = **0.0** (exactly at the mean, average performance) **The T-Score: A More User-Friendly Scale** Negative z-scores confuse parents and elementary-age pupils, who prefer positive numbers. The T-score (not to be confused with the t-test in statistics) rescales z-scores to a mean of 50 and an SD of 10, eliminating negatives: **T = 50 + 10z** Using the examples above: - Ana: T = 50 + 10(2.0) = **70** (far above average) - Ben: T = 50 + 10(−1.0) = **40** (below average) - Cara: T = 50 + 10(0.0) = **50** (average) On the T scale, 50 is average, 40 is 1 SD below average, 60 is 1 SD above average, and so on. This scale is common on standardized tests and educational reports. **Worked Example 2: Comparing Performance Across Subjects (Classic LET Computation)** This is the single most-tested standard score item on the LET. **Scenario:** Carlo took two tests: - **Mathematics:** Raw score 85, class mean 80, SD 5 - **English:** Raw score 88, class mean 86, SD 4 **Question:** In which subject did Carlo perform better relative to his class? **Solution:** Do NOT compare raw scores (85 vs. 88). Compute z-scores to account for different class means and spreads. - **Math z-score:** z = (85 − 80) / 5 = 5 / 5 = **+1.0** - **English z-score:** z = (88 − 86) / 4 = 2 / 4 = **+0.5** **Conclusion:** Carlo's z-score is higher in Mathematics (+1.0) than English (+0.5). He performed **better in Mathematics**, standing a full SD above his class average versus only half an SD in English. Despite the lower raw score in Math, his relative standing is stronger. **Why This Matters for Classroom Practice** Imagine two pupils in your Grade 4 class: - **Maya** scored 92 in a 100-point Reading test (mean 85, SD 4). z = (92 − 85) / 4 = +1.75 (top ~4% of class). - **Jordan** scored 82 in a 100-point Mathematics test (mean 80, SD 2). z = (82 − 80) / 2 = +1.0 (top ~16% of class). On raw score alone, Maya looks stronger (92 > 82). But in relative standing, Maya is in the top 4%, while Jordan is in the top 16%. This does not mean one is 'smarter'—it means Maya excels in a subject where the class is strong overall, while Jordan excels in a subject with less variability. Standard scores reveal the true picture. **Other Standard Score Scales in Use** - **Stanines (1–9):** Mean 5, SD ≈ 2. Common on U.S. standardized tests; less used in Philippines. - **Percentile Rank (0–100):** Position in group (already covered), not a standard score, but often reported alongside z or T. - **IQ-type scales (mean 100, SD 15):** Formula 100 + 15z; used in aptitude and intelligence testing. On the LET, stick with z and T scores; they are the standards you will most often encounter and compute.

Heading

Standard Scores: Comparing Across Tests and Subjects

Examples

  • A pupil scores 75 on a Grade 5 test (class mean 70, SD 5) and 78 on another test (class mean 74, SD 4). First test: z = (75 − 70) / 5 = +1.0. Second test: z = (78 − 74) / 4 = +1.0. Same relative standing; do not mistake the different raw scores for different performance levels.
  • Two Grade 4 pupils' English and Mathematics scores (class means and SDs provided). Pupil A: English 85 (mean 80, SD 3, z = +1.67), Mathematics 88 (mean 85, SD 5, z = +0.6). Pupil B: English 82 (mean 80, SD 3, z = +0.67), Mathematics 90 (mean 85, SD 5, z = +1.0). Pupil A excels in English; Pupil B excels in Mathematics relative to respective class performance.
  • A national standardized test (mean 50, SD 10) reports a pupil's z-score of +1.5. T-score = 50 + 10(1.5) = 65. The pupil is 1.5 SDs above the national average—an advanced performance indicating readiness for gifted programs.

Key Points

  • Raw scores from different tests cannot be compared; standard scores convert them to a common scale.
  • z-score = (X − Mean) / SD; expresses how many SDs a score is from the mean.
  • z = 0 at mean, z > 0 above mean, z < 0 below mean; z ≥ ±2.0 is notably extreme.
  • T = 50 + 10z; mean 50, SD 10, eliminates negatives; parent-friendly.
  • To compare performance across subjects or tests, use z-scores or T-scores, not raw scores.
  • LET loves 'compare two subjects' computation items—always compute z-scores first.

A correlation coefficient (r) quantifies the strength and direction of the linear relationship between two variables. In the classroom, you might ask: Do pupils who study more tend to score higher? Do absences relate to lower grades? Does reading fluency predict reading comprehension? Correlation answers these questions numerically. **Two Key Pieces of Information in a Correlation Coefficient** **1. Direction (Sign)** - **Positive correlation (r > 0):** The variables move together in the same direction. As one increases, the other tends to increase (e.g., study hours and test scores, reading fluency and reading comprehension). - **Negative correlation (r < 0):** The variables move in opposite directions. As one increases, the other tends to decrease (e.g., absences and grades, anxiety and test performance). - **No correlation (r ≈ 0):** No linear relationship; knowing one variable tells you nothing about the other (e.g., shoe size and reading ability). **2. Strength (Magnitude)** The correlation coefficient r ranges from −1.00 to +1.00. The closer |r| (absolute value) is to 1, the stronger the relationship. The sign (+ or −) indicates direction; strength depends only on magnitude. **Rough Conceptual Guide to Strength** | Magnitude | Interpretation | |---|---| | 0.00 to 0.20 | Negligible or no relationship | | 0.21 to 0.40 | Low or weak relationship | | 0.41 to 0.60 | Moderate relationship | | 0.61 to 0.80 | High or strong relationship | | 0.81 to 1.00 | Very high or very strong relationship | **Important:** A correlation of −0.85 is **stronger** than +0.60 because strength is about magnitude, not sign. The negative sign tells you direction (inverse); the 0.85 magnitude tells you strength (near-perfect). **Worked Example 1: Interpreting Correlations** A Grade 5 teacher computes correlations between pupil characteristics and test performance: - **Reading fluency and reading comprehension: r = +0.78** (high positive). Pupils who read fluently tend to comprehend well. Fluency supports comprehension. - **Hours studied and exam score: r = +0.65** (moderate positive). More study time relates to higher scores, but the relationship is not perfect (some pupils study inefficiently; some have background advantages). - **Days absent and exam score: r = −0.72** (high negative). More absences relate to lower scores. Attendance is critical. - **Shoe size and reading ability: r = ≈ 0.05** (negligible). No meaningful relationship; they are independent. **Worked Example 2: Reliability and Validity Coefficients (Test Quality)** Correlations underpin test quality indices required under DepEd assessment standards: - **Test-retest reliability: r = +0.85.** When the same test is given twice a few weeks apart, scores correlate at 0.85 (high). The test is reliable; it measures consistently over time. - **Criterion-related validity: r = +0.72 (correlation between a new reading test and teacher ratings of reading).** The new test correlates moderately with an accepted criterion. The test has acceptable validity; it measures what it claims to measure. These high correlations assure you that the test is dependable and valid—a prerequisite for fair grading under RA 7836's ethical obligation to use valid assessment tools. **The Critical Guard Rail: Correlation ≠ Causation** This is perhaps the most important concept in statistics. A strong correlation between two variables does NOT mean one causes the other. **Classic Example:** Ice cream sales and drowning deaths show a high positive correlation in summer. But ice cream does not cause drowning (they share a common cause: warm weather and increased outdoor activity). **In the Classroom:** A high correlation between study time and exam scores (r = +0.70) might suggest that more studying causes higher scores. But an alternative explanation is that motivated students both study more AND score higher; motivation, not study time itself, drives the correlation. Without random assignment (giving some pupils mandatory extra study and others none), you cannot claim causation. **Practical Application: What Correlations Tell You** You can use classroom correlations to identify support needs: - If pupil behavior (disruptions, off-task) correlates r = −0.68 with reading comprehension, behavior management is a lever for improving reading. - If family reading at home correlates r = +0.74 with vocabulary growth, a family literacy intervention might be effective. - If interactive whiteboard use correlates r = +0.15 with mathematics achievement, the tool is not driving improvement; other factors matter more. Correlations guide your pedagogical decisions. They flag relationships worth investigating further but require careful interpretation (avoiding causation claims) and professional judgment.

Heading

Correlation: Measuring Relationships Between Two Variables

Examples

  • Grade 4 summer school attendance (days attended) correlates r = +0.68 with end-of-summer reading gain. Attendees improved more than non-attendees, but causation is not proven; perhaps motivated families attended more AND read at home more, with home reading driving gain.
  • A teacher finds r = −0.81 between test anxiety (measured via questionnaire) and exam performance. High anxiety relates strongly to lower scores. The teacher institutes relaxation techniques; if anxiety drops, does performance rise? The correlation suggests a link worth testing, but the teacher cannot claim anxiety causes low performance without experimental evidence.
  • Grade 6 pupils' age (months since birth in a single cohort) and reading level show r ≈ 0.15 (negligible). Age within a grade does not predict reading ability; other factors (prior exposure, instruction quality) matter far more.

Key Points

  • Correlation coefficient r measures strength and direction of linear relationship between two variables.
  • Positive r: variables move together; Negative r: variables move oppositely; r ≈ 0: no relationship.
  • r ranges from −1.00 to +1.00; strength depends on magnitude (|r|), not sign.
  • Rough scale: 0–0.20 negligible, 0.21–0.40 low, 0.41–0.60 moderate, 0.61–0.80 high, 0.81–1.00 very high.
  • r = −0.85 is STRONGER than r = +0.60 (strength is about magnitude).
  • Reliability and validity coefficients (test quality) are correlation-based; high correlations assure test dependability.
  • CRITICAL: Correlation does NOT imply causation; shared causes and confounding variables complicate interpretation.

DepEd Order No. 8, s. 2015 is the official policy guideline on classroom assessment for the K-12 Basic Education Program. It governs how you collect, aggregate, and report grades in your elementary classroom. Compliance with this policy is a professional obligation under the Code of Ethics for Professional Teachers (RA 7836), which mandates that teachers "use assessment data to guide and improve the learning and development of learners" and "ensure assessment is valid, fair, and non-discriminatory." **Core Principles of the DepEd K-12 Grading System** The system is **standards-based and competency-based**, meaning grades reflect learner progress toward defined learning competencies, not personality traits or effort alone. Grades communicate: "Can the learner perform this competency?" not "Is the learner a good person?" It is **continuous and cumulative**, collecting evidence throughout the quarter, not relying on a single test. It is **learner-centered**, valuing diverse demonstrations of learning (written tests, performances, projects, portfolios) over pencil-and-paper exams alone. It is **transparent**, with clear learning targets, assessment methods, and grade communication to families. **The Three Components of Quarterly Assessment** Every quarterly grade is built from three components: | Component | What It Includes | Example in Grade 2 | |---|---|---| | **Written Work (WW)** | Formative, low-stakes written outputs; quizzes, short tests, essays, worksheets, journal entries, written reflections | Three weekly phonics quizzes (10 points each), one unit test on reading comprehension (30 points), one written reflection on favorite story (15 points) | | **Performance Tasks (PT)** | Summative demonstrations of learning through doing; projects, experiments, presentations, performances, portfolios, role-plays, demonstrations | Three performance tasks: (1) Read aloud a story fluently to classmates; (2) Conduct a simple science experiment and record observations; (3) Create a portfolio of best writing samples with reflections | | **Quarterly Assessment (QA)** | Summative exam or culminating performance at quarter's end; may be written test, oral exam, or performance | One comprehensive written examination covering all competencies taught that quarter; or, one culminating performance task like a class book fair or writing showcase | Notice that **Performance Tasks carry substantial weight** in the system, reflecting DepEd's learner-centered, competency-based philosophy. Raw knowledge (WW) is necessary but not sufficient; learners must *demonstrate* competencies in authentic, meaningful contexts (PT). **Weight Distribution by Subject (Grades 1–10)** DepEd Order No. 8 specifies weights that vary by subject, reflecting different learning emphases: | Subject Group | Written Work | Performance Tasks | Quarterly Assessment | |---|---|---|---| | **Languages, Araling Panlipunan, EsP** | 30% | 50% | 20% | | **Science, Mathematics** | 40% | 40% | 20% | | **MAPEH (Music, Arts, Physical Education, Health)** | 20% | 60% | 20% | | **EPP/TLE (Entrepreneurship, Technology, Livelihood Education)** | 20% | 60% | 20% | **Why These Weights?** Languages value communication competencies, best demonstrated through performance (reading aloud, oral presentations, dialogues). Mathematics values both procedural fluency (written work) and application (performance tasks like solving real-world problems). MAPEH and TLE emphasize *doing*—singing, dancing, creating, building—so performance tasks dominate. This weight distribution encodes educational priorities into the grading system. **How to Compute a Quarterly Grade: Step-by-Step Process** **Step 1: Collect and organize all evidence (WW, PT, QA scores) during the quarter.** Example for a Grade 3 pupil in Mathematics (WW 40%, PT 40%, QA 20%): - **Written Work:** Quiz 1: 8/10, Quiz 2: 7/10, Unit Test: 28/30, Worksheet: 18/20. Total raw = 8 + 7 + 28 + 18 = 61 points. Highest possible = 10 + 10 + 30 + 20 = 70 points. - **Performance Tasks:** Task 1 (measure classroom objects and calculate perimeter): 9/10. Task 2 (create a data collection chart from a survey): 8/10. Total raw = 17 points. Highest possible = 20 points. - **Quarterly Assessment:** Final exam: 38/50 points. **Step 2: Convert each component to a percentage score (component percentage).** - **WW%:** (61 / 70) × 100 = **87.14%** - **PT%:** (17 / 20) × 100 = **85%** - **QA%:** (38 / 50) × 100 = **76%** **Step 3: Apply the subject weights to calculate weighted scores.** Using Mathematics weights (WW 40%, PT 40%, QA 20%): - **WW weighted:** 87.14 × 0.40 = 34.86 - **PT weighted:** 85 × 0.40 = 34.00 - **QA weighted:** 76 × 0.20 = 15.20 **Step 4: Sum the weighted scores to get the initial grade.** **Initial grade = 34.86 + 34.00 + 15.20 = 84.06** (round to 84) **Step 5: Transmute the initial grade using the transmutation table.** The transmutation table converts the initial grade (0–100 scale) to a reported grade (0–100 scale, but with a minimum passing grade of 75). The table ensures that: - An initial grade of 100 transmutes to 100 (perfect performance). - An initial grade of 60 transmutes to 75 (minimum passing grade). - An initial grade below 60 transmutes to below 75 (not yet passing). - The lowest reported grade that may appear on the report card is 60. **Simplified Transmutation Table (DepEd Guidelines):** | Initial Grade | Reported Grade | |---|---| | 100 | 100 | | 95 | 95 | | 90 | 90 | | 85 | 85 | | 80 | 80 | | 75 | 75 | | 70 | 71 | | 65 | 68 | | 60 | 75 | | Below 60 | (Adjusted based on school policy, typically 60 as minimum) | (Note: Different schools may use slightly different transmutation tables; always check your school's official table.) For an initial grade of **84**, the reported grade is approximately **84** (no major transmutation needed; it falls in the typical range). If the initial grade were **60**, it would transmute to **75** (the minimum passing grade), ensuring that pupils who demonstrate basic competency receive credit. **Step 6: Report the grade and descriptor on the report card.** The pupil's quarterly grade is **84** (or transmuted value), which falls in the **Satisfactory** range: | Grade Range | Descriptor | |---|---| | 90–100 | **Outstanding (O)** | | 85–89 | **Very Satisfactory (VS)** | | 80–84 | **Satisfactory (S)** | | 75–79 | **Fairly Satisfactory (FS)** | | Below 75 | **Did Not Meet Expectations (DNM)** | **Annual and Final Grades** The **final grade** in a subject for the school year is the average of the four quarterly grades. Pupils who fall below 75 (Did Not Meet Expectations) on the final grade are flagged for remediation. Promotion or retention decisions follow DepEd guidelines, which typically allow one failing grade if other subjects are passing, but retention if multiple subjects are failed or if the pupil does not meet critical competencies (e.g., reading in Grade 3). **Grading in Special Contexts** **Special Education (SPED) Pupils:** Grades are based on progress toward the pupil's Individualized Education Plan (IEP), not grade-level standards. A SPED pupil may receive an Outstanding grade for consistent effort and progress, even if absolute performance is below grade level. Fairness demands this approach. **Late or Absent Pupils:** If a pupil was absent much of a quarter, collect evidence when the pupil returns. Do not penalize for absences in the grade itself; instead, flag absences on the attendance record and consider remediation. Under RA 7610 (Special Protection of Children Against Child Abuse, Exploitation and Discrimination Act), using grades punitively violates the child's right to education. **Pupils with Different Learning Needs:** Differentiate assessment. Some pupils may demonstrate competencies through alternative performance tasks (e.g., oral report instead of written essay) while maintaining rigor. Documentation of differentiation protects you professionally and ensures fairness. **Reporting and Communicating Grades to Parents** Your quarterly report card is only the beginning. Under DepEd policy, you must communicate results to parents and learners regularly, not just at quarter's end. Use parent-teacher conferences, progress monitoring reports, and written feedback to: - Celebrate strengths and competencies mastered. - Identify specific areas for improvement (not vague criticism). - Suggest concrete strategies for home support (reading together, practicing math facts). - Listen to parent and learner perspectives (perhaps home factors, health issues, or learning style differences explain performance). This is especially important under the Code of Ethics for Professional Teachers (RA 7836), which obligates you to "inform parents of learner progress and to listen to their suggestions and legitimate grievances." Grades in isolation are incomplete feedback; partnerships with families matter. **Common Grading Pitfalls to Avoid** 1. **Not collecting enough evidence in Written Work.** If you only give one big test per quarter, Written Work is unreliable. Use frequent, low-stakes quizzes and worksheets. 2. **Conflating effort with competency.** A struggling pupil who tries hard deserves recognition, but should not receive an A if competencies are not met. Instead, note effort in a comment and design support to build competency. 3. **Including non-academic factors (behavior, attendance) in the grade.** Behavior and attendance belong on separate report card sections, not in the subject grade. Subject grades reflect competency only. 4. **Double-counting.** Do not count a performance task and then again count a written reflection on that same task as separate evidence. One demonstration of learning, once. 5. **Ignoring the transmutation table.** If policy says an initial 60 transmutes to 75, apply it consistently. Do not arbitrarily change grades. 6. **Not reviewing for bias.** Examine your grades for patterns: Do boys consistently score higher in Math than girls? Do wealthy pupils' grades exceed poor pupils'? Use data to check for unconscious bias and adjust assessment practices.

Heading

The DepEd K-12 Grading System (DepEd Order No. 8, s. 2015)

Examples

  • Grade 2 Language Arts pupil's 1st Quarter (WW 30%, PT 50%, QA 20%): WW% = 85, PT% = 88, QA% = 82. Quarterly grade = (85 × 0.30) + (88 × 0.50) + (82 × 0.20) = 25.5 + 44 + 16.4 = 85.9, rounded to 86. Reported as 86 (Very Satisfactory, VS).
  • Grade 5 Science pupil's 4th Quarter (WW 40%, PT 40%, QA 20%): WW% = 75, PT% = 70, QA% = 65. Quarterly grade = (75 × 0.40) + (70 × 0.40) + (65 × 0.20) = 30 + 28 + 13 = 71. Transmuted to 72 (Fairly Satisfactory, FS). Pupil needs remediation in Science.
  • A Grade 4 pupil attended only 60% of the quarter due to family illness. Collected evidence from days present: WW% = 82, PT% = 78, QA% = 75. Quarterly grade = 79 (Fairly Satisfactory, FS). The grade reflects competency; absences are documented separately. Upon return, the teacher provides 'catch-up' opportunities for missed learning, not as separate grades but as accelerated instruction toward mastery.

Key Points

  • DepEd Order No. 8, s. 2015 governs K-12 grading; it is standards-based, competency-based, continuous, and learner-centered.
  • Three components: Written Work (WW, 20-40%), Performance Tasks (PT, 40-60%), Quarterly Assessment (QA, 20%).
  • Weights vary by subject: Languages/AP/EsP (WW 30%, PT 50%, QA 20%), Math/Science (WW 40%, PT 40%, QA 20%), MAPEH/TLE (WW 20%, PT 60%, QA 20%).
  • Quarterly grade = (WW% × weight) + (PT% × weight) + (QA% × weight), then transmuted to reported grade.
  • Transmutation table: initial 100 → 100, initial 60 → 75 (minimum passing), lowest reported grade is 60.
  • Grade descriptors: 90-100 Outstanding, 85-89 Very Satisfactory, 80-84 Satisfactory, 75-79 Fairly Satisfactory, below 75 Did Not Meet Expectations.
  • Final grade = average of four quarterly grades; determines promotion or remediation.
  • Communicate grades regularly to parents; celebrate strengths and identify specific areas for growth.
  • Do not include behavior or attendance in subject grades; keep them on separate report card sections.
  • Review grades for bias; ensure fair, non-discriminatory assessment (RA 7836 and RA 7610 obligations).

Understanding statistics and the grading system means nothing without skill in interpreting real assessment results and taking action. This section synthesizes all concepts and shows how a reflective teacher uses data to improve learning. **Step 1: Examine the Shape and Spread of Your Class Distribution** After a major assessment (unit test, quarterly exam), plot the scores and examine the distribution visually (via histogram, dot plot, or stem-and-leaf plot). **Question:** Is the distribution normal (symmetrical), positively skewed, or negatively skewed? - **Positively skewed (tail right, hump left):** Most pupils scored low; the test was difficult or the class needs reteaching. **Median and mode better represent the typical student than the mean.** Plan targeted intervention for the majority. - **Negatively skewed (tail left, hump right):** Most pupils scored high; the test was easy or the class mastered the competency. **The class is ready to progress.** Use advanced learners as peer tutors for stragglers. - **Symmetrical (normal):** Balanced difficulty; some pupils mastered, others did not, but the distribution is reasonable. **Use the mean and SD to identify pupils within 1 SD (typical) versus those 2+ SDs away (notably high or low).** **Worked Interpretation 1: Positively Skewed Grade 3 Reading Quiz** Twenty Grade 3 pupils' phonics quiz (out of 20 items): 6, 8, 9, 10, 10, 11, 12, 12, 13, 14, 14, 15, 16, 16, 17, 17, 18, 19, 19, 20. - **Mean:** (6 + 8 + 9 + ... + 20) ÷ 20 = 265 ÷ 20 = 13.25 (66%) - **Median:** (14 + 14) ÷ 2 = 14 (70%) - **Mode:** 10, 14, 16, 17, 19 (bimodal; no single mode; scores cluster at low-mid range) - **Range:** 20 − 6 = 14 points; **SD ≈ 3.8** - **Distribution shape:** Mean (13.25) < Median (14) appears to contradict the positively-skewed rule... let me reconsider. Actually, the hump is at 14-17 and a tail stretches down to 6-9, so the tail is at the LOW end. This is a **negatively skewed** distribution (tail left, hump right). Most pupils scored in the 10-20 range; the quiz was easy or the class learned well. **Teacher's Interpretation:** The class as a whole demonstrated phonics competency (median 70%, mean 66%). The one pupil who scored 6 (30%) is a clear outlier and needs intensive one-on-one phonics intervention. No whole-class reteach is needed; instead, small-group instruction for the 3-4 pupils scoring below 10, and enrichment for those scoring 18+. **Worked Interpretation 2: Positively Skewed Grade 5 Mathematics Unit Test** Twenty-five Grade 5 pupils' 50-item Mathematics test: 15, 18, 20, 22, 25, 26, 28, 30, 31, 32, 32, 33, 34, 35, 35, 36, 37, 38, 40, 42, 43, 44, 45, 46, 48. - **Mean:** (15 + 18 + ... + 48) ÷ 25 = 775 ÷ 25 = 31 (62%) - **Median:** The 13th score is 34 (68%) - **Range:** 48 − 15 = 33 points; **SD ≈ 8.5** - **Distribution shape:** Hump at 30-40s, tail at the low end (15-25). Mean (31) < Median (34), confirming negative skew... wait, that contradicts the positively-skewed rule. Let me recheck: **tail at LOW end = negatively skewed? NO. Tail at LOW end = tail points left = negatively skewed. But scores pile up at high end here... Actually, most scores are 30+, with the tail stretching down to 15. Tail points to the low end (left), scores pile at high end (right). This is NEGATIVELY skewed.** Actually, let me reconsider the distribution name one more time: The **tail points toward the low end (left), so this is negatively skewed.** But that contradicts the mean-median order. Mean < Median suggests negative skew, which checks out. Most students scored 30-48 (60-96%); only a few scored 15-25. The test was within reach; the class performed reasonably well. **Teacher's Interpretation:** The median (68%) and mean (62%) indicate room for growth but respectable performance. Using the SD (≈8.5), pupils within 1 SD of the mean scored 22-39 (44-78%), a wide range. Pupils scoring below 22 (4 pupils) are struggling with key competencies and need small-group reteach focusing on foundational skills. Pupils scoring above 40 (8 pupils) have mastered the unit and are ready for enrichment (challenge problems, peer tutoring, or advanced topics). The bulk of the class (13 pupils, 52-78%) is consolidating understanding and ready to move forward with targeted reinforcement. **Step 2: Identify Individual Pupils Using Standard Scores** Compute the z-score or percentile rank for each pupil to pinpoint those notably high or low. **Criterion:** Pupils with z ≥ +1.5 (top ~7%) merit recognition and enrichment. Pupils with z ≤ −1.5 (bottom ~7%) need intervention and support. **Worked Example: Grade 4 Mathematics Quiz** Class mean = 22, SD = 3 (out of 30 points). - **Pupil A (score 28):** z = (28 − 22) / 3 = +2.0. Top ~2.5%; outstanding performance. Recognize and challenge. - **Pupil B (score 22):** z = (22 − 22) / 3 = 0.0. Average performance. On-track. - **Pupil C (score 16):** z = (16 − 22) / 3 = −2.0. Bottom ~2.5%; struggling. Intensive support needed. **Step 3: Compare Across Subjects for Individual Pupils** Use z-scores to identify relative strengths and weaknesses. **Worked Example: Grade 5 Pupil's Quarterly Performance** - **Language Arts:** Raw 85, class mean 80, SD 4. z = +1.25 (above average, very satisfactory) - **Mathematics:** Raw 75, class mean 78, SD 6. z = −0.50 (slightly below average, satisfactory) - **Science:** Raw 82, class mean 75, SD 5. z = +1.40 (above average, very satisfactory) - **Araling Panlipunan:** Raw 70, class mean 72, SD 3. z = −0.67 (below average, fairly satisfactory) **Interpretation:** The pupil is a strong reader and scientist, average in Math, and weak in AP. Provide tutoring in Math and AP, and encourage the pupil to join the Science club. This profile is more nuanced than just looking at raw scores. **Step 4: Examine Correlation Patterns to Identify Support Levers** Compute correlations between pupil characteristics (attendance, behavior, home reading time via parent report) and academic performance. **Example Correlations (Grade 4 cohort, n=25):** - Daily attendance and Mathematics achievement: r = +0.72 (high positive). **Action:** Prioritize attendance monitoring; work with families on chronic absenteeism. - Minutes spent in small-group reading instruction per week and reading fluency gain: r = +0.68 (moderate positive). **Action:** Increase instructional time for struggling readers to boost fluency. - Use of interactive whiteboard during lessons and engagement (observed on-task behavior): r = +0.15 (negligible). **Action:** The whiteboard is not a magic tool; focus on instructional design, not technology. **Step 5: Use Data to Adjust Instruction** Data informs next steps: **If most pupils scored low (positively skewed):** - Reteach the unit using different strategies (visual learners, kinesthetic activities, small groups). - Slow the pace; do not advance until foundational understanding is solid. - Provide formative assessments frequently (daily exit tickets, quizzes) to monitor progress before the next summative test. **If performance is highly variable (large SD, wide range):** - Implement differentiated instruction: flexible grouping based on mini-assessments. - Provide tiered activities at different complexity levels during independent practice. - Use peer tutoring: advanced pupils help struggling peers. **If some pupils excel while others struggle:** - Identify the advanced pupils and provide enrichment: challenge problems, independent projects, leadership roles. - Identify the struggling pupils and provide intensive intervention: one-on-one or small-group tutoring, shorter practice sessions, simpler materials initially, then gradual increase in complexity. - Do not allow a one-size-fits-all approach; respond to data. **Step 6: Communicate Results Meaningfully to Parents and Learners** A quarterly report card is data; feedback is communication that turns data into understanding and action. **Avoid unclear statements:** "Pupils need to improve in Mathematics." (vague) **Instead, use data:** "Your child mastered addition with regrouping (90%) but still struggles with subtraction with regrouping (62%). I recommend practicing subtraction problems at home using this worksheet. We will provide small-group practice during math centers next week." **Avoid comparing siblings or peers:** "Your daughter did better than her brother did in Grade 4 Math." (unfair, discouraging to the sibling) **Instead, use growth:** "Your daughter has improved from 65% to 78% in reading comprehension over the quarter. She is making excellent progress and is on track to meet grade-level standards by year's end." **Involve the learner:** Ask pupils to reflect on their own data: "Look at your quiz scores this quarter. Which topics are you confident about? Which do you find tricky? What will help you improve?" This metacognition builds agency and responsibility. **Step 7: Document Your Interpretation and Actions** Keep a data journal: - After each major assessment, note the class distribution, mean, SD, and your interpretation. - Record which pupils are excelling (z > +1.5), which are struggling (z < −1.5). - List the instructional adjustments you made (reteach, differentiated groups, enrichment). - Note improvement in subsequent assessments; use this as evidence of the effectiveness of your data-driven adjustments. This documentation serves multiple purposes: - **Reflective practice:** You become a teacher-researcher, continuously improving. - **Professional accountability:** If a parent questions a pupil's grade, you have data and reasoning to support your decisions (required under RA 7836). - **Curriculum improvement:** Over multiple years, patterns emerge; you identify which topics pupils consistently struggle with and redesign instruction. **Fairness and Non-Discrimination: A Critical Note** Under RA 7836 (Code of Ethics for Professional Teachers) and RA 7610 (Special Protection of Children), assessment and grading must be non-discriminatory. Check your interpretation for bias: - Are girls and boys achieving equally, or do patterns emerge by gender? If girls consistently score higher in Language Arts, ensure boys have equal access to oral language practice, not less challenging texts. - Do wealthy pupils' grades exceed poor pupils'? Investigate: Is assessment unfair (assumes prior knowledge, emphasizes styles that favor certain backgrounds)? Or do home factors (parental education, home resources) explain the difference? Address root causes, not the grade itself. - Are pupils with disabilities or learning differences included in assessments (with accommodations per their IEPs) and given equitable feedback and support? Data-driven instruction is powerful, but it is only ethical when applied fairly to all learners.

Heading

Interpreting Assessment Results in the Classroom: Putting It All Together

Examples

  • After a Grade 3 Reading Fluency screening (DIBELS or similar): mean 65 wpm, SD 15, range 25-95. SD is large; the class is heterogeneous. Median = 68. Ten pupils score below 50 wpm (bottom 17%); provide daily small-group fluency instruction. Eight pupils score above 85 wpm (top 17%); pair as peer models and give them more complex texts. The bulks of 20 pupils (68%) span 50-85 and are on track.
  • A Grade 4 pupil's Science performance: Quiz 1 = 15/20, Quiz 2 = 14/20, Performance Task (model ecosystem) = 9/10, Quarterly Exam = 28/50. Converted to percentages: 75%, 70%, 90%, 56%. WW% = (15+14)/40 = 72.5%, PT% = 90%, QA% = 56%. Weights (WW 40%, PT 40%, QA 20%): Grade = (72.5 × 0.4) + (90 × 0.4) + (56 × 0.2) = 29 + 36 + 11.2 = 76.2 (Fairly Satisfactory, FS). Pupil excels in performance (environmental understanding, hands-on investigation) but struggles on written exams. Teacher provides test-taking strategy instruction and ensures upcoming quarterly exam values multiple formats (short-answer, matching, diagram-labeling, not only multiple-choice).
  • Grade 5 end-of-quarter analysis: Class Math mean = 80, median = 82 (negatively skewed, test slightly easy). But SD = 12, indicating heterogeneity. Five pupils scored above 92 (top 10%); they are accelerated to Grade 6 topics. Seven pupils scored below 70 (bottom 15%); they join an after-school Math clinic twice weekly. The middle 18 move forward. Teacher sends home a note: 'Great job on fractions! Most of the class mastered adding and subtracting fractions. A few pupils still need practice; we'll do more in class. If your child was in the clinic group, we'll provide worksheets for home practice and celebrate progress next quarter.'

Key Points

  • After each major assessment, examine the shape of the class distribution (normal, positively skewed, negatively skewed) to guide instructional response.
  • Use z-scores to identify pupils notably high (z > +1.5) or low (z < −1.5) for enrichment and intervention.
  • Compare pupils across subjects using z-scores to identify relative strengths and weaknesses.
  • Compute correlations between pupil characteristics and achievement to identify support levers (attendance, instructional time, family engagement).
  • Adjust instruction based on data: reteach if most scored low, differentiate if spread is wide, enrich advanced learners, intervene with struggling learners.
  • Communicate results meaningfully to parents: specific feedback tied to data, growth over time, concrete suggestions, no unfair comparisons.
  • Document interpretation and actions in a data journal for reflection, accountability, and continuous improvement.
  • Check grades and interpretations for bias; ensure non-discriminatory, equitable assessment and support for all learners (RA 7836, RA 7610).
Loading diagram…
Loading diagram…
Loading diagram…
Loading diagram…
Loading diagram…

Ready to practise for the LET Secondary 2026?

Super Tutor's AI review plan adapts to your weak areas and builds a weekly practice schedule around your target LET Secondary exam date.