LET Secondary Assessment of Learning — Constructing, Administering and Analyzing TestsDetailed Explanation
This is the "office hours" version of Constructing, Administering and Analyzing Tests for the LET Secondary 2026. No shortcuts, no hand-waving — just a full unpacking of why Professional Regulation Commission (PRC) cares about each concept and how the Assessment of Learning section items tend to play out on exam day. Read this once, then hit the practice questions with real understanding.
Exam context
The Licensure Examination for Professional Teachers — Secondary is conducted by Professional Regulation Commission (PRC) and is scheduled for Bi-annual. The Assessment of Learning subtest is marked as "Core" in the official pattern, and Constructing, Administering and Analyzing Tests appears in position 2nd of 5 in the LET Secondary Assessment of Learning review rotation. Passing mark: Weighted average of 75% with no grade below 50%. Recent LET Secondary 2026 papers have drawn roughly a meaningful share of questions from this subject.
Constructing, Administering and Analyzing Tests - Detailed Explanation
One of the most important professional skills of a licensed Filipino teacher is the ability to construct valid, reliable, and fair tests. Under RA 7836 (Philippine Teachers Professionalization Act), teachers are expected to be competent in assessment — from writing items to scoring and interpreting results. This chapter covers three interconnected phases: (1) writing selected-response and constructed-response items, (2) assembling and administering tests fairly, and (3) analyzing items through difficulty and discrimination indices. The LET for Elementary Level regularly includes computation problems on item analysis as well as item-critique questions where you must identify which guideline was violated. Mastering this chapter means you are ready to build tests that serve your Grade 1–6 pupils honestly and professionally, consistent with DepEd's Assessment Guidelines (DO 8, s. 2015) and the Code of Ethics for Professional Teachers.
Concepts
Writing Multiple-Choice Questions (MCQ)
A multiple-choice question (MCQ) consists of two parts: the STEM, which presents the problem, and the OPTIONS, which include one correct answer called the KEY and several wrong answers called DISTRACTORS. The goal of every MCQ is to measure a specific learning competency while giving no unintended hints to the test-taker. STEM RULES: 1. State a single, complete problem in the stem. The student should understand what is being asked BEFORE reading the options. 2. Put as much wording as possible in the stem to keep the options short. 3. Avoid negative stems such as 'Which of the following is NOT...' If you must use a negative, bold or capitalize it (NOT, EXCEPT) to reduce misreading. 4. Avoid grammatical clues — if the stem ends with 'an ___,' only options beginning with a vowel sound are grammatically correct, which is an unintended hint. OPTION RULES: 1. Make all distractors PLAUSIBLE — they must reflect real errors learners make, not obvious nonsense. 2. Keep all options HOMOGENEOUS in content, grammar, and length — the key must not stand out as the longest or most detailed. 3. Ensure ONE clearly best answer exists. 4. Avoid 'all of the above' — a student who recognizes any two correct options can guess the answer by elimination. 5. Avoid absolute words (always, never, all, none) in distractors — experienced test-takers know these are usually wrong. 6. Arrange options logically (alphabetically or numerically) and randomize where the key falls across items. PHILIPPINE CLASSROOM EXAMPLE: A Grade 5 Science teacher writes: 'Which organ pumps blood throughout the body? A. Lungs B. Brain C. Heart D. Kidney.' This item follows the rules — the stem is clear, options are homogeneous (all are body organs), and there is one best answer (C).
Examples
In the LET, item-critique questions ask you to NAME the specific rule that was violated, not just say the item is bad. Practice pinpointing: grammatical clue, length give-away, non-homogeneous options, implausible distractor, or absolute word.
Scenario
Identify the flaw in this item: 'A synonym of benevolent is an ___. A. kind word B. adjective that describes a person who is generous, charitable, and disposed to doing good C. cruel word D. small word'
Solution
This item violates THREE rules: (1) The article 'an' in the stem grammatically points only to option B, which starts with a vowel sound — this is an unintended grammatical clue. (2) Option B is far longer than the others — the key stands out by length. (3) The options are not homogeneous — B is a full definition while A, C, and D are two-word phrases.
Good distractors are based on typical errors of your target learners. For Grade 3 Math, the distractors should attract pupils who made a multiplication mistake — not random numbers.
Scenario
A Grade 3 Math teacher writes: 'What is 7 × 8? A. 54 B. 56 C. 58 D. 64.' Evaluate this item.
Solution
This is a well-constructed item. The stem is clear and complete. Options are homogeneous (all are two-digit numbers near 56). Distractors (54, 58, 64) reflect actual computation errors pupils make (e.g., 7×8=54 is a common mistake; 8×8=64 is a near-miss). There is one correct answer (B. 56).
Applications
- Used in quarterly assessments and periodical exams for Grades 1–6 in all subject areas
- Applied when constructing Table of Specifications (TOS) aligned with K-12 Most Essential Learning Competencies (MELCs)
- Used in DepEd-standardized achievement tests at Grade 3 and Grade 6 levels
- Essential for constructing pre-tests and post-tests in action research and school-based assessment
Misconceptions
- MISCONCEPTION: 'The correct answer should be longer and more detailed than the distractors because it needs to be complete.' TRUTH: Options must be equal in length and detail. An unusually long option signals the key.
- MISCONCEPTION: 'Using always and never makes distractors harder.' TRUTH: Experienced test-takers know that extreme absolutes are almost always false, so these words actually make distractors easier to eliminate.
- MISCONCEPTION: 'All of the above is a good option because it tests if students know all answers.' TRUTH: A student who recognizes just two correct options can deduce that all of the above is correct, rewarding partial knowledge.
Related Concepts
- Table of Specifications (TOS)
- Bloom's Taxonomy — cognitive level of items
- Validity and reliability of tests
- Distractor analysis in item analysis
Common Exam Questions
Example
LET-style: 'Which guideline was violated in this MCQ stem: All of the following are causes of poverty EXCEPT ___? A. unemployment B. lack of education C. corruption D. hardworking citizens' — Answer: Option D is implausible (clearly not a cause of poverty) so it is a non-functioning distractor. Also the negative EXCEPT is not emphasized.
Approach
Read the item carefully. Check the stem for negatives, grammatical clues, and completeness. Check options for homogeneity, length equality, plausibility of distractors, and absence of absolutes. Name the specific rule violated.
Question Type
Item Critique
Example
LET-style: 'Which statement is TRUE about constructing MCQ distractors? A. They should be obviously wrong. B. They should reflect common learner errors. C. They should be shorter than the key. D. They should use absolute words like always.' — Answer: B
Approach
From four options describing MCQ construction guidelines, select the one that correctly states the rule.
Question Type
Best Practice Identification
Key Points To Remember
- MCQ = stem (the problem) + key (correct answer) + distractors (plausible wrong answers)
- The stem must present a complete, single problem before the student reads the options
- Options must be homogeneous — same type, same length, same grammar
- Avoid 'all of the above' because it rewards partial knowledge
- Avoid absolute terms (always, never) in distractors — they signal FALSE to savvy test-takers
- Grammatical clues (e.g., 'an ___') must be eliminated from the stem
- The correct answer should not be obviously longer or more specific than the distractors
Writing True-False and Matching Items
TRUE-FALSE ITEMS measure whether a statement is factually correct or incorrect. They are quick to answer and easy to score but are vulnerable to guessing (50% chance), so they require careful construction. RULES FOR TRUE-FALSE: 1. Test ONE idea per statement — never combine two facts, one true and one false, in a single statement. 2. Avoid SPECIFIC DETERMINERS — words like all, always, never, none lean toward FALSE; words like sometimes, generally, usually lean toward TRUE. Test-savvy students exploit these patterns. 3. Avoid double negatives ('It is not untrue that...') which confuse rather than test. 4. Avoid trivial facts or tricky wording. 5. Keep roughly equal numbers of true and false statements and vary their order randomly. MATCHING ITEMS test the ability to associate related concepts, such as terms with definitions, events with dates, or laws with their provisions. RULES FOR MATCHING: 1. Keep both the PREMISES (left column) and RESPONSES (right column) HOMOGENEOUS — use one consistent basis for matching (e.g., all terms and their definitions, not a mix of terms and questions). 2. Provide MORE RESPONSES than premises so the last answer cannot be gotten by elimination. 3. Keep lists SHORT — about 5 to 8 premises per set is ideal. 4. Place shorter items on the RIGHT (responses column) to reduce reading load. 5. Write clear directions that state the basis for matching and whether responses can be used more than once. PHILIPPINE CLASSROOM EXAMPLE: A Grade 4 Social Studies teacher creates a matching exercise pairing Philippine national heroes (left column, 8 names) with their specific contribution (right column, 10 descriptions — 2 extra). This follows all matching rules.
Examples
A corrected version: 'Rizal wrote his novels Noli Me Tangere and El Filibusterismo in Spanish.' — single idea, no determiners, clearly either true or false based on content knowledge.
Scenario
Evaluate this true-false item: 'Rizal always wrote his novels in Spanish and they were never written in any other language.' T or F?
Solution
This item is FLAWED. First, it combines TWO ideas (language he wrote in AND exclusivity of language). Second, it uses specific determiners 'always' and 'never' which signal that the statement is likely FALSE — test-wise students can guess the answer without actual knowledge.
Extra responses eliminate the 'process of elimination' shortcut that allows students to get the last answer correct without truly knowing it.
Scenario
A Grade 6 teacher has 6 premises and 6 responses in a matching exercise. What is the problem?
Solution
When premises and responses are equal in number (6 and 6), the last match can be answered by elimination without knowing the content. The teacher should add at least 2 extra responses (making it 6 premises, 8 responses) to prevent this.
Applications
- True-false items are useful for quick comprehension checks on factual content in Science and Araling Panlipunan
- Matching items are ideal for vocabulary exercises in English and Filipino language classes
- Both formats are used in formative assessments and unit tests across Grade 1–6
Misconceptions
- MISCONCEPTION: 'Equal numbers of responses and premises in matching makes it balanced and fair.' TRUTH: Equal numbers allow elimination guessing. Always provide more responses than premises.
- MISCONCEPTION: 'Using always and never makes true-false items more precise.' TRUTH: These are specific determiners that signal FALSE to test-wise students, reducing the item's validity.
Related Concepts
- Multiple-choice item construction
- Item analysis — difficulty and discrimination
- Objective testing vs. subjective testing
Common Exam Questions
Example
LET-style: 'What is wrong with this matching exercise that has 5 premises and exactly 5 responses?' — Answer: The number of responses equals the number of premises, allowing the last answer to be gotten by elimination. There should be more responses than premises.
Approach
Identify the specific rule violated in a presented true-false or matching item.
Question Type
Rule Application
Key Points To Remember
- True-false: test ONE idea per statement only
- Specific determiners: always/never/all/none → lean FALSE; sometimes/usually/generally → lean TRUE
- Matching: always provide MORE responses than premises to prevent elimination guessing
- Matching lists should be SHORT (5–8 premises) and HOMOGENEOUS
- True-false items are highly vulnerable to guessing (50-50 chance)
- Directions for matching must state if responses can be used more than once
Writing Constructed-Response Items: Completion and Essay
CONSTRUCTED-RESPONSE ITEMS require students to produce their own answers rather than select from given choices. There are two main types: completion (short answer) and essay. COMPLETION ITEMS (Short Answer): These require a single, brief, correct answer — typically a word, phrase, or number. RULES: 1. Place the BLANK NEAR THE END of the statement, AFTER the problem has been clearly set up. 2. Use BLANKS OF EQUAL LENGTH so the length of the blank does not hint at the answer. 3. Avoid grammatical clues — never write 'a ___' or 'an ___' before the blank because the article reveals whether the answer begins with a vowel or consonant. 4. Require ONE specific answer; avoid items with multiple acceptable completions. ESSAY ITEMS: Essays measure higher-order thinking — synthesis, evaluation, and organization — that objective tests cannot capture. There are two subtypes: 1. RESTRICTED-RESPONSE ESSAY: Limits both the CONTENT and the FORM of the answer. The task is narrow and specific (e.g., 'List three causes of poverty in the Philippines and explain each in two sentences.'). Easier to score with a checklist. Samples more content areas per test. 2. EXTENDED-RESPONSE ESSAY: Gives the student FREEDOM to organize, argue, and demonstrate higher-order thinking (e.g., 'Evaluate the impact of the K-12 Reform on Filipino teachers.'). Richer in measuring deep understanding but harder to score consistently. SCORING ESSAYS RELIABLY: 1. Prepare a MODEL ANSWER or RUBRIC before scoring (not after reading student responses). 2. Score ONE ITEM ACROSS ALL PAPERS before moving to the next item — this is called the point method or item-by-item scoring. It keeps your standard consistent. 3. Score ANONYMOUSLY where possible to reduce HALO EFFECT — the tendency to rate a student's answer based on your overall impression of that student. 4. Use an ANALYTIC RUBRIC (breaks performance into specific criteria with separate scores) for detailed feedback or a HOLISTIC RUBRIC (single overall score) for quick judgment.
Examples
Grammatical clues in completion items reduce validity because students can use grammar knowledge rather than content knowledge to fill in the blank.
Scenario
Evaluate this completion item: 'The national hero of the Philippines, who wrote Noli Me Tangere, is an ___.'
Solution
This item is FLAWED. The article 'an' before the blank reveals that the answer begins with a vowel sound (e.g., 'author' or 'Ilustrado'). A corrected version: 'The Filipino national hero who wrote Noli Me Tangere is ___.' — no article, blank at end, single specific answer.
For elementary pupils, restricted-response essays are generally more appropriate because they provide clear structure, are easier to score reliably, and still measure analytical thinking within a defined scope.
Scenario
A Grade 6 teacher wants to assess students' ability to analyze the causes of deforestation. Which type of essay is more appropriate?
Solution
A RESTRICTED-RESPONSE essay is more appropriate: 'List three causes of deforestation in the Philippines and explain each cause in two sentences.' This limits the scope (three causes, two sentences each), making it easier to score consistently using a checklist.
Applications
- Completion items are used in formative tests for Science, Araling Panlipunan, and Mathematics at Grades 1–6
- Restricted-response essays are used in quarterly Performance Tasks under DepEd DO 8, s. 2015
- Extended-response essays appear in culminating tasks and portfolio assessments in upper grades
- Rubric-based scoring aligns with DepEd's assessment for learning philosophy
Misconceptions
- MISCONCEPTION: 'Placing the blank at the beginning of a completion item makes it clearer.' TRUTH: Placing it at the beginning requires students to keep an unknown in mind while reading the rest, making comprehension harder. Place it at the end.
- MISCONCEPTION: 'Extended-response essays are always better than restricted-response because they measure more.' TRUTH: Extended-response essays are harder to score reliably. Restricted-response essays sample more content and are easier to score consistently.
- MISCONCEPTION: 'Scoring all of one student's essay answers before moving to the next is more efficient.' TRUTH: This leads to halo effect. Score one item across all papers before moving to the next item.
Related Concepts
- Rubric construction — analytic vs. holistic
- Halo effect in assessment
- Performance-based assessment
- DepEd DO 8, s. 2015 — Policy Guidelines on Classroom Assessment
Common Exam Questions
Example
LET-style: 'What is the main problem with scoring essay tests without a rubric prepared in advance?' — Answer: Without a pre-prepared rubric, the teacher's scoring standard shifts as they read more papers, making the scores unreliable and subjective.
Approach
Determine whether a completion or essay item follows construction guidelines. Name the specific violation if any.
Question Type
Item Evaluation
Example
'Explain in your own words, in any number of paragraphs, everything you know about the Philippine Revolution.' — This is an EXTENDED-RESPONSE essay because it gives the student freedom in content selection, length, and organization.
Approach
Identify whether an essay prompt is restricted-response or extended-response based on how much freedom the student has.
Question Type
Classification
Key Points To Remember
- Completion: place blank NEAR THE END after the problem is stated
- Completion: use EQUAL-LENGTH blanks — length must not be a clue
- Completion: avoid grammatical articles (a/an) before the blank
- Restricted-response essay: LIMITS content and form — easier to score, samples more content
- Extended-response essay: GIVES FREEDOM — richer but harder to score reliably
- Score essays item-by-item (one question across all papers) to maintain consistent standards
- Prepare rubric or model answer BEFORE scoring, not after reading papers
- Score anonymously to prevent halo effect
Assembling and Administering the Test
After writing items, the teacher must ASSEMBLE them into a proper test and ADMINISTER it fairly. These steps are professional duties under the Code of Ethics for Professional Teachers and RA 7836, which requires teachers to be competent in their professional practice including assessment. ASSEMBLING THE TEST: 1. GROUP ITEMS BY TYPE — all true-false items together, then all MCQs, then all essays. Mixed formats confuse students and slow down answering. 2. ARRANGE FROM EASY TO DIFFICULT — start with items most students can answer to build confidence and reduce test anxiety. Hard items come last. 3. WRITE CLEAR DIRECTIONS for each section stating: what to do, how to answer, and how points are earned. 4. Keep an ITEM AND ITS OPTIONS ON THE SAME PAGE — never break an MCQ across two pages. 5. Space items adequately for readability — crowded tests increase reading errors. 6. Prepare the ANSWER KEY and scoring plan BEFORE printing and distributing the test. ADMINISTERING THE TEST: 1. Provide a COMFORTABLE, WELL-LIT, QUIET environment. Under RA 7610 (Special Protection of Children Act), teachers must ensure that the assessment environment does not cause undue stress or psychological harm to pupils. 2. Give CLEAR ORAL AND WRITTEN INSTRUCTIONS and state the time limit explicitly before starting. 3. Minimize TEST ANXIETY by maintaining a calm, encouraging atmosphere. 4. Prevent CHEATING through proper seating arrangement, active proctoring, and using multiple test versions if possible. 5. Manage TIMING so all learners have a fair chance to complete the test. CORRECTION FOR GUESSING: Objective tests are sometimes scored with a penalty for guessing using: SCORE = R - W/(k - 1) Where: R = number of right answers W = number of wrong answers (omitted items are NOT penalized) k = number of options per item WORKED EXAMPLE: A 4-option, 50-item test. Student answers 40 correctly, misses 9, omits 1. Score = 40 - 9/(4-1) = 40 - 9/3 = 40 - 3 = 37 RATIONALE: On a 4-option MCQ, pure guessing succeeds 1 out of 4 times. So for every 4 guesses, a student gets 1 right and 3 wrong. Subtracting W/(k-1) = 3/3 = 1 point removes that lucky guess. This formula discourages random guessing but most Philippine classroom tests skip it.
Examples
For every 3 wrong answers on a 4-option test, the formula removes 1 point (3/3 = 1), which is the estimated number of lucky guesses that produced those wrong answers. The student's true score estimate is 31, not 35.
Scenario
A teacher has a 50-item, 4-option multiple-choice test. A student answered 35 correctly, got 12 wrong, and left 3 items blank. Apply the correction for guessing.
Solution
Score = R - W/(k-1) = 35 - 12/(4-1) = 35 - 12/3 = 35 - 4 = 31 The corrected score is 31. The 3 omitted items are not counted as wrong.
Starting with difficult essay items exhausts students mentally before they reach the easier sections, leading to poor performance that does not accurately reflect true ability.
Scenario
A Grade 5 teacher arranged a 60-item test in this order: Items 1–20 are essay questions, Items 21–40 are MCQs, and Items 41–60 are true-false. Is this arrangement correct?
Solution
The ITEM GROUPING is correct (same types are together). However, the arrangement violates the EASY-TO-DIFFICULT principle. Essay questions (items requiring the most thinking and writing) are placed first, which increases test anxiety and uses up cognitive energy before easier items. The recommended order: true-false first (easiest), then MCQs (moderate), then essays (most demanding) — and within each section, from easier to harder items.
Applications
- Applied when preparing quarterly examinations under DepEd's assessment calendar
- Used when organizing Alternative Delivery Mode (ADM) and modular test materials
- Relevant when preparing Table of Specifications (TOS) and corresponding answer keys
- Correction for guessing may be applied in standardized testing contexts
Misconceptions
- MISCONCEPTION: 'Omitted items should be counted as wrong in the correction for guessing.' TRUTH: Only items with WRONG answers (incorrect response chosen) are counted as W. Omitted items are not penalized.
- MISCONCEPTION: 'Mixing item types throughout the test makes it more interesting.' TRUTH: Mixed formats slow down test-takers and reduce efficiency. Group same types together.
- MISCONCEPTION: 'Hard items should come first to challenge the best students.' TRUTH: Easy items first reduces anxiety for all students and allows everyone to demonstrate basic competence before encountering difficult items.
Related Concepts
- Table of Specifications (TOS)
- Test reliability and validity
- Guessing and measurement error
- RA 7836 — Professional competence in assessment
Common Exam Questions
Example
LET-style: 'In a 5-option, 40-item test, a student answered 30 correctly, 8 incorrectly, and omitted 2. What is the corrected score?' Solution: 30 - 8/(5-1) = 30 - 8/4 = 30 - 2 = 28
Approach
Apply the formula Score = R - W/(k-1). Identify R, W, and k from the problem. Remember omitted items are not W.
Question Type
Computation
Example
LET-style: 'In assembling a test, items should be arranged from ___.' — Answer: Easy to difficult
Approach
Choose the option that correctly describes proper test assembly or administration procedure.
Question Type
Best Practice
Key Points To Remember
- Group items by TYPE in the test (all MCQs together, all T-F together, etc.)
- Arrange items from EASY TO DIFFICULT to build student confidence
- Never break an MCQ or matching item across two pages
- Prepare answer key BEFORE printing and distributing the test
- Correction for guessing formula: Score = R - W/(k-1)
- Omitted items are NOT counted as wrong in the guessing correction formula
- k = number of options per item (for 4-option MCQ, k = 4)
- Test environment must be comfortable and low-anxiety, consistent with child protection principles under RA 7610
Item Analysis: Difficulty Index
After scoring a test, a professional teacher conducts ITEM ANALYSIS to determine which items worked well and which need revision. The first index is the DIFFICULTY INDEX (p), which measures how EASY or HARD an item was for the group tested. FORMULA: p = Number of examinees who answered correctly / Total number of examinees The difficulty index ranges from 0.00 to 1.00. IMPORTANT INTERPRETATION RULE: A HIGHER p means an EASIER item (more people got it right). A LOWER p means a HARDER item (fewer people got it right). The word 'difficulty' in the name is counter-intuitive — p is actually an EASE index! DIFFICULTY BANDS: 0.00 – 0.20 = Very Difficult (only a few got it right) 0.21 – 0.40 = Difficult 0.41 – 0.60 = Moderately Difficult / Average (IDEAL RANGE) 0.61 – 0.80 = Easy 0.81 – 1.00 = Very Easy (almost everyone got it right) IDEAL DIFFICULTY: An item with p near 0.50 best separates strong from weak learners. Items with p above 0.85 (too easy) or p below 0.15 (too hard) contribute little to discrimination. WHEN USING UPPER AND LOWER GROUPS: p = (Correct in upper group + Correct in lower group) / Total students in both groups WORKED EXAMPLE 1: In a class of 40 students, 30 answered an item correctly. p = 30/40 = 0.75 → The item is EASY. It is retainable but not highly discriminating because most students got it right. WORKED EXAMPLE 2: Upper group of 10: 8 answered correctly. Lower group of 10: 3 answered correctly. p = (8+3)/(10+10) = 11/20 = 0.55 → Moderately difficult. Ideal range.
Examples
When almost all students answer correctly, the item cannot separate students who truly mastered the competency from those who guessed. It adds little measurement value.
Scenario
In a Grade 6 class of 50 students, 45 answered Item 1 correctly. Compute p and interpret.
Solution
p = 45/50 = 0.90 Interpretation: VERY EASY (0.81–1.00 band) Decision: This item contributes very little to distinguishing high from low achievers. Consider revising to increase difficulty or replacing it.
Using only the upper and lower groups (not the full class) in the formula is standard practice in classroom item analysis, especially when working with the top and bottom 27% or top and bottom halves.
Scenario
From the upper group of 15 students, 6 answered Item 5 correctly. From the lower group of 15 students, 4 answered correctly. Compute p.
Solution
p = (6 + 4) / (15 + 15) = 10/30 = 0.33 Interpretation: DIFFICULT (0.21–0.40 band) Decision: The item is on the harder side but still within an acceptable range. Paired with the discrimination index, it may be worth retaining or slightly revising.
Applications
- Used after quarterly examinations to evaluate item quality and improve the item bank
- Applied when revising test items for the next school year
- Used to identify competencies where pupils performed consistently poorly (very low p) for remediation planning
- Referenced in DepEd School-Based Assessment Reports and Learning Action Cell (LAC) sessions on assessment improvement
Misconceptions
- MISCONCEPTION: 'A high difficulty index means the item is difficult.' TRUTH: A HIGH p means MORE students got it right — so the item is EASY. The name is counter-intuitive. Think of p as a proportion of correctness, not hardness.
- MISCONCEPTION: 'An item with p = 1.00 is a perfect item.' TRUTH: An item everyone gets right has zero discrimination power. It cannot separate high from low achievers.
- MISCONCEPTION: 'The lower the p, the better the item.' TRUTH: Items that are too easy (p > 0.85) AND items that are too difficult (p < 0.15) are both poor discriminators. The ideal is near 0.50.
Related Concepts
- Discrimination index (D)
- Distractor analysis
- Test reliability
- Normal distribution of scores
Common Exam Questions
Example
LET-style: 'In a class of 60, 24 students answered an item correctly. What is the difficulty index and how is it classified?' Solution: p = 24/60 = 0.40. Classification: Difficult (0.21–0.40 band).
Approach
Apply p = correct/total. For group data, use p = (upper correct + lower correct) / total in both groups. Classify the result using the difficulty bands.
Question Type
Computation
Example
LET-style: 'An item has p = 0.92. What does this suggest?' Answer: The item is very easy. Almost all students got it right. It contributes little to discrimination and should be revised or replaced.
Approach
Given a p value, state whether the item is too easy, too hard, or ideal, and what action should be taken.
Question Type
Interpretation
Key Points To Remember
- p = correct / total; ranges from 0.00 to 1.00
- HIGHER p = EASIER item (more people correct)
- LOWER p = HARDER item (fewer people correct)
- IDEAL p is around 0.50 — maximizes discrimination between strong and weak learners
- p below 0.15 = too hard; p above 0.85 = too easy — both are poor discriminators
- When using upper and lower groups: p = (upper correct + lower correct) / total in both groups
- Difficulty band: 0.00-0.20 very difficult; 0.21-0.40 difficult; 0.41-0.60 average; 0.61-0.80 easy; 0.81-1.00 very easy
Item Analysis: Discrimination Index and Distractor Analysis
The DISCRIMINATION INDEX (D) measures how well an item SEPARATES HIGH SCORERS from LOW SCORERS. An item that good students get right but weak students get wrong is doing its job — it discriminates positively. GROUP DEFINITION: The upper group (UG) = top 27% of scorers The lower group (LG) = bottom 27% of scorers The 27% criterion comes from Kelley's research showing this split maximizes contrast while keeping groups large enough to be statistically stable. In a class of 40, each group = 0.27 × 40 ≈ 11 students. In classroom practice, the upper and lower halves or a clean 10 are also commonly used. FORMULA: D = (Correct in Upper Group - Correct in Lower Group) / Number of students in ONE group D ranges from -1.00 to +1.00. POSITIVE D = Good. The upper group outperformed the lower group on this item. ZERO D = The item does not distinguish at all. Upper and lower groups performed equally. NEGATIVE D = DANGER. The lower group outperformed the upper group — the item may be MISKEYED or seriously flawed. DISCRIMINATION BANDS: 0.40 and above = Very Good / Excellent → Retain 0.30 – 0.39 = Good → Retain; minor improvement possible 0.20 – 0.29 = Fair / Marginal → Needs improvement 0.19 and below = Poor → Revise or reject Negative = Defective → Discard; CHECK FOR MISKEYING first DISTRACTOR ANALYSIS: For each wrong option, examine how many upper-group and lower-group students chose it. 1. GOOD DISTRACTOR: Attracts MORE lower-group than upper-group students — it traps those who did not master the content. 2. NON-FUNCTIONAL DISTRACTOR: Chosen by NO ONE — it must be revised or replaced because no student finds it plausible. 3. PROBLEMATIC DISTRACTOR: Chosen by MORE upper-group than lower-group students — this is a RED FLAG suggesting ambiguity in the item or a miskeyed answer. 4. THE KEY: Should always be chosen by MORE upper-group than lower-group students (mirrors positive D). WORKED COMPLETE EXAMPLE: Option | Upper (n=10) | Lower (n=10) A | 1 | 3 B | 0 | 2 *C | 8 | 3 ← KEY D | 1 | 2 p = (8+3)/20 = 0.55 (moderately difficult — ideal) D = (8-3)/10 = 0.50 (very good — retain) Distractor A: 3 lower vs. 1 upper — functioning Distractor B: 2 lower vs. 0 upper — functioning but weak, could be improved Distractor D: 2 lower vs. 1 upper — functioning VERDICT: RETAIN the item.
Examples
When 8 out of 10 high scorers get an item right, but only 3 out of 10 low scorers do, the item is clearly measuring something that distinguishes mastery from non-mastery.
Scenario
From the upper group of 10 students, 8 answered correctly. From the lower group of 10, 3 answered correctly. Compute D and interpret.
Solution
D = (8 - 3) / 10 = 5/10 = 0.50 Interpretation: VERY GOOD / EXCELLENT (D ≥ 0.40) Decision: RETAIN the item. It strongly separates high achievers from low achievers.
A negative D is the strongest signal that something is wrong. More low scorers than high scorers chose the keyed answer. This almost always means the key is wrong (miskeying) or the item is ambiguous.
Scenario
From the upper group of 12, only 3 answered correctly. From the lower group of 12, 9 answered correctly. Compute D and determine what action to take.
Solution
D = (3 - 9) / 12 = -6/12 = -0.50 Interpretation: NEGATIVE D — DEFECTIVE First action: CHECK THE ANSWER KEY. The item may be miskeyed (the teacher may have recorded option C as the key when option B is actually correct). If the key is confirmed correct, discard or completely rewrite the item.
Distractor analysis is essential because it tells you not just HOW MANY students missed the item, but WHY — which wrong option tripped them up, and whether the item itself may be at fault.
Scenario
Distractor analysis for Item 7: Option A (wrong) was chosen by 7 upper-group students and 2 lower-group students. The key was Option B chosen by 3 upper and 9 lower. What is wrong?
Solution
This item shows TWO red flags: (1) Option A was chosen by more upper-group (7) than lower-group (2) students — this distractor is attracting the best students, suggesting it may actually be correct or the item is ambiguous. (2) The key (Option B) was chosen by more lower-group than upper-group students, producing a NEGATIVE D. First action: recheck whether Option A might actually be the correct answer (miskeying). If confirmed, the key should be changed to A.
Applications
- Used after quarterly examinations to maintain and improve a school's item bank
- Applied in DepEd's Learning Action Cell (LAC) sessions focused on assessment quality
- Used to provide diagnostic data on which competencies need reteaching (items with very low p in a domain indicate class-wide non-mastery)
- Referenced when preparing achievement test reports for school administrators and parents
Misconceptions
- MISCONCEPTION: 'A negative D means students did not study.' TRUTH: A negative D means something is WRONG WITH THE ITEM ITSELF — most likely miskeying or ambiguity. It has nothing to do with study habits.
- MISCONCEPTION: 'A distractor that no one chose makes the test easier for the right reasons.' TRUTH: A non-functional distractor (chosen by no one) wastes an option slot. It effectively turns a 4-option MCQ into a 3-option MCQ, increasing guessing probability.
- MISCONCEPTION: 'We should use the entire class to compute D.' TRUTH: D is computed using ONLY the upper group and lower group (top and bottom 27% or halves), NOT the entire class. The middle students are excluded from D computation.
- MISCONCEPTION: 'D and p measure the same thing.' TRUTH: p measures how easy/hard the item is for the whole group. D measures how well the item separates high from low scorers. Both are needed for full item evaluation.
Related Concepts
- Difficulty index (p)
- Standard deviation and score distribution
- Test reliability coefficient
- Formative vs. summative assessment interpretation
Common Exam Questions
Example
LET-style: 'In an item analysis using 10 students each in the upper and lower groups, 6 upper-group and 2 lower-group students answered correctly. What is D and how is it classified?' Solution: D = (6-2)/10 = 4/10 = 0.40. Classification: Very Good / Excellent. Action: Retain.
Approach
Apply D = (Upper correct - Lower correct) / group size. Watch for negative results and interpret accordingly.
Question Type
Computation
Example
LET-style: 'Distractor X was chosen by 0 students in both the upper and lower groups. What should the teacher do?' Answer: Replace or revise Distractor X — it is non-functional because no student found it plausible enough to choose.
Approach
Given a table of response patterns, identify which distractors are functioning, non-functioning, or problematic.
Question Type
Distractor Evaluation
Example
LET-style: 'An item yields D = -0.30. What is the FIRST action the teacher should take?' Answer: Check the answer key for miskeying. If the correct answer was encoded incorrectly, correcting the key may resolve the negative D.
Approach
Given a negative D, state the first corrective action and what it means about the item.
Question Type
Interpretation
Key Points To Remember
- D = (Upper correct - Lower correct) / Size of one group
- D ranges from -1.00 to +1.00
- Positive D = good discrimination; Negative D = defective item (check for miskeying)
- D ≥ 0.40 = Excellent; 0.30–0.39 = Good; 0.20–0.29 = Fair; below 0.20 = Poor
- Upper group = top 27% of scorers; Lower group = bottom 27%
- Good distractor: chosen by MORE lower-group than upper-group students
- Non-functional distractor: chosen by NO ONE — must be revised or replaced
- Negative D first action: CHECK THE ANSWER KEY for miskeying before discarding the item
Practice Problems
p = 0.70 means 70% of pupils answered correctly. The item falls in the Easy band. For a diagnostic test, this might be acceptable for fundamental competencies. For a summative exam testing mastery, slightly more challenge (targeting p near 0.50) would better separate high from low achievers.
Problem
PROBLEM 1 (Computation — Difficulty Index): In a class of 40 Grade 5 pupils, 28 answered Item 3 correctly on the Science quarterly exam. Compute the difficulty index and classify the item.
Solution
p = 28/40 = 0.70 Classification: EASY (0.61–0.80 band) Decision: The item is somewhat easy. It is retainable, especially if it tests a basic competency, but it has limited discrimination power. Consider whether it needs to be made slightly more challenging.
With 7 out of 10 top pupils answering correctly versus only 2 out of 10 bottom pupils, the item is doing exactly what a good item should — it rewards mastery and penalizes non-mastery.
Problem
PROBLEM 2 (Computation — Discrimination Index): Using the top 10 and bottom 10 pupils from a class of 40, the following responses were recorded for Item 6: Upper group: 7 correct; Lower group: 2 correct. Compute D, classify it, and state your action.
Solution
D = (7 - 2) / 10 = 5/10 = 0.50 Classification: VERY GOOD / EXCELLENT (D ≥ 0.40) Action: RETAIN the item. It strongly separates high-performing from low-performing pupils.
An item can have ideal difficulty but poor discrimination. This often happens when distractors are not plausible enough — lower-group students eliminate them easily and guess between the remaining options, accidentally choosing the correct answer.
Problem
PROBLEM 3 (Computation — Both Indices): In an item analysis using the upper 15 and lower 15 students of a class: 9 upper and 6 lower answered correctly. Compute BOTH p and D, classify each, and state your recommendation.
Solution
Difficulty Index: p = (9 + 6) / (15 + 15) = 15/30 = 0.50 Classification: Moderately Difficult / IDEAL Discrimination Index: D = (9 - 6) / 15 = 3/15 = 0.20 Classification: FAIR / MARGINAL (0.20–0.29 band) Recommendation: The difficulty is ideal (p = 0.50), but the discrimination is only marginal (D = 0.20). REVISE the item — examine the distractors. Non-functioning or ambiguous distractors may be reducing discrimination. Strengthen the distractors to better trap lower-group students.
The 2 omitted items are not penalized — the formula only penalizes guessing, not skipping. The student's corrected score is approximately 35 out of 50, compared to the raw score of 38.
Problem
PROBLEM 4 (Correction for Guessing): A student took a 50-item, 4-option multiple-choice test. The student answered 38 correctly, answered 10 incorrectly, and left 2 items blank. Apply the correction for guessing formula and find the corrected score.
Solution
Formula: Score = R - W/(k - 1) R = 38 (correct answers) W = 10 (wrong answers; blank items are NOT counted) k = 4 (options per item) Score = 38 - 10/(4-1) Score = 38 - 10/3 Score = 38 - 3.33 Score ≈ 34.67 or approximately 35 (rounded) Note: The 2 blank items are NOT included in W.
In LET item-critique questions, you must name SPECIFIC violations, not just say 'the item is bad.' Practice matching each flaw to the rule it breaks: (1) non-homogeneous options, (2) length as a cue, (3) non-functional distractor, (4) grammatical clue.
Problem
PROBLEM 5 (Item Critique — MCQ): Identify ALL rule violations in this MCQ: 'Plants make their own food through a process called ___. A. photosynthesis B. It is the process by which green plants use sunlight, water, and carbon dioxide to produce oxygen and energy-rich glucose C. respiration D. it'
Solution
This item violates FOUR rules: 1. BLANK IN STEM: The stem uses a completion format (blank) instead of a clear question format — the article 'a' before the blank may also serve as a grammatical clue. 2. NON-HOMOGENEOUS OPTIONS: Option B is a full paragraph-length definition while A, C, and D are single words. 3. LENGTH GIVE-AWAY: Option B (the intended key) is dramatically longer than the other options — the correct answer stands out by length. 4. NON-PLAUSIBLE DISTRACTOR: Option D ('it') is not a plausible answer — no student would seriously consider it, making it a non-functional distractor.
Distractor analysis adds diagnostic depth beyond the two main indices. An upper-group students choosing a wrong distractor more than lower-group students is the clearest sign of item ambiguity or miskeying.
Problem
PROBLEM 6 (Distractor Analysis): The following response pattern was recorded for an MCQ with key C, using upper 10 and lower 10 students: Option A: Upper=0, Lower=1 Option B: Upper=3, Lower=1 Option C (key): Upper=7, Lower=4 Option D: Upper=0, Lower=4 Evaluate each distractor and state the overall verdict.
Solution
DIFFICULTY: p = (7+4)/20 = 11/20 = 0.55 (Moderately Difficult — ideal) DISCRIMINATION: D = (7-4)/10 = 3/10 = 0.30 (Good — retain with minor improvement possible) DISTRACTOR ANALYSIS: - Option A: Upper=0, Lower=1 — Functioning but very weakly. Only 1 student chose it. Should be revised to attract more lower-group students. - Option B: Upper=3, Lower=1 — PROBLEMATIC. More upper-group (3) than lower-group (1) chose this wrong answer. This is a red flag — Option B may be ambiguous or partially correct. Recheck. - Option D: Upper=0, Lower=4 — Functioning well. Four lower-group students chose it while no upper-group student did. OVERALL VERDICT: The item has acceptable difficulty and good discrimination, but Option B needs revision — it is confusing the better students. Option A is weak and should also be strengthened.
Exam Preparation Tips
- MASTER THE TWO FORMULAS: p = correct/total (or upper+lower correct / both group totals) and D = (upper correct - lower correct) / one group size. Practice at least 10 computation problems each. The LET regularly gives numerical data and asks you to compute and classify.
- MEMORIZE THE DIFFICULTY BANDS: 0.00-0.20=Very Difficult, 0.21-0.40=Difficult, 0.41-0.60=Average/Ideal, 0.61-0.80=Easy, 0.81-1.00=Very Easy. The ideal is near 0.50.
- MEMORIZE THE DISCRIMINATION BANDS: 0.40+=Excellent, 0.30-0.39=Good, 0.20-0.29=Fair, below 0.20=Poor, Negative=Defective. Negative D means: check for miskeying FIRST.
- COUNTER-INTUITIVE ALERT: Higher p = EASIER item (more people correct). Do not confuse 'high difficulty index' with 'a hard item.' Train yourself to say: 'p = 0.90 means very easy' not 'very difficult.'
- ITEM CRITIQUE PRACTICE: For MCQ flaws, always check: (1) Is the stem complete and clear? (2) Are options homogeneous in type, length, grammar? (3) Are there grammatical clues (a/an)? (4) Are there absolutes (always/never) in distractors? (5) Is there a non-functional distractor? Practice naming the SPECIFIC rule violated, not just saying 'the item is bad.'
- DISTRACTOR ANALYSIS PATTERN: GOOD distractor = more lower than upper students chose it. NON-FUNCTIONAL = no one chose it (revise it). PROBLEMATIC = more upper than lower chose a wrong answer (recheck the key or revise the item).
- CORRECTION FOR GUESSING: Score = R - W/(k-1). The key is that OMITTED items are NOT counted as W. Only items with a WRONG answer chosen are penalized. Practice with 4-option (k=4) and 5-option (k=5) formats.
- ESSAY SCORING: Remember the three reliability strategies — score item-by-item (not student-by-student), prepare rubric BEFORE scoring, and score anonymously to prevent halo effect. Know the difference: restricted-response limits content AND form; extended-response gives freedom.
- MATCHING RULE: ALWAYS more responses than premises. This is the most frequently tested matching rule on the LET.
- TRUE-FALSE SPECIFIC DETERMINERS: all/always/never/none = lean FALSE; sometimes/usually/generally = lean TRUE. This is a classic LET question about true-false item construction.
- CONNECT TO RA 7836: Professional competence in assessment is a requirement for licensure. The PRC and DepEd both expect licensed teachers to apply valid, reliable, and fair assessment — not just test 'tradition.'
- LET ITEM TYPE PATTERN: Expect about 3-5 computation items (p and D), 2-3 item-critique MCQs (identify the flaw), and 2-3 conceptual items (best practice identification). Practice all three formats.
- FOR DISTRACTOR ANALYSIS QUESTIONS: Read the table carefully — identify which options are upper group and which are lower group. Then check: does the key attract more upper or more lower? Does each distractor attract more lower or more upper?
- REVIEW CONTEXT: In the Philippine K-12 context, DepEd DO 8, s. 2015 guides classroom assessment. Item analysis is most directly applied in school-based assessments, LAC sessions on assessment quality, and teacher action research. Expect LET questions to frame computation problems within a Grade 1-6 classroom scenario.
In summary
Constructing, administering, and analyzing tests is a core professional skill that every licensed Filipino elementary teacher must master. Under RA 7836 and the Code of Ethics for Professional Teachers, competence in assessment is not optional — it is a legal and ethical obligation to every Grade 1–6 pupil you will serve. This chapter covers three phases that work together: writing valid items (with specific rules for MCQs, true-false, matching, completion, and essays), assembling and administering tests fairly (protecting student welfare consistent with the spirit of RA 7610), and conducting item analysis using the difficulty index (p) and discrimination index (D) to improve test quality. For the LET, master these priorities: (1) Compute p and D from given data and classify using the correct bands. (2) Identify specific MCQ construction rule violations when presented with a flawed item. (3) Interpret negative D immediately as a signal to check for miskeying. (4) Apply the guessing correction formula correctly, remembering omitted items are not penalized. (5) Know when to retain, revise, or reject an item based on both indices together. Beyond the LET, these skills will make you a more professional, fair, and child-centered teacher — one who builds assessments that truly reflect what pupils have learned and uses the data to improve both teaching and learning in alignment with DepEd's vision of quality, equitable, and culture-sensitive education.
Previous chapter
Principles of Assessment and the Table of Specifications
Next chapter
Authentic and Performance-Based Assessment
Ready to practise for the LET Secondary 2026?
Super Tutor's AI review plan adapts to your weak areas and builds a weekly practice schedule around your target LET Secondary exam date.