What Textual Evidence Multiple Choice Worksheets Actually Are and Why They Still Get Used
A Textual Evidence Multiple Choice Worksheet is exactly what it sounds like: a set of questions where students pick the answer that best matches what a passage says. You read the passage. You pick the choice. There is no room for creative interpretation on the answer side. That is the whole point. I spent years building these for middle school English language arts and high school reading intervention classes. The basic version is straightforward to produce, but the version that actually measures comprehension without being a guessing game is harder than most people expect. Most teachers reuse the same generic items they find online and wonder why their students score well on one worksheet and poorly on the next. The gap is not the students. It is the question design.
Textual Evidence Multiple Choice Worksheet
If you are looking for a ready-made Textual Evidence Multiple Choice Worksheet, there are several sources, but I will keep it simple. You can build your own in about twenty minutes if you know what to avoid. A free worksheet generator in Google Docs or a plain text template works fine. The real work is in writing items that actually require evidence retrieval rather than pattern matching. Start with the passage, not the questions. Pick a text that is at or slightly below the target grade level reading complexity. If you use something too dense, the question becomes a decoding test, not an evidence test. That is a mistake I made early on with a nineteenth-century excerpt. My kids could not get past the syntax to answer anything. Write three or four distractors that are plausible but do not have support in the text. Not obviously wrong. That is the difference between a decent item and a weak one. A common bad distractor looks like this: the student picks an answer that is true in the real world but absent from the passage. Those should never be correct answers on an evidence item, but they make good wrong choices when written carefully.
I had a specific problem once where nearly every student chose the wrong answer on a question about a character motivation. The distractor was a line that appeared in the passage but described a different character. The test writer had included three characters in the passage and tangled their actions together. I learned to run a simple mapping exercise: list every character or subject in the passage and check that the evidence for each correct answer links back to the right one. This fixed about half of my flawed items before I ever showed them to a class.
Get the Full Details

Item Types You Should Know
There are a few standard formats. The literal retrieval question asks for something stated directly. You would expect a student to find a sentence that matches the answer. The inference question asks the student to draw a conclusion from stated information. The vocabulary-in-context question asks for meaning based on surrounding text. The author purpose question asks why the passage was written in a particular way. Here is a counter-intuitive point most beginners miss: literal retrieval questions are often harder to write well than inference questions. It sounds backwards. The issue is that easy literal questions become trivial pattern matches, while well-written inference questions actually require reasoning. If your worksheet is all literal retrieval, your data will be noisy. Students guess correctly without reading anything. Anchoring an answer to a specific line or quote is essential for credibility. If a correct choice cannot be traced to a line in the passage, the item is invalid. Always include the line number or phrase reference when you build the answer key. It saves you time during review and stops arguments later.
Common Pitfalls and How to Fix Them
Distractors that repeat words from the passage without matching the meaning are the most frequent error. A student sees the word "danger" in the text and picks the answer that contains "danger," even if the answer describes safety. This happens because the item is testing word recognition instead of evidence understanding. The fix is to check each wrong choice for superficial language overlap. Remove it or rewrite it so the match depends on meaning, not vocabulary echo. Another trap is the two-true-answers problem. Sometimes a second choice is technically defensible because a reader can argue it from the text. If that happens, the item measures ambiguity rather than comprehension. Go back to the passage and adjust the wording so only one choice survives scrutiny, or replace the item entirely. I also recommend avoiding negative stems whenever possible. A question like "Which of the following is NOT supported by the passage?" forces students to evaluate every option for correctness instead of selecting the best evidence. It increases cognitive load without improving the quality of the measurement. You get worse discrimination scores and more fatigue.
Difficulty Calibration
If you want to estimate how hard your worksheet will be before you hand it out, try a quick pilot with five to ten students who resemble your target population. Record how many chose each option. If more than sixty percent select the same distractor, that option is doing too much work and you should revise the item. If fewer than twenty percent get the correct answer, the item is probably too difficult or poorly written. This rough check usually takes fifteen to twenty minutes and prevents you from using broken items in a graded assignment. I used to skip this step and then spend an hour afterward trying to figure out why my class average dropped. The pilot is faster than the cleanup.

Answer Key and Scoring
Build the answer key at the same time as the items. Do not write questions and then realize you cannot support an answer. Each correct response should include the exact passage reference. Include a one-sentence justification for each item. When you review the worksheet later, those justifications let you spot weak items in seconds. Scoring is simple: one point per correct answer, unless you are using a rubric for short constructed response follow-ups. If you plan to follow multiple choice with written explanations, make sure the worksheet instructions state that clearly. Otherwise students will skip the evidence requirement and your data becomes useless.
When This Approach Fails
A Textual Evidence Multiple Choice Worksheet will not measure close reading depth on its own. It measures identification and basic inference. If you need to assess how well a student can synthesize across paragraphs, trace argument structure, or evaluate bias, you need short response items or performance tasks in addition to multiple choice. Relying only on this format gives you a partial picture and can create false confidence in student ability. It also breaks down with very low literacy populations. If students cannot decode the passage at an acceptable rate, the item measures reading fluency instead of evidence use. In those cases, pair the worksheet with read-aloud support or switch to simpler texts until decoding is solid enough for the assessment to be valid.
Practical Template You Can Use
Create a blank sheet with these sections: That structure keeps the worksheet consistent and makes peer review faster. Another teacher can glance at the justification and spot a weak item in ten seconds. If you want a ready-to-use Textual Evidence Multiple Choice Worksheet, you can copy this template into a document, insert your passage, write four questions, fill in the key, and run the five-student pilot. The total time from blank page to usable worksheet is usually under thirty minutes for a single passage with four items.
