Set up your notebook. You will need it at the end of class.
Copy these. Leave space under each one for a definition.
- respondent
- item
- topline
- self-report
- variability
Think about these. Nothing to write.
- How much do you think the person sitting next to you uses AI?
- Have you ever said you do something less than you actually do?
- If someone asked how much time you spend on your phone, would your answer be right?
Copy this into your Lab Notebook and leave room under it. You will answer it at the end of class.
What does a survey answer actually tell you?
A survey looks simple from the outside. Someone asks a group of people a set of questions, collects the answers, and reports what the group said. The reporting is where the difficulty starts, because a pile of answers is not yet information. Turning one into the other requires several small decisions, and every one of them can be made badly.
Start with the pieces. A respondent is one person who answered. An item is one question on the survey. If forty people answer a survey with twenty questions, there are forty respondents and twenty items, and the survey has collected eight hundred separate answers. Those two words get confused constantly, including by adults who should know better, because both of them can be counted and the counts sound similar.
The first thing a researcher builds from those answers is a topline. A topline lists every item, every answer a respondent could have picked, and how many respondents picked each one. It contains no conclusions and no opinions. It is the raw shape of what the group said, and it is deliberately boring. A researcher who skips the topline and jumps straight to conclusions has no way to show anyone else where those conclusions came from.
Every item on a topline carries a number written as n. The n is how many respondents the item is based on. This matters more than it looks like it should. If a survey has forty respondents but six of them skipped a particular question, then that item rests on thirty-four people, not forty. A careful topline shows the skipped answers rather than hiding them, because a question that many people refused to answer is telling you something. Sometimes the refusal is the finding.
Now the harder problem. Almost everything a survey collects is self-report, meaning the only source for the answer is the person giving it. If a survey asks how many hours you slept last night, nobody checked. If it asks how often you argue with your brother, nobody was watching. The answer goes into the topline looking exactly like a measurement, but it is not a measurement. It is a statement.
Self-report is not worthless. For many questions there is no better instrument available, and people are often roughly right about themselves. But self-report has a specific and predictable weakness: people are inconsistent in the direction of their errors. On questions about behavior that carries any social weight, respondents shade their answers toward whatever seems acceptable. They round their exercise up and their screen time down. They do it without deciding to, which is why simply asking them to be honest does not fix it.
There is a second kind of question that looks like self-report and is not. When a survey asks a respondent to estimate what other people do, the answer is a guess about a group. Guesses about groups fail differently than statements about yourself. People tend to assume that whatever they notice most is what happens most, so a respondent who has seen one dramatic example will estimate high. Neither kind of answer is more trustworthy than the other. They are simply wrong in different directions, and knowing which kind you are holding changes how much weight it can carry.
That brings up the property that separates an interesting item from a dull one. Consider two questions asked of the same group. The first asks whether respondents have ever ridden in a car. Nearly all of them say yes. The second asks how many hours they slept last night, and the answers run from four to eleven. The second item has variability, meaning the answers spread out instead of clustering. The first item has almost none.
An item with no variability can still be true and still be useful. It just cannot explain any difference between the people in the group, because on that item there is no difference to explain. If everyone answers the same way, the item cannot tell you why one respondent does something and another does not. Researchers care about variability for this reason. It is where the differences live, and differences are what most questions are actually about.
None of this requires advanced mathematics. It requires reading a topline carefully, noticing which items rest on how many people, keeping track of whether an answer is a statement about the respondent or a guess about somebody else, and paying attention to whether the answers spread out or pile up. Those four habits do most of the work. The mistakes that follow from skipping them are not subtle ones, and they are made constantly by people who never learned that a survey answer is a claim someone made rather than a fact someone verified.
B) 40. The passage says forty people and twenty questions gives you forty respondents and twenty items. A respondent is a person. An item is a question. The 800 is real, but it is the number of separate answers, not the number of respondents.
B. The passage says the refusal is sometimes the finding. Hiding the skips would hide that. This is why every item in our topline lists a Skipped line even when it is only one person.
C. The passage is direct about this. Nobody checked. Nobody was watching. The answer goes into the topline looking like a measurement, but it is a statement.
C. That spread is variability. The passage says an item with no variability cannot explain any difference between the people in the group, because on that item there is no difference to explain.
B) 34. The passage says that item rests on thirty-four people, not forty. The n travels with the item, not with the survey, which is why a careful topline prints it on every line instead of once at the top.
B. And the passage is pointed about what it leaves out: no conclusions and no opinions. It calls the topline deliberately boring, which is the property that makes it useful. A reader can check your work against it.
B. The passage is careful here. Nobody has to be lying for self-report to go wrong. If the shading is not a decision, then an instruction aimed at decisions cannot reach it.
C. Exercise gets rounded up and screen time gets rounded down. A predictable error is a different problem from a random one, because you can say which way an item is likely to be wrong before you even look at it.
B. That is why a guess about a group fails differently from a statement about yourself. Neither one is more trustworthy. They are wrong in different directions, and knowing which kind you are holding changes how much weight it can carry.
B. The passage puts it exactly this way: on that item there is no difference to explain. The item can be perfectly accurate and still be unable to tell you why one respondent does something and another does not.
- Yes6
- No64
- Not sure8
- Skipped1
- 04
- 10
- 22
- 31
- 41
- 56
- 61
- 76
- 89
- 98
- 1040
- Skipped1
Our question is how much. A10 asked do you pay.
It sits near our question without answering it. That is a different failure from being wrong, and it is the one that is easy to miss.
Which one shows more difference between the people in this room?
That spread is variability — the thing the passage described.
A10 piles up. A17 spreads out. An item that piles up cannot explain a difference between people, because on that item there is no difference to explain.
It is about AI. It does not tell us how much this class uses it.
That is what the third bucket is for. Without it, A10 gets stretched into a bucket it does not belong in, and nobody reading the result can see it happened.
Partner work first. Then quiet work time.
- No cellphones
- No sleeping
- No non-academic use of Chromebooks
- A quiet room and music
- No interruptions from the teacher
- If you are caught up on this course's work, you may work on anything legitimate for another class