Do Now · Lab Notebook · 10 minutes

Set up your notebook. You will need it at the end of class.

Terminology and Concepts to look out for

Copy these. Leave space under each one for a definition.

  • bucket
  • univariate
  • claim
Relevance to Me

Think about these. Nothing to write.

  • When you tell someone how often you do something, do you round up or down?
  • Is it easier to guess about yourself, or about other people?
Question of the Day

Copy this into your Lab Notebook and leave room under it. You will answer it at the end of class.

Why can two survey items about the same thing disagree?

Sorting Evidence Before Using It
Day 2 · I Do passage

A person holding a stack of survey results faces a problem before they face any question about what the results mean. The stack is not organized. Some items in it speak directly to whatever the person wants to know. Some speak to it sideways. Some do not speak to it at all, even though they concern the same general subject. Before any of it can be used, it has to be sorted, and how it gets sorted determines what conclusions are available afterward.

The sorting categories are called buckets, and choosing them is the real work. A bucket is not a topic. It is a kind of evidence. Two items can both be about school attendance and still belong in different buckets, because one asks a student how often they come to school and the other asks a teacher how often that student comes to school. Same subject, different kind of evidence, different bucket.

This matters because the buckets are what make disagreement visible. If every item goes into one undifferentiated pile labeled "attendance," and the pile contains both student answers and teacher answers, then a researcher who summarizes the pile will produce an average that hides the fact that the two groups said different things. Sorted into separate buckets, the disagreement becomes the finding rather than being smoothed away.

Every good bucket system needs one bucket that most people forget to build. Some items in the stack will not answer the question at all. They are related, they are interesting, and they are irrelevant to the specific thing being asked. An item recording which brand of shoes students wear is genuinely about students and genuinely useless for a question about attendance. That item needs somewhere to go, and the somewhere cannot be a bucket that implies it counts as evidence. Researchers who do not build a reject bucket end up quietly stretching a good bucket to hold something that does not belong in it, and the stretch is invisible in the final report.

Deciding an item does not answer the question is a judgment, and it is a judgment that has to be defended. The defense usually comes down to a difference between what the item asked and what the question asks. A question about how often something happens is not answered by an item about whether it happens at all. A question about how much is not answered by an item about which. Naming that mismatch out loud is the difference between rejecting an item and dismissing it.

So far, everything described here counts one item at a time. A researcher looks at an item, reads the counts under each answer, and decides what that item shows. This is called univariate work, from a root meaning one variable. Almost all reporting on surveys is univariate. Sixty-two percent said this. Forty said that. Each number describes the answers to a single question.

Univariate work is not a lesser form of analysis. It is the foundation, and most published survey findings never go beyond it. But it has a specific limit that is worth naming early. A univariate count tells you how many people gave each answer. It cannot tell you anything about which people. Two items counted separately show two distributions and nothing about the relationship between them, because counting one item at a time discards the information about who said what.

Once items are sorted and counted, the last step is to say something. That statement is a claim, and a claim is not the same thing as a guess or an opinion. A claim is a statement that specific evidence supports and that could be shown wrong by other evidence. Both halves matter. A statement nothing supports is a guess. A statement nothing could contradict is not a claim either, because a statement that survives every possible piece of evidence is not saying anything about the world.

A well-built claim carries its evidence with it. Instead of asserting that attendance is a problem, it names which items support that and what those items actually asked. Doing so makes the claim checkable, which is the entire point. A reader who disagrees can go look at the same items and argue about them. A reader facing an unsupported assertion can only agree or refuse.

The order is fixed and it does not reverse. Sort first, count second, claim third. A researcher who forms the claim first will sort the evidence toward it without noticing, and will build the buckets that produce the answer they already have. This is not usually dishonesty. It is what happens when the conclusion arrives before the categories do.

Work these together
10 questions on the passage. Pick an answer, then check it. Nothing here is collected.
Question 1 of 10the passage
The passage says two items can both be about school attendance and still belong in different buckets. What makes them different?
Question 2 of 10the passage
According to the passage, what goes wrong when a researcher does not build a reject bucket?
Question 3 of 10the passage
The passage says univariate counting has a specific limit. What is it?
Question 4 of 10the passage
The passage says a claim needs two things. It has to be supported by evidence, and it has to be something other evidence could show wrong. Why does the second part matter?
Question 5 of 10the passage
The passage describes a researcher who puts student answers and teacher answers into one pile labeled "attendance." What does the passage say goes wrong?
Question 6 of 10the passage
The passage says that deciding an item does not answer the question is a judgment that has to be defended. What does that defense usually come down to?
Question 7 of 10the passage
Which of these mismatches does the passage give as an example?
Question 8 of 10the passage
According to the passage, what does a well-built claim carry with it?
Question 9 of 10the passage
The passage says the order is sort, then count, then claim. What does it say happens when a researcher forms the claim first?
Question 10 of 10the passage
How does the passage characterize a researcher who forms the claim before sorting?
From Our Class Survey
Your answers, collected August 11. 79 respondents. Selected items from the full topline.
A06 · Schoolwork FrequencyDocument A
How often do you use AI for schoolwork?
  • Never10
  • Once or twice44
  • About weekly12
  • Almost daily12
n = 79
A17 · Peer EstimateDocument B
Out of 10 students in this class, how many use AI without permission?
  • 04
  • 10
  • 22
  • 31
  • 41
  • 56
  • 61
  • 76
  • 89
  • 98
  • 1040
n = 79
A20 · How CommonDocument C
How common is AI use in this class?
  • Not at all common4
  • Not too common6
  • Somewhat common20
  • Very common48
n = 79
A21 · Compared to OthersDocument C
Compared to other students in this class, how much do you use AI?
  • Much less24
  • Somewhat less40
  • About the same12
  • Somewhat more or much more2
n = 79
A01 · Chatbots and AssistantsDocument D
Which of these chatbots or assistants have you used?
  • ChatGPT70
  • Gemini36
  • Grok7
  • Copilot6
  • None of these4
  • Claude2
  • Perplexity1
  • DeepSeek1
79 respondents answered. You could pick more than one, so these do not add to 79.
All 29 survey items →
The work
Answer the checkable ones together. The discussion questions have no answer today.
Document A
Document A is A06. Which bucket?
Document B
Document B is A17. 40 respondents answered 10. Which bucket?
Document D
Item A01 asks which chatbots you have used. Which bucket?
Class discussionno answer today
A06: 54 of 79 respondents say they use AI never or once or twice.

A17: 40 of 79 say all ten of their classmates use it without permission.

What do you make of that?

We are not answering this today. Write down what you think and bring it back tomorrow.

Naming what we did

Every item so far, we counted one at a time. A06 alone. A17 alone. A20 alone.

That is what the passage called univariate. It tells you how many people gave each answer, and nothing about which people.

Notebook: univariate
The rest of the block

Partner work first. Then quiet work time.

The three rules
  • No cellphones
  • No sleeping
  • No non-academic use of Chromebooks
What you get
  • A quiet room and music
  • No interruptions from the teacher
  • If you are caught up on this course's work, you may work on anything legitimate for another class
If you are picking, pick reading or writing.
Everything we build in this class, we build with words. The tools take writing as input. How well you write is the ceiling on what you can make them do.
Open this now: muggsofcompsci.net/decks/cs1-miniq-day2-wedo