Week 1 Discussion Questions
Week 1 results
Discussion honors
Team honors are based on the class ratings recorded during the discussion.
-
1
Gold
Team 7
Team members
- Melika
- Olivia
- Raina
- Sophia
- Rahaf
- Sarah
-
2
Silver
Team 6
Team members
- Conner
- Haoran
- Lu
- Runting
- Urva
- Wenhao
- Yanxi
-
3
Bronze
Team 4
Team members
- Cameron
- Krishay
- Owen
- Ryan
- Shushi
Top challenger
- Lyndsey Bajgrowicz
These questions were selected and combined from the Week 1 student submissions. Work through the reasoning with your group and be prepared to explain your conclusions.
Fourteen questions are listed below. Questions discussed in class are marked, and contributor NetIDs appear at the end of each question.
Question 01Statistical foundations and model reliability
From Error Distributions to Loss Functions
In the linear model , suppose the errors are independent with a common fixed scale. Why does a Gaussian error model make maximum-likelihood estimation of equivalent to minimizing squared error, while a Laplace error model leads to minimizing absolute error? Compare how the two estimators respond to a single large outlier.
This question is combined from: chahak2.
Question 02Statistical foundations and model reliability
Stable Predictions, Unstable Coefficients (Discussed)
Suppose two columns of are nearly collinear. Why can small changes in the data produce large changes in the individual least-squares coefficient estimates while the fitted values remain nearly unchanged? Use rank, geometry, or the normal equations in your explanation, and discuss what this means for interpreting individual coefficients versus predicting within the observed data region.
This question is combined from: tingyun3, gjia3, slxia2.
Question 03Statistical foundations and model reliability
Why Training R-Squared Rewards Extra Predictors
Two nested ordinary least-squares models are fit to the same response on the same observations, and both include an intercept. Why can adding predictors never increase the training residual sum of squares and therefore never decrease ? Why can this property make training misleading for comparing models of different sizes, and what alternative check would better assess whether the added predictors improve prediction?
This question is combined from: stewary2.
Question 04AI verification and trust
When Working Code Gives the Wrong Answer
An AI provides a confident statistical explanation and code that runs and produces plausible output. What kinds of errors could still remain? Design a short verification workflow that combines at least two independent checks, such as mathematical reasoning, documentation, a fixed-seed simulation, or a deliberately chosen counterexample, and explain what each check can reveal that simply rerunning the code cannot.
This question is combined from: shushim2, cudzich3, calebsg3, heta2.
Question 05AI verification and trust
When LLM Agreement Is Not Independent Verification
Suppose several large language models give the same answer to a statistical question. Under what conditions is that agreement weak evidence because the models may have correlated errors? Distinguish agreement among models from verification using independent evidence, and propose one external check that could overturn an incorrect consensus.
This question is combined from: wenhao7.
Question 06AI verification and trust
When Good Validation Performance Is Misleading (Discussed)
An AI agent recommends a model because it performs well under a chosen validation scheme. How could data leakage, temporal dependence, nonrepresentative data, distribution shift, or a mismatch between the metric and the application make that performance misleading? Choose one application and propose a validation or stress-testing design that better matches its intended use.
This question is combined from: ericc13, uvashi2.
Question 07Reproducibility and diagnostics
What Does a Random Seed Actually Control? (Discussed)
Homework 01 asks everyone to set a random seed before a simulation. What does setting the seed control? Compare what should happen when the same code is rerun in the same software environment with (i) the same seed and (ii) no fixed seed. Why does this matter when another person tries to reproduce the result?
This question is combined from: panico2, bpham9, kyang53, melikah2.
Question 08Reproducibility and diagnostics
Same Seed, Different Results
Two researchers both use seed 43201 but obtain different estimates, perhaps because one uses R and the other Python. Alternatively, the same R analysis produces a different result several months later. Why is matching the seed not sufficient? Develop an ordered diagnostic checklist that considers the random-number generator, software and package versions, data and preprocessing, and numerical methods. What should be saved or reported so another person can reproduce the analysis?
This question is combined from: connerj2, allyw2, ousher2.
Question 09AI agents and system design
AI Agent or Chatbot?
Imagine the same language model in two settings: one only replies to a userβs messages, while the other can use tools or packaged skills, keep track of a multi-step task, and take actions. What criteria would you use to call the second system an AI agent rather than a chatbot? Give one task for which a chatbot is preferable and one for which an agent is preferable, and explain why.
This question is combined from: zexuanj2, cs148.
Question 10AI agents and system design
How Does an Agent Harness Change the Outcome?
Suppose the same model receives the same prompt and skill but runs in two different agent harnesses. One harness offers many tools, repeated retries, and a large reasoning budget; the other exposes only the needed tools and uses a clear stopping rule. How can these choices change completeness, cost, and failure modes? For one simple task and one complex task, choose a harness configuration and justify your design.
This question is combined from: prerith2, amoghu2, lugong2.
Question 11Git and GitHub workflow
Question 12AI ethics, security, and data privacy
Safe Boundaries for Agents and Sensitive Data
Suppose a local AI agent is asked to clean and model confidential human-subjects data while also having access to files and development tools. What is the minimum data and tool access it should receive? Design safeguards involving raw versus deidentified data, file and network permissions, human approval, logging, and containment. What important risk would remain even after these safeguards are applied?
This question is combined from: ruitong7, rruiz35, selsa3.
Question 13AI impact and human skills
Human Skills That Matter More With AI (Discussed)
As AI handles more coding, routine analysis, and written explanation, which human skills become more, rather than less, important in statistical work? Choose two or three skills, such as problem formulation, domain judgment, communication of uncertainty, or responsibility for consequences, and explain why they are difficult to delegate. How should this change what students practice?
This question is combined from: jingyi64.
Question 14Learning strategies and prerequisites
Learning With AI Without Dependence (Discussed)
How can a student use AI to rebuild rusty prerequisites and coding skills while preserving an integrated understanding and independent problem-solving ability? Propose a repeatable learning cycle that specifies what the student should attempt independently, when AI should provide explanation or feedback, and how the student should test the resulting understanding on a new problem. What warning signs would indicate that AI is producing fragmented knowledge or dependence rather than genuine learning?
This question is combined from: runting5, rainas3, maeved2.