Imagine you're learning to drive. You wouldn't start on a highway — you'd start in an empty parking lot, making mistakes where the stakes are zero. This paper asks whether AI can be that parking lot for classroom discussion: a low-stakes rehearsal space where students practice articulating ideas before they have to do it live in front of peers. The committed claim: repeated use of a purpose-built voice-based AI discussion partner causes a durable 31% increase in voluntary class contributions, measured in real MBA sessions after the AI prep sessions end. This is not about test scores or content mastery — it's about whether AI rehearsal changes observable live behavior in the room. The distinction matters because most ed-tech research measures what students know, not what students do. The design is clean for an education study. 759 MBA students across ten sections of the same course were each randomly assigned two of ten class sessions to prepare for with the AI partner — a within-subject design that lets each student serve as their own control. The dependent variable is voluntary contributions per session, coded from class records. The study was preregistered, which is rare in ed-tech and immediately lifts it above the modal paper in this space. The effect persists into sessions where the student was NOT assigned to use the AI, suggesting something transferred — not just better prep for a single topic, but a behavioral shift. The mechanism the authors identify is comfort, not knowledge. Students who used the AI more reported greater comfort speaking up and greater perceived learning, but crucially not greater focus or motivation. This is a specific and falsifiable sub-claim: the AI isn't making students care more or pay more attention; it's lowering the activation energy for speaking. Think of it as desensitization — the same logic behind exposure therapy. You practice the scary thing in a safe context until it stops feeling scary. Where this sits on the ladder: prior work on AI tutoring tools (e.g., Khan Academy's Khanmigo, Carnegie Learning, various chatbot tutors) has focused almost entirely on test-score outcomes, with mixed results. Studies on class participation typically use non-AI interventions like structured discussion protocols. This paper occupies a genuinely under-explored intersection — AI as a behavioral intervention rather than a content-delivery mechanism. The 31% effect size is large for a classroom intervention, but the population (MBA students at a presumably elite program) limits generalizability. MBA students are already selected for willingness to talk. The architecture is straightforward: a voice-based AI discussion partner (not a chatbot — voice matters because the target behavior is verbal participation). The paper doesn't detail the underlying model, which is a gap. Whether this is GPT-4 with a voice wrapper or something custom-built matters for replication and cost scaling. The within-subject random assignment to two of ten sessions is the core methodological innovation — it separates the AI effect from topic difficulty, instructor variation, and student baseline talkativeness. The honest gap: the authors don't report what happens over a full semester of AI use, only after two sessions. The durability claim rests on sessions after those two uses, but we don't know the decay curve. They also don't compare voice AI to simpler interventions — would a structured written reflection exercise produce the same comfort gain at a fraction of the cost? The successor experiment is obvious: run this with a text-based AI partner to isolate whether voice specifically matters, and extend to a full semester to measure decay. My read on why they didn't: they're saving the modality comparison for the next paper, and the semester-long design requires a new IRB cycle and a new cohort.