{
 "ledger": {
  "title": "Claim Ledger: Use AI to Learn Without Getting Worse at It",
  "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/",
  "essay": {
   "title": "Use AI to Learn Without Getting Worse at It",
   "url": "https://durabilitycurve.com/blog/learn-with-ai-without-getting-worse/",
   "published": "2026-09-24"
  },
  "last_checked": "2026-09-27",
  "format": "current",
  "counts": {
   "claims": 70,
   "by_status": {
    "VERIFIED": 57,
    "EXECUTED": 7,
    "REPORTED": 6
   },
   "verified_at_primary": 57,
   "removed": 0
  },
  "license": "https://creativecommons.org/licenses/by/4.0/",
  "attribution": "The Durability Curve (Harry Floyd), https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/"
 },
 "rows": [
  {
   "id": "B1",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-B1",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "In a Turkish high-school trial, a plain ChatGPT-style assistant raised practice scores by 48% and a hint-only tutor by 127%",
   "quote": "\"GPT Base and GPT Tutor would increase performance on the assisted practice sessions by 48% and 127%, respectively\"",
   "removal_reason": "",
   "source": "Bastani et al., PNAS 2025",
   "source_url": "https://www.pnas.org/doi/10.1073/pnas.2422633122",
   "locator": "Main Results; Table 1 col (1)",
   "date_read": "2026-09-24"
  },
  {
   "id": "B2",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-B2",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "Once the AI was taken away, the plain-assistant group scored 17% lower on the exam than students who never had it",
   "quote": "\"GPT Base diminished the average control student's performance on the unassisted exam by 17%.\"",
   "removal_reason": "",
   "source": "Bastani et al., PNAS 2025",
   "source_url": "https://pmc.ncbi.nlm.nih.gov/articles/PMC12232635/",
   "locator": "Main Results; Table 1 col (2)",
   "date_read": "2026-09-24"
  },
  {
   "id": "B3",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-B3",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "The analysis the authors registered in advance finds a smaller exam harm than the headline (−0.035, CI −0.057 to −0.012), still negative",
   "quote": "\"GPT Base vs. Control ... −0.035 [−0.057, −0.012] 0.003\"",
   "removal_reason": "",
   "source": "Bastani et al., PNAS 2025, SI",
   "source_url": "https://pmc.ncbi.nlm.nih.gov/articles/PMC12232635/",
   "locator": "SI Appendix A.6, Table A.2",
   "date_read": "2026-09-24"
  },
  {
   "id": "B4",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-B4",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "The hint-only tutor removed the exam harm but produced no exam gain",
   "quote": "\"this negative effect is essentially eradicated in the GPT Tutor arm, though we still do not observe a positive effect.\"",
   "removal_reason": "",
   "source": "Bastani et al., PNAS 2025",
   "source_url": "https://www.pnas.org/doi/10.1073/pnas.2422633122",
   "locator": "Main Results; Introduction",
   "date_read": "2026-09-24"
  },
  {
   "id": "B5",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-B5",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "The tutor's prompt carried the correct solution and told it not to give the whole solution away",
   "quote": "\"GPT-4 is given a prompt including the solution to each problem (to mitigate hallucinations) as well as instructions to avoid giving away the entire solution\"",
   "removal_reason": "",
   "source": "Bastani et al., PNAS 2025",
   "source_url": "https://www.pnas.org/doi/10.1073/pnas.2422633122",
   "locator": "Footnote †; Experimental Design",
   "date_read": "2026-09-24"
  },
  {
   "id": "B6",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-B6",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "Without the answer key, the plain assistant got the practice problems right only 51% of the time",
   "quote": "\"GPT Base gives a correct answer only 51% of the time on average\"",
   "removal_reason": "",
   "source": "Bastani et al., PNAS 2025",
   "source_url": "https://www.pnas.org/doi/10.1073/pnas.2422633122",
   "locator": "\"GPT Errors vs. Student Performance\"; Fig. 2; SI C.1",
   "date_read": "2026-09-24"
  },
  {
   "id": "B7",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-B7",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "Students used the plain assistant as a crutch, asking for and copying solutions",
   "quote": "\"students often use GPT Base as a \"crutch\" by asking for and copying solutions\"",
   "removal_reason": "",
   "source": "Bastani et al., PNAS 2025",
   "source_url": "https://www.pnas.org/doi/10.1073/pnas.2422633122",
   "locator": "Introduction; Potential Mechanism",
   "date_read": "2026-09-24"
  },
  {
   "id": "B8",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-B8",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "The trial covered nearly 1,000 students",
   "quote": "\"We conducted four 90-min sessions for about fifty 9th, 10th, and 11th-grade classes, comprising nearly 1,000 students.\"",
   "removal_reason": "",
   "source": "Bastani et al., PNAS 2025",
   "source_url": "https://www.pnas.org/doi/10.1073/pnas.2422633122",
   "locator": "Experimental Design",
   "date_read": "2026-09-24"
  },
  {
   "id": "B9",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-B9",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "The trial measured short-term outcomes only",
   "quote": "\"we focus on short-term outcomes due to limitations imposed by our partner school\"",
   "removal_reason": "",
   "source": "Bastani et al., PNAS 2025",
   "source_url": "https://www.pnas.org/doi/10.1073/pnas.2422633122",
   "locator": "Discussion, final paragraph",
   "date_read": "2026-09-24"
  },
  {
   "id": "K1",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-K1",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "In a Harvard physics trial, students' median post-test was 4.5 with the AI tutor against 3.5 in an active-learning class",
   "quote": "\"higher median (M) post-score (M = 4.5, N = 142) compared to those in the in-class active learning group (M = 3.5, N = 174)\"",
   "removal_reason": "",
   "source": "Kestin et al., Scientific Reports 2025",
   "source_url": "https://www.nature.com/articles/s41598-025-97652-6",
   "locator": "Results § Learning gains",
   "date_read": "2026-09-24"
  },
  {
   "id": "K2",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-K2",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "Median learning gains with the AI tutor were more than double the class's",
   "quote": "\"in the AI-tutored group were over double those for students in the in-class active learning group\"",
   "removal_reason": "",
   "source": "Kestin et al., Scientific Reports 2025",
   "source_url": "https://pmc.ncbi.nlm.nih.gov/articles/PMC12179260/",
   "locator": "Results § Learning gains",
   "date_read": "2026-09-24"
  },
  {
   "id": "K3",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-K3",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "The regression effect size was 0.63",
   "quote": "\"While the linear regression suggests an effect size of 0.63, this is an underestimation due to ceiling effect\"",
   "removal_reason": "",
   "source": "Kestin et al., Scientific Reports 2025",
   "source_url": "https://pmc.ncbi.nlm.nih.gov/articles/PMC12179260/",
   "locator": "Results § Linear regression model; Table S1",
   "date_read": "2026-09-24"
  },
  {
   "id": "K4",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-K4",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "The Harvard tutor's instructions released one step at a time and kept replies brief",
   "quote": "\"Only give away ONE STEP AT A TIME, DO NOT give away the full solution in a single message\" / \"Keep responses BRIEF (a few sentences or less) but helpful.\"",
   "removal_reason": "",
   "source": "Kestin et al., Supplementary Material 1",
   "source_url": "https://www.nature.com/articles/s41598-025-97652-6",
   "locator": "Supplement, system prompt",
   "date_read": "2026-09-24"
  },
  {
   "id": "K5",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-K5",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "The Harvard team loaded each question with its step-by-step worked answer",
   "quote": "\"we enriched our prompts with comprehensive, step-by-step answers\"",
   "removal_reason": "",
   "source": "Kestin et al., Scientific Reports 2025",
   "source_url": "https://www.nature.com/articles/s41598-025-97652-6",
   "locator": "Discussion § Designing successful student-AI interactions",
   "date_read": "2026-09-24"
  },
  {
   "id": "K6",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-K6",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "194 students took part",
   "quote": "\"Of the 233 enrolled students, 194 were eligible for inclusion in the study.\"",
   "removal_reason": "",
   "source": "Kestin et al., Scientific Reports 2025",
   "source_url": "https://www.nature.com/articles/s41598-025-97652-6",
   "locator": "Methods § Study population",
   "date_read": "2026-09-24"
  },
  {
   "id": "K9",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-K9",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "Each student learned one lesson with the AI tutor and another in the active-learning class (crossover)",
   "quote": "\"Given that our experiment is a crossover design in which each student experiences both conditions\"",
   "removal_reason": "",
   "source": "Kestin et al., Scientific Reports 2025",
   "source_url": "https://pmc.ncbi.nlm.nih.gov/articles/PMC12179260/",
   "locator": "Results (regression controls); Methods § Study design",
   "date_read": "2026-09-24"
  },
  {
   "id": "A1",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-A1",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "In an Anthropic trial, 52 developers learned a new Python library, half with an AI assistant",
   "quote": "\"In our main study, 52 participants completed the task, 26 for each of the control and treatment groups.\"",
   "removal_reason": "",
   "source": "Shen & Tamkin, arXiv 2601.20245",
   "source_url": "https://arxiv.org/abs/2601.20245",
   "locator": "§5.2.1",
   "date_read": "2026-09-24"
  },
  {
   "id": "A2",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-A2",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "The AI group scored 4.15 points lower on a 27-point quiz, about two grade points",
   "quote": "\"There is a 4.15 point difference between the means ... For a 27-point quiz, this translates into a 17% score difference or 2 grade points.\"",
   "removal_reason": "",
   "source": "Shen & Tamkin, arXiv 2601.20245",
   "source_url": "https://arxiv.org/html/2601.20245v2",
   "locator": "§5.2.2",
   "date_read": "2026-09-24"
  },
  {
   "id": "A3",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-A3",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "Nobody could use AI during the quiz",
   "quote": "\"All participants were not allowed to use AI in the comprehension check.\"",
   "removal_reason": "",
   "source": "Shen & Tamkin, arXiv 2601.20245",
   "source_url": "https://arxiv.org/html/2601.20245v2",
   "locator": "Figure 4 caption",
   "date_read": "2026-09-24"
  },
  {
   "id": "A4",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-A4",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "Using AI did not make them significantly faster",
   "quote": "\"While using AI to complete our coding task did not significantly improve task completion time, the level of skill formation gained ... is significantly reduced.\"",
   "removal_reason": "",
   "source": "Shen & Tamkin, arXiv 2601.20245",
   "source_url": "https://arxiv.org/html/2601.20245v2",
   "locator": "§1 / §5.2.2",
   "date_read": "2026-09-24"
  },
  {
   "id": "A5",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-A5",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "High-scoring ways of using the AI averaged 65% to 86% on the quiz; low-scoring ways 24% to 39% (post-hoc, small groups)",
   "quote": "\"high-scoring interaction patterns (65%-86% quiz score) vs low-scoring interaction patterns (24%-39% quiz score)\"",
   "removal_reason": "",
   "source": "Shen & Tamkin, arXiv 2601.20245",
   "source_url": "https://arxiv.org/abs/2601.20245",
   "locator": "§6; Figure 11",
   "date_read": "2026-09-24"
  },
  {
   "id": "A6",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-A6",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "Asking the AI conceptual questions was a high-scoring pattern, and the fastest of them",
   "quote": "\"On average, this mode was the fastest among high-scoring patterns and second fastest overall after the AI Delegation mode.\"",
   "removal_reason": "",
   "source": "Shen & Tamkin, arXiv 2601.20245",
   "source_url": "https://arxiv.org/abs/2601.20245",
   "locator": "§6; Figure 11 (Conceptual Inquiry)",
   "date_read": "2026-09-24"
  },
  {
   "id": "N1",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-N1",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "In a World Bank trial in Nigeria, students had twelve 90-minute sessions guided by teachers",
   "quote": "\"Those assigned to the intervention attended twelve 90-minute sessions in computer labs, engaging in curriculum-aligned activities guided by teachers.\"",
   "removal_reason": "",
   "source": "De Simone et al., World Bank PRWP 11125",
   "source_url": "https://documents1.worldbank.org/curated/en/099548105192529324/pdf/IDU-c09f40d8-9ff8-42dc-b315-591157499be7.pdf",
   "locator": "§1, p. 2",
   "date_read": "2026-09-24"
  },
  {
   "id": "N2",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-N2",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "English scores rose 0.238 standard deviations, which the authors equate to 1.5 years of ordinary schooling",
   "quote": "\"Our ITT effect of 0.238 standard deviation in English is equivalent to increasing 1.5 years of 'business-as-usual' schooling in Nigeria\"",
   "removal_reason": "",
   "source": "De Simone et al., World Bank PRWP 11125",
   "source_url": "https://documents1.worldbank.org/curated/en/099548105192529324/pdf/IDU-c09f40d8-9ff8-42dc-b315-591157499be7.pdf",
   "locator": "§4.1, p. 18",
   "date_read": "2026-09-24"
  },
  {
   "id": "N3",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-N3",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "In the sample week's starting prompt, the chatbot answered the student's question, then set exercises on which it was told to hint after a wrong reply and give the answer only if the reply was still wrong",
   "quote": "\"provide encouraging words and provide a hint\" … \"if my reply is still incorrect\"",
   "removal_reason": "",
   "source": "De Simone et al., World Bank PRWP 11125",
   "source_url": "https://documents1.worldbank.org/curated/en/099548105192529324/pdf/IDU-c09f40d8-9ff8-42dc-b315-591157499be7.pdf",
   "locator": "Figure 11, PDF p. 51 (Week 7 starting prompt, read from rendered image)",
   "date_read": "2026-09-24"
  },
  {
   "id": "N4",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-N4",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "The control group got no programme at all, so extra time and attention are not separated from the AI",
   "quote": "\"the control group, which did not receive any intervention but continued their regular learning in the classroom\"",
   "removal_reason": "",
   "source": "De Simone et al., World Bank PRWP 11125",
   "source_url": "https://documents1.worldbank.org/curated/en/099548105192529324/pdf/IDU-c09f40d8-9ff8-42dc-b315-591157499be7.pdf",
   "locator": "§2.2, p. 9",
   "date_read": "2026-09-24"
  },
  {
   "id": "N5",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-N5",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "The authors credit the whole package (prompts plus teacher guidance), not the chatbot alone",
   "quote": "\"we interpret that the intervention as a whole -which includes the interaction with the LLM and teacher guidance with specific prompts- is driving the results.\"",
   "removal_reason": "",
   "source": "De Simone et al., World Bank PRWP 11125",
   "source_url": "https://documents1.worldbank.org/curated/en/099548105192529324/pdf/IDU-c09f40d8-9ff8-42dc-b315-591157499be7.pdf",
   "locator": "§4.3, p. 21",
   "date_read": "2026-09-24"
  },
  {
   "id": "N6",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-N6",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "Nigeria's third-term exam (effect 0.206 SD) came after the intervention on content not limited to it; it ran the day after the last session, so it is not a test of whether the learning lasted",
   "quote": "\"the third term exam score, with an effect size of 0.206 standard deviation (SE = 0.067), although this exam was not limited to the intervention's specific content\"",
   "removal_reason": "",
   "source": "De Simone et al., World Bank PRWP 11125",
   "source_url": "https://documents1.worldbank.org/curated/en/099548105192529324/pdf/IDU-c09f40d8-9ff8-42dc-b315-591157499be7.pdf",
   "locator": "§4.1; timing from Appendix timeline (pilot sessions end 7/11/24, third-term exam 7/12/24)",
   "date_read": "2026-09-24"
  },
  {
   "id": "O1",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-O1",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "OpenAI's own trial ran over 300 college students through study sessions of a nominal 40 minutes",
   "quote": "\"we ran a randomized study with over 300 college students preparing for neuroscience and microeconomics exams\" / \"the nominal 40 minute sessions\"",
   "removal_reason": "",
   "source": "OpenAI, \"New tools for understanding AI and learning outcomes\" (2026-03-04)",
   "source_url": "https://openai.com/index/understanding-ai-and-learning-outcomes/",
   "locator": "\"Origins and early research\"; \"Study design\"",
   "date_read": "2026-09-24"
  },
  {
   "id": "O2",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-O2",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "Study mode scored roughly 15% higher in microeconomics",
   "quote": "\"roughly a 15% higher score relative\"",
   "removal_reason": "",
   "source": "OpenAI (2026-03-04)",
   "source_url": "https://openai.com/index/understanding-ai-and-learning-outcomes/",
   "locator": "\"Findings\", bullet 2",
   "date_read": "2026-09-24"
  },
  {
   "id": "O3",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-O3",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "In neuroscience, study mode was not distinguishable from studying with ordinary online resources",
   "quote": "\"results were not distinguishable from students studying with traditional online resources\"",
   "removal_reason": "",
   "source": "OpenAI (2026-03-04)",
   "source_url": "https://openai.com/index/understanding-ai-and-learning-outcomes/",
   "locator": "\"Findings\", bullet 1",
   "date_read": "2026-09-24"
  },
  {
   "id": "O4",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-O4",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "The comparison group used ordinary online resources, not ordinary ChatGPT",
   "quote": "\"a control group studied using traditional online resources such as Google Search and YouTube, with AI generated overview features disabled\"",
   "removal_reason": "",
   "source": "OpenAI (2026-03-04)",
   "source_url": "https://openai.com/index/understanding-ai-and-learning-outcomes/",
   "locator": "\"Study design\", para 1",
   "date_read": "2026-09-24"
  },
  {
   "id": "O5",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-O5",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "OpenAI calls the results early; analysis is still underway",
   "quote": "\"While analysis is still underway, early results give us confidence\"",
   "removal_reason": "",
   "source": "OpenAI (2026-03-04)",
   "source_url": "https://openai.com/index/understanding-ai-and-learning-outcomes/",
   "locator": "\"Origins and early research\"",
   "date_read": "2026-09-24"
  },
  {
   "id": "G1",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-G1",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "Google's own trial of Gemini Guided Learning in Sierra Leone raised maths scores 0.258 standard deviations (CI 0.027 to 0.488)",
   "quote": "\"yielding a gain of +0.258 standard deviations across the Guided Learning classrooms (intent to treat; 95% confidence interval [0.027, 0.488]\"",
   "removal_reason": "",
   "source": "LearnLM Team, Google & Fab AI, tech report (May 2026)",
   "source_url": "https://storage.googleapis.com/deepmind-media/LearnLM/learnLM_sierraleone_may26.pdf",
   "locator": "p. 3; Table C.4",
   "date_read": "2026-09-24"
  },
  {
   "id": "G2",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-G2",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "It enrolled 1,763 students in 48 grade 7 and 8 classrooms",
   "quote": "\"The trial enrolled 𝑁= 1763 students aged 13 or older in 48 grades 7 and 8 classrooms.\"",
   "removal_reason": "",
   "source": "LearnLM Team tech report",
   "source_url": "https://storage.googleapis.com/deepmind-media/LearnLM/learnLM_sierraleone_may26.pdf",
   "locator": "p. 2",
   "date_read": "2026-09-24"
  },
  {
   "id": "G3",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-G3",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "Students shared devices in pairs, in lessons teachers ran",
   "quote": "\"Students accessed the Gemini app on tablets or desktop computers, sharing at a 2:1 student-to-device ratio.\"",
   "removal_reason": "",
   "source": "LearnLM Team tech report",
   "source_url": "https://storage.googleapis.com/deepmind-media/LearnLM/learnLM_sierraleone_may26.pdf",
   "locator": "p. 2",
   "date_read": "2026-09-24"
  },
  {
   "id": "G4",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-G4",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "The gain came from grade 8; the grade 7 estimate was slightly negative (−0.078; grade 8 interaction 0.429)",
   "quote": "\"Arm: Treatment –0.078* (0.037)\"; \"Arm: Treatment × Grade: Grade 8 0.429** (0.142)\"",
   "removal_reason": "",
   "source": "LearnLM Team tech report",
   "source_url": "https://storage.googleapis.com/deepmind-media/LearnLM/learnLM_sierraleone_may26.pdf",
   "locator": "Table C.14, p. 28",
   "date_read": "2026-09-24"
  },
  {
   "id": "G5",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-G5",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "Gemini gave a direct solution in 2.1% of its messages (Gemini-classified)",
   "quote": "\"far more frequently than providing direct solutions (2.1%)\"",
   "removal_reason": "",
   "source": "LearnLM Team tech report",
   "source_url": "https://storage.googleapis.com/deepmind-media/LearnLM/learnLM_sierraleone_may26.pdf",
   "locator": "p. 4",
   "date_read": "2026-09-24"
  },
  {
   "id": "G6",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-G6",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "The report has not been peer reviewed",
   "quote": "\"Please note works submitted as a preprint have not undergone a peer review process.\"",
   "removal_reason": "",
   "source": "LearnLM Team tech report",
   "source_url": "https://storage.googleapis.com/deepmind-media/LearnLM/learnLM_sierraleone_may26.pdf",
   "locator": "p. 8",
   "date_read": "2026-09-24"
  },
  {
   "id": "D1",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-D1",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "OpenAI's own help page says Study mode may still give a direct answer",
   "quote": "\"there may be times when it gives a direct answer\"",
   "removal_reason": "",
   "source": "OpenAI Help Center, \"Using study mode in ChatGPT\"",
   "source_url": "https://help.openai.com/en/articles/11780217-using-study-mode-in-chatgpt",
   "locator": "Article body",
   "date_read": "2026-09-24"
  },
  {
   "id": "D2",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-D2",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "Study mode is not available inside ChatGPT Projects",
   "quote": "\"The Study option is not available in Temporary Chats, GPTs, or Projects.\"",
   "removal_reason": "",
   "source": "OpenAI Help Center",
   "source_url": "https://help.openai.com/en/articles/11780217-using-study-mode-in-chatgpt",
   "locator": "Article body",
   "date_read": "2026-09-24"
  },
  {
   "id": "D3",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-D3",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "A ChatGPT scheduled task made inside a project cannot read that project's files",
   "quote": "\"If you create a task in a project, it cannot access uploaded files or files stored in that project.\"",
   "removal_reason": "",
   "source": "OpenAI Help Center, Tasks",
   "source_url": "https://help.openai.com/en/articles/10291617",
   "locator": "Article body",
   "date_read": "2026-09-24"
  },
  {
   "id": "D4",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-D4",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "Claude's free plan allows up to five Projects",
   "quote": "\"Free users can create a maximum of five projects.\"",
   "removal_reason": "",
   "source": "Claude Help Center",
   "source_url": "https://support.claude.com/en/articles/9519177",
   "locator": "Article body",
   "date_read": "2026-09-24"
  },
  {
   "id": "D5",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-D5",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "Gemini Guided Learning sits under Add files, then More tools",
   "quote": "\"In the text box, click Add Files.\" / \"At the bottom, click More tools\" … \"Guided Learning\"",
   "removal_reason": "",
   "source": "Google Gemini Help",
   "source_url": "https://support.google.com/gemini/answer/16448384",
   "locator": "\"Use Guided Learning\", steps 2–3",
   "date_read": "2026-09-27"
  },
  {
   "id": "X1",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-X1",
   "status": "EXECUTED",
   "status_label": "Executed",
   "removed": false,
   "claim": "Study appears as a composer mode in the ChatGPT web app (Plus)",
   "quote": "The author's screenshot (not published): composer shows \"Study\" mode chip",
   "removal_reason": "",
   "source": "First-party: ChatGPT web app (Plus), the author's run",
   "source_url": "",
   "locator": "",
   "date_read": "2026-09-24"
  },
  {
   "id": "X2",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-X2",
   "status": "EXECUTED",
   "status_label": "Executed",
   "removed": false,
   "claim": "In 8 runs of the final Level 2 prompt (7 across Claude, Gemini and GPT, 1 in the ChatGPT app), the tutor withheld the answer on a direct demand and, after the second wrong try, gave the answer (some models per step) and named the missed step; once (GPT, code) it described the fix in words while withholding the code",
   "quote": "See run record",
   "removal_reason": "",
   "source": "First-party: python3 run.py, run_gemini.py, run_codex.py + the author's ChatGPT run",
   "source_url": "",
   "locator": "",
   "date_read": "2026-09-24"
  },
  {
   "id": "X3",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-X3",
   "status": "EXECUTED",
   "status_label": "Executed",
   "removed": false,
   "claim": "In a simulated project (instructions plus file given to the model directly), the Level 3 setup stuck to the file, said when the file did not cover a question, and scored a no-feedback quiz correctly on all three model families; the Claude and Gemini apps themselves were not tested",
   "quote": "See run record",
   "removal_reason": "",
   "source": "First-party: python3 l3.py, l3x.py",
   "source_url": "",
   "locator": "",
   "date_read": "2026-09-24"
  },
  {
   "id": "X4",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-X4",
   "status": "EXECUTED",
   "status_label": "Executed",
   "removed": false,
   "claim": "In the ChatGPT app, our first Level 3 instructions answered a \"just tell me\" question at once from general knowledge; the revised instructions held it back and answered from the file",
   "quote": "See run record",
   "removal_reason": "",
   "source": "First-party: ChatGPT web (Plus), Project \"Cold Check Test\", the author's run (v2) + Claude via Claude in Chrome (v3)",
   "source_url": "",
   "locator": "",
   "date_read": "2026-09-24"
  },
  {
   "id": "X5",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-X5",
   "status": "EXECUTED",
   "status_label": "Executed",
   "removed": false,
   "claim": "In the ChatGPT app, the cold-check prompt held all feedback to the end and scored a 3-of-6 quiz correctly against the file",
   "quote": "See run record",
   "removal_reason": "",
   "source": "First-party: ChatGPT web (Plus), Project \"Cold Check Test\", driven by Claude via Claude in Chrome",
   "source_url": "",
   "locator": "",
   "date_read": "2026-09-24"
  },
  {
   "id": "D6",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-D6",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "In Claude.ai, the old Styles menu is gone; Anthropic now offers an official \"learn\" skill (Customize → Skills → Discover → learn → Add) whose stated goal is to help the learner answer it themselves, and which does not trigger on coding tasks or factual lookups",
   "quote": "\"The goal is not to answer the learner's question but to help them be able to answer it themselves\" / \"Don't trigger for: Tasks: coding, writing, calculation, translation, factual lookup\"",
   "removal_reason": "",
   "source": "Anthropic \"learn\" skill (updated Sep 14), SKILL.md + Overview, read in the live Claude.ai app",
   "source_url": "https://claude.ai/new#customize/skills/id/learn",
   "locator": "Skill Overview + Contents › SKILL.md \"Learning Mode\"",
   "date_read": "2026-09-24"
  },
  {
   "id": "A7",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-A7",
   "status": "REPORTED",
   "status_label": "Reported",
   "removed": false,
   "claim": "They were Python programmers recruited through a crowd-work platform; 29 of the 52 had seven or more years of coding experience",
   "quote": "Paper Table 1 (experience bands; 7+ years = 29 of 52) / \"the small group size of the 1-3 year participant group (n=4)\"",
   "removal_reason": "",
   "source": "Shen & Tamkin, arXiv 2601.20245",
   "source_url": "https://arxiv.org/abs/2601.20245",
   "locator": "Table 1; §5.2.2",
   "date_read": "2026-09-24"
  },
  {
   "id": "S2",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-S2",
   "status": "EXECUTED",
   "status_label": "Executed",
   "removed": false,
   "claim": "Only two of the seven studies went through journal peer review (PNAS and Scientific Reports); the others are two preprints, a working paper, a vendor tech report and a vendor blog post",
   "quote": "PNAS 122(26); Scientific Reports 15, 17458; arXiv 2601.20245 and arXiv 2409.09047 (preprints); World Bank Policy Research Working Paper 11125; Google tech report \"have not undergone a peer review process\"; OpenAI \"While analysis is still underway\"",
   "removal_reason": "",
   "source": "The seven studies' own publication records",
   "source_url": "",
   "locator": "",
   "date_read": "2026-09-24"
  },
  {
   "id": "D7",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-D7",
   "status": "REPORTED",
   "status_label": "Reported",
   "removed": false,
   "claim": "Microsoft Copilot has a \"Study and learn\" mode, chosen from the Quick response menu under the prompt",
   "quote": "\"Press (or tap) 'Quick response' under your prompt and select an appropriate mode\"",
   "removal_reason": "",
   "source": "Microsoft Support, \"Conversation modes in Microsoft Copilot\"",
   "source_url": "https://support.microsoft.com/en-us/microsoft-copilot/conversation-modes-in-microsoft-copilot",
   "locator": "Article body",
   "date_read": "2026-09-24"
  },
  {
   "id": "D8",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-D8",
   "status": "REPORTED",
   "status_label": "Reported",
   "removed": false,
   "claim": "Claude Code has a Learning output style, switched on with /output-style learning",
   "quote": "\"/output-style learning\" (Learning style: Claude \"stops and waits\" at TODO(human) markers)",
   "removal_reason": "",
   "source": "Claude Code docs, Output styles",
   "source_url": "https://code.claude.com/docs/en/output-styles",
   "locator": "Built-in styles section",
   "date_read": "2026-09-24"
  },
  {
   "id": "K7",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-K7",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "The Harvard tutor's instructions allowed it to give the answer if the student demanded it",
   "quote": "\"DO NOT not tell them the answer UNLESS they demand you to give them the answer\"",
   "removal_reason": "",
   "source": "Kestin et al., Supplementary Material 1",
   "source_url": "https://www.nature.com/articles/s41598-025-97652-6",
   "locator": "Supplement, system prompt",
   "date_read": "2026-09-24"
  },
  {
   "id": "K8",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-K8",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "The Harvard tutor's instructions asked students to try first",
   "quote": "\"encourage them to give it a try first\"",
   "removal_reason": "",
   "source": "Kestin et al., Supplementary Material 1",
   "source_url": "https://www.nature.com/articles/s41598-025-97652-6",
   "locator": "Supplement, system prompt",
   "date_read": "2026-09-24"
  },
  {
   "id": "A8",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-A8",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "The highest-scoring pattern (2 people) had the AI generate code and then asked it questions to understand it, averaging 86%; asking conceptual questions averaged 65%; the lowest patterns, which delegated or relied on it, averaged 24% to 39%",
   "quote": "\"high-scoring interaction patterns (65%-86% quiz score) vs low-scoring interaction patterns (24%-39% quiz score)\"",
   "removal_reason": "",
   "source": "Shen & Tamkin, arXiv 2601.20245",
   "source_url": "https://arxiv.org/abs/2601.20245",
   "locator": "§1.1; §6; Figure 11 (Generation-Then-Comprehension n=2 86%; Conceptual Inquiry n=7 65%)",
   "date_read": "2026-09-24"
  },
  {
   "id": "A10",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-A10",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "The low-scoring patterns were delegation, progressive reliance and iterative AI debugging (relying on the AI to debug or verify their code)",
   "quote": "\"Iterative AI Debugging (n=4): Participants in this group relied on AI to debug or verify their code.\"",
   "removal_reason": "",
   "source": "Shen & Tamkin, arXiv 2601.20245v2",
   "source_url": "https://arxiv.org/html/2601.20245v2",
   "locator": "§6, interaction patterns",
   "date_read": "2026-09-24"
  },
  {
   "id": "L1",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-L1",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "In pre-registered lab experiments on learning to code, giving students an AI chatbot had no overall effect on learning",
   "quote": "\"we find no effect of LLMs on overall learning outcomes\"",
   "removal_reason": "",
   "source": "Lehmann, Cornelius & Sting, arXiv 2409.09047v2",
   "source_url": "https://arxiv.org/abs/2409.09047",
   "locator": "Abstract; Tables 5, 7",
   "date_read": "2026-09-24"
  },
  {
   "id": "L2",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-L2",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "Students who used it to substitute for their own work (e.g. generating solutions to exercises) covered more topics but understood less; students who used it to complement their work (e.g. asking for explanations) understood more (exploratory, not randomised)",
   "quote": "\"Students who substitute some of their learning activities with LLMs (e.g., by generating solutions to exercises) increase the volume of topics\" / \"Students who complement their learning activities with LLMs (e.g., by asking for explanations) do not increase topic volume but do increase their understanding.\"",
   "removal_reason": "",
   "source": "Lehmann et al., arXiv 2409.09047v2",
   "source_url": "https://arxiv.org/abs/2409.09047",
   "locator": "Abstract; §9",
   "date_read": "2026-09-24"
  },
  {
   "id": "L3",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-L3",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "42% of messages asking for a solution were sent without a single attempt",
   "quote": "\"42% of messages asking for a solution were sent without a single attempt\"",
   "removal_reason": "",
   "source": "Lehmann et al., arXiv 2409.09047v2",
   "source_url": "https://arxiv.org/abs/2409.09047",
   "locator": "§9.1 (Study 3)",
   "date_read": "2026-09-24"
  },
  {
   "id": "L4",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-L4",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "The chatbot raised how much students felt they had learned by more than their actual learning explains",
   "quote": "\"LLMs increase perceived learning by more than can be explained by actual differences in learning\"",
   "removal_reason": "",
   "source": "Lehmann et al., arXiv 2409.09047v2",
   "source_url": "https://arxiv.org/abs/2409.09047",
   "locator": "§9.3",
   "date_read": "2026-09-24"
  },
  {
   "id": "D9",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-D9",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "Claude Skills need code execution switched on (Settings > Capabilities > Code execution and file creation on Free, Pro and Max)",
   "quote": "\"This feature requires code execution to be enabled.\"",
   "removal_reason": "",
   "source": "Claude Help Center, \"Use Skills in Claude\"",
   "source_url": "https://support.claude.com/en/articles/12512180-use-skills-in-claude",
   "locator": "Article body",
   "date_read": "2026-09-24"
  },
  {
   "id": "D10",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-D10",
   "status": "REPORTED",
   "status_label": "Reported",
   "removed": false,
   "claim": "ChatGPT's Study mode is chosen by typing @study or from the + menu, and is on every plan",
   "quote": "\"Available across ChatGPT plans globally on web, iOS, and Android\"",
   "removal_reason": "",
   "source": "OpenAI Help Center, \"Using study mode in ChatGPT\"",
   "source_url": "https://help.openai.com/en/articles/11780217-using-study-mode-in-chatgpt",
   "locator": "Article body",
   "date_read": "2026-09-24"
  },
  {
   "id": "X6",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-X6",
   "status": "EXECUTED",
   "status_label": "Executed",
   "removed": false,
   "claim": "The cold check with a topic line kept to the named topics and scored correctly in the ChatGPT app (3 of 6, the planted misses exactly), and kept to topic on 17 of 18 questions across Claude, Gemini and GPT",
   "quote": "See run record",
   "removal_reason": "",
   "source": "First-party: ChatGPT web (Plus) via Claude in Chrome + l3_cc2.py, l3x_cc2.py",
   "source_url": "",
   "locator": "",
   "date_read": "2026-09-24"
  },
  {
   "id": "P1",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-P1",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "Tutor prompts built to make students work before giving answers were published by Ethan and Lilach Mollick in 2023, and the Nigeria programme borrowed from them",
   "quote": "\"Some of the prompt structures were derived from Mollick and Mollick (2023a)\"",
   "removal_reason": "",
   "source": "De Simone et al., World Bank PRWP 11125, fn 3; Mollick & Mollick, \"Assigning AI\" (SSRN 4475995)",
   "source_url": "https://papers.ssrn.com/sol3/papers.cfm?abstract_id=4475995",
   "locator": "§2.1, pp. 7 to 8 (main text)",
   "date_read": "2026-09-24"
  },
  {
   "id": "A9",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-A9",
   "status": "REPORTED",
   "status_label": "Reported",
   "removed": false,
   "claim": "The no-AI group averaged about 65% on the quiz",
   "quote": "Figure 6 plots about 65.4% for the control group (read from marker positions); the Anthropic blog gives 67%",
   "removal_reason": "",
   "source": "Shen & Tamkin, arXiv 2601.20245, Figure 6; Anthropic blog",
   "source_url": "https://arxiv.org/abs/2601.20245",
   "locator": "Figure 6",
   "date_read": "2026-09-24"
  },
  {
   "id": "D11",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-D11",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "ChatGPT project: New project in the sidebar; instructions via the project's settings",
   "quote": "\"Select New project in the sidebar.\" / \"Select the more options menu (•••), then select Project settings to add instructions for the project.\"",
   "removal_reason": "",
   "source": "OpenAI Help Center, Projects in ChatGPT",
   "source_url": "https://help.openai.com/en/articles/10169521",
   "locator": "\"Create a project\"; \"Add project instructions\"",
   "date_read": "2026-09-27"
  },
  {
   "id": "D12",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-D12",
   "status": "VERIFIED",
   "status_label": "Verified",
   "removed": false,
   "claim": "Claude project: Projects, then New Project, with instructions and knowledge files",
   "quote": "\"click “Projects,”\" … \"Click \"+ New Project\" in the upper right corner.\" / \"Anything you upload to this space will be used across all of your chats within that project.\" / \"Click on \"Set project instructions.\"\"",
   "removal_reason": "",
   "source": "Claude Help Center, projects article",
   "source_url": "https://support.claude.com/en/articles/9519177",
   "locator": "\"How to create a project\"; \"Add content to project knowledge\"; \"Add project instructions\"",
   "date_read": "2026-09-27"
  },
  {
   "id": "D13",
   "url": "https://durabilitycurve.com/claims/learn-with-ai-without-getting-worse/#c-D13",
   "status": "REPORTED",
   "status_label": "Reported",
   "removed": false,
   "claim": "Gemini: a Gem (Gems, then New Gem) with files added under Knowledge",
   "quote": "Help-page path: Gems → New Gem → Knowledge → Add files",
   "removal_reason": "",
   "source": "Google Gemini Help",
   "source_url": "https://support.google.com/gemini/answer/15146780",
   "locator": "Article body",
   "date_read": "2026-09-24"
  }
 ]
}