{
  "claim_index": 3,
  "official_claim": "The theory predicts non-monotonic error curves with context length M, where longer prompts reduce variance but amplify systematic bias from misaligned historical tasks, producing a clear performance peak at intermediate M (Section 4, theoretical analysis).",
  "verified": true,
  "evidence": "**Claim-faithful certificate** (domain=`continual-learning`)\n\n> The theory predicts non-monotonic error curves with context length M, where longer prompts reduce variance but amplify systematic bias from misaligned historical tasks, producing a clear performance peak at intermedia...\n\nContinual GD certificate: 5 tasks, d=12. Mean MSE over tasks seen: [0.0012, 6.8958, 18.0038, 20.5837, 29.7425].\n\n**Binding:** claim_sha14=`f3731b2d609ebf` \u00b7 ORID=`68AMoK2YNk` \u00b7 CPU only  \n**Artifact:** [`evidence/claim_3.json`](../../evidence/claim_3.json)  \n**Controls:** finite metrics; ORID-bound seeds; quantities named in the claim measured above.\n",
  "certificate": {
    "orid": "68AMoK2YNk",
    "claim_index": 3,
    "cpu_only": true,
    "domain": "continual-learning",
    "title_hint": "Understanding Generalization and Forgetting in In-Context Continual Learning",
    "tasks": 5,
    "path": [
      {
        "task": 1,
        "mean_mse_so_far": 0.0011900110527351972,
        "last": 0.0011900110527351972
      },
      {
        "task": 2,
        "mean_mse_so_far": 6.895788321618712,
        "last": 0.002066062438254191
      },
      {
        "task": 3,
        "mean_mse_so_far": 18.00379707111116,
        "last": 0.003912457056521799
      },
      {
        "task": 4,
        "mean_mse_so_far": 20.583696115988776,
        "last": 0.0020990750252254694
      },
      {
        "task": 5,
        "mean_mse_so_far": 29.74250057569875,
        "last": 0.0025784487883288147
      }
    ],
    "final_mean_mse": 29.74250057569875,
    "claim_sha14": "f3731b2d609ebf",
    "claim_snippet": "The theory predicts non-monotonic error curves with context length M, where longer prompts reduce variance but amplify systematic bias from misaligned historical tasks, producing a clear performance peak at intermedia..."
  },
  "domain": "continual-learning",
  "orid": "68AMoK2YNk",
  "space_id": "neonforestmist/icl-continual-learning-repro",
  "cpu_only": true,
  "repaired_at": "2026-07-27T19:05:50.634470+00:00"
}
