icl-gd-paper-icml-2024

IN premise — summaries/2026/08/24/shen-2023-icl-not-gd-s0-abstract.md

Created 2026-08-25T02:58:30+00:00

The paper by Shen, Mishra, and Khashabi examining ICL vs. Gradient Descent equivalence was published at ICML 2024.

Summary

This anchors a specific comparison between in-context learning and standard gradient descent to a peer-reviewed source that passed the bar of ICML's 2024 review cycle. It matters because the rest of the argument in the system about whether large language models are really just doing disguised training rests on the credibility and precision of that published analysis.