dai-2023-six-classification-datasets
IN premise — summaries/2026/08/24/dai-2023-icl-gradient-descent-sA-appendix.md
Created 2026-08-24T17:10:54+00:00
Dai et al. 2023 evaluate ICL versus finetuning on six classification benchmarks: SST2, SST5, MR (Movie Reviews), Subj (Subjectivity), AGNews, and CB (CommitmentBank/NLI).
Summary
This pins down exactly which text-classification tasks are being compared, covering sentiment, topic, and entailment, where in-context learning is tested against standard finetuning. It matters because any downstream claim about which method "wins" at classification is only as credible as the specific benchmarks behind it, so knowing the six tasks by name is the baseline for judging whether a conclusion actually applies to the problem you care about.