Tags
2 pages
LLM
GPT-3: In-Context Learning and the Few-Shot Explosion
DeepSeek-R1: Eliciting Reasoning in LLMs via Reinforcement Learning