Slide
LLM Research Directions
Day 3 of the SCIPE workshop on large language models: attention under the hood and open research directions — why Pass@k is unstable, Bayes@N posteriors, rankings with uncertainty, and ranking reasoning LLMs under test-time scaling.