Announcement_11
📄 New preprint from my Thomson Reuters Labs internship: Equal Ranking Quality, Different Decisions: Training Order-Consistent LLM Scorers! We show that LLM scorers with the same ranking quality can still make different decisions, and propose OC-SFT to train order-consistent scorers. Code is available here.