Paper_colm2026
Our paper “In-Situ Behavioral Evaluation for LLM Fairness, Not Standardized-Test Scores” is accepted to COLM 2026. We argue that LLM fairness should be evaluated through in-situ behavioral pattern rather than standardized-test Q&A benchmarks.