A Fairness–Utility Evaluation Framework for Assessing Large Language Models (LLMs)
Large language models are increasingly used in contexts where their outputs can affect people directly, including hiring, admissions, and lending. This growing role makes it important to consider not only how well these models perform, but also whether their behavior is fair. Although many fairness metrics, bias benchm...