conftrace_

Papers

187,652 papers found · 36,278 more still awaiting a processed abstract Show those too
Evaluating Bias in LLMs for Job-Resume Matching: Gender, Race, and Education
Hayate Iso, Pouya Pezeshkpour, Nikita Bhutani et al.
2025 NAACL
Evaluating Compositional Generalisation in VLMs and Diffusion Models
Beth Pearson, Bilal Boulbarss, Michael Wray et al.
2025 EMNLP
Evaluating Compound AI Systems through Behaviors, Not Benchmarks
Pranav Bhagat, K N Ajay Shastry, Pranoy Panda et al.
2025 EMNLP
Evaluating Cultural and Social Awareness of LLM Web Agents
Haoyi Qiu, Alexander Fabbri, Divyansh Agarwal et al.
2025 NAACL
Evaluating Cultural Knowledge and Reasoning in LLMs Through Persian Allusions
Melika Nobakhtian, Yadollah Yaghoobzadeh, Mohammad Taher Pilehvar
2025 EMNLP
2025 COLING
Evaluating distillation methods for data-efficient syntax learning
Takateru Yamakoshi, Thomas L. Griffiths, R. Thomas McCoy et al.
2025 EMNLP
2025 NAACL
Evaluating Evaluation Metrics – The Mirage of Hallucination Detection
Atharva Kulkarni, Yuan Zhang, Joel Ruben Antony Moniz et al.
2025 EMNLP
2025 NAACL
2025 COLING