Similar Items: Ev 2 R: Evaluating Evidence Retrieval in Automated Fact-Checking
- Can LLMs Automate Fact-Checking Article Writing?
- CROC: Evaluating and Training T2I Metrics with Pseudo- and Human-Labeled Contrastive Robustness Checks
- Contrafactives and facts for knowledge
- Interactive metadiscourse in L1 and L2 English: Evidence from editorials
- R ESEARCH QA: Evaluating Scholarly Question Answering at Scale Across 75 Fields with Survey-Mined Questions and Rubrics
- Phonetic Challenges in French Pronunciation: Evidence-Based Error Analysis and Pedagogical Strategies for Turkish A2+ Learners