Similar Items: VoiceBench: Benchmarking LLM-Based Voice Assistants
- IssueBench: Millions of Realistic Prompts for Measuring Issue Bias in LLM Writing Assistance
- From Benchmarks to Skills: Low-Rank Factors for LLM Evaluation
- Between Face and Voice: Semiotic Relationships
- Goal Alignment in LLM-Based User Simulators for Conversational AI
- BP-LLM : Belief Propagation for Binary Feedback in Large Language Model Alignment
- On the Limitations of Language-targeted Pruning: Investigating the Calibration Language Impact in Multilingual LLM Pruning