Similar Items: SupraTok: Cross-Boundary Tokenization for Enhanced Language Model Performance
- SupraTok: Cross-Boundary Tokenization for Enhanced Language Model Performance
- Re-evaluating the Word Token for Bilingual Speech Processing: The Case for Intonation Units
- Cross-layer Attention Sharing for Pre-trained Large Language Models
- Can Language Models Learn Typologically Implausible Languages?
- Enhancing spoken language identification on Indian languages with orthogonal task and related task integration: unveiling new dimensions of accuracy
- Anthropocentric Bias in Language Model Evaluation