Improving Lexical Difficulty Prediction with Context-Aligned Contrastive Learning and Ridge Ensembling Paper • 2605.08950 • Published May 9
ITLC at SemEval-2026 Task 11: Normalization and Deterministic Parsing for Formal Reasoning in LLMs Paper • 2603.02676 • Published May 9
Lius: Translation Model Based Instructional Lingustic Using Continual Instruction Tuning In Kupang Malay Paper • 2606.11786 • Published Jun 10 • 2
Lius: Translation Model Based Instructional Lingustic Using Continual Instruction Tuning In Kupang Malay Paper • 2606.11786 • Published Jun 10 • 2
CommonLID: Re-evaluating State-of-the-Art Language Identification Performance on Web Data Paper • 2601.18026 • Published Jan 25
UniSkill: A Dataset for Matching University Curricula to Professional Competencies Paper • 2603.03134 • Published Mar 3
WorkRB: A Community-Driven Evaluation Framework for AI in the Work Domain Paper • 2604.13055 • Published Mar 17
CroCo: Cross-Lingual Contrastive Preference Tuning on Self-Generations Paper • 2605.26293 • Published May 25 • 6
CroCo: Cross-Lingual Contrastive Preference Tuning on Self-Generations Paper • 2605.26293 • Published May 25 • 6
ScholarChemQA: Unveiling the Power of Language Models in Chemical Research Question Answering Paper • 2407.16931 • Published Jul 24, 2024
AutoLLMResearch: Training Research Agents for Automating LLM Experiment Configuration -- Learning from Cheap, Optimizing Expensive Paper • 2605.11518 • Published May 12 • 4
AutoLLMResearch: Training Research Agents for Automating LLM Experiment Configuration -- Learning from Cheap, Optimizing Expensive Paper • 2605.11518 • Published May 12 • 4
Global PIQA: Evaluating Physical Commonsense Reasoning Across 100+ Languages and Cultures Paper • 2510.24081 • Published Oct 28, 2025 • 24
CommonLID: Re-evaluating State-of-the-Art Language Identification Performance on Web Data Paper • 2601.18026 • Published Jan 25
Anthropogenic Regional Adaptation in Multimodal Vision-Language Model Paper • 2604.11490 • Published Apr 13 • 16
Sailor2: Sailing in South-East Asia with Inclusive Multilingual LLMs Paper • 2502.12982 • Published Feb 18, 2025 • 19
Seeing Culture: A Benchmark for Visual Reasoning and Grounding Paper • 2509.16517 • Published Sep 20, 2025 • 3
Can Large Language Models Understand, Reason About, and Generate Code-Switched Text? Paper • 2601.07153 • Published Jan 12
M4-RAG: A Massive-Scale Multilingual Multi-Cultural Multimodal RAG Paper • 2512.05959 • Published Dec 5, 2025