Time Present and Time Past: Benchmarking Large Language Models on Temporally Evolving Document Understanding Paper โข 2608.08512 โข Published 17 days ago
Do Small Models Use the Law You Give Them? Context-Injected Fine-Tuning for Legal QA in Bangladesh Paper โข 2607.23446 โข Published Jul 26 โข 1
Many Dialects, Many Languages, One Cultural Lens: Evaluating Multilingual VLMs for Bengali Culture Understanding Across Historically Linked Languages and Regional Dialects Paper โข 2603.21165 โข Published Mar 22
Ready to Translate, Not to Represent? Bias and Performance Gaps in Multilingual LLMs Across Language Families and Domains Paper โข 2510.07877 โข Published Oct 9, 2025
MathMist: A Parallel Multilingual Benchmark Dataset for Mathematical Problem Solving and Reasoning Paper โข 2510.14305 โข Published Oct 16, 2025 โข 2
Watch, Listen, Understand, Mislead: Tri-modal Adversarial Attacks on Short Videos for Content Appropriateness Evaluation Paper โข 2507.11968 โข Published Jul 16, 2025
VisText-Mosquito: A Multimodal Dataset and Benchmark for AI-Based Mosquito Breeding Site Detection and Reasoning Paper โข 2506.14629 โข Published Jun 17, 2025 โข 1
BanglishRev: A Large-Scale Bangla-English and Code-mixed Dataset of Product Reviews in E-Commerce Paper โข 2412.13161 โข Published Dec 17, 2024
A Bidirectional Siamese Recurrent Neural Network for Accurate Gait Recognition Using Body Landmarks Paper โข 2412.03498 โข Published Dec 4, 2024 โข 2
MathMist: A Parallel Multilingual Benchmark Dataset for Mathematical Problem Solving and Reasoning Paper โข 2510.14305 โข Published Oct 16, 2025 โข 2
A Bidirectional Siamese Recurrent Neural Network for Accurate Gait Recognition Using Body Landmarks Paper โข 2412.03498 โข Published Dec 4, 2024 โข 2