Risk Under Pressure: Compute-Aware Evaluation of Adversarial Robustness in Language Models Paper • 2606.11409 • Published Jun 9 • 9
BigDocs: An Open and Permissively-Licensed Dataset for Training Multimodal Models on Document and Code Tasks Paper • 2412.04626 • Published Dec 5, 2024 • 15
Chitrarth: Bridging Vision and Language for a Billion People Paper • 2502.15392 • Published Feb 21, 2025
LitLLMs, LLMs for Literature Review: Are we there yet? Paper • 2412.15249 • Published Dec 15, 2024 • 2
IndicVisionBench: Benchmarking Cultural and Multilingual Understanding in VLMs Paper • 2511.04727 • Published Nov 6, 2025
VoiceAgentBench: Are Voice Assistants ready for agentic tasks? Paper • 2510.07978 • Published Oct 9, 2025