Publications
Highlighted papers are representative. Full list on Google Scholar.
2026
Where You Backpropagate Matters: Token Hypothesis for Memory-Efficient Fine-Tuning
Md Kowsher, Chen Chen
NeurIPS 2026
LiME: Lightweight Mixture of Experts for Efficient Multimodal Multi-task Learning
Md Kowsher, Haris Mansoor, Nusrat Jahan Prottasha, Ozlem Garibay, Victor Zhu, Zhengping Ji, Chen Chen
ICML 2026 Spotlight
SliceFine: The Universal Winning-Slice Hypothesis for Pretrained Networks
Md Kowsher, Ali O. Polat, Prasanth Murali, Ehsan Mohammady Ardehaly, Chen Chen
ICML 2026
Monkey Jump: MoE-Style PEFT for Efficient Multi-Task Learning
Nusrat Jahan Prottasha, Md Kowsher, Chun-Nam Yu, Chen Chen, Ozlem Garibay
EMNLP 2026 (Main) Oral
FlowNIB: An Information Bottleneck Analysis of Bidirectional vs. Unidirectional Language Models
Md Kowsher, Nusrat Jahan Prottasha, Shiyun Xu, Shetu Mohanto, Chen Chen, Niloofar Yousefi, Ozlem Garibay
ICLR 2026
2025
Predicting Through Generation: Why Generation Is Better for Prediction
Md Kowsher, Nusrat Jahan Prottasha, Prakash Bhat, Chun-Nam Yu, Mojtaba Soltanalian, Ivan Garibay, Ozlem Garibay, Chen Chen, Niloofar Yousefi
ACL 2025 (Main)
RoCoFT: Efficient Finetuning of Large Language Models with Row-Column Updates
Md Kowsher, Tara Esmaeilbeig, Chun-Nam Yu, Chen Chen, Mojtaba Soltanalian, Niloofar Yousefi
ACL 2025 (Main); NeurIPS 2024 FITML Workshop Oral
TituLLMs: A Family of Bangla LLMs with Comprehensive Benchmarking
Shahriar Kabir Nahin, Rabindra Nath Nandi, Sagor Sarker, Quazi Sarwar Muhtaseem, Md Kowsher, Apu Chandraw Shill, Md Ibrahim, Mehadi Hasan Menon, Tareq Al Muntasir, Firoj Alam
ACL 2025 (Findings)
LLM-Mixer: Multiscale Mixing in LLMs for Time Series Forecasting
Md Kowsher, Md Shohanur Islam Sobuj, Nusrat Jahan Prottasha, E. Alejandro Alanis, Ozlem Ozmen Garibay, Niloofar Yousefi
ACL 2025 TRL Workshop
Does Self-Attention Need Separate Weights in Transformers?
Md Kowsher, Nusrat Jahan Prottasha, Chun-Nam Yu, Ozlem Ozmen Garibay, Niloofar Yousefi
NAACL 2025
BnTTS: Few-Shot Speaker Adaptation in Low-Resource Setting
Mohammad Jahid Ibna Basher, Md Kowsher, Md Saiful Islam, Rabindra Nath Nandi, Nusrat Jahan Prottasha, Mehadi Hasan Menon, Tareq Al Muntasir, Shammur Absar Chowdhury, Firoj Alam, Niloofar Yousefi, Ozlem Ozmen Garibay
NAACL 2025
Propulsion: Steering LLM with Tiny Fine-Tuning
Md Kowsher, Nusrat Jahan Prottasha, Prakash Bhat
COLING 2025 Oral
Infinite Reservoir Transformer
Jia Xu, Md Kowsher
US Patent Application 18/780,055 (Pub. No. US20250200362A1)
2024
Parameter-Efficient Fine-Tuning of Large Language Models using Semantic Knowledge Tuning
Nusrat Jahan Prottasha, Asif Mahmud, Md Shohanur Islam Sobuj, Prakash Bhat, Md Kowsher, Niloofar Yousefi, Ozlem Ozmen Garibay
Scientific Reports (Nature Portfolio), 2024
Token Trails: Navigating Contextual Depths in Conversational AI with ChatLLM
Md Kowsher, R. Panditi, Nusrat Jahan Prottasha, Prakash Bhat, Anupam Kumar Bairagi, Mohammad Shamsul Arefin
NLDB 2024
L-Tuning: Synchronized Label Tuning for Prompt and Prefix in Large Language Models
Md Kowsher, Md Shohanur Islam Sobuj, Asif Mahmud, Nusrat Jahan Prottasha, Prakash Bhat
ICLR 2024 Tiny Papers
2023
Contrastive Learning for Universal Zero-Shot NLI with Cross-Lingual Sentence Embeddings
Md Kowsher, Md Shohanur Islam Sobuj, Nusrat Jahan Prottasha, Mohammad Shamsul Arefin, Yasuhiko Morimoto
EMNLP 2023 Workshop (MRL)
NAM: From Hybrid Dialogers to Neural Responders
João Luís Lins, Nikhil Reddy, Abdul Rafae Khan, Md Kowsher, Abhijeet Gusain, Yeshwanth Reddy, Xuting Tang, Preet Jhanglani, Nusrat Zahan, Mengjiao Zhang, Yu Yu, Darshil Shah, Jia Xu
Alexa Prize SocialBot Grand Challenge 5 Proceedings (2nd place)
Impact Learning: A Learning Method from Feature's Impact and Competition
Nusrat Jahan Prottasha, Saydul Akbar Murad, Md Kowsher, Apurba Adhikary, Sujit Biswas, Anupam Kumar Bairagi
Journal of Computational Science (Elsevier)
2022
Bangla-BERT: Transformer-Based Efficient Model for Transfer Learning and Language Understanding
Md Kowsher, Abdullah As Sami, Nusrat Jahan Prottasha, Mohammad Shamsul Arefin, Pranab Kumar Dhar, Takeshi Koshiba
IEEE Access, 2022
CARAN: A Context-Aware Recency-Based Attention Network for Point-of-Interest Recommendation
Md Billal Hossain, Mohammad Shamsul Arefin, Iqbal H. Sarker, Md Kowsher, Pranab Kumar Dhar, Takeshi Koshiba
IEEE Access, 2022