Rafid Ahmed
CS Ph.D. Student · Former AI Engineer
I am a Computer Science Ph.D. student at the University of Central Florida, where I am co-supervised by Dr. Mubarak Shah and Dr. Yuzhang Shang at the Center for Research in Computer Vision (CRCV). Previously, I worked as an AI Engineer at Penta Global Limited, where we collaborated with Neustring to build agentic solutions for major telecommunication companies.
I began my research journey through NLP and VLMs, motivated by the potential to make AI systems more inclusive, interpretable, and impactful. My current work centers on vision-language-action (VLA) models and world action models (WAM), along with vision-language models, image diffusion, diffusion LLMs, and preference optimization. Earlier, my work spanned low-resource learning, the application of LLMs in the medical domain, parameter-efficient fine-tuning and Reinforcement Learning, with a particular emphasis on improving the reasoning capabilities and explainability of both VLMs and LLMs.
Research Interests
- Vision-Language-Action (VLA) Models: grounding perception and language in embodied action
- World Action Models (WAM): learning predictive world models for action and planning
- Vision-Language Models: reasoning, explainability, and parameter-efficient fine-tuning (LoRA)
- Image Diffusion: generative modeling for visual understanding and synthesis
- Diffusion LLMs: diffusion-based language modeling and generation
- Preference Optimization: aligning models with human preferences (RLHF, DPO)
- Medical AI: applying LLMs and VLMs to healthcare and clinical reasoning
- Low-Resource Learning: building inclusive AI for under-represented languages
News
Exciting news! Our work on “Evaluating Large Vision Language Models on Bangla Medical Visual Question Answering” has been accepted at the 2026 Meeting of the Association for Computational Linguistics (ACL), San Diego, California.
Our work on “Benchmarking Large Language Models on Bangla Dialect Translation and Dialectal Sentiment Analysis” got the Best Paper Award in BLP workshop at AACL-IJCNLP 2025.
Our work on “How Good LLMs Are at Answering Bangla Medical Visual Questions? Dataset and Benchmarking” Got accepted in AAAI 2026 AIMedHealth.
Our work on “Benchmarking Large Language Models on Bangla Dialect Translation and Dialectal Sentiment Analysis” got accepted in BLP workshop at AACL-IJCNLP 2025.
Our work on “PentaML at BLP-2025 Task 1: Linear Probing of Pre-trained Transformer-based Models for Bangla Hate Speech Detection” got accepted in BLP workshop at AACL-IJCNLP 2025.
Selected Publications
- Benchmarking Large Language Models on Bangla Dialect Translation and Dialectal Sentiment AnalysisIn Proceedings of the IJCNLP 2025 Second Workshop on Bangla Language Processing (BLP-2025), Dec 2025