Published | ISAI 2025
Knee osteoarthritis detection and categorization
A deep learning study of knee X-rays using the Kellgren-Lawrence grading scale. The work received a Best Paper Award at ISAI 2025.
Machine learning | Computer vision | Research
Machine Learning Engineer
and Computer Vision Researcher
I am a machine learning engineer focused on visual intelligence and practical AI systems. I move between model development, applied retrieval, and the product work needed to make those systems useful.
My current interests include multimodal learning, medical image analysis, retrieval quality, and reliable LLM workflows.
Peer-reviewed work, accepted research, and independent study.
Published | ISAI 2025
A deep learning study of knee X-rays using the Kellgren-Lawrence grading scale. The work received a Best Paper Award at ISAI 2025.
Accepted | ICADCML 2026
A CBAM-enhanced U-Net approach for binary and multi-class knee osteoporosis detection, using texture-based feature extraction.
Preprint | Independent research
An astronomy vision-language model that aligns frozen visual features with a language model, then adds instruction tuning with LoRA.
An action-focused version of CLAD that keeps visual observations and endpoint latents as context. Trained and evaluated on NIV and CrossTask, with checkpoints, logs, and five-seed results published.
View project ->A two-stage astronomy vision-language model. It aligns frozen CLIP ViT-L/14 features with Qwen2.5-1.5B-Instruct, then adds LoRA instruction tuning for visual question answering on astronomical imagery.
View project ->A vision-language model for aerial and satellite imagery. It adapts the AstraQ-VL training recipe to VRSBench with a frozen CLIP encoder, Qwen2.5-3B, connector alignment, and LoRA instruction tuning.
View project ->A document-grounded question-answering CLI with retrieval-augmented generation, support for PDFs and text files, and an included evaluation step for relevance and hallucinations.
View project ->An end-to-end deep learning pipeline that reconstructs CT images from sinograms, developed for sparse-view and low-dose acquisition scenarios.
View project ->A conversational AI application with query routing, sub-query decomposition, NDJSON streaming, resumable sessions, and optional web grounding.
View project ->A hybrid retrieval system with semantic reranking, vector deduplication, and routing across multiple indexes.
View GitHub ->A document question-answering pipeline with multi-format parsing, OCR, hybrid retrieval, and LLM generation.
View GitHub ->2025 - Present
Axtria, Bengaluru
Working on multi-agent LLM workflows, retrieval-augmented analytics, and real-time conversational systems for enterprise use.
May 2025 - Sept 2025
Indian Institute of Technology, Kharagpur
Built a modified U-Net for CT image reconstruction on 20,000+ sinogram-image pairs, reaching PSNR above 35 dB and SSIM above 0.9 for sparse-view and low-dose scenarios.
May 2024 - July 2024
Axtria, Bengaluru
Built and validated time-series forecasting models and automated recurring analytics for planning work.
2023 - 2025
IIT (ISM) Dhanbad
Thesis on a multi-stage deep learning framework for knee osteoarthritis and osteoporosis detection, producing two peer-reviewed papers including a Best Paper Award at ISAI 2025.
2019 - 2023
IIEST Shibpur
Coursework included signal processing, digital systems, programming, and linear algebra.
ML blog
The archive covers topics such as meta-stacking, adversarial validation, model evaluation, deep learning, and optimization.
Browse articles ->04 | Contact