I’m broadly interested in Medical Imaging and Computer Vision, with my work spanning multimodal learning, domain generalization, text recognition & spotting, document analysis, and cross-domain visual understanding. I’m particularly interested in developing learning systems that generalize reliably across challenging domains and modalities.
At the Computer Vision and Pattern Recognition Unit, Indian Statistical Institute, Kolkata, I have worked with Dr. Umapada Pal on text-centric computer vision, including domain-agnostic text recognition, domain-independent text spotting, and cross-domain scene understanding across challenging visual settings. I have also worked with Dr. Palash Ghosal on multimodal learning for integrated tumor segmentation and survival prediction.
Alongside research, I work as an AI & Cloud Engineer at NCompass, building agentic AI systems and ML infrastructure across GCP and AWS. I completed my B.Tech. in Computer Science and Engineering from Sikkim Manipal Institute of Technology.