Skip to content
View rana-rishith's full-sized avatar

Block or report rana-rishith

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
rana-rishith/README.md

Hey, I'm Rana 👋 B.Tech Information Technology @ Manipal University Jaipur, India I build things at the intersection of vision and language — multimodal AI, deep learning, and NLP. Currently exploring parameter-efficient fine-tuning methods for adapting large models to specialized domains.

🔬 What I'm Working On

Medical Image Captioning — Multimodal system combining ViT-Base + Phi-2 (2.7B) with LoRA fine-tuning to generate radiology image descriptions. Trained on a single RTX 4090 in ~5 hours.

🛠️ Tech Stack Languages: Python, Java, SQL, C ML/DL: PyTorch, HuggingFace Transformers, PEFT/LoRA Tools: Git, Linux, RunPod, Jupyter, VScode Interests: Multimodal LLMs, AI Engineering, AI Automations, Computer Vision, NLP, Efficient Fine-Tuning

📫 Let's Connect LinkedIn: https://www.linkedin.com/in/rana-rishith-musunuri/ Email: musunuri.2430030148@muj.manipal.edu Email: ranarishith.24@gmail.com

Open to research , technical internships, AI engineering roles, and collaborations in deep learning and multimodal AI.

Pinned Loading

  1. medical-image-captioning medical-image-captioning Public

    Medical Image Captioning with ViT-Base + Phi-2 + LoRA on ROCOv2 radiology dataset

    Python

  2. rana-rishith rana-rishith Public