Skip to content
View nirbhay41120003's full-sized avatar

Highlights

  • Pro

Block or report nirbhay41120003

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
nirbhay41120003/README.md

Nirbhay Gupta

AI/ML Engineer · Edge AI & LLM Systems

Portfolio · Email · LinkedIn


About

B.Tech student in Electronics and Communication Engineering at IIIT Dharwad (2023–2027), pursuing a minor in Generative AI and Large Language Models. I build production-oriented systems across three areas: edge AI deployment, LLM-based products, and data engineering.

  • 🔭 Currently building ListenMe — an always-on, local, real-time audio transcription app with RAG-based long-term memory
  • 🧠 Working across RAG pipelines, fine-tuning, model editing, and agentic workflows
  • ⚡ Deploying models to edge hardware (NVIDIA Jetson Orin Nano) with TensorRT/ONNX
  • 🛠️ Prior internship: geospatial data pipelines at the School of Planning and Architecture, Bhopal

Selected Projects

MLops-Network-Security End-to-end MLOps pipeline for phishing URL detection. FastAPI · MLflow · DagsHub · AWS S3 · Docker · ECR/EC2.

Personal RAG Assistant — live at chat.nirbhay.engineer Deployed retrieval-augmented generation assistant.

ListenMe Always-on local real-time audio transcription app using NVIDIA NeMo/Nemotron ASR, with ChromaDB for long-term memory, SQLite for raw transcripts, and a fine-tuned Gemma model for end-of-day summarization.

RoastMate QLoRA fine-tuning of Gemma-2 2B with an integrated RAG pipeline; deployed via Gradio on Lightning AI.

ROME on GPT-2 Implementation of Rank-One Model Editing for direct factual knowledge edits in a transformer.


Tech Stack

Languages: Python, Bash, SQL ML/AI: PyTorch, TensorFlow, Hugging Face, LoRA/QLoRA fine-tuning, RAG (ChromaDB, Pinecone) Edge & Deployment: NVIDIA Jetson Orin Nano, TensorRT, ONNX, Docker Infra & Data: FastAPI, MLflow, AWS (S3, EC2, ECR), n8n, MySQL Other: Git, Linux, computer networking (IP, DHCP, NAT, BGP, SDN)


GitHub Stats

GitHub Stats GitHub Streak


nirbhaygupta.me · nirbhaygupta4113@gmail.com

Popular repositories Loading

  1. EchoMind EchoMind Public

    Local-first voice memory assistant — capture speech, transcribe live, and ask a private RAG chatbot grounded in what you've actually said. No cloud required.

    Python 80 2

  2. ROME-Rank-One-Model-Editing ROME-Rank-One-Model-Editing Public

    Jupyter Notebook 1

  3. EmotionSense-CNN EmotionSense-CNN Public

    This project implements a Convolutional Neural Network (CNN) to detect emotions from facial expressions. It utilizes real-time video footage to classify emotions such as happiness, sadness, anger, …

    Python

  4. AttendanceSystem AttendanceSystem Public

    Attendance System on Raspberry pi and Using Neural Network For Face Recognition

    Python

  5. dbms-chat-assistant dbms-chat-assistant Public

    Python

  6. nirbhay41120003 nirbhay41120003 Public