Skip to content
View abhaydwived's full-sized avatar

Organizations

@CUK-COMMIT

Block or report abhaydwived

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
abhaydwived/README.md

Hi, I'm Abhay Narayan Dwivedi

Robotics Software · Reinforcement Learning · AI/ML B.Tech Mathematics & Computing, Central University of Karnataka (2023–2027)

IEEE ICC 2025 xTerra Robotics IIT Mandi


About Me

I work at the intersection of robotics, reinforcement learning, and intelligent control, with hands-on experience in bipedal locomotion, inverse kinematics, motion planning, robot simulation, and learning-based control.

I have worked with MuJoCo and PyBullet, developing controllers and RL systems for both simulated and physical robotic systems.

Currently exploring: Reinforcement Learning · Robot Learning · Optimal Control · MPC · Intelligent Robotics


Experience

xTerra Robotics — Intern

May 2026 – July 2026

  • Implemented PD and model-based controllers in MuJoCo for CartPole, Furuta Pendulum, and Reaction Wheel Pendulum.
  • Deployed PD balancing on physical hardware, maintaining upright stability for 10+ seconds under real-world noise and actuator delays.
  • Trained a PPO agent with a custom reward to learn bipedal sit-to-stand from IK-generated reference trajectories.
  • Developed IK pipelines for deep squat-to-stand, sit-to-stand, and fallen-to-stand transitions with smooth joint trajectories.
  • Engineered a PD torque controller for dynamic deep squat-to-stand motion.

IIT Mandi — Research Intern

May 2025 – December 2025

  • Developed 3+ RL-based biped locomotion systems using custom Gymnasium and PyBullet environments.
  • Benchmarked SAC, TD3, and DDPG on a 6–8 DOF underactuated biped over 10M+ training steps.
  • Achieved 94% navigation success and 2% fall rate under stochastic disturbances.
  • Integrated A* global planning with SAC-based locomotion for hierarchical obstacle avoidance, achieving 87.6% goal success in cluttered environments.
  • Designed progressive waypoint-based reward shaping for stable, goal-directed locomotion.

YBI Foundation — Data Analytics Intern

October 2024 – December 2024

  • Completed 3+ end-to-end data analytics projects involving EDA, feature engineering, and predictive modeling.
  • Built predictive pipelines using Python, Pandas, and Scikit-learn, achieving 75–85% model accuracy.

Publications

Learning Multi-Skill Locomotion in Underactuated Biped: A Waypoint-Based Reward Shaping Approach

Published — IEEE ICC 2025


Featured Projects

LLM-Guided Reward Shaping for Bipedal Locomotion

Automated pipeline using Gemini 2.5 Flash to iteratively generate and refine reward functions for a SAC-trained biped across flat, uneven, slope, and stair terrains.

  • 151+ unit course completion vs. −0.76 for the vanilla PPO baseline.
  • Gait symmetry index of 0.023 and torso tilt of 0.086 rad.
  • Evaluated 21 gait-specific metrics.
  • Integrated human-guided constraints for naturalistic locomotion.
  • Implemented automated training resume and rollback on generated-code failures.

Gemini 2.5 Flash SAC PyBullet Gymnasium NumPy Matplotlib

RL Biped Locomotion & Obstacle Avoidance — IIT Mandi

6–8 DOF underactuated biped trained with SAC, TD3, and DDPG over 10M+ steps, combined with A* hierarchical planning and waypoint-based reward shaping.

94% navigation success · 2% fall rate

PyBullet SAC TD3 DDPG A* Gymnasium

Biped Sit-to-Stand & Squat-to-Stand — xTerra Robotics

Developed IK-generated reference trajectories, PPO-based learning, and PD torque control for dynamic biped stand-up transitions.

MuJoCo PPO Inverse Kinematics PD Control

Face Recognition Attendance System

Real-time contactless attendance system with 128-D face encoding, liveness detection, and a secure web dashboard with live camera streaming and CSV export.

Python Flask OpenCV dlib ONNX SQLite

MNIST Neural Network From Scratch

Fully connected neural network implemented using NumPy with manual forward/backward propagation, softmax, cross-entropy loss, and mini-batch gradient descent.

NumPy Matplotlib

Hand Tracking Virtual Painter

Real-time gesture-controlled drawing system supporting drawing, erasing, brush control, and color selection.

OpenCV MediaPipe NumPy


Skills

Languages Python · C++ · JavaScript · SQL · R

ML / Deep Learning PyTorch · TensorFlow · Keras · Scikit-learn · NumPy · Pandas

Robotics / RL / Simulation Gymnasium · PyBullet · MuJoCo · MjLab · SAC · PPO · TD3 · DDPG · Inverse Kinematics · Motion Planning · PD Control

Computer Vision OpenCV · MediaPipe · dlib · ONNX

Tools Git · TensorBoard · Flask · SQLite · Power BI


Connect

Open to internship opportunities in Software Robotics, ML Engineering, and Data Science, as well as research collaborations.

📧 abhaydwivedi10122005@gmail.com

LinkedIn Portfolio

Pinned Loading

  1. Learning_Multi-Skill-Locomotion-in-Underactuated-Biped Learning_Multi-Skill-Locomotion-in-Underactuated-Biped Public

    Benchmarking SAC, TD3, and DDPG on multi-skill bipedal locomotion using progressive waypoint-based reward shaping in PyBullet. Published at IEEE ICC 2025.

    Python 1

  2. Underactuated-Biped-Obstacle-Avoidance-using-Deep-Reinforcement-Learning Underactuated-Biped-Obstacle-Avoidance-using-Deep-Reinforcement-Learning Public

    This project presents a robust and energy-efficient obstacle avoidance framework for an 8-DOF bipedal robot using Deep Reinforcement Learning (Soft Actor-Critic). By tightly integrating an A* plann…

    Python 1

  3. Face-Recognition-Attendance-system Face-Recognition-Attendance-system Public

    A real-time face recognition-based attendance system built with Flask, OpenCV, and face_recognition. This project enables automatic attendance marking, user management, live monitoring, and reporti…

    Python 2 1

  4. LLM-Guided-Reinforcement-Learning-for-BipedalWalker-v3 LLM-Guided-Reinforcement-Learning-for-BipedalWalker-v3 Public

    Automated LLM-Guided Reinforcement Learning Testbed. This project leverages the modern BipedalWalker-v3 environment from Gymnasium to orchestrate a continuous cycle of agent training and intellige…

    Python 1