Skip to content
View abhaydwived's full-sized avatar

Organizations

@CUK-COMMIT

Block or report abhaydwived

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
abhaydwived/README.md

Hi, I'm Abhay Narayan Dwivedi

Robotics Software · Reinforcement Learning · AI/ML

B.Tech Mathematics & Computing, Central University of Karnataka (2023–2027)


About Me

I work at the intersection of robotics, reinforcement learning, and intelligent control, with a focus on robot learning, bipedal locomotion, motion planning, and simulation-based control.

My work spans from robot modeling and inverse kinematics to reinforcement learning, trajectory generation, and control, with experience in both simulation and physical robotic systems.

I primarily work with MuJoCo, PyBullet, Gymnasium, and Python, and I am interested in building learning-based robotic systems that can adapt to complex environments and real-world constraints.

Currently exploring: Robot Learning · Reinforcement Learning · Whole-Body Control · MPC · Optimal Control · Sim-to-Real


Research & Technical Interests

Robotics & AI
├── Robot Learning
│   ├── Reinforcement Learning
│   ├── Imitation Learning
│   ├── Reward Shaping
│   └── Sim-to-Real
│
├── Humanoid & Biped Robotics
│   ├── Bipedal Locomotion
│   ├── Multi-Skill Locomotion
│   ├── Motion Generation
│   └── Whole-Body Control
│
├── Control & Optimization
│   ├── PD / PID Control
│   ├── Model-Based Control
│   ├── MPC
│   └── Trajectory Optimization
│
└── Planning & Perception
    ├── Inverse Kinematics
    ├── Motion Planning
    ├── A* / Sampling-Based Planning
    └── Computer Vision

Tech Stack

Category Technologies
Languages Python, C++, C, SQL, JavaScript, R
Robotics & RL Gymnasium, Stable-Baselines3, SAC, PPO, TD3, DDPG
Control PD/PID, Model-Based Control, MPC, Whole-Body Control
Robotics Inverse Kinematics, Motion Planning, Robot Dynamics
Simulation MuJoCo, PyBullet, MjLab, Isaac Lab
Deep Learning PyTorch, TensorFlow, Keras
Computer Vision OpenCV, MediaPipe, dlib, ONNX
Tools Git, Linux, TensorBoard, Weights & Biases

Research

IEEE ICC 2025 | Published

Learning Multi-Skill Locomotion in Underactuated Biped: A Waypoint-Based Reward Shaping Approach

A 6-DOF underactuated biped in PyBullet and benchmarked SAC, TD3 and DDPG on five locomotion tasks (standing, push recovery, walking, uneven terrain, stair descent) using progressive waypoint reward shaping.


Currently Exploring

  • Model-Based Reinforcement Learning
  • Multi-Agent Reinforcement Learning
  • Whole-Body Control for humanoid robots
  • Model Predictive Control
  • Sim-to-Real Robot Learning
  • Learning-based motion planning

Connect

I’m interested in robotics research, open-source robotics projects, and internship opportunities in: Robotics Software · Reinforcement Learning · Robot Learning · ML Engineering

📧 abhaydwivedi10122005@gmail.com

LinkedIn Portfolio

Pinned Loading

  1. Learning_Multi-Skill-Locomotion-in-Underactuated-Biped Learning_Multi-Skill-Locomotion-in-Underactuated-Biped Public

    Built a 6-DOF underactuated biped in PyBullet and benchmarked SAC, TD3 and DDPG on five locomotion tasks (standing, push recovery, walking, uneven terrain, stair descent) using progressive waypoint…

    Python 1

  2. Underactuated-Biped-Obstacle-Avoidance-using-Deep-Reinforcement-Learning Underactuated-Biped-Obstacle-Avoidance-using-Deep-Reinforcement-Learning Public

    This project presents a robust and energy-efficient obstacle avoidance framework for an 8-DOF bipedal robot using Deep Reinforcement Learning (Soft Actor-Critic). By tightly integrating an A* plann…

    Python 1

  3. Face-Recognition-Attendance-system Face-Recognition-Attendance-system Public

    A real-time face recognition-based attendance system built with Flask, OpenCV, and face_recognition. This project enables automatic attendance marking, user management, live monitoring, and reporti…

    Python 2 1

  4. LLM-Guided-Reinforcement-Learning-for-BipedalWalker-v3 LLM-Guided-Reinforcement-Learning-for-BipedalWalker-v3 Public

    Automated LLM-Guided Reinforcement Learning Testbed. This project leverages the modern BipedalWalker-v3 environment from Gymnasium to orchestrate a continuous cycle of agent training and intellige…

    Python 1

  5. Mental-Health-Score Mental-Health-Score Public

    Mental Health Score Predictor is a full-stack web application designed to evaluate and predict a user's mental well-being score based on their daily digital behavior and lifestyle factors.

    Jupyter Notebook

  6. Unitree-G1_Wbc_Balancing_Sit-to-Stand Unitree-G1_Wbc_Balancing_Sit-to-Stand Public

    Implementation of a Whole-Body Controller (WBC) based on Quadratic Programming (QP) for the Unitree G1 humanoid robot. The framework utilizes MuJoCo for simulating physics and kinematics, focusing…

    Python