Skip to content
View mturan33's full-sized avatar

Highlights

  • Pro

Block or report mturan33

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
mturan33/README.md

Mehmet Turan Yardimci

Robot Learning | Reinforcement Learning | Humanoids

Incoming M.Sc. student at Karlsruhe Institute of Technology (KIT). Computer Engineering graduate, Cukurova University. I work on RL fine tuning of vision language action policies on consumer hardware, and hierarchical control for humanoid robots.

mehmetturanyardimci@hotmail.com | LinkedIn | Google Scholar | arXiv | Portfolio


Projects

Online RL fine tuning of a flow matching VLA policy on a single 12GB consumer GPU, following the pi_RL approach. A verification first substrate: what trains is a logged fact, not an assumption. Project page

G1 Vision Language Action pipeline (in progress, repository not public yet)

RL to IL to VLA pipeline for the Unitree G1: expert demonstrations from trained RL policies, distilled into end to end visuomotor policies (ACT, Diffusion Policy, GR00T N1.6).

VLM task planning (Qwen3 VL) over PPO locomotion and arm policies on the Unitree G1: walk, reach, grasp, drawer, pick and place.

Multi stage PPO curriculum for whole body G1 locomotion in Isaac Lab: velocity tracking, terrain, torso, arm coordination.

Also: BARN local planner benchmark (under review).


Open to research collaborations in humanoid robotics and robot learning.

Pinned Loading

  1. smolvla_flow_rl smolvla_flow_rl Public

    Online reinforcement learning fine tuning for a flow matching vision language action policy, on a single 12GB consumer GPU.

    Python 1

  2. isaac-g1-ulc isaac-g1-ulc Public

    Low Level RL Controller for G1

    Python 16 1

  3. isaac-g1-hierarchical isaac-g1-hierarchical Public

    VLM-RL Hierarchical Loco-Manupilation For Long-Horizon Tasks With G1 robot in Isaac Lab/Sim

    Python 13 1

  4. isaaclab-anymal-locomotion isaaclab-anymal-locomotion Public

    A legged locomotion project

    Python 3

  5. mujoco-ant-ppo mujoco-ant-ppo Public

    Training a MuJoCo Ant agent to walk using PPO from scratch.

    Python 5

  6. benchmark-local-path-planners-barn-challenge benchmark-local-path-planners-barn-challenge Public

    A Framework for BARN of classical and learning-based local path planners in BARN Challenge navigation benchmark.

    HTML 1