Skip to content
View yulinlina's full-sized avatar
🎯
Focusing
🎯
Focusing

Highlights

  • Pro

Block or report yulinlina

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
yulinlina/README.md
Typing


👋 About Me

I'm Wang Haolin, an AI major at Sichuan University (SCU) in Chengdu. I build LLM agents and the harnesses that drive them, and push embodied AI from simulation onto real legged robots and robot arms — with a side track in RL for high-frequency markets.

  • 🎓 Artificial Intelligence @ Sichuan University · Chengdu, China
  • 🤖 LLM agents & harnesses — multi-agent pipelines, MCP tooling, skill systems, context engineering
  • 🦿 Embodied AI / VLA — RL locomotion (PPO, imitation, AMP), π0.5 VLA deployment on edge devices
  • 📈 RL for markets — market making & microstructure-based direction prediction at tick level
  • 🧠 Methods I use daily: PPO · imitation learning · domain randomization · teacher-student distillation · sim2real
  • 📫 Reach me: whl@scu.edu.cn

🎯 What I'm Working On

GitHub Trophies

⚡ Tech Arsenal

Core Languages

LLM Agents & Harnesses

Robotics, VLA & RL

Infrastructure


🚀 Featured Projects

🤖 Agents & Harnesses

start_rise

Autonomous multi-agent harness

LLM-driven rounds: trend analysis → spec evaluation → code generation with test-driven self-healing → deployment → engagement, orchestrated end-to-end by a scheduled pipeline.

Stars Last commit

LLM Agents Orchestration Self-Healing

⭐ Star | 🔗 Visit

📉 mcpx

MCP Context eXterminator

Transparent proxy that compresses MCP tool definitions & responses — saves up to 80% of context-window tokens for AI coding agents.

Stars Last commit

MCP Context Engineering Proxy

⭐ Star | 🔗 Visit

🏭 clawforge

Claude Code Skill Factory & Marketplace

One command to create, test, publish and install agent skills, hooks and sub-agent configs — the npm for AI agent capabilities.

Stars Last commit

Claude Code Skills Marketplace

⭐ Star | 🔗 Visit

🎮 simforge

Plain English → robot simulation

Describe a robot in plain English, get a simulation-ready MuJoCo scene with physics, assets and RL training scripts in seconds.

Stars Last commit

MuJoCo Codegen Robotics

⭐ Star | 🔗 Visit

🦿 Locomotion & VLA

  • π0.5 VLA real-robot deployment — LoRA-finetuned π0.5 (PaliGemma 2B + action expert) running closed-loop on NVIDIA Jetson Thor × AgileX Piper arm: natural-language pick-and-place with 10-step action chunks, full perception → inference → control loop
  • DreamWaQ for wheel-legged robots — adapted DreamWaQ (PPO + β-VAE implicit terrain imagination) to 4×4 wheel-legged platforms: up to 1 m step climbing, Isaac Gym → ONNX → MuJoCo sim2sim → C++ deployment on real robots
  • Backflip on a 60 kg quadruped — DeepMimic-style imitation on ZJ-T30-V2: trajectory-optimization reference + PPO tracking, phase-driven control, motor torque-speed envelope
  • AMP & CAMP multi-gait control — adversarial motion priors for natural trot, rough-terrain curriculum and fall-recovery; skill-conditioned CAMP learns walk / trot / pronk / bound in a single policy
  • Extreme Parkour — end-to-end depth-vision parkour on Unitree Go2: teacher-student privileged RL, depth-latent + heading distillation, ROS 2 deployment
  • HIMLoco for Go2W — human-imitation locomotion (ICLR 2024) re-implemented on the wheel-legged Go2W: Isaac Gym training + MuJoCo sim2sim

🧠 AI Fundamentals

  • AI by Hand (Excel) — deep learning rebuilt from scratch in spreadsheets: backprop, RNN/LSTM/xLSTM, ResNet, full-stack Transformer, Mamba, all the way to AlphaFold — the best way to really understand the math

💹 HFT & Market Microstructure

  • RLMM — reinforcement-learning market making on Binance ETH-USDT perps: tick-level hftbacktest Gym environment, RecurrentPPO + LSTM with a 4-phase curriculum over ~900 days of L2 orderbook data; plus a 66-dim microstructure direction model (84.5% validation accuracy, profitable on 93/100 out-of-sample days)

🧪 Currently Exploring

  • 🤖 Agent harness engineering — long-running loops, context budgeting, skill ecosystems
  • 🧠 VLA on the edge — shrinking π0-class models onto Jetson-class devices with usable control rates
  • 🦿 Wheel-legged sim-to-real — taking Go2W RL policies from Isaac Gym onto the real robot
  • 🏞️ Agile skills — imitation & contrastive learning for parkour-class maneuvers

📊 GitHub Analytics


📈 Contribution Heatmap


📬 Let's Connect


⚡ "知行合一 — knowledge and action as one." ⚡

Pinned Loading

  1. Cuda-Note Cuda-Note Public

    the code about cuda of GPU course

    C 2

  2. Mechine-Learing-Note Mechine-Learing-Note Public

    the code about mechine learning

    Jupyter Notebook 1

  3. Algorithm-Design Algorithm-Design Public

    The code about the basic algorithem

    C++ 1

  4. Mytorch Mytorch Public

    devise a network frame

    Jupyter Notebook 3

  5. yulinlina.github.io yulinlina.github.io Public

    Github Pages template for academic personal websites, forked from mmistakes/minimal-mistakes

    JavaScript 1

  6. funNLP funNLP Public

    Forked from fighting41love/funNLP

    NLP集大成

    Python 1