Skip to content

Navigation Menu

Sign in
Sign up
@wanglezz
wanglezz
Follow

Fang Chengjie wanglezz

  • University of Science and Technology of China
  • Hefei
  • 18:47 (UTC +08:00)

Highlights

  • Pro

Block or report wanglezz

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
wanglezz /README.md

Hi there 👋,I'm Fang Chengjie.

✨ About Me

🎓 I'm a graduate student majoring in Software Engineering at the University of Science and Technology of China.

  • 🔬 I work on LLM post-training — SFT, GRPO/PPO, and RL for multi-turn tool-using agents.
  • 🧪 Currently building an agentic RL pipeline on τ2-bench: failure attribution → trajectory distillation → SFT → GRPO, with ablations on reward shaping and KL anchoring.
  • ⚙️ Also interested in the systems side: vLLM rollout, FSDP training, weight sync, and memory optimization.
  • 💬 Ask me about GRPO/PPO internals, agent evaluation & failure analysis, Python, C++, algorithms.
  • 📫 Reach me at fangchengjie@outlook.com.

Pinned Loading

  1. tau2-rl tau2-rl Public

    Python

  2. wanglezz wanglezz Public

    1

  3. remote-sensing remote-sensing Public

    Python

  4. Paddle Paddle Public

    Forked from PaddlePaddle/Paddle

    PArallel Distributed Deep LEarning: Machine Learning Framework from Industrial Practice (『飞桨』核心框架,深度学习&机器学习高性能单机、分布式训练和跨平台部署)

    C++ 1

AltStyle によって変換されたページ (->オリジナル) /