Skip to content
View WenzheWang's full-sized avatar
🎯
Focusing
🎯
Focusing

Block or report WenzheWang

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
WenzheWang/README.md

Hi there, I'm Wenzhe Wang! 👋

I'm a PhD graduate from Zhejiang University, working at the intersection of AI infrastructure and computer vision. I contribute to large-scale training frameworks (VeOmni, verl) while publishing research at top-tier venues (NeurIPS, ICCV, IJCAI, ECAI, IEEE TMI).

🔬 Research

My research spans distributed training systems, vision-language models, and computational imaging.

Selected Publications:

📖 Full publication list: Google Scholar

💻 Open Source

I'm a contributor to open-source AI training frameworks, including VeOmni (model-centric distributed training recipe zoo) and verl / verl-omni (multimodal RL post-training frameworks).

Selected Contributions:

Research directions:

  • Parallelism & memory optimization — packed sequence packing, ChunkMBS, FSDP2/hsdp integration
  • Training-time validation — distributed metric aggregation and evaluation pipeline
  • Hardware portability — NPU training enablement and cross-platform validation
  • Multimodal reward engineering — batched reward inference and reward model integration
  • Rollout & training pipeline — async semantics, rollout-train consistency, group management
  • Efficient post-training — LoRA/FSDP for VLMs, distributed reward handling

🛠️ Tech Stack

Languages      Python · C++ · CUDA · Shell
Frameworks     PyTorch · VeOmni · verl · FSDP/FSDP2 · DeepSpeed
Distributed    DDP · HSDP · FSDP · Tensor Parallel · Expert Parallel · ChunkMBS
Infrastructure Slurm · Ray · wandb · Docker
Vision/RL       Diffusion Models · VLMs · Video Retrieval · Medical Imaging

📫 Connect


💫 "From research papers to production training frameworks — bridging the gap between algorithms and infrastructure."

Pinned Loading

  1. verl-project/verl verl-project/verl Public

    verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

    Python 23.3k 4.5k

  2. ByteDance-Seed/VeOmni ByteDance-Seed/VeOmni Public

    VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zoo

    Python 2.2k 269

  3. verl-project/verl-omni verl-project/verl-omni Public

    Multimodal RL training framework for diffusion & omni models

    Python 957 180