Skip to content
View yuchenwang3's full-sized avatar
:octocat:
Focusing
:octocat:
Focusing
  • Sunnyvale/Mountain View, CA
  • 10:35 (UTC -07:00)

Block or report yuchenwang3

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
yuchenwang3/README.md
Yuchen (Ean) Wang — agentic post-training and ML systems

M.S. CS @ UIUC · Research intern @ Alibaba Accio · PKU Zhi Class.

Website CV Scholar LinkedIn Email

Research

Occamy-1.0: official logo

35B-A3B agent model for long-horizon tool use.
Execution-grounded data · Agentic post-training

Report Demo Model Code

CineFlow: figure from the paper

Dependency-driven parallel video generation.
1.7–5.5× end-to-end speedup in the reported evaluation

Project Paper

Dynamic Prefill: figure from the project report

Adaptive batching and prompt packing for LLM serving.
Up to 20% lower TTFT on the reported traces

Report Code

CUDA Attention RL for Legal Reasoning

Open-source contributions

9 merged · 1 adopted solution · 15 open · 12 projects

modelscope organization avatar
ms-swift

#9598 · merged Add order-preserving packing

#9602 · merged Warm up NCCL before training

#9599 · merged Pass through Muon Nesterov settings

#9591 · merged Expose Muon coefficient selection

flashinfer-ai organization avatar
FlashInfer

#4984 · merged Calibrate FP8 attention

vllm-project organization avatar
vime

#337 · merged Forward recompute flags; fix hybrid models

NVIDIA-NeMo organization avatar
Emerging Optimizers

#230 · merged Keep Muon scale-invariant

NVIDIA-NeMo organization avatar
NeMo Gym

#2726 · merged Preserve HTTP errors across process boundaries
Co-author

Dao-AILab organization avatar
FlashAttention

#2507 · merged Stabilize backward JIT keys without CPU–GPU sync
Solution adopted by the PR author

vllm-project organization avatar
vLLM

#54699 · merged Convert MoE weights in place

NVIDIA organization avatar
Megatron-LM

#5396 · open Fuse GDN Q/K normalization

#5463 · open Recompute Mamba selectively

#5400 · open Route GDN input projections to Adam

#5431 · open Exclude GDN input projections from global clipping

#5395 · open Skip gradient clipping for Muon

NVIDIA-NeMo organization avatar
NeMo RL

#3943 · open Move references, not teacher payloads

#2962 · open Sanitize non-finite async log probabilities

sgl-project organization avatar
SGLang

#39765 · open Fix Mamba cache publication under overlap scheduling

#38063 · open Explain cold MXFP4 JIT startup

#31621 · open Honor weight-check exclusions during reset

verl-project organization avatar
verl

#7597 · open Validate actor FSDP strategy

NousResearch organization avatar
Hermes Agent

#113511 · open Control partial-stream continuation for batch evaluation

#113538 · open Clarify API retry budgets and streaming defaults

#100693 · open Resolve nested tool schemas

#102549 · open Make SSH reconnects race-safe

All contributions ↗ Engineering notes ↗

Pinned Loading

  1. vllm-project/vllm vllm-project/vllm Public

    A high-throughput and memory-efficient inference and serving engine for LLMs

    Python 92.1k 22.4k

  2. sgl-project/sglang sgl-project/sglang Public

    SGLang is a high-performance serving framework for large language models and multimodal models.

    Python 36.1k 9k

  3. NVIDIA/Megatron-LM NVIDIA/Megatron-LM Public

    Ongoing research training transformer models at scale

    Python 17.9k 4.5k

  4. modelscope/ms-swift modelscope/ms-swift Public

    Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, …

    Python 15.7k 1.7k

  5. NousResearch/hermes-agent NousResearch/hermes-agent Public

    The agent that grows with you

    Python 247k 51.8k

  6. verl-project/verl verl-project/verl Public

    verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

    Python 23.5k 4.6k