Skip to content
View ZhuWenjie98's full-sized avatar

Highlights

  • Pro

Block or report ZhuWenjie98

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
ZhuWenjie98/README.md

Wenjie Zhu

Trustworthy Multimodal Intelligence · Vision-Language Models · Open-World Learning

Website Google Scholar GitHub Email

About Me

I am a fourth-year Ph.D. candidate in the Visual Computing Lab at The Hong Kong Polytechnic University, advised by Prof. Lei Zhang and Prof. Wenjun Zeng. Before my Ph.D., I received my M.S. from New York University and my B.S. from Central South University.

My research focuses on building reliable multimodal systems that can recognize uncertainty, adapt under distribution shifts, and reason safely in open-world environments.

  • Reliable Agent: Self-Evolving Agents, Recursive Self-Improvement Agents
  • Trustworthy Perception: OOD detection and test-time adaptation with VLMs and MLLMs
  • MLLM Evaluation: out-of-context understanding, robustness, and trustworthy evaluation
  • Medical AI: Medical MLLM, Medical Agent

Research Highlights

Project Focus Venue Community
ANTS Test-time MLLM understanding and adaptive negative textual spaces for OOD detection CVPR 2026 Oral Stars
DDE Training-free dual distribution estimation for zero-shot noisy test-time adaptation ECCV 2026 Stars
KRNFT Knowledge-regularized negative feature tuning for VLM-based OOD detection ACM MM 2025 Oral Stars
MMOOC A benchmark for out-of-context evaluation in multimodal large language models Benchmark Stars
OpenOOD-VLM Unified benchmarking for generalized OOD detection with vision-language models Open Source Stars

Selected Publications

  • Dual Distribution Estimation for Zero-shot Noisy Test-Time Adaptation with VLMs
    ECCV 2026 · Project · Paper · Code

  • ANTS: Adaptive Negative Textual Space Shaping for OOD Detection via Test-Time MLLM Understanding and Reasoning
    CVPR 2026 Oral · Project · Paper · Code

  • Knowledge Regularized Negative Feature Tuning of Vision-Language Models
    ACM MM 2025 Oral · Paper · Code

  • MMOOC: A Comprehensive Benchmark for Out-of-Context Evaluation in Multimodal Large Language Models
    Paper · Code

  • LAPT: Label-driven Automated Prompt Tuning for OOD Detection with Vision-Language Models
    ECCV 2024 · Paper

GitHub at a Glance

Wenjie Zhu's GitHub contribution summary
GitHub statistics Repositories by language

GitHub language statistics reflect public repository composition, not overall proficiency.

Tools & Methods

Python PyTorch Hugging Face OpenCV LaTeX Linux Git

Let's Connect

I am interested in research collaborations around trustworthy vision-language models, open-world recognition, and test-time adaptation. For publications, code, and current projects, visit my research homepage or reach me by email.

Researching reliable multimodal intelligence for the open world.

Pinned Loading

  1. ANTS ANTS Public

    (CVPR2026 Oral) ANTS: Adaptive Negative Textual Space Shaping for OOD Detection via Test-Time MLLM Understanding and Reasoning

    Python 94 7

  2. DDE DDE Public

    (ECCV2026) Dual Distribution Estimation for Zero-shot Noisy Test-Time Adaptation with VLMs

    17

  3. PolyU-VCLab/OpenOOD-VLM PolyU-VCLab/OpenOOD-VLM Public

    ECCV24, NeurIPS24, CVPR26*2, ECCV26, Benchmarking Generalized Out-of-Distribution Detection with Vision-Language Models

    Python 59 10

  4. MMOOC MMOOC Public

    MMOOC: A Comprehensive Benchmark for Out-of-Context Evaluation in Multimodal Large Language Models

    Python 31