Skip to content
View wanghao9610's full-sized avatar

Highlights

  • Pro

Organizations

@omdsh-plugins

Block or report wanghao9610

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
wanghao9610/README.md

Hi there πŸ‘‹

I'm Hao Wang (ηŽ‹θ±ͺ), a Ph.D. candidate at HCP Lab, Sun Yat-sen University, and Pengcheng Laboratory, advised by Prof. Xiaodan Liang and Assoc. Prof. Xiangyuan Lan.

πŸ”¬ My research centers on open-ended computer vision and multimodal large language models, and I'm increasingly exploring multimodal agentic models.

πŸŽ“ I'll graduate in December 2026 and am actively looking for research roles in industry β€” I'm also open to research collaborations on interesting projects.

πŸ“§ Reach me via Email: wanghao9610@gmail.com or WeChat: wangh9610.

Research Projects

  • πŸ”₯ Any segmentation in images and videos: X2SAM
  • πŸ”₯ From segment anything to any segmentation: X-SAM
  • Unified open-vocabulary detection: OV-DINO
  • Temporal memory attention for video semantic segmentation: TMANet

Hobby Projects

  • πŸ”₯ Systematic Toolchain for Organizing Research over Years (Harness, WIP): STORY
  • πŸ”₯ Systematic Toolchain for Authoring, Guiding, and Editing (Harness, WIP): STAGE
  • πŸ”₯ Systematic Toolchain for AI Research (Harness, WIP): STAR
  • πŸ”₯ Oh My DeepSeek Harness Plugins: omdsh-plugins
  • Template paper for arXiv or any conference: arXivTeX
  • Run Codex on a remote server: Codex Remote Connector
  • Run Claude on multiple AI providers: Claude Model Proxy

Pinned Loading

  1. OV-DINO OV-DINO Public

    OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

    Python 414 34

  2. X-SAM X-SAM Public

    [AAAI2026] X-SAM: From Segment Anything to Any Segmentation

    Python 392 17

  3. X2SAM X2SAM Public

    [ECCV2026] X2SAM: Any Segmentation in Images and Videos

    Python 104 5

  4. STAR STAR Public

    STAR: Systematic Toolchain for AI Research (Harness, WIP)

    Shell 52 1

  5. arXivTeX arXivTeX Public

    Template the paper for arXiv or any conference

    TeX 33

  6. Claude-Model-Proxy Claude-Model-Proxy Public

    Run Claude on multiple AI providers

    JavaScript 16 1