Skip to content
@FudanCVL

FudanCVL

Fudan Computer Vision Laboratory. A research group at Fudan University focusing on Computer Vision and Generative AI.

FudanCVL

Welcome to the official GitHub organization of FudanCVL, the Computer Vision Lab at Fudan University. 👋

Led by Prof. Henghui Ding, our research focuses on computer vision, multimodal learning, generative AI, world models, and embodied intelligence. We develop algorithms, benchmarks, and open-source resources for understanding and generating visual intelligence across images, videos, 3D scenes, and multimodal environments.

Pinned Loading

  1. EffectErase EffectErase Public

    [CVPR 2026 Highlight] EffectErase: Joint Video Object Removal and Insertion for High-Quality Effect Erasing

    Python 160 8

  2. GlyphPrinter GlyphPrinter Public

    [CVPR 2026 Highlight] GlyphPrinter: Region-Grouped Direct Preference Optimization for Glyph-Accurate Visual Text Rendering

    Python 104 11

  3. PSDesigner PSDesigner Public

    [CVPR 2026] PSDesigner: Automated Graphic Design with a Human-Like Creative Workflow

    156 8

  4. AnyI2V AnyI2V Public

    [ICCV 2025] AnyI2V: Animating Any Conditional Image with Motion Control Generation

    Python 123 6

  5. OmniAVS OmniAVS Public

    [ICCV 2025] Towards Omnimodal Expressions and Reasoning in Referring Audio-Visual Segmentation

    Python 91 2

  6. SAM2Matting SAM2Matting Public

    [ECCV 2026] SAM2Matting: Generalized Image and Video Matting

    Python 130 9

Repositories

Showing 10 of 29 repositories
  • .github Public
    FudanCVL/.github's past year of commit activity
    0 0 0 0 Updated Jul 24, 2026
  • SAM-MT Public

    [ECCV 2026] Real-Time Interactive Multi-Target Video Segmentation

    FudanCVL/SAM-MT's past year of commit activity
    Python 66 4 1 0 Updated Jul 10, 2026
  • APRS Public

    [ECCV 2026] Seek to Segment: Active Perception for Panoramic Referring Segmentation

    FudanCVL/APRS's past year of commit activity
    Python 5 MIT 1 0 0 Updated Jul 9, 2026
  • Unison Public

    [ICML 2026] Unison: Benchmarking Unified Multimodal Models via Synergistic Understanding and Generation

    FudanCVL/Unison's past year of commit activity
    Python 14 MIT 1 0 0 Updated Jun 30, 2026
  • SAM2Matting Public

    [ECCV 2026] SAM2Matting: Generalized Image and Video Matting

    FudanCVL/SAM2Matting's past year of commit activity
    Python 130 9 0 0 Updated Jun 30, 2026
  • FeVOS Public

    [ECCV2026] FeVOS: Foresight Expression Video Object Segmentation

    FudanCVL/FeVOS's past year of commit activity
    Python 11 0 0 0 Updated Jun 25, 2026
  • AVTrack Public

    [ICML 2026] AVTrack: Audio-Visual Tracking in Human-centric Complex Scenes

    FudanCVL/AVTrack's past year of commit activity
    Python 26 MIT 0 0 0 Updated Jun 20, 2026
  • AVI-Bench Public

    [ICML'26] Toward Human-like Audio-Visual Intelligence of Omni-MLLMs

    FudanCVL/AVI-Bench's past year of commit activity
    Python 18 0 0 0 Updated Jun 20, 2026
  • OcclusionFormer Public

    [ICML 2026] OcclusionFormer: Arranging Z-Order for Layout-Grounded Image Generation

    FudanCVL/OcclusionFormer's past year of commit activity
    Python 22 0 0 0 Updated May 21, 2026
  • ROSE Public

    [CVPR 2026 Findings] ROSE: Retrieval-Oriented Segmentation Enhancement

    FudanCVL/ROSE's past year of commit activity
    8 0 1 0 Updated Apr 16, 2026

Top languages

Loading…

Most used topics

Loading…