Sunqi Fan

self-image.png

I am a first year CS PhD student at Tsinghua University, advised by Prof. Shi-Min Hu. My research interest lies in computer vision, with a focus on designing intelligent visual agents to understand the dynamic real world.

I obtained my Bachelor’s degree in Computer Science from Tsinghua University. You may find my CV here: Sunqi’s Curriculum Vitae. Feel free to reach out to me via email if you have any questions or just want to chat!

Email: stephensunqifan@gmail.com

Google Scholar / Github / Twitter / LinkedIn

news

Jun 30, 2026 We release GUICrafter, a weakly-supervised and self-evolving GUI agent.
Jun 21, 2026 One paper on video understanding is accpeted by ECCV’26! Check out this project page.
Jun 20, 2026 A wonderful trip to Inner Mongolia. [Photo 1] [Photo 2]
Dec 06, 2025 Attend NeurIPS 2025 in San Diego. [Photo 1] [Photo 2] [Photo 3]
Oct 22, 2025 Attend CNCC 2025 in Harbin.
Sep 18, 2025 Our paper on tool-augmented VideoQA is accpted by NeurIPS 2025!
Sep 01, 2025 Begin my journey to pursue a CS PhD at Tsinghua University.
Jun 06, 2025 Attend VALSE 2025 in Zhuhai.
Dec 01, 2023 FlexKBQA is accepted by AAAI 2024 as oral presentation.

selected publications

  1. GUICrafter.png
    GUICrafter: Weakly-Supervised GUI Agent Leveraging Massive Unannotated Screenshots
    Sunqi Fan, Lingshan Chen, Runqi Yin, and 4 more authors
    2026
  2. vg-gui-TASKER.png
    Bridging VideoQA and Video-Guided Agentic Tasks via Generalized Keyframe Extraction
    Sunqi Fan, Qingle Liu, Runqi Yin, and 2 more authors
    In ECCV, 2026
  3. VideoTool.png
    Tool-Augmented Spatiotemporal Reasoning for Streamlining Video Question Answering Task
    Sunqi Fan, Jiashuo Cui, Meng-Hao Guo, and 1 more author
    In NeurIPS, 2025
  4. oral
    FlexKBQA.png
    FlexKBQA: a flexible LLM-powered framework for few-shot knowledge base question answering
    Zhenyu Li*Sunqi Fan*, Yu Gu, and 5 more authors
    In AAAI, 2024