2026 Safe, or Simply Incapable? Rethinking Safety Evaluation for Phone-Use Agents Zhengyang Tang, Yi Zhang, Chenxin Li, and 18 more authors 2026 arXiv Code PhoneWorld: Scaling Phone-Use Agent Environments Zhengyang Tang, Yuxuan Liu, Xin Lai, and 21 more authors 2026 arXiv PhoneHarness: Harnessing Phone-Use Agents through Mixed GUI, CLI, and Tool Actions Chenxin Li, Zhengyao Fang, Zhengyang Tang, and 18 more authors 2026 arXiv Code Website PhoneBuddy: Training Open Models for Agentic Phone Use Zhengyang Tang, Xin Lai, Pengyuan Lyu, and 23 more authors 2026 arXiv Code Website GUICrafter: Weakly-Supervised GUI Agent Leveraging Massive Unannotated Screenshots Sunqi Fan, Lingshan Chen, Runqi Yin, and 4 more authors 2026 arXiv Code Bridging VideoQA and Video-Guided Agentic Tasks via Generalized Keyframe Extraction Sunqi Fan, Qingle Liu, Runqi Yin, and 2 more authors In ECCV, 2026 arXiv Code Website HyMobileAgent: Data-Environment Co-Scaling for Efficient GUI Agents Hy Vision Team, Huawen Shen, Zhengyang Tang, and 20 more authors 2026 arXiv 2025 Tool-Augmented Spatiotemporal Reasoning for Streamlining Video Question Answering Task Sunqi Fan, Jiashuo Cui, Meng-Hao Guo, and 1 more author In NeurIPS, 2025 arXiv Code Agentic Keyframe Search for Video Question Answering Sunqi Fan, Meng-Hao Guo, and Shuojin Yang 2025 arXiv Code 2024 oral FlexKBQA: a flexible LLM-powered framework for few-shot knowledge base question answering Zhenyu Li*, Sunqi Fan*, Yu Gu, and 5 more authors In AAAI, 2024 arXiv Code Optimization Techniques for Unsupervised Complex Table Reasoning via Self-Training Framework Zhenyu Li, Xiuxing Li, Sunqi Fan, and 1 more author IEEE Transactions on Knowledge and Data Engineering, 2024 arXiv Code 2023 FAAC: Facial Animation Generation with Anchor Frame and Conditional Control for Superior Fidelity and Editability Linze Li, Sunqi Fan, Hengjun Pu, and 6 more authors 2023 arXiv Code Website A Survey of Video Generation with Diffusion Models Linze Li, and Sunqi Fan 2023 Code Website