default search action
David Fan 0001
Person information
- affiliation: Meta Fundamental AI Research (FAIR), New York, NY, USA
- affiliation (former): Amazon Prime Video, Seattle, WA, USA
- affiliation (former): Princeton University, Vision and Learning Lab, Princeton, NJ, USA
Other persons with the same name
- David Fan — disambiguation page
- David Fan 0002 — University of Dayton, Research Institute (UDRI), Dayton, OH, USA
Refine list
refinements active!
zoomed in on ?? of ?? records
view refined list in
2020 – today
- 2026
- [i16]Basile Terver, Randall Balestriero, Megi Dervishi, David Fan, Quentin Garrido, Tushar Nagarajan, Koustuv Sinha, Wancong Zhang, Mike Rabbat, Yann LeCun, Amir Bar:
A Lightweight Library for Energy-Based Joint-Embedding Predictive Architectures. CoRR abs/2602.03604 (2026) - [i15]Shengbang Tong, David Fan, John Nguyen, Ellis Brown, Gaoyue Zhou, Shengyi Qian, Boyang Zheng, Théophane Vallaeys, Junlin Han, Rob Fergus, Naila Murray, Marjan Ghazvininejad, Mike Lewis, Nicolas Ballas, Amir Bar, Michael Rabbat, Jakob Verbeek, Luke Zettlemoyer, Koustuv Sinha, Yann LeCun, Saining Xie:
Beyond Language Modeling: An Exploration of Multimodal Pretraining. CoRR abs/2603.03276 (2026) - 2025
- [c10]David Fan, Shengbang Tong, Jiachen Zhu, Koustuv Sinha, Zhuang Liu, Xinlei Chen, Michael Rabbat, Nicolas Ballas, Yann LeCun, Amir Bar, Saining Xie:
Scaling Language-Free Visual Representation Learning. ICCV 2025: 1-13 - [c9]Shengbang Tong, David Fan, Jiachen Zhu, Yunyang Xiong, Xinlei Chen, Koustuv Sinha, Michael Rabbat, Yann LeCun, Saining Xie, Zhuang Liu:
MetaMorph: Multimodal Understanding and Generation via Instruction Tuning. ICCV 2025: 17001-17012 - [c8]Yicheng Wang, Zhikang Zhang, Jue Wang, David Fan, Zhenlin Xu, Linda Liu, Xiang Hao, Vimal Bhat, Xinyu Li:
GEXIA: Granularity Expansion and Iterative Approximation for Scalable Multi-Grained Video-Language Learning. WACV 2025: 4725-4735 - [c7]Seon-Ho Lee, Jue Wang, David Fan, Zhikang Zhang, Linda Liu, Xiang Hao, Vimal Bhat, Xinyu Li:
Now you see Me: Context-Aware Automatic Audio Description. WACV 2025: 5530-5539 - [i14]David Fan, Shengbang Tong, Jiachen Zhu, Koustuv Sinha, Zhuang Liu, Xinlei Chen, Michael Rabbat, Nicolas Ballas, Yann LeCun, Amir Bar, Saining Xie:
Scaling Language-Free Visual Representation Learning. CoRR abs/2504.01017 (2025) - [i13]Mido Assran, Adrien Bardes, David Fan, Quentin Garrido, Russell Howes, Mojtaba Komeili, Matthew J. Muckley, Ammar Rizvi, Claire Roberts, Koustuv Sinha, Artem Zholus, Sergio Arnaud, Abha Gejji, Ada Martin, Francois Robert Hogan, Daniel Dugas, Piotr Bojanowski, Vasil Khalidov, Patrick Labatut, Francisco Massa, Marc Szafraniec, Kapil Krishnakumar, Yong Li, Xiaodong Ma, Sarath Chandar, Franziska Meier, Yann LeCun, Michael Rabbat, Nicolas Ballas:
V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning. CoRR abs/2506.09985 (2025) - [i12]Junlin Han, Shengbang Tong, David Fan, Yufan Ren, Koustuv Sinha, Philip Torr, Filippos Kokkinos:
Learning to See Before Seeing: Demystifying LLM Visual Priors from Language Pre-training. CoRR abs/2509.26625 (2025) - [i11]Raktim Gautam Goswami, Amir Bar, David Fan, Tsung-Yen Yang, Gaoyue Zhou, Prashanth Krishnamurthy, Michael Rabbat, Farshad Khorrami, Yann LeCun:
World Models Can Leverage Human Videos for Dexterous Manipulation. CoRR abs/2512.13644 (2025) - 2024
- [c6]David Fan
, Jue Wang
, Shuai Liao, Zhikang Zhang
, Vimal Bhat, Xinyu Li
:
Text-Guided Video Masked Autoencoder. ECCV (5) 2024: 282-298 - [c5]Seon-Ho Lee, Jue Wang, Zhikang Zhang, David Fan, Xinyu Li:
Video Token Merging for Long Video Understanding. NeurIPS 2024 - [i10]David Fan, Jue Wang, Shuai Liao, Zhikang Zhang, Vimal Bhat, Xinyu Li:
Text-Guided Video Masked Autoencoder. CoRR abs/2408.00759 (2024) - [i9]Seon-Ho Lee, Jue Wang, Zhikang Zhang, David Fan, Xinyu Li:
Video Token Merging for Long-form Video Understanding. CoRR abs/2410.23782 (2024) - [i8]Yicheng Wang, Zhikang Zhang, Jue Wang, David Fan, Zhenlin Xu, Linda Liu, Xiang Hao, Vimal Bhat, Xinyu Li:
GEXIA: Granularity Expansion and Iterative Approximation for Scalable Multi-grained Video-language Learning. CoRR abs/2412.07704 (2024) - [i7]Seon-Ho Lee, Jue Wang, David Fan, Zhikang Zhang, Linda Liu, Xiang Hao, Vimal Bhat, Xinyu Li:
NowYouSee Me: Context-Aware Automatic Audio Description. CoRR abs/2412.10002 (2024) - [i6]Shengbang Tong, David Fan, Jiachen Zhu, Yunyang Xiong, Xinlei Chen, Koustuv Sinha, Michael Rabbat, Yann LeCun, Saining Xie, Zhuang Liu:
MetaMorph: Multimodal Understanding and Generation via Instruction Tuning. CoRR abs/2412.14164 (2024) - 2023
- [c4]David Fan
, Jue Wang, Shuai Liao, Yi Zhu, Vimal Bhat, Hector J. Santos-Villalobos, Rohith MV, Xinyu Li:
Motion-Guided Masking for Spatiotemporal Representation Learning. ICCV 2023: 5596-5606 - [c3]Najmeh Sadoughi, Xinyu Li, Avijit Vajpayee, David Fan, Bing Shuai, Hector J. Santos-Villalobos, Vimal Bhat, Rohith MV:
MEGA: Multimodal Alignment Aggregation and Distillation For Cinematic Video Segmentation. ICCV 2023: 23274-23283 - [i5]David Fan, Deyu Yang, Xinyu Li, Vimal Bhat, Rohith MV:
Nearest-Neighbor Inter-Intra Contrastive Learning from Unlabeled Videos. CoRR abs/2303.07317 (2023) - [i4]Najmeh Sadoughi, Xinyu Li, Avijit Vajpayee, David Fan, Bing Shuai, Hector J. Santos-Villalobos, Vimal Bhat, Rohith MV:
MEGA: Multimodal Alignment Aggregation and Distillation For Cinematic Video Segmentation. CoRR abs/2308.11185 (2023) - [i3]David Fan, Jue Wang, Shuai Liao, Yi Zhu, Vimal Bhat, Hector J. Santos-Villalobos, Rohith MV, Xinyu Li:
Motion-Guided Masking for Spatiotemporal Representation Learning. CoRR abs/2308.12962 (2023) - 2021
- [c2]Shixing Chen, Xiaohan Nie, David Fan
, Dongqing Zhang, Vimal Bhat, Raffay Hamid:
Shot Contrastive Self-Supervised Learning for Scene Boundary Detection. CVPR 2021: 9796-9805 - [i2]Shixing Chen, Xiaohan Nie, David Fan, Dongqing Zhang, Vimal Bhat, Raffay Hamid:
Shot Contrastive Self-Supervised Learning for Scene Boundary Detection. CoRR abs/2104.13537 (2021) - 2020
- [c1]Weifeng Chen, Shengyi Qian, David Fan
, Noriyuki Kojima, Max Hamilton
, Jia Deng:
OASIS: A Large-Scale Dataset for Single Image 3D in the Wild. CVPR 2020: 676-685 - [i1]Weifeng Chen, Shengyi Qian, David Fan
, Noriyuki Kojima, Max Hamilton, Jia Deng:
OASIS: A Large-Scale Dataset for Single Image 3D in the Wild. CoRR abs/2007.13215 (2020)
Coauthor Index
manage site settings
To protect your privacy, all features that rely on external API calls from your browser are turned off by default. You need to opt-in for them to become active. All settings here will be stored as cookies with your web browser. For more information see our F.A.Q.
Unpaywalled article links
Add open access links from to the list of external document links (if available).
Privacy notice: By enabling the option above, your browser will contact the API of unpaywall.org to load hyperlinks to open access articles. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the Unpaywall privacy policy.
Archived links via Wayback Machine
For web page which are no longer available, try to retrieve content from the of the Internet Archive (if available).
Privacy notice: By enabling the option above, your browser will contact the API of archive.org to check for archived content of web pages that are no longer available. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the Internet Archive privacy policy.
Reference lists
Add a list of references from ,
, and
to record detail pages.
load references from crossref.org and opencitations.net
Privacy notice: By enabling the option above, your browser will contact the APIs of crossref.org, opencitations.net, and semanticscholar.org to load article reference information. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the Crossref privacy policy and the OpenCitations privacy policy, as well as the AI2 Privacy Policy covering Semantic Scholar.
Citation data
Add a list of citing articles from and
to record detail pages.
load citations from opencitations.net
Privacy notice: By enabling the option above, your browser will contact the API of opencitations.net and semanticscholar.org to load citation information. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the OpenCitations privacy policy as well as the AI2 Privacy Policy covering Semantic Scholar.
OpenAlex data
Load additional information about publications from .
Privacy notice: By enabling the option above, your browser will contact the API of openalex.org to load additional information. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the information given by OpenAlex.
last updated on 2026-08-09 23:32 CEST by the dblp team
all metadata released as open data under CC0 1.0 license
see also: Terms of Use | Privacy Policy | Imprint