awesome-egocentric-vision
A curated list of egocentric (first-person) vision and related area resources
Last updated Jul 30, 2026
336
Stars
34
Forks
3
Issues
0
Stars/day
Attention Score
84
Topics
Language breakdown
No language data available.
▸ Files
click to expand
README
Awesome Egocentric Vision 
A curated list of egocentric vision resources.
Egocentric (first-person) vision is a sub-field of computer vision that analyses image/video data obtained using a wearable camera simulating a person's visual field.
Getting Started
New to egocentric vision? A few landmark resources already in this list are a good entry point:
- Start here: An Outlook into the Future of Egocentric Vision (IJCV 2024) is a broad survey of the field's tasks, datasets, and open challenges
- Foundational datasets: Ego4D, EPIC-Kitchens 2020, Ego-Exo4D
Contents
> Clustered into various problem statements. - Action/Activity Recognition - Object/Hand Recognition - Action/Gaze Anticipation - Localization - Clustering - Video Summarization - Social Interactions - Pose Estimation - Human Object Interaction - Temporal Boundary Detection - Privacy in Egocentric Videos - Multiple Egocentric Tasks - Task Understanding - Ego-Exo Cross-View Learning - Egocentric Video-Language Models & Question Answering - Egocentric Video Generation & World Models - 3D Scene Reconstruction & Mapping - Assistive & Navigation - Miscellaneous (New Tasks)> Clustered according to the conferences. - CVPR - ECCV - ICCV - WACV - BMVC - NeurIPS
Papers
Clustered in various problem statements.
Action/Activity Recognition
Show papers (44)
- ProbRes: Probabilistic Jump Diffusion for Open-World Egocentric Activity Recognition - Sanjoy Kundu, Shanmukha Vellamcheti, and Sathyanarayanan N. Aakur. In ICCV 2025.
- Understanding Multi-Task Activities from Single-Task Videos - Yuhan Shen and Ehsan Elhamifar. In CVPR 2025.
- Test-Time Adaptation for Combating Missing Modalities in Egocentric Videos - Merey Ramazanova, Alejandro Pardo, Bernard Ghanem, and Motasem Alfarra. In ICLR 2025.
- On the Utility of 3D Hand Poses for Action Recognition - Md Salman Shamil, Dibyadip Chatterjee, Fadime Sener, Shugao Ma, and Angela Yao. In ECCV 2024. [[project page]](https://s-shamil.github.io/HandFormer/)
- SoundingActions: Learning How Actions Sound from Narrated Egocentric Videos - Changan Chen, Kumar Ashutosh, Rohit Girdhar, David Harwath, and Kristen Grauman. In CVPR 2024. [[project page]](https://vision.cs.utexas.edu/projects/soundingactions)
- X-MIC: Cross-Modal Instance Conditioning for Egocentric Action Generalization - Anna Kukleva, Fadime Sener, Edoardo Remelli, Bugra Tekin, Eric Sauser, Bernt Schiele, and Shugao Ma. In CVPR 2024. [[code]](https://github.com/annusha/xmic)
- Progress-Aware Online Action Segmentation for Egocentric Procedural Task Videos - Yuhan Shen and Ehsan Elhamifar. In CVPR 2024. [[code]](https://github.com/Yuhan-Shen/ProTAS)
- TIM: A Time Interval Machine for Audio-Visual Action Recognition - Jacob Chalk, Jaesung Huh, Evangelos Kazakos, Andrew Zisserman, and Dima Damen. In CVPR 2024. [[project page]](https://jacobchalk.github.io/TIM-Project) [[code]](https://github.com/JacobChalk/TIM)
- Multimodal Distillation for Egocentric Action Recognition - Gorjan Radevski, Dusan Grujicic, Matthew Blaschko, Marie-Francine Moens, and Tinne Tuytelaars. In ICCV 2023. [[code]](https://github.com/gorjanradevski/multimodal-distillation)
- What can a cook in Italy teach a mechanic in India? Action Recognition Generalisation Over Scenarios and Locations - Chiara Plizzari, Toby Perrett, Barbara Caputo, and Dima Damen. In ICCV 2023. [[project page]](https://web.archive.org/web/20241209215715/https://chiaraplizz.github.io/what-can-a-cook/) [[code]](https://github.com/Chiaraplizz/ARGO1M-What-can-a-cook)
- MMG-Ego4D: Multimodal Generalization in Egocentric Action Recognition - Xinyu Gong, Sreyas Mohan, Naina Dhingra, Jean-Charles Bazin, YILEI LI, Zhangyang Wang, Rakesh Ranjan. In CVPR 2023.
- Therbligs In Action: Video Understanding through Motion Primitives - Eadom Dessalene, Michael Maynord, Cornelia Fermu ̈ller, Yiannis Aloimonos. In CVPR 2023. [[project page]](https://prg.cs.umd.edu/Therbligs)
- Learning Video Representations from Large Language Models - Yue Zhao, Ishan Misra, Philipp Krähenbühl, Rohit Girdhar. In CVPR 2023. [[project page]](https://facebookresearch.github.io/LaViLa/) [[code]](https://github.com/facebookresearch/LaViLa) [[demo]](https://huggingface.co/spaces/nateraw/lavila)
- Learning State-Aware Visual Representations from Audible Interactions - Himangi Mittal, Pedro Morgado, Unnat Jain, Abhinav Gupta. In NeurIPS 2022. [[Code]](https://github.com/HimangiM/RepLAI) [[Video]](https://www.youtube.com/watch?v=hn5P8BPrPZ4)
- Egocentric Activity Recognition and Localization on a 3D Map - Miao Liu, Lingni Ma, Kiran Somasundaram, Yin Li, Kristen Grauman, James M. Rehg, Chao Li. In ECCV 2022.
- SOS! Self-supervised Learning Over Sets Of Handled Objects In Egocentric Action Recognition - Victor Escorcia, Ricardo Guerrero, Xiatian Zhu, Brais Martinez. In ECCV 2022.
- E2(GO)MOTION: Motion Augmented Event Stream for Egocentric Action Recognition - Chiara Plizzari, Mirco Planamente, Gabriele Goletto, Marco Cannici, Emanuele Gusso, Matteo Matteucci, Barbara Caputo. In CVPR 2022.
- Domain Generalization through Audio-Visual Relative Norm Alignment in First Person Action Recognition - Mirco Planamente, Chiara Plizzari, Emanuele Alberti, and Barbara Caputo. In WACV 2022.
- With a Little Help from my Temporal Context: Multimodal Egocentric Action Recognition - Evangelos Kazakos, Jaesung Huh, Arsha Nagrani, Andrew Zisserman, and Dima Damen. In BMVC 2021. [[project page]](https://ekazakos.github.io/MTCN-project/) [[code]](https://github.com/ekazakos/MTCN)
- Stacked Temporal Attention: Improving First-person Action Recognition by Emphasizing Discriminative Clips - Lijin Yang, Yifei Huang, Yusuke Sugano, and Yoichi Sato. In BMVC 2021. [[project page]](https://www.bmvc2021-virtualconference.com/conference/papers/paper0243.html)
- Interactive Prototype Learning for Egocentric Action Recognition - Xiaohan Wang, Linchao Zhu, Heng Wang, and Yi Yang. In ICCV 2021.
- Multi-Modal Domain Adaptation for Fine-Grained Action Recognition - Jonathan Munro and Dima Damen. In CVPR 2020. [[project page]](https://jonmun.github.io/mmsada/) [[code]](https://github.com/jonmun/MM-SADA-code)
- Integrating Human Gaze Into Attention for Egocentric Activity Recognition - Kyle Min, Jason J. Corso. In WACV 2021. [[code]](https://github.com/MichiganCOG/Gaze-Attention)
- EPIC-Fusion: Audio-Visual Temporal Binding for Egocentric Action Recognition - Evangelos Kazakos, Arsha Nagrani, Andrew Zisserman, and Dima Damen. In ICCV 2019. [[code]](https://github.com/ekazakos/temporal-binding-network) [[project page]](https://ekazakos.github.io/TBN/)
- LSTA: Long Short-Term Attention for Egocentric Action Recognition - Swathikiran Sudhakaran, Sergio Escalera, and Oswald Lanz. In CVPR 2019. [[code]](https://github.com/swathikirans/LSTA)
- Egocentric Activity Recognition on a Budget - Rafael Possas, Sheila Pinto Caceres, and Fabio Ramos. In CVPR 2018. [[demo]](https://youtu.be/GBo4sFNzhtU)
- From Lifestyle VLOGs to Everyday Interaction - David F. Fouhey, Weicheng Kuo, Alexei A. Efros, and Jitendra Malik. In CVPR 2018. [[project page]](https://web.archive.org/web/20241102024857/https://web.eecs.umich.edu/~fouhey/2017/VLOG/index.html)
- Actor and Observer: Joint Modeling of First and Third-Person Videos - Gunnar A. Sigurdsson, Abhinav Gupta, Cordelia Schmid, Ali Farhadi, and Karteek Alahari. In CVPR 2018. [[code]](https://github.com/gsig/actor-observer)
- In the eye of beholder: Joint learning of gaze and actions in first person video - Yin Li, Miao Liu, and James M. Rehg. In ECCV 2018.
- Privacy-Preserving Human Activity Recognition from Extreme Low Resolution - Michael S. Ryoo, Brandon Rothrock, Charles Fleming, and Hyun Jong Yang. In AAAI 2017.
- Jointly Recognizing Object Fluents and Tasks in Egocentric Videos - Yang Liu, Ping Wei, and Song-Chun Zhu. In ICCV 2017.
- Trajectory Aligned Features For First Person Action Recognition - Suriya Singh, Chetan Arora, and C.V. Jawahar. In Pattern Recognition 2017.
- First Person Action Recognition Using Deep Learned Descriptors - Suriya Singh, Chetan Arora, and C.V. Jawahar. In CVPR 2016. [[project page]](http://cvit.iiit.ac.in/research/projects/cvit-projects/first-person-action-recognition) [[code]](https://github.com/suriyasingh/EgoConvNet)
- Understanding Hand-Object Manipulation with Grasp Types and Object Attributes - Minjie Cai, Kris M. Kitani, and Yoichi Sato. In Robotics: Science and Systems 2016.
- Delving into egocentric actions - Yin Li, Zhefan Ye, and James M. Rehg. In CVPR 2015.
- Pooled Motion Features for First-Person Videos - Michael S. Ryoo, Brandon Rothrock, and Larry H. Matthies. In CVPR 2015.
- Generating Notifications for Missing Actions: Don't forget to turn the lights off! - Bilge Soran, Ali Farhadi, and Linda Shapiro. In ICCV 2015.
- First-Person Activity Recognition: What Are They Doing to Me? - M. S. Ryoo and Larry Matthies. In CVPR 2013.
- Detecting activities of daily living in first-person camera views - Hamed Pirsiavash and Deva Ramanan. In CVPR 2012.
- Learning to recognize daily actions using gaze - Alireza Fathi, Yin Li, and James M. Rehg. In ECCV 2012.
- Learning to recognize objects in egocentric activities - Alireza Fathi, Xiaofeng Ren, and James M. Rehg. In CVPR 2011.
- Fast unsupervised ego-action learning for first-person sports videos - Kris M. Kitani, Takahiro Okabe, Yoichi Sato, and Akihiro Sugimoto. In CVPR 2011. [[project page]](https://www.ri.cmu.edu/publications/fast-unsupervised-ego-action-learning-for-first-person-sports-videos/)
- Temporal segmentation and activity classification from first-person sensing - Ekaterina H. Spriggs, Fernando De La Torre, and Martial Hebert. In CVPR Workshops 2009.
- Wearable hand activity recognition for event summarization - W.W. Mayol and D.W. Murray. In IEEE International Symposium on Wearable Computers, 2005.
Object/Hand Recognition
Show papers (24)
- Towards Stable Self-Supervised Object Representations in Unconstrained Egocentric Video - Yuting Tan, Xilong Cheng, Yunxiao Qin, Zhengnan Li, and Jingjing Zhang. In CVPR 2026.
- EgoXtreme: A Dataset for Robust Object Pose Estimation in Egocentric Views under Extreme Conditions - Taegyoon Yoon, Yegyu Han, Seojin Ji, Jaewoo Park, Sojeong Kim, Taein Kwon, and Hyung-Sin Kim. In CVPR 2026. [[project page]](https://taegyoun88.github.io/EgoXtreme/) [[code]](https://github.com/taegyoun88/EgoXtreme)
- Robust Egocentric Referring Video Object Segmentation via Dual-Modal Causal Intervention - Haijing Liu, Zhiyuan Song, Hefeng Wu, Tao Pu, Keze Wang, and Liang Lin. In NeurIPS 2025.
- Is Tracking Really More Challenging in First Person Egocentric Vision? - Matteo Dunnhofer, Zaira Manigrasso, and Christian Micheloni. In ICCV 2025. [[project page]](https://machinelearning.uniud.it/datasets/vista/) [[code]](https://github.com/matteo-dunnhofer/fpv-tracking-toolkit)
- HaWoR: World-Space Hand Motion Reconstruction from Egocentric Videos - Jinglei Zhang, Jiankang Deng, Chao Ma, and Rolandos Alexandros Potamias. In CVPR 2025. [[project page]](https://hawor-project.github.io/)
- HOT3D: Hand and Object Tracking in 3D from Egocentric Multi-View Videos - Prithviraj Banerjee, Sindi Shkodrani, Pierre Moulon, Shreyas Hampali, Shangchen Han, Fan Zhang, et al. In CVPR 2025. [[project page]](https://facebookresearch.github.io/hot3d/)
- ActionVOS: Actions as Prompts for Video Object Segmentation - Liangyang Ouyang, Ruicong Liu, Yifei Huang, Ryosuke Furuta, and Yoichi Sato. In ECCV 2024. [[code]](https://github.com/ut-vision/ActionVOS)
- Instance Tracking in 3D Scenes from Egocentric Videos - Yunhan Zhao, Haoyu Ma, Shu Kong, and Charless Fowlkes. In CVPR 2024. [[code]](https://github.com/IT3DEgo/IT3DEgo)
- Learning to Segment Referred Objects from Narrated Egocentric Videos - Yuhan Shen, Huiyu Wang, Xitong Yang, Matt Feiszli, Ehsan Elhamifar, Lorenzo Torresani, and Effrosyni Mavroudi. In CVPR 2024.
- EgoTracks: A Long-term Egocentric Visual Object Tracking Dataset - Hao Tang, Kevin J Liang, Kristen Grauman, Matt Feiszli, and Weiyao Wang. In NeurIPS 2023. [[dataset]](https://ego4d-data.org/docs/data/egotracks/)
- Self-Supervised Object Detection from Egocentric Videos - Peri Akiva, Jing Huang, Kevin J Liang, Rama Kovvuri, Xingyu Chen, Matt Feiszli, Kristin Dana, and Tal Hassner. In ICCV 2023.
- EgoObjects: A Large-Scale Egocentric Dataset for Fine-Grained Object Understanding - Chenchen Zhu, Fanyi Xiao, Andres Alvarado, Yasmine Babaei, Jiabo Hu, Hichem El-Mohri, Sean Culatana, Roshan Sumbaly, and Zhicheng Yan. In ICCV 2023. [[project page]](https://research.facebook.com/blog/2023/3/egoobjects-large-scale-egocentric-dataset-for-category-and-instance-level-object-understanding/) [[code]](https://github.com/facebookresearch/EgoObjects)
- Hierarchical Temporal Transformer for 3D Hand Pose Estimation and Action Recognition from Egocentric RGB Videos - Yilin Wen, Hao Pan, Lei Yang, Jia Pan, Taku Komura, Wenping Wang. In CVPR 2023. [[Code]](https://github.com/fylwen/HTT)
- Generative Adversarial Network for Future Hand Segmentation from Egocentric Video - Wenqi Jia, Miao Liu, James M. Rehg. In ECCV 2022.
- Whose Hand Is This? Person Identification From Egocentric Hand Gestures - Satoshi Tsutsui, Yanwei Fu, and David J. Crandall. In WACV 2021.
- Generalizing Hand Segmentation in Egocentric Videos with Uncertainty-Guided Model Adaptation - Minjie Cai, Feng Lu, and Yoichi Sato. In CVPR 2020. [[code]](https://github.com/cai-mj/UMA)
- H+O: Unified Egocentric Recognition of 3D Hand-Object Poses and Interactions - Bugra Tekin, Federica Bogo, and Marc Pollefeys. In CVPR 2019. [[video]](https://youtu.be/ko6kNZ9DuAk?t=3240)
- Analysis of Hand Segmentation in the Wild - Aisha Urooj Khan and Ali Borji. In CVPR 2018.
- First-Person Hand Action Benchmark with RGB-D Videos and 3D Hand Pose Annotations - Guillermo Garcia-Hernando, Shanxin Yuan, Seungryul Baek, and Tae-Kyun Kim. In CVPR 2018. [[project page]](https://guiggh.github.io/publications/first-person-hands/) [[code]](https://github.com/guiggh/handpose_action)
- Egocentric Gesture Recognition Using Recurrent 3D Convolutional Neural Networks with Spatiotemporal Transformer Modules - Congqi Cao, Yifan Zhang, Yi Wu, Hanqing Lu, and Jian Cheng. In ICCV 2017.
- Lending a hand: Detecting hands and recognizing activities in complex egocentric interactions - Sven Bambach, Stefan Lee, David J. Crandall, and Chen Yu. In ICCV 2015.
- Detecting Snap Points in Egocentric Video with a Web Photo Prior - Bo Xiong and Kristen Grauman. In ECCV 2014. [[project page]](http://vision.cs.utexas.edu/projects/egosnappoints/) [[code]](http://vision.cs.utexas.edu/projects/ego_snappoints/#code)
- Pixel-level hand detection in ego-centric videos - Cheng Li and Kris M. Kitani. In CVPR 2013. [[video]](https://youtu.be/N756YmLpZyY) [[code]](https://github.com/irllabs/handtrack)
- Context-based vision system for place and object recognition - Antonio Torralba, Kevin P. Murphy, William T. Freeman, Mark A. Rubin. In ICCV 2003. [[project page]](https://www.cs.ubc.ca/~murphyk/Vision/placeRecognition.html)
Action/Gaze Anticipation
Show papers (24)
- Gaze Beyond the Frame: Forecasting Egocentric 3D Visual Span - Heeseung Yun, Joonil Na, Jaeyeon Kim, Calvin Murdock, and Gunhee Kim. In NeurIPS 2025.
- HOIGaze: Gaze Estimation During Hand-Object Interactions in Extended Reality Exploiting Eye-Hand-Head Coordination - Zhiming Hu, Daniel Haeufle, Syn Schmitt, and Andreas Bulling. In SIGGRAPH 2025. [[project page]](https://zhiminghu.net/hu25hoigaze.html) [[code]](https://github.com/CraneHzm/HOIGaze)
- Test-time Ego-Exo-centric Adaptation for Action Anticipation via Multi-Label Prototype Growing and Dual-Clue Consistency - Zhaofeng Shi, Heqian Qiu, Lanxiao Wang, Qingbo Wu, Fanman Meng, Lili Pan, and Hongliang Li. In CVPR 2026. [[code]](https://github.com/ZhaofengSHI/DCPGN)
- Forecasting 3D Scanpaths in Egocentric Video - Fiona Ryan, Ishwarya Ananthabhotla, Yijun Qian, Judy Hoffman, James M. Rehg, Vamsi Krishna Ithapu, and Calvin Murdock. In CVPR 2026.
- FIction: 4D Future Interaction Prediction from Video - Kumar Ashutosh, Georgios Pavlakos, and Kristen Grauman. In CVPR 2025. [[code]](https://github.com/thechargedneutron/FIction)
- Listen to Look into the Future: Audio-Visual Egocentric Gaze Anticipation - Bolin Lai, Fiona Ryan, Wenqi Jia, Miao Liu, and James M. Rehg. In ECCV 2024. [[project page]](https://bolinlai.github.io/CSTS-EgoGazeAnticipation/)
- AFF-ttention! Affordances and Attention models for Short-Term Object Interaction Anticipation - Lorenzo Mur-Labadia, Ruben Martinez-Cantin, Jose J. Guerrero, Giovanni Maria Farinella, and Antonino Furnari. In ECCV 2024. [[code]](https://github.com/lmur98/AFFttention)
- PALM: Predicting Actions through Language Models - Sanghwan Kim, Daoji Huang, Yongqin Xian, Otmar Hilliges, Luc Van Gool, and Xi Wang. In ECCV 2024.
- Summarize the Past to Predict the Future: Natural Language Descriptions of Context Boost Multimodal Object Interaction Anticipation - Razvan-George Pasca, Alexey Gavryushin, Muhammad Hamza, Yen-Ling Kuo, Kaichun Mo, Luc Van Gool, Otmar Hilliges, and Xi Wang. In CVPR 2024. [[project page]](https://eth-ait.github.io/transfusion-proj/)
- Uncertainty-aware State Space Transformer for Egocentric 3D Hand Trajectory Forecasting - Wentao Bao, Lele Chen, Libing Zeng, Zhong Li, Yi Xu, Junsong Yuan, and Yu Kong. In ICCV 2023. [[project page]](https://web.archive.org/web/20231211050832/https://actionlab-cv.github.io/EgoHandTrajPred/) [[code]](https://github.com/oppo-us-research/USST)
- Intention-Conditioned Long-Term Human Egocentric Action Forecasting - Esteve Valls Mascaro, Hyemin Ahn, and Dongheui Lee. In WACV 2023.
- A Hybrid Egocentric Activity Anticipation Framework via Memory-Augmented Recurrent and One-shot Representation Forecasting - Tianshan Liu and Kin-Man Lam. In CVPR 2022.
- Learning to Anticipate Egocentric Actions by Imagination - Yu Wu, Linchao Zhu, Xiaohan Wang, Yi Yang, and Fei Wu. In TIP 2021.
- Forecasting Human-Object Interaction: Joint Prediction of Motor Attention and Actions in First Person Video - Miao Liu, Siyu Tang, Yin Li, and James M. Rehg. In ECCV 2020. [[project page]](https://aptx4869lm.github.io/ForecastingHOI/)
- How Can I See My Future? FvTraj: Using First-person View for Pedestrian Trajectory Prediction - Huikun Bi, Ruisi Zhang, Tianlu Mao, Zhigang Deng, and Zhaoqi Wang. In ECCV 2020. [[presentation video]](https://youtu.be/HcsyH7zMHAw) [[summary video]](https://youtu.be/X1cSNWT6Gr0)
- Multimodal Future Localization and Emergence Prediction for Objects in Egocentric View With a Reachability Prior - Osama Makansi, Ozgun Cicek, Kevin Buchicchio, and Thomas Brox. In CVPR 2020. [[demo]](https://youtu.be/_9Ml5IFwbSY) [[code]](https://github.com/lmb-freiburg/FLN-EPN-RPN) [[project page]](https://lmb.informatik.uni-freiburg.de/Publications/2020/MCBB20/)
- EGO-TOPO: Environment Affordances from Egocentric Video - Tushar Nagarajan, Yanghao Li, Christoph Feichtenhofer, and Kristen Grauman. In CVPR 2020. [[project page]](http://vision.cs.utexas.edu/projects/ego-topo/) [[demo]](http://vision.cs.utexas.edu/projects/ego-topo/demo.html)
- What Would You Expect? Anticipating Egocentric Actions with Rolling-Unrolling LSTMs and Modality Attention - Antonino Furnari and Giovanni Maria Farinella. In ICCV 2019 [[code]](https://github.com/fpv-iplab/rulstm) [[demo]](https://youtu.be/buIEKFHTVIg)
- Digging Deeper into Egocentric Gaze Prediction - Hamed R. Tavakoli, Esa Rahtu, Juho Kannala, and Ali Borji. In WACV 2019.
- Predicting Gaze in Egocentric Video by Learning Task-dependent Attention Transition - Yifei Huang, Minjie Cai, Zhenqiang Li, and Yoichi Sato. In ECCV 2018 [[code]](https://github.com/hyf015/egocentric-gaze-prediction)
- First-Person Activity Forecasting with Online Inverse Reinforcement Learning - Nicholas Rhinehart and Kris M. Kitani. In ICCV 2017. [[video]](https://youtu.be/rvVoW3iuq-s)
- Deep future gaze: Gaze anticipation on egocentric videos using adversarial networks - Mengmi Zhang, Keng Teck Ma, Joo Hwee Lim, Qi Zhao, and Jiashi Feng. In CVPR 2017. [[code]](https://github.com/Mengmi/deepfuturegazegan)
- Going deeper into first-person activity recognition - Minghuang Ma, Haoqi Fan, and Kris M. Kitani. In CVPR 2016.
- Learning to predict gaze in egocentric video - Yin Li, Alireza Fathi, and James M. Rehg. In ICCV 2013.
Localization
Show papers (13)
- Beyond Caption-Based Queries for Video Moment Retrieval - David Pujol-Perich, Albert Clapés, Dima Damen, Sergio Escalera, and Michael Wray. In CVPR 2026.
- PRVQL: Progressive Knowledge-guided Refinement for Robust Egocentric Visual Query Localization - Bing Fan, Yunhe Feng, Yapeng Tian, James Chenhao Liang, Yuewei Lin, Yan Huang, and Heng Fan. In ICCV 2025. [[code]](https://github.com/fb-reps/PRVQL)
- Egocentric Action-aware Inertial Localization in Point Clouds with Vision-Language Guidance - Mingfang Zhang, Ryo Yonetani, Yifei Huang, Liangyang Ouyang, Ruicong Liu, and Yoichi Sato. In ICCV 2025.
- Spatial Cognition from Egocentric Video: Out of Sight, Not Out of Mind - Chiara Plizzari, Shubham Goel, Toby Perrett, Jacob Chalk, Angjoo Kanazawa, and Dima Damen. In 3DV 2025. [[project page]](https://dimadamen.github.io/OSNOM/)
- Online Episodic Memory Visual Query Localization with Egocentric Streaming Object Memory - Zaira Manigrasso, Matteo Dunnhofer, Antonino Furnari, Moritz Nottebaum, Antonio Finocchiaro, Davide Marana, Rosario Forte, Giovanni Maria Farinella, and Christian Micheloni. In WACV 2026.
- Spherical World-Locking for Audio-Visual Localization in Egocentric Videos - Heeseung Yun, Ruohan Gao, Ishwarya Ananthabhotla, Anurag Kumar, Jacob Donley, Chao Li, Gunhee Kim, Vamsi Krishna Ithapu, and Calvin Murdock. In ECCV 2024. [[project page]](https://hs-yn.github.io/SWL/)
- EgoLoc: Revisiting 3D Object Localization from Egocentric Videos with Visual Queries - Jinjie Mai, Abdullah Hamdi, Silvio Giancola, Chen Zhao, and Bernard Ghanem. In ICCV 2023. [[code]](https://github.com/Wayne-Mai/EgoLoc)
- Hand-Priming in Object Localization for Assistive Egocentric Vision - Kyungjun Lee, Abhinav Shrivastava, and Hernisa Kacorri. In WACV 2020.
- Egocentric Shopping Cart Localization - Emiliano Spera, Antonino Furnari, Sebastiano Battiato, and Giovanni Maria Farinella. In ICPR 2018.
- Recognizing personal locations from egocentric videos - Antonino Furnari, Giovanni Maria Farinella, and Sebastiano Battiato. In IEEE Transactions on Human-Machine Systems 2017.
- Personal-Location-Based Temporal Segmentation of Egocentric Video for Lifelogging Applications - Antonino Furnari, Sebastiano Battiato, and Giovanni Maria Farinella. In Journal of Visual Communication and Image Representation 2017. [[demo]](https://youtu.be/URM0EdYuKEw) [[project page]](https://web.archive.org/web/20251111135710/https://iplab.dmi.unict.it/EgocentricShoppingCartLocalization/)
- Egocentric Future Localization - Hyun Soo Park, Jyh-Jing Hwang, Yedong Niu, and Jianbo Shi. In CVPR 2016. [[demo]](https://youtu.be/i9CTMZ60zc)
- Real-time localization and mapping with wearable active vision - A.J. Davison, W.W. Mayol, and D.W. Murray. In The Second IEEE and ACM International Symposium 2003.
Clustering
Show papers (2)
- Sr-clustering: Semantic regularized clustering for egocentric photo streams segmentation - Mariella Dimiccoli, Marc Bolanosa, Estefania Talavera Maedeh Aghaei, Stavri G. Nikolov, and Petia Radeva. In Computer Vision and Image Understanding 2017.
- Summarization and Classification of Wearable Camera Streams by Learning the Distributions over Deep Features of Out-of-Sample Image Sequences - Alessandro Perina, Sadegh Mohammadi, Nebojsa Jojic, and Vittorio Murino. In ICCV 2017.
Video Summarization
Show papers (4)
- Query-focused video summarization: Dataset, evaluation, and a memory network based approach - Aidean Sharghi, Jacob S. Laurel and Boqing Gong. In CVPR 2017.
- Toward storytelling from visual lifelogging: An overview - Marc Bolanos, Mariella Dimiccoli, and Petia Radeva. In IEEE Transactions on Human-Machine Systems 2017.
- Story-Driven Summarization for Egocentric Video - Zheng Lu and Kristen Grauman. In CVPR 2013 [[project page]](http://vision.cs.utexas.edu/projects/egocentric/storydriven.html)
- Discovering Important People and Objects for Egocentric Video Summarization - Yong Jae Lee, Joydeep Ghosh, and Kristen Grauman. In CVPR 2012. [[project page]](http://vision.cs.utexas.edu/projects/egocentric/index.html)
Social Interactions
Show papers (6)
- Seeing Conversations: Communication Context Identification in Egocentric Video - Tobias Dorszewski and Jens Hjortkjær. In CVPR 2026.
- Ex2Eg-MAE: A Framework for Adaptation of Exocentric Video Masked Autoencoders for Egocentric Social Role Understanding - Minh Tran, Yelin Kim, Che-Chun Su, Cheng-Hao Kuo, Min Sun, and Mohammad Soleymani. In ECCV 2024.
- The Audio-Visual Conversational Graph: From an Egocentric-Exocentric Perspective - Wenqi Jia, Miao Liu, Hao Jiang, Ishwarya Ananthabhotla, James M. Rehg, Vamsi Krishna Ithapu, and Ruohan Gao. In CVPR 2024. [[project page]](https://vjwq.github.io/AV-CONV/)
- EgoCom: A Multi-person Multi-modal Egocentric Communications Dataset - Curtis G. Northcutt, Shengxin Zha, Steven Lovegrove, and Richard Newcombe. In PAMI 2020.
- Deep Dual Relation Modeling for Egocentric Interaction Recognition - Haoxin Li, Yijun Cai, and Wei-Shi Zheng. In CVPR 2019.
- Recognizing Micro-Actions and Reactions from Paired Egocentric Videos - Ryo Yonetani, Kris M. Kitani, and Yoichi Sato. In CVPR 2016.
Pose Estimation
Show papers (50)
- EgoHumans: An Egocentric 3D Multi-Human Benchmark - Rawal Khirodkar, Aayush Bansal, Lingni Ma, Richard Newcombe, Minh Vo, and Kris Kitani. In ICCV 2023 (Oral). [[code]](https://github.com/rawalkhirodkar/egohumans)
- Towards in-the-wild Egocentric 3D Hand-Object Pose Estimation - Siddhant Bansal, Zhifan Zhu, Shashank Tripathi, Jiahe Zhao, Michael J. Black, and Dima Damen. In ECCV 2026. [[project page]](https://sid2697.github.io/epic-contact/) [[code]](https://github.com/Sid2697/HOPformer)
- E-3DPSM: A State Machine for Event-Based Egocentric 3D Human Pose Estimation - Mayur Deshmukh, Hiroyasu Akada, Helge Rhodin, Christian Theobalt, and Vladislav Golyanik. In CVPR 2026. [[project page]](https://4dqv.mpi-inf.mpg.de/E-3DPSM/)
- Egocentric Visibility-Aware Human Pose Estimation - Peng Dai, Yu Zhang, Yiqiang Feng, Zhen Fan, and Yang Zhang. In CVPR 2026.
- EgoPoseFormer v2: Accurate Egocentric Human Motion Estimation for AR/VR - Zhenyu Li, Sai Kumar Dwivedi, Filip Maric, Carlos Chacon, Nadine Bertsch, Filippo Arcadu, et al. In CVPR 2026. [[project page]](https://zhyever.github.io/EgoPoseFormerv2/)
- Towards Egocentric 3D Hand Pose Estimation in Unseen Domains - Wiktor Mucha, Michael Wray, and Martin Kampel. In WACV 2026.
- UniEgoMotion: A Unified Model for Egocentric Motion Reconstruction, Forecasting, and Generation - Chaitanya Patel, Hiroki Nakamura, Yuta Kyuragi, Kazuki Kozuka, Juan Carlos Niebles, and Ehsan Adeli. In ICCV 2025. [[project page]](https://chaitanya100100.github.io/UniEgoMotion/) [[code]](https://github.com/chaitanya100100/UniEgoMotion)
- Head2Body: Body Pose Generation from Multi-sensory Head-mounted Inputs - Minh Tran, Hongda Mao, Qingshuang Chen, and Yelin Kim. In ICCV 2025.
- Bring Your Rear Cameras for Egocentric 3D Human Pose Estimation - Hiroyasu Akada, Jian Wang, Vladislav Golyanik, and Christian Theobalt. In ICCV 2025. [[project page]](https://4dqv.mpi-inf.mpg.de/EgoRear/)
- Fish2Mesh Transformer: 3D Human Mesh Recovery from Egocentric Vision - Tianma Shen, Aditya Puranik, James Vong, Vrushabh Abhijit Deogirikar, Ryan Fell, Julianna Dietrich, Maria Kyrarini, Christopher Kitts, and David C. Jeong. In ICCV 2025. [[project page]](https://fish2mesh.github.io/)
- EgoMusic-driven Human Dance Motion Estimation with Skeleton Mamba - Quang Nguyen, Nhat Le, Baoru Huang, Minh Nhat Vu, Chengcheng Tang, Van Nguyen, Ngan Le, Thieu Vo, and Anh Nguyen. In ICCV 2025.
- REWIND: Real-Time Egocentric Whole-Body Motion Diffusion with Exemplar-Based Identity Conditioning - Jihyun Lee, Weipeng Xu, Alexander Richard, Shih-En Wei, Shunsuke Saito, Shaojie Bai, Te-Li Wang, Minhyuk Sung, Tae-Kyun Kim, and Jason Saragih. In CVPR 2025. [[project page]](https://jyunlee.github.io/projects/rewind/)
- FRAME: Floor-aligned Representation for Avatar Motion from Egocentric Video - Andrea Boscolo Camiletto, Jian Wang, Eduardo Alvarado, Rishabh Dabral, Thabo Beeler, Marc Habermann, and Christian Theobalt. In CVPR 2025. [[project page]](https://vcai.mpi-inf.mpg.de/projects/FRAME/) [[code]](https://github.com/abcamiletto/frame)
- EgoLM: Multi-Modal Language Model of Egocentric Motions - Fangzhou Hong, Vladimir Guzov, Hyo Jin Kim, Yuting Ye, Richard Newcombe, Ziwei Liu, and Lingni Ma. In CVPR 2025. [[project page]](https://hongfz16.github.io/projects/EgoLM)
- Estimating Body and Hand Motion in an Ego-sensed World - Brent Yi, Vickie Ye, Maya Zheng, Yunqi Li, Lea Müller, Georgios Pavlakos, Yi Ma, Jitendra Malik, and Angjoo Kanazawa. In CVPR 2025. [[project page]](https://egoallo.github.io/)
- Dyn-HaMR: Recovering 4D Interacting Hand Motion from a Dynamic Camera - Zhengdi Yu, Stefanos Zafeiriou, and Tolga Birdal. In CVPR 2025. [[project page]](https://dyn-hamr.github.io/)
- EgoPressure: A Dataset for Hand Pressure and Pose Estimation in Egocentric Vision - Yiming Zhao, Taein Kwon, Paul Streli, Marc Pollefeys, and Christian Holz. In CVPR 2025. [[project page]](https://yiming-zhao.github.io/EgoPressure/)
- Ego4o: Egocentric Human Motion Capture and Understanding from Multi-Modal Input - Jian Wang, Rishabh Dabral, Diogo Luvizon, Zhe Cao, Lingjie Liu, Thabo Beeler, and Christian Theobalt. In CVPR 2025. [[project page]](https://jianwang-mpi.github.io/ego4o/)
- EgoCast: Forecasting Egocentric Human Pose in the Wild - Maria Escobar, Juanita Puentes, Cristhian Forigua, Jordi Pont-Tuset, Kevis-Kokitsi Maninis, and Pablo Arbelaez. In WACV 2025. [[code]](https://github.com/BCV-Uniandes/EgoCast)
- Social EgoMesh Estimation - Luca Scofano, Alessio Sampieri, Edoardo De Matteis, Indro Spinelli, and Fabio Galasso. In WACV 2025. [[code]](https://github.com/L-Scofano/SEEME)
- Estimating Ego-Body Pose from Doubly Sparse Egocentric Video Data - Seunggeun Chi, Pin-Hao Huang, Enna Sachdeva, Hengbo Ma, Karthik Ramani, and Kwonjoon Lee. In NeurIPS 2024. [[project page]](https://sgchi.github.io/dsposer/)
- EgoSim: An Egocentric Multi-view Simulator and Real Dataset for Body-worn Cameras during Motion and Activity - Dominik Hollidt, Paul Streli, Jiaxi Jiang, Yasaman Haghighi, Changlin Qian, Xintong Liu, and Christian Holz. In NeurIPS 2024. [[project page]](https://siplab.org/projects/EgoSim)
- Nymeria: A Massive Collection of Egocentric Multi-modal Human Motion in the Wild - Lingni Ma, Yuting Ye, Fangzhou Hong, Vladimir Guzov, Yifeng Jiang, et al. In ECCV 2024. [[project page]](https://www.projectaria.com/datasets/nymeria/)
- EgoPoseFormer: A Simple Baseline for Stereo Egocentric 3D Human Pose Estimation - Chenhongyi Yang, Anastasia Tkach, Shreyas Hampali, Linguang Zhang, Elliot J. Crowley, and Cem Keskin. In ECCV 2024. [[code]](https://github.com/ChenhongyiYang/egoposeformer)
- EgoPoser: Robust Real-Time Egocentric Pose Estimation from Sparse and Intermittent Observations Everywhere - Jiaxi Jiang, Paul Streli, Manuel Meier, and Christian Holz. In ECCV 2024. [[project page]](https://siplab.org/projects/EgoPoser)
- 3D Hand Pose Estimation in Everyday Egocentric Images - Aditya Prakash, Ruisen Tu, Matthew Chang, and Saurabh Gupta. In ECCV 2024. [[project page]](https://ap229997.github.io/projects/hands/)
- EgoBody3M: Egocentric Body Tracking on a VR Headset using a Diverse Dataset - Amy Zhao, Chengcheng Tang, Lezi Wang, Yijing Li, Mihika Dave, Lingling Tao, Christopher D. Twigg, and Robert Y. Wang. In ECCV 2024. [[dataset]](https://github.com/facebookresearch/EgoBody3M)
- Benchmarks and Challenges in Pose Estimation for Egocentric Hand Interactions with Objects - Zicong Fan, Takehiko Ohkawa, Linlin Yang, Nie Lin, Zhishan Zhou, Shihao Zhou, et al. In ECCV 2024.
- EventEgo3D: 3D Human Motion Capture from Egocentric Event Streams - Christen Millerdurai, Hiroyasu Akada, Jian Wang, Diogo Luvizon, Christian Theobalt, and Vladislav Golyanik. In CVPR 2024. [[project page]](https://4dqv.mpi-inf.mpg.de/EventEgo3D/)
- 3D Human Pose Perception from Egocentric Stereo Videos - Hiroyasu Akada, Jian Wang, Vladislav Golyanik, and Christian Theobalt. In CVPR 2024. [[code]](https://github.com/hiroyasuakada/3D-Human-Pose-Perception-from-Egocentric-Stereo-Videos)
- Egocentric Whole-Body Motion Capture with FisheyeViT and Diffusion-Based Motion Refinement - Jian Wang, Zhe Cao, Diogo Luvizon, Lingjie Liu, Kripasindhu Sarkar, Danhang Tang, Thabo Beeler, and Christian Theobalt. In CVPR 2024. [[code]](https://github.com/jianwang-mpi/egowholemocap)
- Attention-Propagation Network for Egocentric Heatmap to 3D Pose Lifting - Taeho Kang and Youngki Lee. In CVPR 2024. [[code]](https://github.com/tho-kn/EgoTAP)
- Single-to-Dual-View Adaptation for Egocentric 3D Hand Pose Estimation - Ruicong Liu, Takehiko Ohkawa, Mingfang Zhang, and Yoichi Sato. In CVPR 2024. [[code]](https://github.com/ut-vision/S2DHand)
- Real-Time Simulated Avatar from Head-Mounted Sensors - Zhengyi Luo, Jinkun Cao, Rawal Khirodkar, Alexander Winkler, Jing Huang, Kris Kitani, and Weipeng Xu. In CVPR 2024. [[project page]](https://www.zhengyiluo.com/SimXR/)
- Mocap Everyone Everywhere: Lightweight Motion Capture With Smartwatches and a Head-Mounted Camera - Jiye Lee and Hanbyul Joo. In CVPR 2024. [[project page]](https://jiyewise.github.io/projects/MocapEvery/)
- Spectral Graphormer: Spectral Graph-Based Transformer for Egocentric Two-Hand Reconstruction using Multi-View Color Images - Tze Ho Elden Tse, Franziska Mueller, Zhengyang Shen, Danhang Tang, Thabo Beeler, Mingsong Dou, Yinda Zhang, Sasa Petrovic, Hyung Jin Chang, Jonathan Taylor, and Bardia Doosti. In ICCV 2023. [[project page]](https://eldentse.github.io/Spectral-Graphormer/)
- Probabilistic Human Mesh Recovery in 3D Scenes from Egocentric Views - Siwei Zhang, Qianli Ma, Yan Zhang, Sadegh Aliakbarian, Darren Cosker, and Siyu Tang. In ICCV 2023. [[project page]](https://sanweiliti.github.io/egohmr/egohmr.html) [[code]](https://github.com/sanweiliti/EgoHMR)
- AssemblyHands: Towards Egocentric Activity Understanding via 3D Hand Pose Estimation - Takehiko Ohkawa, Kun He, Fadime Sener, Tomas Hodan, LUAN TRAN, Cem Keskin. In CVPR 2023.
- Scene-aware Egocentric 3D Human Pose Estimation - Jian Wang, Diogo Luvizon, Weipeng Xu, Lingjie Liu, Kripasindhu Sarkar, Christian Theobalt. In CVPR 2023.
- Ego-Body Pose Estimation via Ego-Head Pose Estimation - Jiaman Li · Karen Liu · Jiajun Wu. In CVPR 2023.
- EgoBody: Human Body Shape and Motion of Interacting People from Head-Mounted Devices - Siwei Zhang, Qianli Ma, Yan Zhang, Zhiyin Qian, Taein Kwon, Marc Pollefeys, Federica Bogo, Siyu Tang. In ECCV 2022. [[project page]](https://sanweiliti.github.io/egobody/egobody.html) [[dataset]](https://egobody.inf.ethz.ch/) [[code]](https://github.com/sanweiliti/EgoBody)
- UnrealEgo: A New Dataset for Robust Egocentric 3D Human Motion Capture - Hiroyasu Akada, Jian Wang, Soshi Shimada, Masaki Takahashi, Christian Theobalt, Vladislav Golyanik. In ECCV 2022. [[project page]](https://4dqv.mpi-inf.mpg.de/UnrealEgo/) [[code]](https://github.com/hiroyasuakada/UnrealEgo) [[dataset]](https://4dqv.mpi-inf.mpg.de/UnrealEgo/) [[demo]](https://4dqv.mpi-inf.mpg.de/UnrealEgo/data/unrealegodistribution.mp4)
- Estimating Egocentric 3D Human Pose in the Wild with External Weak Supervision - Jian Wang, Lingjie Liu, Weipeng Xu, Kripasindhu Sarkar, Diogo Luvizon, Christian Theobalt. In CVPR 2022. [[project page]](https://web.archive.org/web/20240726094113/https://people.mpi-inf.mpg.de/~jianwang/projects/egopw/)
- Estimating Egocentric 3D Human Pose in Global Space - Jian Wang, Lingjie Liu, Weipeng Xu, Kripasindhu Sarkar, Christian Theobalt. In ICCV 2021. [[project page]](https://web.archive.org/web/20240423122243/https://people.mpi-inf.mpg.de/~jianwang/projects/globalegomocap/)
- Automatic Calibration of the Fisheye Camera for Egocentric 3D Human Pose Estimation From a Single Image - Yahui Zhang, Shaodi You, and Theo Gevers. In WACV 2021.
- You2Me: Inferring Body Pose in Egocentric Video via First and Second Person Interactions - Evonne Ng, Donglai Xiang, Hanbyul Joo, and Kristen Grauman. In CVPR 2020. [[demo]](http://vision.cs.utexas.edu/projects/you2me/demo.mp4) [[project page]](http://vision.cs.utexas.edu/projects/you2me/) [[dataset]](https://github.com/facebookresearch/you2me/tree/master/data#) [[code]](https://github.com/facebookresearch/you2me#)
- Ego-Pose Estimation and Forecasting as Real-Time PD Control - Ye Yuan and Kris Kitani. In ICCV 2019. [[code]](https://github.com/Khrylx/EgoPose) [[project page]](https://www.ye-yuan.com/ego-pose) [[demo]](https://youtu.be/968IIDZeWE0)
- xR-EgoPose: Egocentric 3D Human Pose From an HMD Camera - Denis Tome, Patrick Peluse, Lourdes Agapito, and Hernan Badino. In ICCV 2019. [[demo]](https://youtu.be/zem03fZWLrQ) [[dataset]](https://github.com/facebookresearch/xR-EgoPose)
- Seeing Invisible Poses: Estimating 3D Body Pose from Egocentric Video - Hao Jiang and Kristen Grauman. In CVPR 2017.
- First-Person Pose Recognition using Egocentric Workspaces - Gregory Rogez, James S. Supancic, and Deva Ramanan. In CVPR 2015.
Human Object Interaction
Show papers (20)
- MEgoHand: Multimodal Egocentric Hand-Object Interaction Motion Generation - Bohan Zhou, Yi Zhan, Zhongbin Zhang, and Zongqing Lu. In NeurIPS 2025. [[project page]](https://beingbeyond.github.io/MEgoHand/) [[code]](https://github.com/BeingBeyond/MEgoHand)
- Perceiving and Acting in First-Person: A Dataset and Benchmark for Egocentric Human-Object-Human Interactions - Liang Xu, Chengqun Yang, Zili Lin, Fei Xu, Yifan Liu, Congsheng Xu, et al. In ICCV 2025. [[project page]](https://liangxuy.github.io/InterVLA/)
- Learning Precise Affordances from Egocentric Videos for Robotic Manipulation - Gen Li, Nikolaos Tsagkas, Jifei Song, Ruaridh Mon-Williams, Sethu Vijayakumar, Kun Shao, and Laura Sevilla-Lara. In ICCV 2025. [[project page]](https://reagan1311.github.io/affgrasp)
- ForeHOI: Feed-forward 3D Object Reconstruction from Daily Hand-Object Interaction Videos - Yuantao Chen, Jiahao Chang, Chongjie Ye, Chaoran Zhang, Zhaojie Fang, Chenghong Li, and Xiaoguang Han. In CVPR 2026. [[project page]](https://tao-11-chen.github.io/projectpages/ForeHOI/) [[code]](https://github.com/Tao-11-chen/ForeHOI)
- EgoFlow: Gradient-Guided Flow Matching for Egocentric 6DoF Object Motion Generation - Abhishek Saroha, Huajian Zeng, Xingxing Zuo, Daniel Cremers, and Xi Wang. In CVPR 2026. [[project page]](https://abhi-rf.github.io/egoflow/) [[code]](https://github.com/abhi-rf/egoflow)
- ParaHome: Parameterizing Everyday Home Activities Towards 3D Generative Modeling of Human-Object Interactions - Jeonghwan Kim, Jisoo Kim, Jeonghyeon Na, and Hanbyul Joo. In CVPR 2025. [[code]](https://github.com/canoneod/ParaHome)
- Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision - Tomoya Yoshida, Shuhei Kurita, Taichi Nishimura, and Shinsuke Mori. In CVPR 2025. [[project page]](https://biscue5.github.io/egoscaler-project-page/) [[code]](https://github.com/Biscue5/EgoScaler)
- ANNEXE: Unified Analyzing, Answering, and Pixel Grounding for Egocentric Interaction - Yuejiao Su, Yi Wang, Qiongyang Hu, Chuang Yang, and Lap-Pui Chau. In CVPR 2025. [[project page]](https://yuggiehk.github.io/annexe/)
- EgoChoir: Capturing 3D Human-Object Interaction Regions from Egocentric Views - Yuhang Yang, Wei Zhai, Chengfeng Wang, Chengjun Yu, Yang Cao, and Zheng-Jun Zha. In NeurIPS 2024. [[project page]](https://yyvhang.github.io/EgoChoir/) [[code]](https://github.com/yyvhang/EgoChoirrelease)
- Are Synthetic Data Useful for Egocentric Hand-Object Interaction Detection? - Rosario Leonardi, Antonino Furnari, Francesco Ragusa, and Giovanni Maria Farinella. In ECCV 2024. [[project page]](https://fpv-iplab.github.io/HOI-Synth/) [[code]](https://github.com/fpv-iplab/HOI-Synth)
- Fine-grained Affordance Annotation for Egocentric Hand-Object Interaction Videos - Zecheng Yu, Yifei Huang, Ryosuke Furuta, Takuma Yagi, Yusuke Goutsu, and Yoichi Sato. In WACV 2023.
- EgoPCA: A New Framework for Egocentric Hand-Object Interaction Understanding - Yue Xu, Yong-Lu Li, Zhemin Huang, Michael Xu Liu, Cewu Lu, Yu-Wing Tai, and Chi-Keung Tang. In ICCV 2023. [[project page]](https://mvig-rhos.com/egopca)
- ARCTIC: A Dataset for Dexterous Bimanual Hand-Object Manipulation - Zicong Fan, Omid Taheri, Dimitrios Tzionas, Muhammed Kocabas, Manuel Kaufmann, Michael J. Black, Otmar Hilliges. In CVPR 2023. [[code]](https://github.com/zc-alexfan/arctic)
- Fine-Grained Egocentric Hand-Object Segmentation: Dataset, Model, and Applications - Lingzhi Zhang, Shenghao Zhou, Simon Stent, Jianbo Shi. In ECCV 2022. [[project page]](https://web.archive.org/web/20230422000343/https://www.seas.upenn.edu/~shzhou2/projects/eosdataset/) [[code]](https://github.com/owenzlz/EgoHOS) [[dataset]](https://github.com/owenzlz/EgoHOS)
- HOI4D: A 4D Egocentric Dataset for Category-Level Human-Object Interaction - Yunze Liu, Yun Liu, Che Jiang, Kangbo Lyu, Weikang Wan, Hao Shen, Boqiang Liang, Zhoujie Fu, He Wang, Li Yi. In CVPR 2022. [[project page]](https://hoi4d.github.io/) [[video]](https://youtu.be/yzNqm0JISU0)
- Hand-Object Contact Prediction via Motion-Based Pseudo-Labeling and Guided Progressive Label Correction - Takuma Yagi, Md Tasnimul Hasan, and Yoichi Sato. In BMVC 2021. [[project page]](https://www.bmvc2021-virtualconference.com/conference/papers/paper0096.html) [[code]](https://github.com/takumayagi/handobjectcontact_prediction/)
- The MECCANO Dataset: Understanding Human-Object Interactions from Egocentric Videos in an Industrial-like Domain - Francesco Ragusa, Antonino Furnari, Salvatore Livatino, and Giovanni Maria Farinella. In WACV 2021. [[project page]](https://iplab.dmi.unict.it/MECCANO/)
- Forecasting Human-Object Interaction: Joint Prediction of Motor Attention and Actions in First Person Video - Miao Liu, Siyu Tang, Yin Li, and James M. Rehg. In ECCV 2020. [[project page]](https://aptx4869lm.github.io/ForecastingHOI/)
- You-Do, I-Learn: Discovering Task Relevant Objects and their Modes of Interaction from Multi-User Egocentric Video - Dima Damen, Tessid Leelasawassuk, Osian Haines, Andrew Calway,and Walterio Mayol-Cuevas. In BMVC 2014 [[project page]](http://www.bmva.org/bmvc/2014/papers/paper059/index.html)
- Automated capture and delivery of assistive task guidance with an eyewear computer: the GlaciAR system - Teesid Leelasawassuk, Dima Damen, and Walterio Mayol-Cuevas. In Augmented Human International Conference, ACM 2017.
Temporal Boundary Detection
Show papers (4)
- Streaming Detection of Queried Event Start - Cristóbal Eyzaguirre, Eric Tang, Shyamal Buch, Adrien Gaidon, Jiajun Wu, and Juan Carlos Niebles. In NeurIPS 2024. [[project page]](https://sdqesdataset.github.io/)
- Ego-Only: Egocentric Action Detection without Exocentric Transferring - Huiyu Wang, Mitesh Kumar Singh, and Lorenzo Torresani. In ICCV 2023.
- Trespassing the Boundaries: Labeling Temporal Bounds for Object Interactions in Egocentric Video - Davide Moltisanti, Michael Wray, Walterio Mayol-Cuevas, and Dima Damen. In ICCV 2017.
- Temporal segmentation of egocentric videos -Yair Poleg, Chetan Arora, and Shmuel Peleg. In CVPR 2014.
Privacy in Egocentric Videos
Show papers (4)
🔗 More in this category