CaptainCook4D: A Dataset for Understanding Errors in Procedural Activities
Abstract
Following step-by-step procedures is an essential component of various activities carried out by individuals in their daily lives. These procedures serve as a guiding framework that helps to achieve goals efficiently, whether it is assembling furniture or preparing a recipe. However, the complexity and duration of procedural activities inherently increase the likelihood of making errors. Understanding such procedural activities from a sequence of frames is a challenging task that demands an accurate interpretation of visual information and the ability to reason about the structure of the activity. To this end, we collect a new egocentric 4D dataset, CaptainCook4D, comprising 384 recordings (94.5 hours) of people performing recipes in real kitchen environments. This dataset consists of two distinct types of activity: one in which participants adhere to the provided recipe instructions and another in which they deviate and induce errors. We provide 5.3K step annotations and 10K fine-grained action annotations and benchmark the dataset for the following tasks: error recognition, multistep localization and procedure learning.
Cite
Text
Peddi et al. "CaptainCook4D: A Dataset for Understanding Errors in Procedural Activities." Neural Information Processing Systems, 2024. doi:10.52202/079017-4307Markdown
[Peddi et al. "CaptainCook4D: A Dataset for Understanding Errors in Procedural Activities." Neural Information Processing Systems, 2024.](https://mlanthology.org/neurips/2024/peddi2024neurips-captaincook4d/) doi:10.52202/079017-4307BibTeX
@inproceedings{peddi2024neurips-captaincook4d,
title = {{CaptainCook4D: A Dataset for Understanding Errors in Procedural Activities}},
author = {Peddi, Rohith and Arya, Shivvrat and Challa, Bharath and Pallapothula, Likhitha and Vyas, Akshay and Gouripeddi, Bhavya and Zhang, Qifan and Wang, Jikai and Komaragiri, Vasundhara and Ragan, Eric and Ruozzi, Nicholas and Xiang, Yu and Gogate, Vibhav},
booktitle = {Neural Information Processing Systems},
year = {2024},
doi = {10.52202/079017-4307},
url = {https://mlanthology.org/neurips/2024/peddi2024neurips-captaincook4d/}
}