{"url":"/dataset/stanford-ecm","name":"Stanford-ECM","full_name":"Stanford-ECM","description_markdown":"**Stanford-ECM** is an egocentric multimodal dataset which comprises about 27 hours of egocentric video augmented with heart rate and acceleration data. The lengths of the individual videos cover a diverse range from 3 minutes to about 51 minutes in length. A mobile phone was used to collect egocentric video at 720x1280 resolution and 30 fps, as well as triaxial acceleration at 30Hz. The mobile phone was equipped with a wide-angle lens, so that the horizontal field of view was enlarged from 45 degrees to about 64 degrees. A wrist-worn heart rate sensor was used to capture the heart rate every 5 seconds. The phone and heart rate monitor was time-synchronized through Bluetooth, and all data was stored in the phone’s storage. Piecewise cubic polynomial interpolation was used to fill in any gaps in heart rate data. Finally, data was aligned to the millisecond level at 30 Hz.\n\nSource: [http://ai.stanford.edu/~syyeung/ecm_dataset/egocentric_multimodal.html](http://ai.stanford.edu/~syyeung/ecm_dataset/egocentric_multimodal.html)\nImage Source: [http://ai.stanford.edu/~syyeung/ecm_dataset/egocentric_multimodal.html](http://ai.stanford.edu/~syyeung/ecm_dataset/egocentric_multimodal.html)","description_withheld":null,"homepage":"http://ai.stanford.edu/~syyeung/ecm_dataset/egocentric_multimodal.html","introduced_date":"2017-01-01","introduced_date_note":null,"introduced_by":{"paper":"/paper/jointly-learning-energy-expenditures-and","title":"Jointly Learning Energy Expenditures and Activities Using Egocentric Multimodal Signals","first_author":"Katsuyuki Nakamura","url":null},"license":null,"modalities":[{"name":"Videos","url":"/datasets/modality/videos"},{"name":"Audio","url":"/datasets/modality/audio"}],"tasks":[{"name":"Scene Understanding","url":"/task/scene-understanding","datasets_with_task":"/datasets/task/scene-understanding"},{"name":"Video Understanding","url":"/task/video-understanding","datasets_with_task":"/datasets/task/video-understanding"}],"languages":[],"variants":["Stanford-ECM"],"data_loaders":[],"num_papers_in_archive":2,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}