{"url":"/dataset/smarty4covid","name":"Smarty4covid","full_name":"The smarty4covid dataset and knowledge base: a framework enabling interpretable analysis of audio signals","description_markdown":"Harnessing the power of Artificial Intelligence (AI) and m-health towards detecting new bio-markers indicative of the onset\r\nand progress of respiratory abnormalities/conditions has greatly attracted the scientific and research interest especially during\r\nCOVID-19 pandemic. The smarty4covid dataset contains audio signals of cough (4,676), regular breathing (4,665), deep\r\nbreathing (4,695) and voice (4,291) as recorded by means of mobile devices following a crowd-sourcing approach. Other self\r\nreported information is also included (e.g. COVID-19 virus tests), thus providing a comprehensive dataset for the development\r\nof COVID-19 risk detection models. The smarty4covid dataset is released in the form of a web-ontology language (OWL)\r\nknowledge base enabling data consolidation from other relevant datasets, complex queries and reasoning. It has been utilized\r\ntowards the development of models able to: (i) extract clinically informative respiratory indicators from regular breathing\r\nrecords, and (ii) identify cough, breath and voice segments in crowd-sourced audio recordings. A new framework utilizing\r\nthe smarty4covid OWL knowledge base towards generating counterfactual explanations in opaque AI-based COVID-19 risk\r\ndetection models is proposed and validated.","description_withheld":null,"homepage":"https://arxiv.org/pdf/2307.05096.pdf","introduced_date":null,"introduced_date_note":null,"introduced_by":null,"license":null,"modalities":[],"tasks":[],"languages":[],"variants":["Smarty4covid"],"data_loaders":[],"num_papers_in_archive":1,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}