Your browser doesn't support javascript.
loading
BD2K Training Coordinating Center's ERuDIte: the Educational Resource Discovery Index for Data Science.
Ambite, José Luis; Fierro, Lily; Gordon, Jonathan; Burns, Gully A; Geigl, Florian; Lerman, Kristina; Van Horn, John D.
Affiliation
  • Ambite JL; University of Southern California's Information Sciences Institute (ISI), Marina del Rey, CA 90292.
  • Fierro L; University of Southern California's Information Sciences Institute (ISI), Marina del Rey, CA 90292.
  • Gordon J; University of Southern California's Information Sciences Institute (ISI), Marina del Rey, CA 90292.
  • Burns GA; University of Southern California's Information Sciences Institute (ISI), Marina del Rey, CA 90292.
  • Geigl F; performed as a visiting Ph.D. student at ISI.
  • Lerman K; University of Southern California's Information Sciences Institute (ISI), Marina del Rey, CA 90292.
  • Van Horn JD; University of SouthernCalifornia's Stevens Neuroimaging and Informatics Institute, Los Angeles, CA 90033.
IEEE Trans Emerg Top Comput ; 9(1): 316-328, 2021.
Article in En | MEDLINE | ID: mdl-35548703
ABSTRACT
Data science is a field that has developed to enable efficient integration and analysis of increasingly large data sets in many domains. In particular, big data in genetics, neuroimaging, mobile health, and other subfields of biomedical science, promises new insights, but also poses challenges. To address these challenges, the National Institutes of Health launched the Big Data to Knowledge (BD2K) initiative, including a Training Coordinating Center (TCC) tasked with developing a resource for personalized data science training for biomedical researchers. The BD2K TCC web portal is powered by ERuDIte, the Educational Resource Discovery Index, which collects training resources for data science, including online courses, videos of tutorials and research talks, textbooks, and other web-based materials. While the availability of so many potential learning resources is exciting, they are highly heterogeneous in quality, difficulty, format, and topic, making the field intimidating to enter and difficult to navigate. Moreover, data science is rapidly evolving, so there is a constant influx of new materials and concepts. We leverage data science techniques to build ERuDIte itself, using data extraction, data integration, machine learning, information retrieval, and natural language processing to automatically collect, integrate, describe, and organize existing online resources for learning data science.
Key words

Full text: 1 Collection: 01-internacional Database: MEDLINE Language: En Journal: IEEE Trans Emerg Top Comput Year: 2021 Document type: Article

Full text: 1 Collection: 01-internacional Database: MEDLINE Language: En Journal: IEEE Trans Emerg Top Comput Year: 2021 Document type: Article
...