Πηγαίνετε εκτός σύνδεσης με την εφαρμογή Player FM !
Joys and Challenges with Big Research Data
Manage episode 343889236 series 2556771
Ana Trisovic is a Research Associate at Harvard School of Public Health and a Sloan Fellow at the Institute for Quantitative Social Science. Effectively, she does data engineering for her research group and works on reproducible data and software dissemination. First, Ana speaks of her background, from her first job at Microsoft Development Center Serbia, to CERN, UChicago, and Harvard. She shares what inspired her to pursue projects relating to open-source software, open data, and open science. Her work focuses on big data workflows and research reproducibility, and she shares her experiences working with particle physics experimental data, geospatial and climate data, and sensitive medical data. As a member of Consortium of Scientific Software Registries and Repositories (SciCodes), she contributes to research data and software sharing and preservation efforts. Her study shows that research software and code scripts are frequently shared with data, and she is working on better supporting those in the Dataverse data repository. We discuss data engineering roles in the broader RSE scope and recognize them as undervalued yet critical for research groups working on secondary data analysis. Ana speaks of the joys and challenges of working with diverse datasets and the value of open-source software, reusable data workflows, and adequate documentation. She shares recommendations for publishing research data with software and emphasizes the role of data repositories. We end the conversion with community engagement topic ideas.
144 επεισόδια
Manage episode 343889236 series 2556771
Ana Trisovic is a Research Associate at Harvard School of Public Health and a Sloan Fellow at the Institute for Quantitative Social Science. Effectively, she does data engineering for her research group and works on reproducible data and software dissemination. First, Ana speaks of her background, from her first job at Microsoft Development Center Serbia, to CERN, UChicago, and Harvard. She shares what inspired her to pursue projects relating to open-source software, open data, and open science. Her work focuses on big data workflows and research reproducibility, and she shares her experiences working with particle physics experimental data, geospatial and climate data, and sensitive medical data. As a member of Consortium of Scientific Software Registries and Repositories (SciCodes), she contributes to research data and software sharing and preservation efforts. Her study shows that research software and code scripts are frequently shared with data, and she is working on better supporting those in the Dataverse data repository. We discuss data engineering roles in the broader RSE scope and recognize them as undervalued yet critical for research groups working on secondary data analysis. Ana speaks of the joys and challenges of working with diverse datasets and the value of open-source software, reusable data workflows, and adequate documentation. She shares recommendations for publishing research data with software and emphasizes the role of data repositories. We end the conversion with community engagement topic ideas.
144 επεισόδια
Όλα τα επεισόδια
×Καλώς ήλθατε στο Player FM!
Το FM Player σαρώνει τον ιστό για podcasts υψηλής ποιότητας για να απολαύσετε αυτή τη στιγμή. Είναι η καλύτερη εφαρμογή podcast και λειτουργεί σε Android, iPhone και στον ιστό. Εγγραφή για συγχρονισμό συνδρομών σε όλες τις συσκευές.