Artwork

Inhoud geleverd door Vanessa Sochat. Alle podcastinhoud, inclusief afleveringen, afbeeldingen en podcastbeschrijvingen, wordt rechtstreeks geüpload en geleverd door Vanessa Sochat of hun podcastplatformpartner. Als u denkt dat iemand uw auteursrechtelijk beschermde werk zonder uw toestemming gebruikt, kunt u het hier beschreven proces https://nl.player.fm/legal volgen.
Player FM - Podcast-app
Ga offline met de app Player FM !

Joys and Challenges with Big Research Data

34:01
 
Delen
 

Manage episode 343889236 series 2556771
Inhoud geleverd door Vanessa Sochat. Alle podcastinhoud, inclusief afleveringen, afbeeldingen en podcastbeschrijvingen, wordt rechtstreeks geüpload en geleverd door Vanessa Sochat of hun podcastplatformpartner. Als u denkt dat iemand uw auteursrechtelijk beschermde werk zonder uw toestemming gebruikt, kunt u het hier beschreven proces https://nl.player.fm/legal volgen.

Ana Trisovic is a Research Associate at Harvard School of Public Health and a Sloan Fellow at the Institute for Quantitative Social Science. Effectively, she does data engineering for her research group and works on reproducible data and software dissemination. First, Ana speaks of her background, from her first job at Microsoft Development Center Serbia, to CERN, UChicago, and Harvard. She shares what inspired her to pursue projects relating to open-source software, open data, and open science. Her work focuses on big data workflows and research reproducibility, and she shares her experiences working with particle physics experimental data, geospatial and climate data, and sensitive medical data. As a member of Consortium of Scientific Software Registries and Repositories (SciCodes), she contributes to research data and software sharing and preservation efforts. Her study shows that research software and code scripts are frequently shared with data, and she is working on better supporting those in the Dataverse data repository. We discuss data engineering roles in the broader RSE scope and recognize them as undervalued yet critical for research groups working on secondary data analysis. Ana speaks of the joys and challenges of working with diverse datasets and the value of open-source software, reusable data workflows, and adequate documentation. She shares recommendations for publishing research data with software and emphasizes the role of data repositories. We end the conversion with community engagement topic ideas.

  continue reading

144 afleveringen

Artwork
iconDelen
 
Manage episode 343889236 series 2556771
Inhoud geleverd door Vanessa Sochat. Alle podcastinhoud, inclusief afleveringen, afbeeldingen en podcastbeschrijvingen, wordt rechtstreeks geüpload en geleverd door Vanessa Sochat of hun podcastplatformpartner. Als u denkt dat iemand uw auteursrechtelijk beschermde werk zonder uw toestemming gebruikt, kunt u het hier beschreven proces https://nl.player.fm/legal volgen.

Ana Trisovic is a Research Associate at Harvard School of Public Health and a Sloan Fellow at the Institute for Quantitative Social Science. Effectively, she does data engineering for her research group and works on reproducible data and software dissemination. First, Ana speaks of her background, from her first job at Microsoft Development Center Serbia, to CERN, UChicago, and Harvard. She shares what inspired her to pursue projects relating to open-source software, open data, and open science. Her work focuses on big data workflows and research reproducibility, and she shares her experiences working with particle physics experimental data, geospatial and climate data, and sensitive medical data. As a member of Consortium of Scientific Software Registries and Repositories (SciCodes), she contributes to research data and software sharing and preservation efforts. Her study shows that research software and code scripts are frequently shared with data, and she is working on better supporting those in the Dataverse data repository. We discuss data engineering roles in the broader RSE scope and recognize them as undervalued yet critical for research groups working on secondary data analysis. Ana speaks of the joys and challenges of working with diverse datasets and the value of open-source software, reusable data workflows, and adequate documentation. She shares recommendations for publishing research data with software and emphasizes the role of data repositories. We end the conversion with community engagement topic ideas.

  continue reading

144 afleveringen

Todos os episódios

×
 
Loading …

Welkom op Player FM!

Player FM scant het web op podcasts van hoge kwaliteit waarvan u nu kunt genieten. Het is de beste podcast-app en werkt op Android, iPhone en internet. Aanmelden om abonnementen op verschillende apparaten te synchroniseren.

 

Korte handleiding