Report proposes standards for sharing data and code used in computational studies

Reporting new research results involves detailed descriptions of methods and materials used in an experiment. But when a study uses computers to analyze data, create models or simulate things that can’t be tested in a lab, how can other researchers see what steps were taken or potentially reproduce results?

A new report by prominent leaders in computational methods and reproducibility lays out recommendations for ways researchers, institutions, agencies and journal publishers can work together to standardize sharing of data sets and software code. The paper "Enhancing reproducibility for computational methods" appears in the journal Science.

"We have a real issue in disclosure and reporting standards for research that involves computation – which is basically all research today," said Victoria Stodden, a University of Illinois professor of information science and the lead author of the paper. "The standards for putting enough information out there with your findings so that other researchers in the area are able to understand and potentially replicate your work were developed before we used computers."

[video:https://youtu.be/94qM6tnDtcQ]

"It is becoming increasingly accepted for researchers to value open data standards as an essential part of modern scholarship, but it is nearly impossible to reproduce results from original data without the authors' code," said Marcia McNutt, the president of the National Academy of Sciences and a co-corresponding author of the study. "This policy forum makes recommendations to enable practical and useful code sharing."

Sharing complete computational methods – data, code, parameters and the specific steps taken to arrive at the results – is difficult for researchers because there are no standards or guides to refer to, Stodden said. It's an extra step for busy researchers to incorporate into their reporting routine, and even if someone wants to share their data or code, there are questions of how to format and document it, where to store it and how to make it accessible.

The report makes seven specific recommendations, such as documenting digital objects and making them retrievable, open licensing, placing links to datasets and workflows in scientific articles, and reproducibility checks before publication in a scholarly journal.

The authors hope that disclosing computational methods will not only allow other researchers to verify and reproduce results, but also to build upon studies that have been done, such as performing different analyses with a dataset or using an established workflow with new data.

"Things like how you prepped your data – what you did with outliers or how you normalized variables, all the things that are standard in data analysis – can make a big impact on results," Stodden said. "Some researchers make code and data accessible on point of principle, so it's possible. But it takes time. We know it's hard, but in this report we're trying to say in a very productive and positive way that data, code and workflows need to be part of what gets disclosed as a scientific finding."

Research Areas:
Tags:
Updated on
Backto the news archive

Related News

iSchool represented at Charleston Conference

iSchool adjunct and affiliate faculty will participate in virtual and in-person sessions of the 2024 Charleston Conference. The conference is an annual gathering that draws librarians, publishers, vendors, and others to discuss issues relating to the acquisition and publication of books and serials. 

Schneider group to present at ASIS&T workshop

Members of Associate Professor Jodi Schneider’s group will present their research at the Association for Information Science and Technology (ASIS&T) Workshop on Informetric, Scientometric, and Scientific and Technical Information Research, which will be held virtually on November 6 and 13. The MET-STI 2024 Workshop is collaboratively hosted by the Special Interest Group for Metrics (SIG-MET) and Special Interest Group for Scientific and Technical Information (SIG-STI) of ASIS&T.

Jodi Schneider

Wong co-edits new edition of Reference and Information Services

Adjunct Lecturer Melissa Wong (MSLIS '94) and Laura Saunders, professor of library and information science at Simmons University, are the co-editors of Reference and Information Services: An Introduction, Seventh Edition, which was recently published by Bloomsbury Libraries Unlimited. The textbook provides a comprehensive update to the previous edition, also co-edited by Wong and Saunders, and serves as an essential resource for LIS students and practitioners alike.

Melissa Wong

iSchool researchers to present at ASSETS 2024

iSchool faculty and students will present their research at the 26th International Association for Computing Machinery (ACM) Special Interest Group (SIG) ACCESS Conference on Computers and Accessibility (ASSETS 2024), which will be held on October 28-30 in St. John's, Newfoundland and Labrador, Canada. The conference is the premier forum for presenting research on design, evaluation, use, and education related to computing for people with disabilities and older adults.

iSchool well represented at ASIS&T 2024

iSchool faculty, staff, and students will participate in the 87th Annual Meeting of the Association for Information Science and Technology (ASIS&T), which will be held on October 25-29 in Calgary, Canada. The theme of this year's conference is "Putting People First: Responsibility, Reciprocity, and Care in Information Research and Practice." The meeting is the premier international conference dedicated to the study of information, people, and technology in contemporary society.

iSchool Building