Sharing Big Video Data: Ethics, Methods, and Technology

Saved in:
Bibliographic Details
Title: Sharing Big Video Data: Ethics, Methods, and Technology
Language: English
Authors: Joanne W. Golann (ORCID 0000-0001-9337-4674), Lori Bougher, Richard Hall, Thomas J. Espenshade
Source: Sociological Methods & Research. 2026 55(1):340-372.
Availability: SAGE Publications. 2455 Teller Road, Thousand Oaks, CA 91320. Tel: 800-818-7243; Tel: 805-499-9774; Fax: 800-583-2665; e-mail: journals@sagepub.com; Web site: https://sagepub.com
Peer Reviewed: Y
Page Count: 33
Publication Date: 2026
Sponsoring Agency: National Science Foundation (NSF)
Contract Number: 2214309
Document Type: Journal Articles
Reports - Research
Descriptors: Video Technology, Databases, Information Management, Information Security, Information Storage, Access to Information, Technological Advancement, Data Use
Geographic Terms: New Jersey
DOI: 10.1177/00491241241277524
ISSN: 0049-1241
1552-8294
Abstract: Data sharing and transparency are becoming more common across the social sciences. In this article, we provide an overview of ethical, methodological, and technological considerations and challenges when developing large video-based datasets intended to be shared across researchers. We cover data security, storage, and access as well as data documentation, tagging, and transcription. Our discussions are framed by our own efforts to create a secure and user-friendly database for the New Jersey Families Study, a two-week, in-home video study of 21 families with a 2- to 4-year-old child. In collecting over 11,470 hours of video data, the New Jersey Families Study is one of the very few large-scale video projects in the field of sociology. This project has provided us with a unique opportunity to explore video data management and data sharing techniques, particularly in light of a host of cutting-edge developments in data science.
Abstractor: As Provided
Entry Date: 2026
Accession Number: EJ1496175
Database: ERIC
Full text is not displayed to guests.
FullText Links:
  – Type: pdflink
    Url: https://content.ebscohost.com/cds/retrieve?content=AQICAHj0k_4E0hTGH8RJwT4gCJyBsGNe_WN95AvKlDbXJGqwxwHXOHmE6x6uE2mBFgoFYOCGAAAA4zCB4AYJKoZIhvcNAQcGoIHSMIHPAgEAMIHJBgkqhkiG9w0BBwEwHgYJYIZIAWUDBAEuMBEEDD7FesU055hEJaJzpQIBEICBmzHP_W1JrLxhXNR6Hf_VMSY-EugM4JDWU4v5X5glZbQhA4RTTTM93UrppKVb-BIXLgXkn1c72tK-39ho_QRmdkDNdA5Ye9He3Uf4ASos-guBb9nFzIhUGorlQUKzsERhklgQyVdgF7hM-mbwBz-NGludzzofG5Rk9HO4n5yFcO62YSZDQmTeZgno4JJmlWlA1VmMoLoVTACAwjI0
Text:
  Availability: 1
  Value: <anid>AN0190929153;som01feb.26;2026Jan20.00:54;v2.2.500</anid> <title id="AN0190929153-1">Sharing Big Video Data: Ethics, Methods, and Technology </title> <p>Data sharing and transparency are becoming more common across the social sciences. In this article, we provide an overview of ethical, methodological, and technological considerations and challenges when developing large video-based datasets intended to be shared across researchers. We cover data security, storage, and access as well as data documentation, tagging, and transcription. Our discussions are framed by our own efforts to create a secure and user-friendly database for the New Jersey Families Study, a two-week, in-home video study of 21 families with a 2- to 4-year-old child. In collecting over 11,470 hours of video data, the New Jersey Families Study is one of the very few large-scale video projects in the field of sociology. This project has provided us with a unique opportunity to explore video data management and data sharing techniques, particularly in light of a host of cutting-edge developments in data science.</p> <p>Keywords: video; methodology; data sharing; big data; qualitative</p> <p>In 1967, the sociologist Frederick Erickson conducted his first video study of small-group discussions using a camera that weighed 25 pounds and a recording reel measuring nearly 16 inches in diameter. His study of classroom interactions in the 1970s made use of a portable camera and a wireless microphone that was so expensive that it was shared between Erickson's project in Boston and the sociologist Hugh Mehan's in San Diego. To transport the microphone between the two teams, researchers received permission from an airline to keep the microphone in the pilot's cockpit and graduate students dropped it off and picked it up at the airports ([<reflink idref="bib26" id="ref1">26</reflink>]). Research using these technologies has advanced substantially since those early days of audio and video recording. In our study of how parents support their children's early learning, for example, we installed in families' homes up to eight small, unobtrusive cameras, each weighing 0.55 pounds and equipped with high-definition video and built-in audio.</p> <p>The accessibility and ubiquity of video recording devices today open new opportunities for the field of sociology. Video data are well suited for studying topics as diverse as crimes, protests, communication, clinician-patient interactions, learning, linguistics, and family life (e.g., [<reflink idref="bib17" id="ref2">17</reflink>]; [<reflink idref="bib65" id="ref3">65</reflink>]; [<reflink idref="bib78" id="ref4">78</reflink>]; [<reflink idref="bib89" id="ref5">89</reflink>]; [<reflink idref="bib104" id="ref6">104</reflink>]; [<reflink idref="bib106" id="ref7">106</reflink>]; [<reflink idref="bib113" id="ref8">113</reflink>]). However, except in certain subfields like conversation or interaction analysis where video is a central mode of data collection ([<reflink idref="bib27" id="ref9">27</reflink>]; [<reflink idref="bib105" id="ref10">105</reflink>]), sociology has been slower than such other disciplines as psychology, anthropology, and learning sciences to embrace these methods. Recently, sociologists have begun to use computational methods to analyze visual data sourced from social media, police body cameras, and news footage ([<reflink idref="bib70" id="ref11">70</reflink>]; [<reflink idref="bib79" id="ref12">79</reflink>]; [<reflink idref="bib110" id="ref13">110</reflink>]; [<reflink idref="bib119" id="ref14">119</reflink>]; [<reflink idref="bib120" id="ref15">120</reflink>]), highlighting the potential of big video data. Reflecting a growing interest in video methods, <emph>Sociological Research and Methods</emph> published a special issue in August 2023 on "The Present and Future of Video-based Social Science Research" (Volume 52, Issue 3).</p> <p>Big video data, while promising, introduce unique challenges related to size and sensitivity. In this article, we offer sociologists practical guidance for working with massive amounts of video data based on our review of existing literature and our experiences with the New Jersey Families Study, a large-scale original video study in which we collected 11,470 h of video footage from inside the homes of 21 families with young children over a two-week period. Earlier guidance on video data collection—from starting out to subject recruitment, and to the selection of technology—was presented by [<reflink idref="bib36" id="ref16">36</reflink>]. Here, we focus on the intermediate step between data collection and analysis: data preparation. Specifically, we discuss important considerations around how to share video data with a wider research community. We divide our discussion into two parts: leveraging research infrastructure for secure data sharing and constructing a user-friendly interface. In the first section, we cover data classification, storage, and access. In the second, we discuss data documentation, cleaning, and navigation. Interwoven through the more practical considerations are issues related to the ethics of data sharing. Our aim is broad in scope, so rather than take a deep dive into a single issue, we offer researchers a roadmap for the kinds of questions and considerations they might encounter when preparing a large video dataset for analysis.</p> <p>This article provides an overview of ethical, methodological, and technological considerations and challenges when developing large video-based datasets intended to be used by multiple researchers. The detailed and self-documenting nature of video makes it particularly well suited for data sharing, but video data are rarely shared (Gilmore et al. 2016b). As data sharing becomes more common and expected across social sciences ([<reflink idref="bib25" id="ref17">25</reflink>]; [<reflink idref="bib12" id="ref18">12</reflink>]), researchers increasingly will need to take into account how they plan to share their data. The lessons we have learned can be translated to researchers working with different types of video data as well as those interested in sharing other kinds of sensitive data like ethnographic or qualitative data.</p> <hd id="AN0190929153-2">The New Jersey Families Study</hd> <p>The New Jersey Families Study is a video-ethnographic examination of how families support their children's early learning. We accumulated 11,470 h of video footage from 21 families in New Jersey that agreed to a two-week naturalistic observation of their daily lives, behaviors, and activities while interacting with their children at home. Each family had a child aged two to four. Otherwise, the final sample is highly diverse in terms of race and ethnicity, social class, family structure, and place of residence. As many as eight high-definition video cameras with microphones were located strategically in up to four rooms in participants' homes—rooms where most parent-child interactions occur. Cameras with motion sensors were activated simultaneously and continuously for two weeks. Survey and interview data collected during six additional points of contact with families supplement the video data.</p> <p>The New Jersey Families Study was motivated by the growing recognition of the significance of the early years in children's development ([<reflink idref="bib42" id="ref19">42</reflink>]; [<reflink idref="bib61" id="ref20">61</reflink>]), the power of parenting in shaping children's outcomes ([<reflink idref="bib7" id="ref21">7</reflink>]), and the need for more observational data on children's environments and early experiences at home. There have been few intensive observational studies of families' home lives in the field of sociology since [<reflink idref="bib59" id="ref22">59</reflink>] ground-breaking work, which is now 25 years old. Homes can be difficult for researchers to access, and video can enable studies of otherwise private domains. Additionally, standalone videos like the ones used in the New Jersey Families Study can reduce reactivity and observer bias, as they do not require the researcher to be present.[<reflink idref="bib6" id="ref23">6</reflink>] Our study is similar to The Center of Everyday Lives of Families (CELF) study, which collected over 1,500 h of in-home video footage on 32 middle-class families in the Los Angeles area in the early 2000s ([<reflink idref="bib89" id="ref24">89</reflink>]), but differs in focusing on young children and including both working-class and middle-class families. It also resembles Deb Roy's study, in which he recorded his own home for three years following the birth of his son, amassing 90,000 h of video data and 140,000 h of audio data. With these data, his team was able to show that children learn words through an accumulation of parent-child interactions grounded in specific spatial, temporal, and linguistic contexts ([<reflink idref="bib99" id="ref25">99</reflink>]).</p> <p>Video data offer several advantages over traditional observational studies for studying family life. In the fields of education and the learning sciences, video has been widely employed to study learning and socialization because of its ability to capture micro-interactions and contextual elements, which may be outside the scope of observation in real time ([<reflink idref="bib20" id="ref26">20</reflink>]). Facial expressions, gestures, intonation, timing in reaction, spacing between participants, location, and background objects are just some of the many possible detailed focal points that video data offer ([<reflink idref="bib6" id="ref27">6</reflink>]; [<reflink idref="bib79" id="ref28">79</reflink>], [<reflink idref="bib80" id="ref29">80</reflink>]). While video has found a niche in the study of micro-interactions, video data also have the potential to stimulate new and different types of analyses. The ability to share raw video data allows researchers with different perspectives to work together to interpret data and assess conclusions ([<reflink idref="bib20" id="ref30">20</reflink>]), which is especially important when studying families from diverse racial/ethnic and socioeconomic groups. Rapid advances in automated transcription and machine learning, as we will discuss, also open possibilities of using big data techniques to study family life. While video has its limitations (see [<reflink idref="bib43" id="ref31">43</reflink>]; [<reflink idref="bib93" id="ref32">93</reflink>]), it offers exciting possibilities for deepening our understanding of family dynamics and children's socialization experiences.</p> <p>The New Jersey Families Study is different and more challenging than typical video studies in its magnitude of data and the sensitivity of the data collected. Our experiences working with a staggering amount of data can inform efforts to integrate big data techniques into data preparation. The private nature of our data—recording the intimate domain of the home, typically a highly restricted setting—also elicits important ethical considerations that can help guide future researchers working with sensitive video data. In the next sections, we walk through key ethical, methodological, and technical considerations that underlie the sharing of large or sensitive video data. Our discussions are framed by our own efforts to share video data from the New Jersey Families Study, but we review the state-of-the-art in different areas that can also assist researchers working with different types of video data.</p> <hd id="AN0190929153-3">Qualitative and Video Data Sharing</hd> <p>Data sharing is becoming and soon will be the norm. In 2022, the Office of Science and Technology Policy (OSTP) released the "Nelson Memo," which requires all federal agencies to develop guidelines by 2025 to make data for federally-funded research freely and publicly available ([<reflink idref="bib82" id="ref33">82</reflink>]). Making data publicly accessible is now required by a number of academic journals and granting agencies ([<reflink idref="bib34" id="ref34">34</reflink>]). The National Science Foundation (NSF) and the National Institutes of Health (NIH), for example, both require data management plans that detail how researchers will store and share their data, with the expectation that researchers will "maximize the appropriate sharing of scientific data" ([<reflink idref="bib84" id="ref35">84</reflink>]). Although data sharing in sociology has been more aspirational than normative, the American Sociological Association (ASA) encourages the sharing of data:</p> <p>As a regular practice, sociologists share data and pertinent documentation as an integral part of a research plan. Sociologists generally make their data available after completion of a project or its major publications, except where proprietary agreements with employers, contractors, or clients preclude such accessibility or when it is impossible to share data and protect the confidentiality of the research participants (e.g., field notes or detailed information from ethnographic interviews). ([<reflink idref="bib3" id="ref36">3</reflink>]: 16)</p> <p>Calls for data sharing and transparency are becoming more common across the social sciences ([<reflink idref="bib12" id="ref37">12</reflink>]; [<reflink idref="bib25" id="ref38">25</reflink>]; [<reflink idref="bib31" id="ref39">31</reflink>]; [<reflink idref="bib77" id="ref40">77</reflink>]), though the social sciences still lag behind the natural sciences in terms of making data publicly available ([<reflink idref="bib109" id="ref41">109</reflink>]).</p> <p>Although data-sharing policies typically provide exceptions for sensitive data because of privacy and confidentiality concerns, there are increasing calls for even qualitative and video researchers to make their data accessible ([<reflink idref="bib10" id="ref42">10</reflink>]; [<reflink idref="bib77" id="ref43">77</reflink>]; [<reflink idref="bib19" id="ref44">19</reflink>]). With proper consent and/or removal of personally identifiable information, risks associated with data sharing can be reduced ([<reflink idref="bib29" id="ref45">29</reflink>]; [<reflink idref="bib111" id="ref46">111</reflink>]). In the last decade, the creation of qualitative and video data repositories like Databrary and Qualitative Data Repository (QDR) have facilitated the process of data sharing, allowing researchers to more easily store and share their data. Currently, Databrary, based at New York University, has amassed over 100,000 h of video recordings, while QDR, housed at Syracuse University, has collected over 130 qualitative or multi-method datasets. International efforts to share qualitative data predate those in the U.S. and are growing ([<reflink idref="bib12" id="ref47">12</reflink>]); the UK Data Service, for example, archives nearly 9,000 qualitative or mixed-methods studies. Related to children's learning, CHILDES (Child Language Data Exchange System), based at Carnegie Mellon University, is a long-standing repository of transcripts of children's speech, many of which are linked to audio and video files ([<reflink idref="bib16" id="ref48">16</reflink>]).</p> <p>Data sharing has clear benefits for the research community and the advancement of the field ([<reflink idref="bib12" id="ref49">12</reflink>]; [<reflink idref="bib4" id="ref50">4</reflink>]; [<reflink idref="bib10" id="ref51">10</reflink>]; [<reflink idref="bib31" id="ref52">31</reflink>]; [<reflink idref="bib71" id="ref53">71</reflink>]; [<reflink idref="bib29" id="ref54">29</reflink>]; [<reflink idref="bib25" id="ref55">25</reflink>]). Data collection is time-consuming and costly, and data sharing allows data to be used by a wider set of researchers, expanding the range and depth of scholarship ([<reflink idref="bib66" id="ref56">66</reflink>]). In this way, it maximizes research participants' contributions, reduces future respondent burdens, and promotes accountability for publicly funded research. Data sharing also opens data to diverse perspectives and interpretations, which can guard against narrow or culturally insensitive interpretations. In fact, video researchers often participate in "collaboratories" where groups of researchers come together to collectively view and analyze data ([<reflink idref="bib37" id="ref57">37</reflink>]). With concerns growing over the validity and replicability of research findings, data sharing also increases transparency, permitting others to verify and reproduce research studies ([<reflink idref="bib30" id="ref58">30</reflink>]). Finally, data sharing can have pedagogical uses, as video often is used for training and instructional purposes ([<reflink idref="bib103" id="ref59">103</reflink>]) and students can benefit from working on projects that use real and well-documented data.</p> <p>At the same time, there are unique challenges to sharing qualitative and video data, particularly related to ethics, methods, and technology ([<reflink idref="bib4" id="ref60">4</reflink>]; [<reflink idref="bib10" id="ref61">10</reflink>]; [<reflink idref="bib41" id="ref62">41</reflink>]; [<reflink idref="bib50" id="ref63">50</reflink>]; [<reflink idref="bib71" id="ref64">71</reflink>]). Ethical concerns include violations of confidentiality and informed consent, breaches of trust with research participants, and the risk of misrepresentation ([<reflink idref="bib11" id="ref65">11</reflink>]; [<reflink idref="bib91" id="ref66">91</reflink>]; [<reflink idref="bib115" id="ref67">115</reflink>]; [<reflink idref="bib62" id="ref68">62</reflink>]). Methodological concerns relate to context and fit, that is, whether outside researchers using the data have context-specific information with which to interpret the data and whether data collected for one purpose are well suited for answering other research questions ([<reflink idref="bib51" id="ref69">51</reflink>]; [<reflink idref="bib41" id="ref70">41</reflink>]; [<reflink idref="bib32" id="ref71">32</reflink>]). Finally, sharing video data poses technical challenges because video files can be large, may contain multiple camera angles, and come in a diversity of file formats that may not be compatible with different kinds of software ([<reflink idref="bib34" id="ref72">34</reflink>]; [<reflink idref="bib19" id="ref73">19</reflink>]). Curating, managing, and storing video data requires a certain degree of technical expertise.</p> <p>These costs and benefits must be weighed when considering whether and how to share project data. Furthermore, certain data may not be appropriate for sharing, given potential social, economic, or legal harms, such as videos of natural disasters, sexual encounters, childbirth, and deviant or criminal acts ([<reflink idref="bib79" id="ref74">79</reflink>]; [<reflink idref="bib62" id="ref75">62</reflink>]). Here we offer guidance to researchers to increase the potential value of sharing data where appropriate and to mitigate potential risks.</p> <hd id="AN0190929153-4">Leveraging Infrastructure for Secure Data Sharing</hd> <p>First, we present our work in identifying a secure research infrastructure for sharing, constructing, and maintaining video data. The conflict between data privacy and the desire for wider accessibility continues to be a challenge for projects with personally identifiable data ([<reflink idref="bib21" id="ref76">21</reflink>]; [<reflink idref="bib67" id="ref77">67</reflink>]). Specifically, there is an inherent trade-off between data privacy and conducting beneficial research using sensitive data ([<reflink idref="bib21" id="ref78">21</reflink>]). As such, "reasonable" data privacy and confidentiality must be maintained while concurrently a system must be in place to grant data access to researchers who have been thoroughly vetted. Where data will be stored, how they will be protected, who will have access, and the nature of data-sharing agreements are all crucial elements of data security. Although our focus is on secure data with restricted access, the process of data classification and decisions of storage and access are relevant for all forms of data sharing.</p> <hd id="AN0190929153-5">Data Classification</hd> <p>Data classification is the placement of data into appropriate categories based on levels of sensitivity and risk of harm from disclosure. Classification determines how the data should be stored and accessed. While category names may vary across institutions, all classification systems typically share a distinction between public data and data that have increasingly sensitive levels of personally identifiable information (PII) that require increasing levels of protection. For example, Princeton University uses four data classification categories: Restricted (highest protection), Confidential (high protection), Unrestricted (medium protection), and Public (low protection) ([<reflink idref="bib97" id="ref79">97</reflink>]). Restricted data contain PII like social security numbers or protected health information, while confidential data include any data that could have an adverse effect on the individuals, university, public, or other entities. Unrestricted data can be shared freely only throughout an organization or university (e.g., training videos), while public data are free for all to access.</p> <p>Classification of the New Jersey Families Study data was relatively straightforward. Because we collected a significant amount of intimate data on just a few families with children within the privacy of their own home, our data are highly sensitive. However, video data can be complicated to classify for several reasons. First, in video data, the gap between consent and awareness may be heightened. Even if participants in a video study provide informed consent, they may lack awareness of the full video content to which they consented for collection. For example, in our study, because participants were likely to habituate to the cameras over the two weeks (see [<reflink idref="bib89" id="ref80">89</reflink>]), they are more likely to do things they did not remember or did unconsciously. They also may have engaged in behaviors that posed harms they did not consider (e.g., racist remarks that could lead to social exclusion or job loss). Informed consent does not relieve researchers of the ethical responsibility to appropriately classify their data.</p> <p>Second, it can be difficult to distinguish public from private in video data. The classification of video data as public just because it is published on a publicly accessible platform (e.g., Facebook, Instagram, TikTok, YouTube, or Twitter) poses an ethical dilemma. Some individuals may not have intended for the broad dissemination of such videos or their use for research purposes ([<reflink idref="bib14" id="ref81">14</reflink>]; [<reflink idref="bib83" id="ref82">83</reflink>]). What a person "seeks to preserve as private, even in an area accessible to the public" is considered private under United States federal law ([<reflink idref="bib52" id="ref83">52</reflink>]). This includes private conversations held in public (e.g., a restaurant), regardless of whether they are sensitive in nature or not. In turn, whatever a person "knowingly exposes to the public, even in his own home or office" is considered public (Ibid). A video filmed in one's home may therefore be public.</p> <p>For video data posted on public platforms, sometimes it is not clear if the uploader is the person in the video or if the person being filmed "knowingly" agreed to the video's posting ([<reflink idref="bib80" id="ref84">80</reflink>]). The public's general lack of awareness in big data collection and the absence of informed consent generate further ethical uncertainty ([<reflink idref="bib101" id="ref85">101</reflink>]:282). A lack of control over information flows and possible amplified exposure can pose additional risks that further complicate the classification of online "public" data ([<reflink idref="bib80" id="ref86">80</reflink>]; [<reflink idref="bib101" id="ref87">101</reflink>]). Such harm can take various forms, such as a woman getting fired after a racist rant recorded in her home gains increased exposure on TikTok or the propagation of discrimination with viral videos of looting during protests. Public opinion can help resolve ambiguity in defining privacy when no clear guidelines exist ([<reflink idref="bib38" id="ref88">38</reflink>]). Context also matters. [<reflink idref="bib86" id="ref89">86</reflink>] places evaluations of what is appropriate information sharing in a framework of contextual integrity, wherein contextual-relative informational norms prescribe the types of information to be shared, between which agents (i.e., subject of information, sender, and receiver), through which transmission principle (e.g., consent and notice vs. reciprocity), and under which contexts.</p> <p>Researchers should consider how sharing video data can create harm or unwanted exposure, and whether some groups, like children or racial minorities, are more at risk if data are shared. Additional questions regarding consent, awareness, and the definition of privacy, together with the loss of anonymity in video data, are just a few of the ethical challenges that warrant specialized rules specific to the review of video research ([<reflink idref="bib20" id="ref90">20</reflink>]:36; [<reflink idref="bib35" id="ref91">35</reflink>]).</p> <hd id="AN0190929153-6">Data Storage</hd> <p>Classifying the sensitivity of the data importantly guides both data storage and access. Those who collect video data first-hand should first abide by any data storage and protection promises specified in the informed consent agreement (if obtained). Because the New Jersey Families Study data were classified as highly sensitive, they were first stored on external hard drives in a locked file cabinet in a locked office. Access was only available through approval of the Institutional Review Board (IRB) and data were viewed under the supervision of an office manager on a computer with no internet connection. Today, the data are stored on a secure server that meets federal data security regulations. Approved users have remote access to a virtual machine where they can view video clips on their computer screens. They are unable to download, copy, or remove data from this secure environment, but they can have access to free and open-source video coding software.</p> <p>This type of physical infrastructure can be costly in administration and maintenance, and is not available at all universities. Cloud-computing also offers remote access and has become an attractive option for the storage and processing of sensitive and big data ([<reflink idref="bib28" id="ref92">28</reflink>]). The National Institute of Health recently announced funding opportunities to address complex computational and data management needs through cloud computing ([<reflink idref="bib85" id="ref93">85</reflink>]). As the world's largest public funder of biomedical research, NIH often influences the policies and practices of other federal grantmaking agencies. One caveat is that outsourcing data to cloud providers can entail a relative loss of control over secure servers ([<reflink idref="bib74" id="ref94">74</reflink>]), but the cloud can often support more computationally intensive processing, including any machine learning in the processing of video data. As we move away from the most sensitive data, additional options are available, including the use of open repositories, which accommodate both data storage needs and accessibility.</p> <hd id="AN0190929153-7">Open Data and Existing Repositories</hd> <p>The rise of digital research has strengthened the call for open data and greater transparency. In 2016, the FAIR (Findable, Accessible, Interoperable, and Reuse) Data Principles were proposed to provide general guidance to promote data sharing ([<reflink idref="bib116" id="ref95">116</reflink>]). These principles facilitate the discovery, reuse, citation, and knowledge integration of scholarly data and have been adopted by numerous research institutions and data repositories. Within this framework, data are assigned a unique identifier (e.g., a Digital Object Identifier [DOI]) upon publication and are described using rich metadata (e.g., study title, purpose, design, data format, dates, etc.). The unique identifier and metadata make data easier to find (Findability), provided they have been properly indexed (i.e., detected by a crawler of a search engine, like Google, and stored as a potential result that will be displayed after a relevant inquiry). Metadata should remain retrievable by the identifier even when the data are no longer available (Accessibility); they should contain FAIR terminology and reference other metadata when appropriate (Interoperability); and they should detail any issues related to licensing, community standards, or ownership (Reuse). Data can be posted with supplemental material, including code, analytic files, and so on.</p> <p>The FAIR principles should be viewed as a continuum that also apply to sensitive and confidential data, with some caveats ([<reflink idref="bib9" id="ref96">9</reflink>]). While sensitive data itself cannot be freely shared, the dataset should still have a persistent identifier. Metadata should still be findable and accessible, providing details on how the data can be accessed, for which purposes, and any additional conditions for accessibility. The Web of Science Data Citation Index provides a central point of access for nearly 450 repositories from around the world, more than 80 for the social sciences.[<reflink idref="bib7" id="ref97">7</reflink>] Examining usage and surveying researchers' needs further improve the utility and structure of metadata ([<reflink idref="bib54" id="ref98">54</reflink>]; [<reflink idref="bib55" id="ref99">55</reflink>]). Databrary is a notable example of a public, special-purpose repository, dedicated to housing video data ([<reflink idref="bib33" id="ref100">33</reflink>], [<reflink idref="bib34" id="ref101">34</reflink>]). Developed for researchers who study development and learning, Databrary embraces open access and cultivates a community of researchers who use video data and agree to abide by a code of conduct ([<reflink idref="bib1" id="ref102">1</reflink>]; [<reflink idref="bib31" id="ref103">31</reflink>]). The site offers easy navigation with a customized open-source tool, Datavyu, for coding, annotating, exploring, and analyzing video content. Researchers can search by keywords in video descriptions, age, and file type. Datasets are assigned a DOI and the full data citation is included on the project page. Researchers can apply to host identifiable data for free. Databrary supports various levels of data sensitivity and purposes such as teaching versus research. Explicit permission from participants and IRB approval are required.</p> <p>Videos can also be archived for dissemination with the Inter-university Consortium for Political and Social Research (ICPSR) at the University of Michigan through their virtual data enclave. It hosts large-scale video projects such as the Measures of Effective Teaching Project (MET). Similar to our system, users access restricted data in a secure environment through a remote virtual machine, where data cannot be downloaded, copied, or otherwise moved outside of the system. Dedicated platforms such as ICPSR help safeguard data against obsoletion by ensuring long-term access, systematic archival with proper versioning control, and backward compatibility in file formats. Large data centers reduce transaction costs and more effectively facilitate reproducibility ([<reflink idref="bib38" id="ref104">38</reflink>]). However, investing in a well-supported infrastructure is costly. In February 2022, the NSF awarded the University of Michigan $38 million to further enhance its research data infrastructure under the Institute for Social Research ([<reflink idref="bib87" id="ref105">87</reflink>]).</p> <p>Repositories with fewer technical safeguards that do not restrict data download rely more on "good faith" with researchers. While the use of air-gapped systems (i.e., no internet connection) and secure access to virtual machines make it more difficult to violate restrictions on sensitive data, bad actors can still film video clips with their phones, share their screens, or allow unauthorized users in the room. Restricted access and data use agreements help mitigate the threat of data breach and unauthorized access.</p> <hd id="AN0190929153-8">Data Access</hd> <p>It is important to balance risk versus utility in the release of sensitive data, adopting the most appropriate privacy safeguards, whether technical, procedural, educational, or legal ([<reflink idref="bib2" id="ref106">2</reflink>]). Fair Information Practice Principles (FIPPS) offer general guidelines for balancing security, privacy and fairness: Data should be collected with consent, individuals should be notified of research purposes and procedures, access should only be granted to those who have a demonstrated research purpose for the data (purpose justification) and only access to content that is absolutely necessary (data minimization), and only for as long as is required (data retention). Our study demonstrates a "walled garden approach" ([<reflink idref="bib101" id="ref107">101</reflink>]: 313–4) to data access, where videos are largely unaltered, maximizing the research value of rich content, but access is heavily restricted to those who are members of our multi-institutional New Jersey Families Study research team. Our current data use agreement prohibits users from sharing confidential information, disclosing video or audio data, and copying, altering, or downloading files. We also include a provision that, as specified by state law, requires data users who observe any instances of child abuse or neglect in any of the video data to report such instances to the Principal Investigator.</p> <p>Video subjects should be informed to the extent possible that their data will be shared, how it will be used, with whom it will be shared, where it will be stored, and for how long. Obtaining consent for data sharing and archiving at the time of data collection is recommended ([<reflink idref="bib11" id="ref108">11</reflink>]; [<reflink idref="bib35" id="ref109">35</reflink>]). The UK's Economic and Social Research Council, for example, states that "most data can be curated and shared ethically provided researchers pay attention right from the planning stages of research" to including consent for data sharing, anonymizing data when needed, and addressing access restrictions before starting research ([<reflink idref="bib24" id="ref110">24</reflink>]:4). QDR's Working with Sensitive Research Data (WSRD) initiative has developed sample informed consent language regarding data sharing for IRB use ([<reflink idref="bib98" id="ref111">98</reflink>]). In cases where researchers have not included data sharing in their original consent forms, they can attempt to re-consent participants. Databrary has developed and made available a set of sample templates and scripts for helping investigators seek permission to share data (https://databrary.org/support/irb.html). Although some concern has been raised about the feasibility of reconsenting, a few studies have found that rates of re-consent among research participants are high ([<reflink idref="bib58" id="ref112">58</reflink>]; [<reflink idref="bib76" id="ref113">76</reflink>]; [<reflink idref="bib112" id="ref114">112</reflink>]). An additional factor to consider and include in consent documents is whether research participants can later revoke permission to share their data. The European GDPR, for example, includes a "right to be forgotten" that allows individuals to withdraw their consent and have their personal data removed from research studies ([<reflink idref="bib95" id="ref115">95</reflink>]).</p> <p>All researchers requesting access to restricted data should be required to sign a data use agreement. At a minimum, they must agree to adhere to all of the same terms of anonymity and confidentiality documented in the original informed consent forms, hold appropriate institutional affiliation, and provide proof of human subjects research training. In addition, the agreement should include a project description with IRB approval, purpose specification for the data requested (i.e., which data and how they will be used), prohibition of disclosure to anyone not identified in the project application (new team members must submit an application), and details on data retention, disposal, or access expiration date. The necessity of these latter elements will depend on the level of security required. Relatedly, some agreements will require a secure data storage plan. Many universities offer boilerplate language for data use agreements more broadly through their legal counsel. Standardization of such agreements and intellectual property rules across institutions helps reduce transaction costs while facilitating data sharing ([<reflink idref="bib35" id="ref116">35</reflink>]; [<reflink idref="bib56" id="ref117">56</reflink>]:720).</p> <p>To access secondary, restricted data, Databrary, for example, requires user registration as well as approval by an investigator and the institution's Authorized Organizational Representative. Only then can users be added to restricted data classified under "authorized users," or videos classified as "private," which are only available to members of the research team. Users must agree to a data access agreement and a statement of rights and responsibilities.</p> <p>Replication requirements and user violations are two final considerations regarding data access. How can access for replication be effectively given, at what point in the publication process, who should have access (e.g., reviewers, editors, additional external researchers) to which data (e.g., what is the minimum necessary), and for how long? Notwithstanding the heated debate over whether replicability standards should be applied to qualitative research (Pratt et al. 2020; [<reflink idref="bib69" id="ref118">69</reflink>]), replication in qualitative and video research poses unique challenges because users need to interact with the data to reproduce results (as opposed to running a code). In their video study of how a young child learns to speak from birth to age three, [<reflink idref="bib99" id="ref119">99</reflink>] do not make the full video and audio dataset available in order to safeguard the privacy of the child and the family being recorded. Instead, they provide aggregate data about individual words available via the GitHub repository (github.com/bcroy/HSP_wordbirth).</p> <p>Secure data enclaves reduce the chance of mistakes that result in data breaches but cannot prevent malfeasance. How do we detect and sanction violations of the data use agreement? Citing the dataset's unique identifier (as earlier described under the FAIR Data Principles) would enable researchers to track the data's usage or misusage ([<reflink idref="bib92" id="ref120">92</reflink>]). Violations and any evidence of misconduct should be reported to the offender's IRB and access to the data could immediately be revoked. Because violations may be discovered only after harm has been inflicted, data use agreements could specify stricter penalties for infractions, especially if the potential for harm is high. Violations and enforcement may be less clear when data are collected from a public platform for research purposes.</p> <hd id="AN0190929153-9">Preparing a User-Friendly Dataset</hd> <p>User-friendliness is an essential priority in creating a shareable dataset. The barriers to entry should not be too high if an innovative dataset hopes to attract users ([<reflink idref="bib116" id="ref121">116</reflink>]). Several considerations should be taken into account to make video data accessible. These include data documentation, cleaning, and navigation.</p> <hd id="AN0190929153-10">Data Documentation</hd> <p>Data documentation is essential for keeping track of the data and for providing secondary users with project details and context. Data documentation involves three primary steps: data inventory, data provenance, and data governance. Data inventory involves creating an inventory of all data collected including recruitment materials, instruments, audio and visual files, and survey and interview responses. Data provenance includes explanations as to the rationale for the collection of each piece of data and how it was collected. Finally, data governance includes maintaining an awareness of where the data are located and who has access to each component of data. For the New Jersey Families Study, we created an extensive data documentation manual for the data collection process. To promote and facilitate data documentation, efforts like the Data Documentation Initiative (DDI) provide standardized tools and training for social and behavioral scientists to document their data ([<reflink idref="bib18" id="ref122">18</reflink>]). ICPSR, for example, follows DDI standards for data documentation and its study records include details such as data type, study purpose, design, data sources, sampling procedures, data format, restrictions, and version history ([<reflink idref="bib46" id="ref123">46</reflink>]).</p> <hd id="AN0190929153-11">Data Cleaning</hd> <p>Data cleaning involves any steps to get the data in the desired format prior to sharing. We focus on two aspects, privacy protection and efficiency in processing, to demonstrate how effective cleaning involves adopting the perspective of the participant and the end-user. Data cleaning can offer maximum privacy protection through anonymization. When it comes to big data, responsibility may come in the form of less, not more data ([<reflink idref="bib101" id="ref124">101</reflink>]). Faces, background setting, body movements, and voices are just a few pieces of information that could be used to identify individuals. However, even though researchers should operate with as little identifiable information as possible for the research question at hand, anonymization through the removal of these elements (and possible over-cleaning) may come at the cost of the utility of the data and may not be practical for research needs ([<reflink idref="bib29" id="ref125">29</reflink>]; [<reflink idref="bib75" id="ref126">75</reflink>]). Faces can be blurred, the speech of non-consenting participants can be muted, and voices can be altered, but this results in loss of information like facial expressions and potentially significant parts of interactions ([<reflink idref="bib29" id="ref127">29</reflink>]; [<reflink idref="bib10" id="ref128">10</reflink>]). Databrary's policy is not to alter research videos to maximize their re-use potential, but they share identifiable data only with the explicit permission of research participants (Gilmore et al. 2016a).</p> <p>Failure to adequately clean the data and a rushed release could unnecessarily increase the risk of harm, sometimes severe harm. An important part of the data cleaning process for the New Jersey Families Study is locating clips that contain images of child nudity and hiding the relevant portions of those frames as well as deleting clips that families did not want included in the data archive. The magnitude of data and the time-intensive nature of this process, combined with a high need for accuracy in correctly classifying instances of nudity, require close collaboration with computer scientists.</p> <p>In addition to ethical considerations, data cleaning should involve a practical assessment of the data's usage. This may involve removing clips that are not part of the research study, stitching together separate clips for more streamlined viewing, removing duplicate or corrupt files, reorganizing file structures, finding the optimal resolution to balance processing speed and image quality, or compressing files for a smoother user experience.</p> <hd id="AN0190929153-12">Data Navigation</hd> <p>Creating a user-friendly interface that enables users to search and filter results is critical when working with large video data collections like the New Jersey Families Study, which has 504,000 discrete clips. Suppose that a user is interested, for example, in clips where the target child is crying. It would take about 16 months of continuous viewing to wade through all of the data, and no user is likely to want to expend such time and effort. To enable users to easily search the data using keywords or filters, one needs to tag the data and create a queryable interface. Effective navigation would tell users who is in the video, what they are doing, what they are saying, where it is taking place, when, and more. In the New Jersey Families Study, we plan to tag each video clip with metadata related to household characteristics, video recording parameters (namely, camera numbers and date + time stamps), participants in the video clips, and content labels for activities and behaviors. Information for tags can come from a mix of sources, including questionnaires, interviews, and the videos themselves. Not least, we can tag the race and social class of parents, family structure, household size, and age of target child for each New Jersey Families Study clip using responses from a short Interest Survey completed by every family at the beginning of the recruitment process. Each video has a camera number indicator that can be linked with a room as well as a timestamp that provides the recording's date and time, which would allow a user to filter by day or location. Metatags such as these are not intended to provide detailed information on the content or interactions in the clip.</p> <p>The time and labor costs required to add further details and identify specific behaviors contained within each clip on a mass scale are significant. Two strategies we used to narrow our focus in creating the next tier of tags were, first, to seek guidance from researchers who we thought would be the most likely to use our data and, second, to consider the types of information most amenable to automated coding.</p> <p>Ten early childhood experts participated in two focus groups, and they strongly encouraged us to code the participants in each New Jersey Families Study video clip ([<reflink idref="bib53" id="ref129">53</reflink>]). Binary codes that report simply whether there are people in a given clip can be automated comparatively easily and produce reasonably accurate results. But, a more elaborate coding scheme that embraces both the detection and identification of all faces would satisfy more demanding research needs. Determining who the participants are in video clips is a two-step process that involves facial detection followed by facial identification. Researchers have developed software programs to deal with both of these issues. In their study of plenary attendance, [<reflink idref="bib88" id="ref130">88</reflink>] used Tiny Face architecture ([<reflink idref="bib44" id="ref131">44</reflink>]) for face detection and an ImageNet pretrained ResNet-18 model (He et al. 2016) for face identification. In some cases, pretrained models are finetuned using data from the target task. [<reflink idref="bib88" id="ref132">88</reflink>] assigned facial identities using a training dataset based on photos and video clips of legislators. Another possibility for participant identification is to identify specific persons by analyzing gait and other characteristics ([<reflink idref="bib13" id="ref133">13</reflink>]). OpenFace ([<reflink idref="bib5" id="ref134">5</reflink>]) and OpenPose ([<reflink idref="bib15" id="ref135">15</reflink>]) are additional open-source libraries for, respectively, identifying facial behavior (e.g., eye gaze, head orientation, and facial expression) and nonverbal communication in image and video data. See also the study by [<reflink idref="bib8" id="ref136">8</reflink>] for additional computer vision resources for identifying individuals, small groups, and crowds.</p> <p>The New Jersey Families Study focus group participants emphasized the need to code certain activities and behaviors that are particularly salient to early childhood researchers. These include access to media and screen time, mealtime, grooming, movement, playtime and cognitive stimulation, sleeping, speech, physical touch, and emotions and emotive action ([<reflink idref="bib53" id="ref137">53</reflink>]). Applications of computer vision modeling in social science are developing ([<reflink idref="bib22" id="ref138">22</reflink>]; [<reflink idref="bib23" id="ref139">23</reflink>]; [<reflink idref="bib81" id="ref140">81</reflink>]; [<reflink idref="bib110" id="ref141">110</reflink>]; [<reflink idref="bib119" id="ref142">119</reflink>]; [<reflink idref="bib120" id="ref143">120</reflink>]) though "the analysis of moving images" in social science trails that of computer science ([<reflink idref="bib88" id="ref144">88</reflink>]:5).</p> <p>It is generally acknowledged by computer scientists that activity recognition is a more difficult task than object detection or facial recognition, complicating the automated coding of activities and behaviors ([<reflink idref="bib88" id="ref145">88</reflink>]). For example, the distinction between playing and standing as isolated movements in the New Jersey Families Study clips can be problematic, and playing can take different physical forms, making it difficult to identify a consistent and predictive pattern through machine learning, further demonstrating the need for some manual coding. Even if the broader pattern of a movement is classified (e.g., sitting), researchers must still inspect the videos and analyze their context to interpret the situational meaning of the behavior (e.g., social ritual). Researchers have used crowdsourcing to help label activities in videos to train fuller models in computer vision ([<reflink idref="bib64" id="ref146">64</reflink>]; [<reflink idref="bib102" id="ref147">102</reflink>]; [<reflink idref="bib114" id="ref148">114</reflink>]). Focus group members proposed a more localized version of crowdsourcing in the form of a dynamic data archive, where as a condition for using New Jersey Families Study data, researchers would be required to archive their more granular codes on the project's server. The research community could then have access to that material, see what has already been done, and build on existing work over time. Databrary allows users to add keyword tags to videos that can be searched. Users can also upload their coding files linked to the videos so that other researchers can see their coding schemes (Gilmore et al. 2016b: 6). [<reflink idref="bib4" id="ref149">4</reflink>] suggest that researchers building a cumulative coding archive should use the same analytic tools, coding categories, and coding schemes to facilitate a collective effort amongst primary and secondary data users. Additionally, all researchers should have access to archived contextual information gathered during the data collection processes in order to minimize the gap in contextual knowledge between researchers. Cross-checking granular interpretations across researchers along with crowdsourced results and predictions based on machine learning offers a multi-method opportunity to identify biases and better gauge reliability in labeling.</p> <p>Audio transcription is another navigation tool to identify what is both being said and done in video data. Transcriptions make it easier to follow what is happening in the video, provide a record that can be shared and analyzed, and allow researchers to search the corpus of data using key words. Even more, they can be used to study broader topics of conversation, which can then be used as metatags. Structural topic models use the co-occurrence of words to figure out topics found within texts ([<reflink idref="bib40" id="ref150">40</reflink>]). For example, if a transcript contains such words as ball, bat, homerun, stadium, and shut out, these suggest that the topic being discussed is "baseball." With topic modeling, topics emerge inductively, allowing users to filter videos for wider themes (e.g., "work," "parenting," and "vacation") that extend beyond the literal terms included in the transcript and identify the most salient and frequently discussed topics.</p> <p>While researchers can manually transcribe their own data, this can amount to an enormous time investment for a large project. For example, in the CELF video project, a large team of researchers worked for nearly a decade to transcribe 1,170,629 lines of talk ([<reflink idref="bib89" id="ref151">89</reflink>]: 265). Automated transcription services offer relatively high levels of accuracy (90% plus), performing less well when the audio is of poor quality, includes multiple people speaking at once, or contains a significant amount of background noise. For the New Jersey Families Study, capturing young children's voices and not yet fully formed speech may be particularly challenging. In addition, researchers who seek to employ external transcription companies should take care to examine their terms of service, as these may involve transferring rights to the data or even copies of the data, which may violate IRB promises. One potential solution that we have explored is to bring open-source automatic speech recognition (ASR) packages (e.g., OpenAI's Whisper) into the secure data environment. But, the strict protocols of the air-gapped system and limited computational resources restrict the capacity for machine learning, demonstrating a trade-off between privacy requirements and computational needs.</p> <p>While automated methods to differentiate between speakers and other sounds, including background noise, continue to progress for transcription ([<reflink idref="bib118" id="ref152">118</reflink>]), audio data also offer the opportunity for phonetic analysis. This integrates finer points of speech analysis and synthesis, enabling researchers to capture the emotion and intensity behind conversational speech. Praat is a free linguistics tool that can be used for a variety of measurements and tasks, including opening sound files, measuring duration, formants, pitch, intensity, voice breaks, source-filter resynthesis, nasality measurement, and formula manipulation of sounds. It can identify different speech patterns among various participants and categorize clips by emotion based on characteristics such as pitch and intensity. Decibel readings can be used to detect instances of raised voices or other loud noises ([<reflink idref="bib96" id="ref153">96</reflink>]). The <emph>Parcelmouth</emph> library provides a user-friendly Python interface for Praat ([<reflink idref="bib48" id="ref154">48</reflink>]). Alternatively, the R package <emph>communication</emph> extracts and preprocesses a number of features from audio files into text that allows one to examine tone, emphasis, energy, and other structural elements of speech, including how it flows over time ([<reflink idref="bib57" id="ref155">57</reflink>]). The transformation of audio and video data to this type of readable, non-video (or tabular) data opens up the capacity to share otherwise sensitive data as identifiable data are removed. When combined with video data, audio data also offers a multimodal opportunity to improve learning models for activity recognition ([<reflink idref="bib107" id="ref156">107</reflink>]) and speech diarization ([<reflink idref="bib117" id="ref157">117</reflink>]).</p> <hd id="AN0190929153-13">Discussion</hd> <p>In collecting over 11,470 h of video data, the New Jersey Families Study is one of the very few large-scale video projects in the field of sociology. This project has provided us with a unique opportunity to explore video data management and data sharing techniques, particularly in light of a host of cutting-edge developments in data science. In this article, we have walked through the steps we have taken to facilitate secure and accessible data sharing. We have provided considerations regarding data security, storage, and access as well as data documentation, tagging, and transcription. Technological advances and evolving modes of data sharing will continue to change the methodological terrain of video data, but also bring new ethical challenges. For example, in family-based research, new big-data projects like 1kD will track children over their first 1000 days ([<reflink idref="bib60" id="ref158">60</reflink>]), and Play and Learning Across a Year (PLAY) plans to create a shareable video database of infant-mother play in over 900 homes ([<reflink idref="bib94" id="ref159">94</reflink>]). PLAY has established a comprehensive "hyperactive" curation workflow where each step of the data collection process is shared (e.g., study-wide materials, training protocols, transcriptions, behavioral annotations) and data curation occurs as data are being collected ([<reflink idref="bib103" id="ref160">103</reflink>]). Expanded opportunities for comparative research internationally also elicit new challenges as different countries have different policies governing privacy and data sharing ([<reflink idref="bib100" id="ref161">100</reflink>]). As social scientists begin to take advantage of the ubiquity of video data and heed growing calls to make their data publicly available, they can save time by learning from our efforts.</p> <p>The principal lesson we have learned so far is that the New Jersey Families Study has turned out to be much larger, more complex, and more challenging than any of us imagined. Data collection was a relatively smooth and inexpensive part of the project. However, data management, analysis, and sharing have taken significant time, effort, and resources. It would have been helpful to anticipate from the beginning how costly data curation would be and to assemble the needed resources from the beginning. Staffing needs also changed once data collection was over, and it would have been beneficial to assemble a research team that included computer scientists and early childhood specialists even if those skills were not heavily used during data collection. Finally, we failed to appreciate soon enough the advantages of opening the data to researchers beyond a single university. Doing so, it seems in retrospect, would have been the best way to maximize the return on the investment of time, money, and effort involved in collecting and curating the data.</p> <p>One key takeaway from our experience is the need to consider the life cycle of data from the start. When a new project is launched, data collection concerns are at the fore but plans for data management, maintenance, and sharing should also be considered, particularly with video data, where each of these stages can be far more complicated than is typical. Plans for data sharing are best considered in the early stages of the project and incorporated into the project design. Knowing that you plan to share your data can impact the type and format of the data you collect, the types of documentation you keep, and importantly, the assurances and consent forms given to research subjects. Allowing for some flexibility in terms of where data can be stored and who can be granted access can be helpful even if you do not have firm plans for data sharing. Given increased calls for transparency in social science research, researchers would benefit from constructing a data management and sharing plan at their project's inception. [<reflink idref="bib103" id="ref162">103</reflink>] helpfully outline the "five Ws" (why, what, where, when, and who) that researchers can ask themselves when preparing at a project's outset for data curation and sharing.</p> <p>Over the course of developing a user-friendly interface for our video data, we found the largest limiting factor to be the time-consuming nature of cleaning and tagging data. Advances in automation, computer vision, and speech recognition can augment traditional techniques to mitigate the enormity of video data and the labor intensity of manual coding or manual transcription. However, visual techniques may not be sufficiently advanced to do many of the tasks a research team may be interested in such as recognition of certain kinds of activities or behaviors. Moreover, a computer vision model is prone to a certain degree of error, identifying negative cases or failing to identify positive cases. However, sociologists are already experimenting with new methods of video data curation and analysis to address some of these limitations ([<reflink idref="bib63" id="ref163">63</reflink>]; [<reflink idref="bib45" id="ref164">45</reflink>]; [<reflink idref="bib8" id="ref165">8</reflink>]). And the field of AI is growing so quickly that problems we see now may not be issues in the near future. Tools for analyzing video are quickly evolving, and developments in generative AI and ChatGPT are already being applied to innovatively describe and summarize video data ([<reflink idref="bib68" id="ref166">68</reflink>]).</p> <p>A final takeaway for researchers who are interested in sharing large amounts of sensitive data is to consider the tradeoffs between privacy and computational capacity. Air-gapped systems typically do not offer the same processing capacity as high-performance computing and the lack of access to the Internet can slow machine learning processes, for example, if packages are not already loaded onto virtual machines. Security can therefore come at the cost of computational intensity or requires investment in additional hardware, which needs to be updated every few years. Building one's own secure infrastructure also entails administrative costs. Even individual servers sitting in an office are prone to security threats if they are not managed by informed system administrators. This is true for even smaller datasets. Cloud computing has become an increasingly appealing way to have security plus computational power, but costs can quickly get out of control, especially when the user base grows. To avoid costly mistakes, researchers should consider the full data pipeline when selecting a storage solution and additional safeguards. Considerations include how the data will be used (e.g., qualitative analysis versus heavy processing), the number of users, journal requirements and reproducibility protocols, and long-term maintenance.</p> <p>Because video captures the complexity of the real world, it can push the methodological boundaries in both the social sciences and computer science, facilitating innovation in sociological inquiry. Video data contain important advantages compared with more traditional survey and interview data. They eliminate recall bias and reduce social desirability bias. Using unobtrusive technologies mitigates interviewer effects. Viewing individuals in their daily routines has the potential of serendipitously revealing important events and behaviors that investigators might not have thought to ask about otherwise. Comparing what people say they do with what they actually do can shed new light on the reliability of responses in conventional surveys and interviews ([<reflink idref="bib49" id="ref167">49</reflink>]). Finally, the ability to replay and recode video segments opens the data to a wider set of disciplinary and cultural perspectives and interpretations. Video data are unique in that they provide access to finely detailed audio and visual data, and create a record that can be revisited by multiple researchers with diverse perspectives ([<reflink idref="bib37" id="ref168">37</reflink>]; [<reflink idref="bib20" id="ref169">20</reflink>]). As the social sciences face increasing demands for transparency, reliability, and reproducibility, video data are uniquely suited for addressing these concerns. The time- and labor-intensive nature of working with video data is not insignificant, but our experiences have shown us that new tools and techniques can make the process more manageable and offer opportunities for cross-disciplinary collaboration.</p> <hd id="AN0190929153-14">Acknowledgments</hd> <p>The research assistance of Maria Maria Castillo, Ana Delgado, Kara Mitchell, Cecilia Kim, Shihe Luan, and Erin Smith is gratefully acknowledged. We also thank the reviewers for their insightful comments and time taken to help strengthen our manuscript.</p> <ref id="AN0190929153-15"> <title> References </title> <blist> <bibl id="bib1" idref="ref102" type="bt">1</bibl> <bibtext> Adolph Karen. 2016. " Video as Data: From Transient Behavior to Tangible Recording." APS Observer. 29(3): 23-25.</bibtext> </blist> <blist> <bibl id="bib2" idref="ref106" type="bt">2</bibl> <bibtext> Altman Micah, Wood Alexandra, O'Brien David R., Vadhan Salil, Gasser Urs. 2015. " Towards a Modern Approach to Privacy-Aware Government Data Releases." Berkeley Technology Law Journal. 30(3):1967–2072.</bibtext> </blist> <blist> <bibl id="bib3" idref="ref36" type="bt">3</bibl> <bibtext> American Sociological Association. 2018. " ASA Code of Ethics. " https://<ulink href="http://www.asanet.org/sites/default/files/asa%5fcode%5fof%5fethics-june2018a.pdf">www.asanet.org/sites/default/files/asa%5fcode%5fof%5fethics-june2018a.pdf</ulink>.</bibtext> </blist> <blist> <bibl id="bib4" idref="ref50" type="bt">4</bibl> <bibtext> Andersson Emilia, Sørvik Gard Ove. 2013. " Reality Lost? Re-use of Qualitative Data in Classroom Video Studies." Forum: Qualitative Social Research. 14(3):1–25.</bibtext> </blist> <blist> <bibl id="bib5" idref="ref134" type="bt">5</bibl> <bibtext> Baltrusaitis Tadas, Zadeh Amir, Lim Yao Chong, Morency Louis-Phillippe. 2018. 13th IEEE International Conference on Automatic Face & Gesture Recognition (FG 2018).</bibtext> </blist> <blist> <bibl id="bib6" idref="ref23" type="bt">6</bibl> <bibtext> Barron Brigid. 2007 "Video as a Tool to Advance Understanding of Learning and Development in Peer, Family, and Other Informal Learning Contexts." Pp. 159–87 in Video Research in the Learning Sciences, edited by Goldman Ricki, Pea Roy, Barron Brigid, Derry Sharon J. Mahwah, NJ: Lawrence Erlbaum Associates, Inc.</bibtext> </blist> <blist> <bibl id="bib7" idref="ref21" type="bt">7</bibl> <bibtext> Belsky Jay, Vandell Deborah Lowe, Burchinal Margaret, Clarke-Stewart K. Alison, McCartney Kathleen, Owen Margaret Tresch and The NICHD Early Child Care Research Network. 2007. " Are There Long-Term Effects of Early Child Care? " Child Development. 78(2):681–701.</bibtext> </blist> <blist> <bibl id="bib8" idref="ref136" type="bt">8</bibl> <bibtext> Bernasco Wim, Hoeben Evelien M., Koelma Dennis, Liebst Lasse Suonperä, Thomas Josephine, Appelman Joska, Snoek Cees G. M., Lindegaard Marie Rosenkrantz. 2023. " Promise Into Practice: Application of Computer Vision in Empirical Research on Social Distancing." Sociological Methods & Research. 52(3):1239–87.</bibtext> </blist> <blist> <bibl id="bib9" idref="ref96" type="bt">9</bibl> <bibtext> Betancort Cabrera Noemi, Bongartz Elke C., Dörrenbächer Nora, Goebel Jan, Kaluza Harald, Siegers Pascal. 2020. " White Paper on Implementing the FAIR Principles for Data in the Social, Behavioural, and Economic Sciences. " RatSWD Working Paper, No. 274. Rat für Sozial- und Wirtschaftsdaten, Berlin, Germany.</bibtext> </blist> <blist> <bibtext> Bishop Libby. 2007. " A Reflexive Account of Reusing Qualitative Data: Beyond Primary/Secondary Dualism." Sociological Research Online. 12(3):43–56.</bibtext> </blist> <blist> <bibtext> Bishop Libby. 2009. " Ethical Sharing and Reuse of Qualitative Data." Australian Journal of Social Issues. 44(3):255–72.</bibtext> </blist> <blist> <bibtext> Bishop Libby, Kuula-Luumi Arja. 2017. " Revisiting Qualitative Data Reuse: A Decade On." SAGE Open. 7(1):2158244016685136.</bibtext> </blist> <blist> <bibtext> Bouchrika Imed, Goffredo Michaela, Carter John, Nixon Mark. 2011. " On Using Gait in Forensic Biometrics." Journal of Forensic Sciences. 56(4):882–89.</bibtext> </blist> <blist> <bibtext> Boyd Danah, Crawford Kate. 2012. " Critical Questions for Big Data." Information, Communication & Society. 15(5):662–79.</bibtext> </blist> <blist> <bibtext> Cao Zhe, Hidalgo Gines, Simon Tomas, Wei Shih-En, Sheikh Yaser. 2019. " OpenPose: Realtime Multi-Person 2D Post Estimation Using Part Affinity Fields." IEEE Transactions on Pattern Analysis and Machine Intelligence. 43(1):172–86.</bibtext> </blist> <blist> <bibtext> CHILDES. 2022. " CHILDES. " https://childes.talkbank.org/.</bibtext> </blist> <blist> <bibtext> Collins Randall. 2009. Violence: A Micro-Sociological Theory. Princeton, NJ: Princeton University Press.</bibtext> </blist> <blist> <bibtext> DDI (Data Documentation Initiative). 2022. " What is DDI? " https://ddialliance.org/learn/what-is-ddi.</bibtext> </blist> <blist> <bibtext> Derry Sharon J. 2007. Guidelines for Video Research in Education: Recommendations from an Expert Panel. Chicago, IL: Data Research and Development Center, NORC at the University of Chicago.</bibtext> </blist> <blist> <bibtext> Derry Sharon J., Pea Roy D., Barron Brigid, Engle Randi A., Erickson Frederick, Goldman Ricki, Hall Rogers, Koschmann Timothy, Lemke Jay L., Sherin Miriam Gamoran. 2010. " Conducting Video Research in the Learning Sciences: Guidance on Selection, Analysis, Technology, and Ethics." The Journal of the Learning Sciences. 19(1):3–53.</bibtext> </blist> <blist> <bibtext> Desai Tanvi, Ritchie Felix, Welpton Richard. 2016. " Five Safes: Designing Data Access for Research. " University of the West of England Economics Working Paper Series 1601. Bristol, England.</bibtext> </blist> <blist> <bibtext> Dietrich Bryce J. 2021. " Using Motion Detection to Measure Social Polarization in the U.S. House of Representatives." Political Analysis. 29(2):250–59.</bibtext> </blist> <blist> <bibtext> Dietrich Bryce J., Enos Ryan D., Sen Maya. 2019. " Emotional Arousal Predicts Voting on the U.S. Supreme Court." Political Analysis. 27(2):237–43.</bibtext> </blist> <blist> <bibtext> Economic and Social Research Council. 2018. " ESRC Research Data Policy. " https://<ulink href="http://www.ukri.org/wp-content/uploads/2021/07/ESRC-200721-ResearchDataPolicy.pdf">www.ukri.org/wp-content/uploads/2021/07/ESRC-200721-ResearchDataPolicy.pdf</ulink>.</bibtext> </blist> <blist> <bibtext> Elman Colin, Kapiszewski Diana, Vinuela Lorena. 2010. " Qualitative Data Archiving: Rewards and Challenges." PS: Political Science & Politics. 43(1):23–7.</bibtext> </blist> <blist> <bibtext> Erickson Frederick. 2011. " Uses of Video in Social Research: A Brief History." International Journal of Social Research Methodology. 14(3):179–89.</bibtext> </blist> <blist> <bibtext> Erickson Frederick, Schultz Jeffrey. 1982. The Counselor as Gatekeeper: Social Interaction in Interviews. New York: Academic Press.</bibtext> </blist> <blist> <bibtext> Foster Ian. 2018. " Research Infrastructure for the Safe Analysis of Sensitive Data." Annals of the American Academy of Political and Social Science. 675(1):102–20.</bibtext> </blist> <blist> <bibtext> Frank Rebecca D., Tyler Allison R. B., Gault Anna, Suzuka Kara, Yakel Elizabeth. 2018. " Privacy Concerns in Qualitative Video Data Reuse." International Journal of Digital Curation. 13(1):47–72.</bibtext> </blist> <blist> <bibtext> Gennetian Lisa A., Frank Michael C., Tamis-LeMonda Catherine S. 2022. " Open Science in Developmental Science." Annual Review of Developmental Psychology. 4(1):377–97.</bibtext> </blist> <blist> <bibtext> Gennetian Lisa A., Tamis-LeMonda Catherine S., Frank Michael C. 2020. " Advancing Transparency and Openness in Child Development Research: Opportunities." Child Development Perspectives. 14(1):3–8.</bibtext> </blist> <blist> <bibtext> Gillies Val, Edwards Rosalind. 2005. " Secondary Analysis in Exploring Family and Social Change: Addressing the Issue of Context." Forum Qualitative Sozialforschung/Forum: Qualitative Social Research. 6(1).</bibtext> </blist> <blist> <bibtext> Gilmore Rick O., Adolph Karen E., Millman David S. 2016a. " Curating Identifiable Data for Sharing: The Databrary Project. " Pp. 1–6in 2016 New York Scientific Data Summit (NYSDS).</bibtext> </blist> <blist> <bibtext> Gilmore Rick O., Adolph Karen E., Millman David S., Gordon Andrew. 2016b. " Transforming Education Research Through Open Video Data Sharing." Advances in Engineering Education. 5(2).</bibtext> </blist> <blist> <bibtext> Gilmore Rick O., Cole Pamela M., Verma Suman, Van Aken Marcel AG, Worthman Carol M. 2020. " Advancing Scientific Integrity, Transparency, and Openness in Child Development Research: Challenges and Possible Solutions." Child Development Perspectives. 14(1):9–14.</bibtext> </blist> <blist> <bibtext> Golann Joanne W., Mirakhur Zitsi, Espenshade Thomas J. 2019. " Collecting Ethnographic Video Data for Policy Research." American Behavioral Scientist. 63(3):387–403.</bibtext> </blist> <blist> <bibtext> Goldman Ricki, Pea Roy, Barron Bridget, Derry Sharon. 2007. Video Research in the Learning Sciences. Mahwah, NJ: Lawrence Erlbaum Associates, Inc.</bibtext> </blist> <blist> <bibtext> Goroff Daniel, Polonetsky Jules, Tene Omer. 2018. " Privacy Protective Research: Facilitating Ethically Responsible Access to Administrative Data." The ANNALS of the American Academy of Political and Social Science. 675(1):46–66.</bibtext> </blist> <blist> <bibtext> Gregory Katherine. 2020. " The Video Camera Spoiled my Ethnography: A Critical Approach." Qualitative and Mixed Methods Failures. 19:1–9.</bibtext> </blist> <blist> <bibtext> Grimmer Justin, Roberts Margaret E., Stewart Brandon M. 2022. Text as Data: A New Framework for Machine Learning and the Social Sciences. Princeton, NJ: Princeton University Press.</bibtext> </blist> <blist> <bibtext> Hammersley Martyn. 2010. " Can We Re-Use Qualitative Data Via Secondary Analysis? Notes on Some Terminological and Substantive Issues." Sociological Research Online. 15(1):47–53.</bibtext> </blist> <blist> <bibtext> Heckman James J. 2011. " The Economics of Inequality: The Value of Early Childhood Education." American Educator. 35(1):31–5.</bibtext> </blist> <blist> <bibtext> Holliday Ruth. 2000. " We've Been Framed: Visualising Methodology." The Sociological Review. 48(4):503–21.</bibtext> </blist> <blist> <bibtext> Hu Peiyun, Ramanan Deva. 2017. "Finding Tiny Faces." Pp. 951–59 in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition.</bibtext> </blist> <blist> <bibtext> Hwang Jackelyn, Dahir Nima, Sarukkai Mayuka, Wright Gabby. 2023. " Curating Training Data for Reliable Large-Scale Visual Data Analysis: Lessons from Identifying Trash in Street View Imagery." Sociological Methods & Research. 52(3):1155-200.</bibtext> </blist> <blist> <bibtext> ICPSR (Inter-university Consortium for Political and Social Research). 2022. " Metadata. " https://<ulink href="http://www.icpsr.umich.edu/web/pages/datamanagement/lifecycle/metadata.html">www.icpsr.umich.edu/web/pages/datamanagement/lifecycle/metadata.html</ulink>.</bibtext> </blist> <blist> <bibtext> Jacob Theodore, Tennenbaum Daniel, Seilhamer Ruth Ann, Bargiel Kay, Sharon Tanya. 1994. " Reactivity Effects During Naturalistic Observation of Distressed and Nondistressed Families." Journal of Family Psychology. 8(3):354–63.</bibtext> </blist> <blist> <bibtext> Jadoul Yannick, Thompson Bill, de Boer Bart. 2018. " Introducing Parselmouth: A Python Interface to Praat." Journal of Phonetics. 71:1–15.</bibtext> </blist> <blist> <bibtext> Jerolmack Colin, Khan Shamus. 2014. " Talk Is Cheap: Ethnography and the Attitudinal Fallacy." Sociological Methods & Research. 43(2):178–209.</bibtext> </blist> <blist> <bibtext> Jerolmack Colin, Murphy Alexandra K. 2019. " The Ethical Dilemmas and Social Scientific Trade-Offs of Masking in Ethnography." Sociological Methods & Research. 48(4):801–27.</bibtext> </blist> <blist> <bibtext> Joyce Jack B., Douglass Tom, Benwell Bethan, Rhys Catrin S., Parry Ruth, Simmons Richard, Kerrison Adrian. 2022. " Should We Share Qualitative Data? Epistemological and Practical Insights from Conversation Analysis." International Journal of Social Research Methodology: 1–15.</bibtext> </blist> <blist> <bibtext> Katz v. United States. 1967. 389 U.S. 347.</bibtext> </blist> <blist> <bibtext> Kim Cecilia H. 2021. A Report on the New Jersey Families Study's Focus Groups from December 2020: Office of Population Research. Princeton University.</bibtext> </blist> <blist> <bibtext> Kim Jihyun, Suzuka Kara, Yakel Elizabeth. 2020. " Reusing Qualitative Video Data: Matching Reuse Goals and Criteria for Selection." Aslib Journal of Information Management. 72(3):395–419.</bibtext> </blist> <blist> <bibtext> Kindel Alexander T., Bansal Vieneet, Catena Kristin D., Hortshorne Thomas H., Jaeger Kate, Koffmann Dawn, McLanahan Sara, et al.2019. " Improving Metadata Infrastructure for Complex Surveys: Insights from the Fragile Families Challenge." Socius: Sociological Research for a Dynamic World. 5:1–24.</bibtext> </blist> <blist> <bibtext> King Gary. 2011. " Ensuring the Data-Rich Future of the Social Sciences." Science. 331(6018):719–21.</bibtext> </blist> <blist> <bibtext> Knox Dean, Lucas Christopher. 2021. " A Dynamic Model of Speech for the Social Sciences." American Political Science Review. 115(2):649–66.</bibtext> </blist> <blist> <bibtext> Kuula Arja. 2011. " Methodological and Ethical Dilemmas of Archiving Qualitative Data." IASSIST Quarterly. 34(3–4):12–12.</bibtext> </blist> <blist> <bibtext> Lareau Annette. 2003. Unequal Childhoods: Class, Race, and Family Life. Berkeley, CA: University of California Press.</bibtext> </blist> <blist> <bibtext> Leap. 2022. " The First 1000 Days. " https://wellcomeleap.org/1kd/.</bibtext> </blist> <blist> <bibtext> Lee Valerie E., Burkam David T. 2002. Inequality at the Starting Gate: Social Background Differences in Achievement as Children Begin School. Washington, DC: Economic Policy Institute.</bibtext> </blist> <blist> <bibtext> Legewie Nicolas, Nassauer Anne. 2018. " YouTube, Google, Facebook: 21st Century Online Video Research and Research Ethics." Forum Qualitative Sozialforschung / Forum: Qualitative Social Research. 19(3).</bibtext> </blist> <blist> <bibtext> Legewie Nicolas M., Nassauer Anne. 2023. " Current and Future Debates in Video Data Analysis." Sociological Methods & Research. 52(3):1107–19.</bibtext> </blist> <blist> <bibtext> Legewie Nicolas M., Nassauer Anne, Stuerznickel Malika. 2019. " Opportunities for Analyzing Visual Data in the 21st Century. Report from the 2019 Blankensee Colloquium on Capturing and Analyzing Social Change." SSRN. https://papers.ssrn.com/sol3/papers.cfm?abstract_id=3489981.</bibtext> </blist> <blist> <bibtext> Levine Mark, Taylor Paul J., Best Rachel. 2011. " Third Parties, Violence, and Conflict Resolution: The Role of Group Size and Collective Action in the Micro-Regulation of Violence." Psychological Science. 22(3):406–12.</bibtext> </blist> <blist> <bibtext> Logan Jessica A. R., Hart Sara A., Schatschneider Christopher. 2021. " Data Sharing in Education Science." AERA Open. 7:10.1177/23328584211006475.</bibtext> </blist> <blist> <bibtext> Lundberg Ian, Narayanan Arvind, Levy Karen, Salganik Matthew J. 2019. " Privacy, Ethics, and Data Access: A Case Study of the Fragile Families Challenge." Socius: Sociological Research for a Dynamic World. 5:1–25.</bibtext> </blist> <blist> <bibtext> Maaz Muhammad, Rasheed Hanoona, Khan Salman, Khan Fahad Shahbaz. 2023. "Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models." ArXiv:2306.05424v1.</bibtext> </blist> <blist> <bibtext> Makel Matthew C., Meyer Melanie S., Simonsen Mary A., Roberts Anne M., Plucker Jonathan A. 2022. " Replication Is Relevant to Qualitative Research." Educational Research and Evaluation. 27(1–2):215–19.</bibtext> </blist> <blist> <bibtext> Makin David A., Willits Dale W., Koslicki Wendy, Brooks Rachael, Dietrich Bryce J., Bailey Rachel L. 2019. " Contextual Determinants of Observed Negative Emotional States in Police–Community Interactions." Criminal Justice and Behavior. 46(2):301–18.</bibtext> </blist> <blist> <bibtext> Mauthner Natasha S., Parry Odette, Backett-Milburn Kathryn. 1998. " The Data Are Out There, or Are They? Implications for Archiving and Revisiting Qualitative Data." Sociology. 32(4):733–45.</bibtext> </blist> <blist> <bibtext> McDermott Ray P., Gospodinoff Kenneth, Aron Jeffrey. 1978. " Criteria for an Ethnographically Adequate Description of Concerted Activities and Their Contexts." Semiotica. 24(3–4):245–76.</bibtext> </blist> <blist> <bibtext> Mehan Hugh. 1979. Learning Lessons: Social Organization in the Classroom. Cambridge, MA: Harvard University Press.</bibtext> </blist> <blist> <bibtext> Mehmood Abid, Natgunanathan Iynkaran, Xiang Yong, Hua Guang, Guo Song. 2016. " Protection of Big Data Privacy." IEEE. 4:1821–34.</bibtext> </blist> <blist> <bibtext> Moore Niamh. 2007. " (Re)Using Qualitative Data? " Sociological Research Online. 12(3):1–13.</bibtext> </blist> <blist> <bibtext> Mozersky Jessica, Parsons Meredith, Walsh Heidi, Baldwin Kari, McIntosh Tristan, DuBois James M. 2020. " Research Participant Views Regarding Qualitative Data Sharing." Ethics & Human Research. 42(2):13–27.</bibtext> </blist> <blist> <bibtext> Murphy Alexandra K., Jerolmack Colin, Smith DeAnna. 2021. " Ethnography, Data Transparency, and the Information Age." Annual Review of Sociology. 47(1):41–61.</bibtext> </blist> <blist> <bibtext> Nassauer Anne. 2016. " From Peaceful Marches to Violent Clashes: A Micro-Situational Analysis." Social Movements. 15(5):1–16.</bibtext> </blist> <blist> <bibtext> Nassauer Anne, Legewie Nicolas M. 2021. " Video Data Analysis: A Methodological Frame for a Novel Research Trend." Sociological Methods & Research. 50(1):135–74.</bibtext> </blist> <blist> <bibtext> Nassauer Anne, Legewie Nicolas M. 2022. Video Data Analysis: How to Use 21st Century Video in the Social Sciences. London, UK: Sage Publications.</bibtext> </blist> <blist> <bibtext> Nelson Laura K. 2020. " Computational Grounded Theory: A Methodological Framework." Sociological Methods & Research. 49(1):3–42.</bibtext> </blist> <blist> <bibtext> Nelson Alondra. 2022. Memorandum for the Heads of Executive Departments and Agencies.Office of Science and Technology Policy.</bibtext> </blist> <blist> <bibtext> Neuhaus Fabian, Webmoor Timothy. 2012. " Agile Ethics for Massified Research and Visualization." Information, Communication, & Society. 15(1):43–6.</bibtext> </blist> <blist> <bibtext> NIH (National Institutes of Health). 2023a. " Final NIH Policy for Data Management and Sharing. " https://grants.nih.gov/grants/guide/notice-files/NOT-OD-21-013.html.</bibtext> </blist> <blist> <bibtext> NIH (National Institutes of Health). 2023b. " Notice of Special Interest (NOSI): Administrative Supplements to Support the Exploration of Cloud in NIH-supported Research. " https://grants.nih.gov/grants/guide/notice-files/NOT-OD-23-070.html.</bibtext> </blist> <blist> <bibtext> Nissenbaum Helen. 2010. Privacy in Context: Technology, Policy, and the Integrity of Social Life. Stanford, CA: Stanford University Press.</bibtext> </blist> <blist> <bibtext> NSF (National Science Foundation). 2022. " New Data Infrastructure Initiative Will Accelerate the Advancement and Impacts of Social and Behavioral Research. " https://<ulink href="http://www.nsf.gov/news/special%5freports/announcements/020422.jsp">www.nsf.gov/news/special%5freports/announcements/020422.jsp</ulink>.</bibtext> </blist> <blist> <bibtext> Nyhuis Dominic, Ringwald Tobias, Rittmann Oliver, Gschwend Thomas, Stiefelhagen Rainer. 2022. "Automated Video Analysis for Social Science Research." Pp. 1–7 in Handbook of Computational Social Science: Theory, Case Studies and Ethics, vol. 1, edited by Engel U., Quan-Haase A., Liu S. X., Lyberg L. London, UK: Routledge.</bibtext> </blist> <blist> <bibtext> Ochs Elinor, Kremer-Sadlik Tamar. 2013. Fast-Forward Family: Home, Work, And Relationships in Middle-Class America. Berkeley: University of California Press.</bibtext> </blist> <blist> <bibtext> Office of Science and Technology Policy, Executive Office of the President. https://<ulink href="http://www.whitehouse.gov/wp-content/uploads/2022/08/08-2022-OSTP-Public-access-Memo.pdf">www.whitehouse.gov/wp-content/uploads/2022/08/08-2022-OSTP-Public-access-Memo.pdf</ulink>.</bibtext> </blist> <blist> <bibtext> Parry Odette, Mauthner Natasha S. 2004. " Whose Data Are They Anyway?: Practical, Legal and Ethical Issues in Archiving Qualitative Research Data." Sociology. 38(1):139–52.</bibtext> </blist> <blist> <bibtext> Peng Kenny, Mathur Arunesh, Narayanan Arvind. 2021. "Mitigating Data Harms Requires Stewardship: Lessons from 1000 Papers." in 35th Conference on Neural Information Processing Systems (NeurIPS 2021) Track on Datasets and Benchmarks. https://arxiv.org/pdf/2108.02922.pdf?tpcc=nleyeonai.</bibtext> </blist> <blist> <bibtext> Pink Sarah. 2001. " More Visualising, More Methodologies: On Video, Reflexivity and Qualitative Research." The Sociological Review. 49(4):586–99.</bibtext> </blist> <blist> <bibtext> PLAY. 2022. " Play and Learning Across a Year. " https://<ulink href="http://www.play-project.org/index.html">www.play-project.org/index.html</ulink>.</bibtext> </blist> <blist> <bibtext> Politou Eugenia, Alepis Efthimios, Patsakis Constantinos. 2018. " Forgetting Personal Data and Revoking Consent Under the GDPR: Challenges and Proposed Solutions." Journal of Cybersecurity. 4(1).</bibtext> </blist> <blist> <bibtext> Praat. 2022. "Praat: Doing Phonetics by Computer." https://<ulink href="http://www.fon.hum.uva.nl/praat/">www.fon.hum.uva.nl/praat/</ulink>.</bibtext> </blist> <blist> <bibtext> Princeton University. 2022. Classify Your Information. Princeton, NJ: Princeton University. Retrieved September 23, 2022. https://protectingourinfo.princeton.edu/classify.</bibtext> </blist> <blist> <bibtext> Qualitative Data Repository. 2022. " Working with Sensitive Research Data. " https://qdr.syr.edu/working-with-sensitive-research-data.</bibtext> </blist> <blist> <bibtext> Roy Brandon C., Frank Michael C., DeCamp Philip, Miller Matthew, Roy Deb. 2015. " Predicting the Birth of a Spoken Word." Proceedings of the National Academy of Sciences. 112(41):12663–68.</bibtext> </blist> <blist> <bibtext> Rutanen Niina, de Souza Amorim Kátia, Marwick Helen, White Jayne. 2018. " Tensions and Challenges Concerning Ethics on Video Research with Young Children – Experiences from an International Collaboration among Seven Countries." Video Journal of Education and Pedagogy. 3(1):7.</bibtext> </blist> <blist> <bibtext> Salganik Matthew J. 2018. Bit by Bit: Social Research in the Digital Age. Princeton, NJ: Princeton University Press.</bibtext> </blist> <blist> <bibtext> Sigurdsson Gunnar, Russakovsky Olga, Farhadi Ali, Laptev Ivan, Abhinav Gupta. 2016. "Much Ado About Time: Exhaustive Annotation of Temporal Data." Pp. 219–28 in Proceedings of the AAAI Conference on Human Computation and Crowdsourcing, vol. 4(1) https://ojs.aaai.org/index.php/HCOMP/article/view/13290.</bibtext> </blist> <blist> <bibtext> Soska Kasey C., Xu Melody, Gonzalez Sandy L., Herzberg Orit, Tamis-LeMonda Catherine S., Gilmore Rick O., Adolph Karen E. 2021. " (Hyper)Active Data Curation: A Video Case Study from Behavioral Science." Journal of EScience Librarianship. 10:3.</bibtext> </blist> <blist> <bibtext> Stanley Steven, Smith Robin James, Ford Eleanor, Jones Joshua. 2020. " Making Something Out of Nothing: Breaching Everyday Life by Standing Still in a Public Place." The Sociological Review. 68(6):1250–72.</bibtext> </blist> <blist> <bibtext> Stivers Tanya, Sidnell Jack. 2012. The Handbook of Conversation Analysis. John Wiley & Sons.</bibtext> </blist> <blist> <bibtext> Stivers Tanya, Timmermans Stefan. 2020. " Medical Authority Under Siege: How Clinicians Transform Patient Resistance into Acceptance." Journal of Health and Social Behavior. 61(1):60–78.</bibtext> </blist> <blist> <bibtext> Sun Zehua, Ke Qiuhong, Rahmani Hossein, Bennamoun Mohammed, Wang Gang, Liu Jun. 2020. " Human Action Recognition from Various Data Modalities: A Review." IEEE Transaction on Patterns Analysis and Machine Intelligence.</bibtext> </blist> <blist> <bibtext> Sun Zehua, Ke Qiuhong, Rahmani Hossein, Bennamoun Mohammed, Wang Gang, Liu Jun. 2022. " Human Action Recognition from Various Data Modalities: A Review." IEEE Transactions on Pattern Analysis and Machine Intelligence.</bibtext> </blist> <blist> <bibtext> Tenopir Carol, Allard Suzie, Douglass Kimberly, Aydinoglu Arsev Umur, Wu Lei, Read Eleanor, Manoff Maribeth, Frame Mike. 2011. " Data Sharing by Scientists: Practices and Perceptions." PLoS ONE. 6(6):e21101.</bibtext> </blist> <blist> <bibtext> Torres Michelle, Cantú Francisco. 2021. " Learning to See: Convolutional Neural Networks for the Analysis of Social Science Data." Political Analysis. 30(1):1–19.</bibtext> </blist> <blist> <bibtext> Tsai Alexander C., Kohrt Brandon A., Matthews Lynn T., Betancourt Theresa S., Lee Jooyoung K., Papachristos Andrew V., Weiser Sheri D., Dworkin Shari L. 2016. " Promises and Pitfalls of Data Sharing in Qualitative Research." Social Science & Medicine. 169:191–98.</bibtext> </blist> <blist> <bibtext> VandeVusse Alicia, Mueller Jennifer, Karcher Sebastian. 2022. " Qualitative Data Sharing: Participant Understanding, Motivation, and Consent." Qualitative Health Research. 32(1):182–91.</bibtext> </blist> <blist> <bibtext> Voigt Rob, Camp Nicholas P., Prabhakaran Vinodkumar, Hamilton William L., Hetey Rebecca C., Griffiths Camilla M., Jurgens David, Jurafsky Dan, Eberhardt Jennifer L. 2017. " Language from Police Body Camera Footage Shows Racial Disparities in Officer Respect." Proceedings of the National Academy of Sciences. 114(25):6521–6.</bibtext> </blist> <blist> <bibtext> Vondrick Carl, Patterson Donald, Ramanan Deva. 2013. " Efficiently Scaling up Crowdsourced Video Annotation: A Set of Best Practices for High Quality, Economical Video Labeling." International Journal of Computer Vision. 101(1):184–204.</bibtext> </blist> <blist> <bibtext> Weller Susie. 2023. " Fostering Habits of Care: Reframing Qualitative Data Sharing Policies and Practices." Qualitative Research. 23(4):1022–41.</bibtext> </blist> <blist> <bibtext> Wilkinson Mark D., Dumontier Michel, Aalbersberg Ijsbrand Jan, Appleton Gabrielle, Axton Myles, Baak Arie, Blomberg Niklas, et al.2016. " The FAIR Guiding Principles for Scientific Data Management and Stewardship." Scientific Data. 3(1):1–9.</bibtext> </blist> <blist> <bibtext> Xu Eric Zhongcong, Song Zeyang, Tsutsui Satoshi, Feng Chao, Ye Mang, Shou Mike Zheng. 2022. "AVA-AVD: Audio-Visual Speaker Diarization in the Wild." In Proceedings of the 30th ACM International Conference on Multimedia.</bibtext> </blist> <blist> <bibtext> Yoshioka Takuya, Abramovski Igor, Aksoylar Cem, Chen Zhuo, David Moshe, Dimitriadis Dimitrios, Gong Yifan, et al. 2019. "Advances in Online Audio-Visual Meeting Transcription." Pp. 276–83 in 2019 IEEE Automatic Speech Recognition and Understanding Workshop (ASRU). Singapore: IEEE.</bibtext> </blist> <blist> <bibtext> Zhang Han, Pan Jennifer. 2019. " Casm: A Deep-Learning Approach for Identifying Collective Action Events with Text and Image Data from Social Media." Sociological Methodology. 49(1):1–57.</bibtext> </blist> <blist> <bibtext> Zhang Han, Peng Yilang. 2021. " Image Clustering: An Unsupervised Approach to Categorize Visual Data in Social Science Research. " SocArXiv. doi: https://doi.org/10.31235/osf.io/mw57x.</bibtext> </blist> </ref> <ref id="AN0190929153-16"> <title> Footnotes </title> <blist> <bibtext> The paper does not include any original data analysis or reference any code.</bibtext> </blist> <blist> <bibtext> The authors declared no potential conflicts of interest with respect to the research, authorship, and/or publication of this article.</bibtext> </blist> <blist> <bibtext> The authors disclosed receipt of the following financial support for the research, authorship, and/or publication of this article: Partial support for this article has been provided by the Data Driven Social Science Initiative and the Office of Population Research at Princeton University, the Discovery Grant and Seeding Success Grant at Vanderbilt University, the Fund for the Advancement of the Discipline at the American Sociological Association, and the National Science Foundation (#2214309).</bibtext> </blist> <blist> <bibtext> Joanne W. Golann https://orcid.org/0000-0001-9337-4674</bibtext> </blist> <blist> <bibtext> Data sharing is not applicable to this article as no datasets were generated or analyzed during the current study.</bibtext> </blist> <blist> <bibtext> While families still may act differently knowing that they are being recorded ([39]), the New Jersey Families Study reduced the comparative risk of reactivity by recording families continuously for an unprecedented two weeks, increasing the odds of habituation where participants forget about the cameras and resume their normal routines (see [89]). Habituation typically occurs because families are too busy, find it difficult to override longstanding behaviors, and have little motivation for modifying such behaviors ([47]: 361).</bibtext> </blist> <blist> <bibtext> For the full list of master data repositories, see https://clarivate.com/webofsciencegroup/master-data-repository-list/.</bibtext> </blist> </ref> <aug> <p>By Joanne W. Golann; Lori Bougher; Richard Hall and Thomas J. Espenshade</p> <p>Reported by Author; Author; Author; Author</p> <p></p> <p>Joanne W. Golann is an associate professor of Public Policy and Education at Peabody College, Vanderbilt University. Her research focuses on how schools and families transmit cultural skills and behaviors to children.</p> <p>Lori Bougher is the director of Research and Strategy at the Initiative for Data-Driven Social Science at Princeton University.</p> <p>Richard Hall is a senior assessment specialist and adjunct lecturer at the School for International Training in Brattleboro, Vermont. His background is in K-12 education as both a researcher and practitioner which informs his measurement and evaluation work in his current role.</p> <p>Thomas J. Espenshade is a senior scholar in the Office of Population Research at Princeton University and the Founding Director of the New Jersey Families Study.</p> </aug> <nolink nlid="nl1" bibid="bib26" firstref="ref1"></nolink> <nolink nlid="nl2" bibid="bib17" firstref="ref2"></nolink> <nolink nlid="nl3" bibid="bib65" firstref="ref3"></nolink> <nolink nlid="nl4" bibid="bib78" firstref="ref4"></nolink> <nolink nlid="nl5" bibid="bib89" firstref="ref5"></nolink> <nolink nlid="nl6" bibid="bib104" firstref="ref6"></nolink> <nolink nlid="nl7" bibid="bib106" firstref="ref7"></nolink> <nolink nlid="nl8" bibid="bib113" firstref="ref8"></nolink> <nolink nlid="nl9" bibid="bib27" firstref="ref9"></nolink> <nolink nlid="nl10" bibid="bib105" firstref="ref10"></nolink> <nolink nlid="nl11" bibid="bib70" firstref="ref11"></nolink> <nolink nlid="nl12" bibid="bib79" firstref="ref12"></nolink> <nolink nlid="nl13" bibid="bib110" firstref="ref13"></nolink> <nolink nlid="nl14" bibid="bib119" firstref="ref14"></nolink> <nolink nlid="nl15" bibid="bib120" firstref="ref15"></nolink> <nolink nlid="nl16" bibid="bib36" firstref="ref16"></nolink> <nolink nlid="nl17" bibid="bib25" firstref="ref17"></nolink> <nolink nlid="nl18" bibid="bib12" firstref="ref18"></nolink> <nolink nlid="nl19" bibid="bib42" firstref="ref19"></nolink> <nolink nlid="nl20" bibid="bib61" firstref="ref20"></nolink> <nolink nlid="nl21" bibid="bib59" firstref="ref22"></nolink> <nolink nlid="nl22" bibid="bib99" firstref="ref25"></nolink> <nolink nlid="nl23" bibid="bib20" firstref="ref26"></nolink> <nolink nlid="nl24" bibid="bib80" firstref="ref29"></nolink> <nolink nlid="nl25" bibid="bib43" firstref="ref31"></nolink> <nolink nlid="nl26" bibid="bib93" firstref="ref32"></nolink> <nolink nlid="nl27" bibid="bib82" firstref="ref33"></nolink> <nolink nlid="nl28" bibid="bib34" firstref="ref34"></nolink> <nolink nlid="nl29" bibid="bib84" firstref="ref35"></nolink> <nolink nlid="nl30" bibid="bib31" firstref="ref39"></nolink> <nolink nlid="nl31" bibid="bib77" firstref="ref40"></nolink> <nolink nlid="nl32" bibid="bib109" firstref="ref41"></nolink> <nolink nlid="nl33" bibid="bib10" firstref="ref42"></nolink> <nolink nlid="nl34" bibid="bib19" firstref="ref44"></nolink> <nolink nlid="nl35" bibid="bib29" firstref="ref45"></nolink> <nolink nlid="nl36" bibid="bib111" firstref="ref46"></nolink> <nolink nlid="nl37" bibid="bib16" firstref="ref48"></nolink> <nolink nlid="nl38" bibid="bib71" firstref="ref53"></nolink> <nolink nlid="nl39" bibid="bib66" firstref="ref56"></nolink> <nolink nlid="nl40" bibid="bib37" firstref="ref57"></nolink> <nolink nlid="nl41" bibid="bib30" firstref="ref58"></nolink> <nolink nlid="nl42" bibid="bib103" firstref="ref59"></nolink> <nolink nlid="nl43" bibid="bib41" firstref="ref62"></nolink> <nolink nlid="nl44" bibid="bib50" firstref="ref63"></nolink> <nolink nlid="nl45" bibid="bib11" firstref="ref65"></nolink> <nolink nlid="nl46" bibid="bib91" firstref="ref66"></nolink> <nolink nlid="nl47" bibid="bib115" firstref="ref67"></nolink> <nolink nlid="nl48" bibid="bib62" firstref="ref68"></nolink> <nolink nlid="nl49" bibid="bib51" firstref="ref69"></nolink> <nolink nlid="nl50" bibid="bib32" firstref="ref71"></nolink> <nolink nlid="nl51" bibid="bib21" firstref="ref76"></nolink> <nolink nlid="nl52" bibid="bib67" firstref="ref77"></nolink> <nolink nlid="nl53" bibid="bib97" firstref="ref79"></nolink> <nolink nlid="nl54" bibid="bib14" firstref="ref81"></nolink> <nolink nlid="nl55" bibid="bib83" firstref="ref82"></nolink> <nolink nlid="nl56" bibid="bib52" firstref="ref83"></nolink> <nolink nlid="nl57" bibid="bib101" firstref="ref85"></nolink> <nolink nlid="nl58" bibid="bib38" firstref="ref88"></nolink> <nolink nlid="nl59" bibid="bib86" firstref="ref89"></nolink> <nolink nlid="nl60" bibid="bib35" firstref="ref91"></nolink> <nolink nlid="nl61" bibid="bib28" firstref="ref92"></nolink> <nolink nlid="nl62" bibid="bib85" firstref="ref93"></nolink> <nolink nlid="nl63" bibid="bib74" firstref="ref94"></nolink> <nolink nlid="nl64" bibid="bib116" firstref="ref95"></nolink> <nolink nlid="nl65" bibid="bib54" firstref="ref98"></nolink> <nolink nlid="nl66" bibid="bib55" firstref="ref99"></nolink> <nolink nlid="nl67" bibid="bib33" firstref="ref100"></nolink> <nolink nlid="nl68" bibid="bib87" firstref="ref105"></nolink> <nolink nlid="nl69" bibid="bib24" firstref="ref110"></nolink> <nolink nlid="nl70" bibid="bib98" firstref="ref111"></nolink> <nolink nlid="nl71" bibid="bib58" firstref="ref112"></nolink> <nolink nlid="nl72" bibid="bib76" firstref="ref113"></nolink> <nolink nlid="nl73" bibid="bib112" firstref="ref114"></nolink> <nolink nlid="nl74" bibid="bib95" firstref="ref115"></nolink> <nolink nlid="nl75" bibid="bib56" firstref="ref117"></nolink> <nolink nlid="nl76" bibid="bib69" firstref="ref118"></nolink> <nolink nlid="nl77" bibid="bib92" firstref="ref120"></nolink> <nolink nlid="nl78" bibid="bib18" firstref="ref122"></nolink> <nolink nlid="nl79" bibid="bib46" firstref="ref123"></nolink> <nolink nlid="nl80" bibid="bib75" firstref="ref126"></nolink> <nolink nlid="nl81" bibid="bib53" firstref="ref129"></nolink> <nolink nlid="nl82" bibid="bib88" firstref="ref130"></nolink> <nolink nlid="nl83" bibid="bib44" firstref="ref131"></nolink> <nolink nlid="nl84" bibid="bib13" firstref="ref133"></nolink> <nolink nlid="nl85" bibid="bib15" firstref="ref135"></nolink> <nolink nlid="nl86" bibid="bib22" firstref="ref138"></nolink> <nolink nlid="nl87" bibid="bib23" firstref="ref139"></nolink> <nolink nlid="nl88" bibid="bib81" firstref="ref140"></nolink> <nolink nlid="nl89" bibid="bib64" firstref="ref146"></nolink> <nolink nlid="nl90" bibid="bib102" firstref="ref147"></nolink> <nolink nlid="nl91" bibid="bib114" firstref="ref148"></nolink> <nolink nlid="nl92" bibid="bib40" firstref="ref150"></nolink> <nolink nlid="nl93" bibid="bib118" firstref="ref152"></nolink> <nolink nlid="nl94" bibid="bib96" firstref="ref153"></nolink> <nolink nlid="nl95" bibid="bib48" firstref="ref154"></nolink> <nolink nlid="nl96" bibid="bib57" firstref="ref155"></nolink> <nolink nlid="nl97" bibid="bib107" firstref="ref156"></nolink> <nolink nlid="nl98" bibid="bib117" firstref="ref157"></nolink> <nolink nlid="nl99" bibid="bib60" firstref="ref158"></nolink> <nolink nlid="nl100" bibid="bib94" firstref="ref159"></nolink> <nolink nlid="nl101" bibid="bib100" firstref="ref161"></nolink> <nolink nlid="nl102" bibid="bib63" firstref="ref163"></nolink> <nolink nlid="nl103" bibid="bib45" firstref="ref164"></nolink> <nolink nlid="nl104" bibid="bib68" firstref="ref166"></nolink> <nolink nlid="nl105" bibid="bib49" firstref="ref167"></nolink>
Header DbId: eric
DbLabel: ERIC
An: EJ1496175
AccessLevel: 3
PubType: Academic Journal
PubTypeId: academicJournal
PreciseRelevancyScore: 0
IllustrationInfo
Items – Name: Title
  Label: Title
  Group: Ti
  Data: Sharing Big Video Data: Ethics, Methods, and Technology
– Name: Language
  Label: Language
  Group: Lang
  Data: English
– Name: Author
  Label: Authors
  Group: Au
  Data: <searchLink fieldCode="AR" term="%22Joanne+W%2E+Golann%22">Joanne W. Golann</searchLink> (ORCID <externalLink term="https://orcid.org/0000-0001-9337-4674">0000-0001-9337-4674</externalLink>)<br /><searchLink fieldCode="AR" term="%22Lori+Bougher%22">Lori Bougher</searchLink><br /><searchLink fieldCode="AR" term="%22Richard+Hall%22">Richard Hall</searchLink><br /><searchLink fieldCode="AR" term="%22Thomas+J%2E+Espenshade%22">Thomas J. Espenshade</searchLink>
– Name: TitleSource
  Label: Source
  Group: Src
  Data: <searchLink fieldCode="SO" term="%22Sociological+Methods+%26+Research%22"><i>Sociological Methods & Research</i></searchLink>. 2026 55(1):340-372.
– Name: Avail
  Label: Availability
  Group: Avail
  Data: SAGE Publications. 2455 Teller Road, Thousand Oaks, CA 91320. Tel: 800-818-7243; Tel: 805-499-9774; Fax: 800-583-2665; e-mail: journals@sagepub.com; Web site: https://sagepub.com
– Name: PeerReviewed
  Label: Peer Reviewed
  Group: SrcInfo
  Data: Y
– Name: Pages
  Label: Page Count
  Group: Src
  Data: 33
– Name: DatePubCY
  Label: Publication Date
  Group: Date
  Data: 2026
– Name: SourceSuprt
  Label: Sponsoring Agency
  Group: SrcSuprt
  Data: National Science Foundation (NSF)
– Name: NumberContract
  Label: Contract Number
  Group: NumCntrct
  Data: 2214309
– Name: TypeDocument
  Label: Document Type
  Group: TypDoc
  Data: Journal Articles<br />Reports - Research
– Name: Subject
  Label: Descriptors
  Group: Su
  Data: <searchLink fieldCode="DE" term="%22Video+Technology%22">Video Technology</searchLink><br /><searchLink fieldCode="DE" term="%22Databases%22">Databases</searchLink><br /><searchLink fieldCode="DE" term="%22Information+Management%22">Information Management</searchLink><br /><searchLink fieldCode="DE" term="%22Information+Security%22">Information Security</searchLink><br /><searchLink fieldCode="DE" term="%22Information+Storage%22">Information Storage</searchLink><br /><searchLink fieldCode="DE" term="%22Access+to+Information%22">Access to Information</searchLink><br /><searchLink fieldCode="DE" term="%22Technological+Advancement%22">Technological Advancement</searchLink><br /><searchLink fieldCode="DE" term="%22Data+Use%22">Data Use</searchLink>
– Name: Subject
  Label: Geographic Terms
  Group: Su
  Data: <searchLink fieldCode="DE" term="%22New+Jersey%22">New Jersey</searchLink>
– Name: DOI
  Label: DOI
  Group: ID
  Data: 10.1177/00491241241277524
– Name: ISSN
  Label: ISSN
  Group: ISSN
  Data: 0049-1241<br />1552-8294
– Name: Abstract
  Label: Abstract
  Group: Ab
  Data: Data sharing and transparency are becoming more common across the social sciences. In this article, we provide an overview of ethical, methodological, and technological considerations and challenges when developing large video-based datasets intended to be shared across researchers. We cover data security, storage, and access as well as data documentation, tagging, and transcription. Our discussions are framed by our own efforts to create a secure and user-friendly database for the New Jersey Families Study, a two-week, in-home video study of 21 families with a 2- to 4-year-old child. In collecting over 11,470 hours of video data, the New Jersey Families Study is one of the very few large-scale video projects in the field of sociology. This project has provided us with a unique opportunity to explore video data management and data sharing techniques, particularly in light of a host of cutting-edge developments in data science.
– Name: AbstractInfo
  Label: Abstractor
  Group: Ab
  Data: As Provided
– Name: DateEntry
  Label: Entry Date
  Group: Date
  Data: 2026
– Name: AN
  Label: Accession Number
  Group: ID
  Data: EJ1496175
PLink https://search.ebscohost.com/login.aspx?direct=true&site=eds-live&db=eric&AN=EJ1496175
RecordInfo BibRecord:
  BibEntity:
    Identifiers:
      – Type: doi
        Value: 10.1177/00491241241277524
    Languages:
      – Text: English
    PhysicalDescription:
      Pagination:
        PageCount: 33
        StartPage: 340
    Subjects:
      – SubjectFull: Video Technology
        Type: general
      – SubjectFull: Databases
        Type: general
      – SubjectFull: Information Management
        Type: general
      – SubjectFull: Information Security
        Type: general
      – SubjectFull: Information Storage
        Type: general
      – SubjectFull: Access to Information
        Type: general
      – SubjectFull: Technological Advancement
        Type: general
      – SubjectFull: Data Use
        Type: general
      – SubjectFull: New Jersey
        Type: general
    Titles:
      – TitleFull: Sharing Big Video Data: Ethics, Methods, and Technology
        Type: main
  BibRelationships:
    HasContributorRelationships:
      – PersonEntity:
          Name:
            NameFull: Joanne W. Golann
      – PersonEntity:
          Name:
            NameFull: Lori Bougher
      – PersonEntity:
          Name:
            NameFull: Richard Hall
      – PersonEntity:
          Name:
            NameFull: Thomas J. Espenshade
    IsPartOfRelationships:
      – BibEntity:
          Dates:
            – D: 01
              M: 02
              Type: published
              Y: 2026
          Identifiers:
            – Type: issn-print
              Value: 0049-1241
            – Type: issn-electronic
              Value: 1552-8294
          Numbering:
            – Type: volume
              Value: 55
            – Type: issue
              Value: 1
          Titles:
            – TitleFull: Sociological Methods & Research
              Type: main
ResultId 1