Active Computing & AI Society, Politics & Law

From Human to Machine: The Ethics of How AI Is Reshaping Data in Scientific Research

In plain English

AI plain-English summary

AI systems are now collecting and generating the data that scientists rely on—replacing human observation, patient interviews, and manual data cleaning with algorithms. This project asks what ethical problems arise when the raw material of research shifts from human interaction to machine mediation. The problem is that most ethical guidelines for AI focus on consumer or clinical uses, not on how AI reshapes the scientific process itself. Yet data quality determines research quality. If AI introduces hidden biases, strips away contextual understanding, or erodes participant trust, the resulting science could be flawed—even if the AI appears efficient. Healthcare is a key case: AI chatbots may gather more honest disclosures from patients than human clinicians, but they may also miss subtle cues or misinterpret responses. If this research succeeds, it will produce practical guidelines for policymakers, funders, and researchers on when and how to use AI in data collection responsibly. The impact would be felt across medical research, biobanks, and clinical databases—systems that quietly underpin diagnostics, drug development, and public health decisions. Without such guidance, the growing reliance on AI-mediated data risks undermining the very validity of the science it is meant to accelerate.

View original technical description
There is a large body of literature on the ethical issues linked to the use of AI applications; however, less attention has been given to the ethical issues linked to using AI in scientific discovery, despite these systems being increasingly used. A particularly underresearched and important area to investigate is the role and the workings of AI in data collection, as the quality of data fundamentally determines the quality of research. Scientific research has traditionally relied on data collected through human interaction. Increasingly, however, AI systems are transforming data collection across disciplines. This shift alters the epistemic nature of research: data are less shaped by human judgment and relational context, and more by algorithmic design and interaction constraints. Given the importance of data in scientific research, it is important to understand precisely which types of AI systems can be used at the data collection or generation stage and what the associated ethical implications are. Healthcare illustrates this trend, as AI systems can now gather clinical information that was traditionally obtained by clinicians through conversations with patients. AI can also generate new data and clean and categorise existing data. These outputs can then be used to feed databases for researchers, such as national biobanks, secure analytics platforms, or clinical record search systems. In some cases, AI-mediated data could be of higher quality than data collected or generated by human researchers. For instance, in psychotherapy, people may disclose more to AI chatbots than to humans, enabling more accurate and comprehensive data. Well-designed AI also has the potential to reduce biases that might arise in human judgement. Furthermore, using AI to “clean” and categorise data may contribute to more efficient scientific research. Yet, there is potential for significant ethical concern. Like AI applications used outside of the research context, issues of bias introduced through inadequate design may persist. AI systems used to collect data may lack the ability to acquire contextual understanding and thereby lead to incomplete or misleading data. Furthermore, research participants may struggle to trust AI systems or may be uncertain about how their data will be used. As AI becomes more integrated into everyday life, scientific research is progressively drawing on data acquired and/or managed through AI-mediated interactions. This shift transforms the nature of the data itself, affecting its structure, context, and interpretability, and may have significant implications for the validity, ethics, and reliability of research findings. The central challenge, then, is to understand how AI alters the very nature of research data and what this means for the knowledge claims that follow. This research project, therefore, asks: how does the use of AI change the nature of data used in scientific research, and what are the associated ethical consequences? Using an empirical bioethics methodology and drawing on medical research as a case study, this study will generate novel guidelines to support policymakers, funders, and researchers in promoting responsible scientific research.

View the original record at the funder ↗

Researchers

Aurelia Sauerbrei (Principal Investigator)

Related Research

Grants with similar aims, by meaning.

How Humans Shape AI for Life Sciences Research
Synthetic Metascience: Tracing Artificial Intelligence-generated epistemic shifts in scientific research practice and cultures
Antecedents and Consequences of Trust in Artificial Agents
AI in Criminology Research: Mapping Methodological Shifts and Epistemic Risks
Developing an evidence-based framework for reducing epistemic trespassing when using generative artificial intelligence: a mixed methods study

Original classification

Fellowship

Plain English summaries and category classifications on this site are generated by AI and may not perfectly reflect the original research.