Political Research Ethics: The Researcher in the War Room: Effective and ethical, in politics

Article P6-09

Who protects voters when a campaign tests a persuasive message on them, and who is left to check the work?

In brief

Campaign message tests on voters sit outside the American human-subjects rule, which reaches institutions and federal money rather than campaigns. The professional codes that do apply mostly govern asking people questions, not intervening in their political world. That leaves the three-gate screen of transparency, welfare and legitimacy, a house heuristic rather than a published standard anyone enforces. The welfare gate has little to weigh because general-election persuasion effects are close to zero, though small is not none. The ethical duty still lands on the person running the test.

How to use this

Before a campaign message test runs, write down your transparency answer, so proceeding is a decision rather than a habit. Put three questions to the design: would the people affected agree if they knew what was being done; are they left worse off; does the test damage political processes? Weigh welfare honestly even though general-election persuasion effects are close to zero, because small is not none. Keep the legitimacy question in the room while the message is designed. For each choice, ask whether it serves the decision you are making or the conclusion you already prefer. If you are a regulator, first measure how much campaign-side research is done, then publish which of it anyone reviews.

What the story is about

In 2024 and 2025, a research team ran undisclosed AI accounts inside r/ChangeMyView, a public forum on Reddit for arguments about contested political questions. Moderators disclosed what was happening, ethics specialists objected over consent and study design, and the work was stopped, as Science's news pages reported (O'Grady, 2025). Reddit then authorised release of the corpus, which Jaidka and Ahmed (2026) analysed in a conference paper: identity adoption in over two-thirds of comments, authority claims in nearly all of them, and a systematic inversion of how the forum's human arguers behave. For anyone who tests political messages, the lesson sits elsewhere. The controls that should have caught this study all sat upstream of it, and none of them engaged, because nothing required anyone to register it before it ran. That is one documented case, not a measured pattern.

American law does require an ethics check for a great deal of research on people. A review board reads the study plan before it starts and can insist on changes to protect the people taking part. The rule behind that duty reaches research "conducted, supported, or otherwise subject to regulation by any Federal department or agency" (United States, 2026). A consultancy testing messages for a candidate is none of those three things, so the duty never arrives. That boundary has no principled basis. Mollen (2024a), in a peer-reviewed journal, finds research-ethics demands inconsistent across categories of real-world technology research, argues that "there are no meaningful differences to justify it", and warns that "it creates the possibility of regulatory evasion at the cost of populations' due protection".

The professional bodies have codes, and the ones that matter here govern asking people questions rather than doing something to them. WAPOR's code protects a respondent's privacy and confidentiality and the right to withdraw, and it is candid about its own limits: "Membership implies no guarantee of qualifications or Code compliance, but it does imply acceptance of the Code by the member." (World Association for Public Opinion Research, 2021). In that code's published text, the words experiment, intervention, treatment, informed consent, persuasion and debrief appear zero times. AAPOR's code, reissued in June 2026, claims a wider reach, carries a duty to refuse work that conflicts with it, and blocks campaigning dressed up as research (American Association for Public Opinion Research, 2026). Neither document hands anyone a gate to walk a message experiment through before it runs.

Something has to fill the gap. This article uses a three-gate screen: transparency, welfare, legitimacy. It comes from our companion series on behavioural science, where it was built to judge nudges, and it is turned here on campaign research. Its transparency gate draws on Bovens (2009), who separates nudges that would still work if people could see what was being done from nudges that would not. The three questions become: would the people involved agree if they knew what was being done to them? Are they left worse off? Does the test damage political processes themselves? The screen is our own tool, not a published standard, and nothing in this literature formalises it. The closest published relative is a list of prompting questions for online A/B tests, the small trials that show different versions of a page to different visitors, from Polonioli et al. (2023), who note that "the ethical dimension of A/B testing has been neglected".

If the people on the receiving end knew what was happening, would they go along with it? Field experiments make that question harder, because the people affected are not a small enrolled group. McDermott and Hatemi (2020), writing in PNAS, argue that "experimenters now target and affect whole societies, releasing interventions into a living public, often without sufficient review or controls", and that debriefing is routinely skipped. Desposato (2018) showed two hypothetical field-experiment designs to subjects and to scholars, varying the details, and reported that "Both scholars and subjects reacted negatively to deception and to experiments without informed consent".

The welfare gate asks whether the people in a study end up worse off than they would have been otherwise. Desposato (2022), in a peer-reviewed essay, works through audits of public officials and finds "there are a number of potential harms of such studies which are generally not captured by the standard human subjects framework". He names aggregate harms, which land on a group rather than on one person, and response-delay harms: the constituent whose letter waits behind a researcher's fake one. The best field evidence says general-election persuasion effects are close to zero, so this gate has very little to weigh. That cuts against the urgency of the harm claim, and it should be said plainly.

The legitimacy gate asks whether the research damages political processes themselves. Political scientists have written the nearest thing to a rule. The American Political Science Association's ethics guide, updated in May 2023, says researchers "should not compromise the integrity of political processes for research purposes without the consent of individuals that are directly engaged by the research process" (American Political Science Association, 2023). The same guide refuses the rulebook framing: "These principles are not intended to be rules, requirements, or prohibitions, and they are not presented as a checklist for ethical research." Barnfield (2023) supplies the reason the ethical answer and the methodological answer agree. Asking what is wrong with giving participants false information about the world, he answers that "misinformation is bad for inference too": "Misinformation moves us away from answering questions about the political world effectively". Tell the truth, and the research improves as well.

So who protects the voter in a campaign message test? The American human-subjects rule stops at institutions and federal money. The professional codes cover asking, and they cover it well; they say next to nothing about intervening. The three-gate screen is the only instrument here that reaches the work, and it is ours, not a standard anyone enforces. Outside the room, nobody checks.

So what

External review is missing, and the ethical duty is still there. It has nowhere to go but the person running the test. A party strategist and a consultant are, between them, the entire review process for a message test. How much campaign research goes unchecked is unknown. As far as this research could establish, there is no published estimate of the share of campaign-side experiments that receive ethical review of any kind. That absence is itself the finding worth carrying out of this article: the question is live, and as far as this research could establish, nobody has counted how often anyone answers it.

For political parties

Parties and their consultants cannot reach for a published screen for campaign research, because none exists. The instruments that come closest were written for someone else: Polonioli et al. (2023) for companies running online A/B tests, the American Political Science Association (2023) for university researchers. That leaves the three gates, and they cost less than the fieldwork. Write your transparency answer down before the test runs, so proceeding is a decision rather than a habit. Weigh welfare honestly even though persuasion effects are small, since small is not the same as none. Keep the legitimacy question in the room while the message is being designed, not after it has aired. Then apply one test to each choice: does this serve the decision you are making, or the conclusion you already prefer?

For government

Government could require registration or ethical review for campaign message tests. Any such rule would be built on an unmeasured gap. Nothing here establishes how many campaign tests happen, how many reach voters without their knowledge, or how many would fail a review if one existed. There is also no enforcement record to learn from: no complaint under the polling code and no proceeding under Europe's AI Act has settled a campaign-research question. An absent record is not an absent case. A regulator who wants to act has two honest first steps. Measure how much campaign-side research is done at all. Then publish which of it, if any, anyone reviews.

References

American Association for Public Opinion Research (2026) AAPOR code of professional ethics and practices. Alexandria, VA: AAPOR, June. Available at: https://aapor.org/wp-content/uploads/2026/06/AAPOR-2026-Code-of-Ethics-Final-.pdf (Accessed: 14 September 2026).

American Political Science Association (2023) A guide to professional ethics in political science. 2nd edn, updated May 2023. Washington, DC: APSA. Available at: https://apsanet.org/Portals/54/diversity%20and%20inclusion%20prgms/Ethics/APSA-Ethics-Guide-Updated-May2023.pdf (Accessed: 14 September 2026).

Barnfield, M. (2023) 'Misinformation in experimental political science', Perspectives on Politics, 21(4), pp. 1210–1220. Available at: https://doi.org/10.1017/s1537592722003115 (Accessed: 14 September 2026).

Bovens, L. (2009) 'The ethics of nudge', in Grüne-Yanoff, T. and Hansson, S.O. (eds) Preference Change: Approaches from Philosophy, Economics and Psychology. Dordrecht: Springer. Available at: https://doi.org/10.1007/978-90-481-2593-7_10 (Accessed: 30 September 2026).

Desposato, S. (2018) 'Subjects and scholars' views on the ethics of political science field experiments', Perspectives on Politics, 16(3), pp. 739–750. Available at: https://doi.org/10.1017/s1537592717004297 (Accessed: 14 September 2026).

Desposato, S. (2022) 'Public impacts from elite audit experiments: aggregate and response delay harms', Political Studies Review, 20(2), pp. 217–227. Available at: https://doi.org/10.1177/14789299211059657 (Accessed: 14 September 2026).

Jaidka, K. and Ahmed, S. (2026) 'How far did they go? The persuasive tactics of covert LLM agents in a discontinued field experiment', in Proceedings of the Language Resources and Evaluation Conference. Available at: https://doi.org/10.63317/477ns4y77c92 (Accessed: 14 September 2026).

McDermott, R. and Hatemi, P.K. (2020) 'Ethics in field experimentation: a call to establish new standards to protect the public from unwanted manipulation and real harms', Proceedings of the National Academy of Sciences, 117(48), pp. 30014–30021. Available at: https://doi.org/10.1073/pnas.2012021117 (Accessed: 14 September 2026).

Mollen, J. (2024a) 'Towards a research ethics of real-world experimentation with emerging technology', Journal of Responsible Technology, 20, 100098. Available at: https://doi.org/10.1016/j.jrt.2024.100098 (Accessed: 14 September 2026).

O'Grady, C. (2025) ''Unethical' AI research on Reddit under fire', Science, 388(6747), 8 May, pp. 570–571. Available at: https://doi.org/10.1126/science.ady8074 (Accessed: 14 September 2026).

Polonioli, A., Ghioni, R., Greco, C., Juneja, P., Tagliabue, J., Watson, D. and Floridi, L. (2023) 'The ethics of online controlled experiments (A/B testing)', Minds and Machines, 33(4), pp. 667–693. Available at: https://doi.org/10.1007/s11023-023-09644-y (Accessed: 14 September 2026).

United States (2026) Protection of human subjects, 45 CFR 46, current edition. Washington, DC: Office of the Federal Register. Available at: https://www.ecfr.gov/current/title-45/subtitle-A/subchapter-A/part-46 (Accessed: 14 September 2026).

World Association for Public Opinion Research (2021) WAPOR code of professional ethics and practices, adopted 17 September 2021. Available at: https://wapor.org/about-wapor/code-of-ethics/ (Accessed: 14 September 2026).

Explore the idea

Let’s talk

Invisible forces shape your world — until you hire Latenta®

Contact