Alemohammad, S., Casco-Rodriguez, J., Luzi, L., Humayun, A. I., Babaei, H., LeJeune, D., Siahkoohi, A., & Baraniuk, R. G. (2023). Self-Consuming Generative Models GO MAD. arXiv.org. https://arxiv.org/abs/2307.01850
Amershi, S., Chickering, D. M., Drucker, S. M., Lee, B., Simard, P., & Suh, J. (2019). ModelTracker: Redesigning performance analysis tools for machine learning. Proceedings of the ACM Conference on Human Factors in Computing Systems (CHI), 1-12. https://doi.org/10.1145/3290605.3300873
Baldwin, R., Cave, M., & Lodge, M. (2012). Understanding regulation: Theory, strategy, and practice. Oxford University Press.
Bansal, G., Nushi, B., Kamar, E., Lasecki, W. S., Weld, D. S., & Horvitz, E. (2019). Beyond accuracy: The role of Mental Models in Human-AI Team performance. Proceedings of the AAAI Conference on Human Computation and Crowdsourcing, 7, 2–11. https://doi.org/10.1609/hcomp.v7i1.5285
Bansal, G., Nushi, B., Kamar, E., Weld, D. S., Lasecki, W. S., & Horvitz, E. (2019). Updates in Human-AI Teams: Understanding and Addressing the Performance/Compatibility Tradeoff. Proceedings of the AAAI Conference on Artificial Intelligence, 33(01), 2429–2437. https://doi.org/10.1609/aaai.v33i01.33012429
Bengio, Y., de Leon Ferreira de Carvalho, A. C. P., Fox B., Nemer, M., Rivera, R. P., Zeng, Y., … & Khan, S. M. (2025). International AI Safety Report 2025. UK Department for Science, Innovation and Technology and AI Safety Institute. https://www.gov.uk/government/publications/international-ai-safety-report-2025
Bengio, Y., Hinton, G., Yao, A., Song, D., Abbeel, P., Darrell, T., Harari, Y. N., Zhang, Y., Xue, L., Shalev-Shwartz, S., Hadfield, G., Clune, J., Maharaj, T., Hutter, F., Baydin, A. G., McIlraith, S., Gao, Q., Acharya, A., Krueger, D., . . . Mindermann, S. (2024b). Managing extreme AI risks amid rapid progress. Science, 384(6698), 842–845. https://doi.org/10.1126/science.adn0117
Bengio, Y., Mindermann, S., Privitera, D., Besiroglu, T., Bommasani, R., Casper, S., … & Zhang, Y. (2024). International Scientific Report on the Safety of Advanced AI (INTERIM Report). arXiv preprint. https://arxiv.org/abs/2412.05282
Binns, R. (2018). Fairness in machine learning: Lessons from political philosophy. Proceedings of the 2018 Conference on Fairness, Accountability, and Transparency. Available: https://proceedings.mlr.press/v81/binns18a.html [2025, January 31]
Brundage, M., Avin, S., Wang, J., Belfield, H., Krueger, G., Hadfield, G., Khlaaf, H., Yang, J., Toner, H., Fong, R., Maharaj, T., Koh, P. W., Hooker, S., Leung, J., Trask, A., Bluemke, E., Lebensold, J., O’Keefe, C., Koren, M., . . . Anderljung, M. (2020, April 15). Toward trustworthy AI Development: Mechanisms for supporting verifiable claims. arXiv.org. https://arxiv.org/abs/2004.07213
Brown, T. B., Mann, B., Ryder, N., Subbiah, M., Kaplan, J., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., Agarwal, S., Herbert-Voss, A., Krueger, G., Henighan, T., Child, R., Ramesh, A., Ziegler, D. M., Wu, J., Winter, C., . . . Amodei, D. (2020). Language Models are Few-Shot Learners. arXiv.org. https://arxiv.org/abs/2005.14165
Bryson, J.J. (2020). The artificial intelligence of the ethics of artificial intelligence: an introductory overview for law and regulation. In: Dubber, M.D., Pasquale, F., Das, S. (eds.) The Oxford Handbook of Ethics of AI. Oxford University Press, New York. https://doi.org/10.1093/oxfordhb/9780190067397.013.1
Bundy J. (2021). When companies say “sorry,” it doesn’t always help their reputation | ASU News. Available: https://news.asu.edu/20210706-discoveries-when-companies-say-sorry-it-doesnt-always-help-their-reputation [2025, January 31]
Calo, R. (2021). Remark at AI debate 2: Artificial intelligence policy: not just a matter of principles. Available: https://www.youtube.com/watch?v=XoYYpLIoxf0 [2025, January 31]
Carlini, N., Athalye, A., Papernot, N., Brendel, W., Rauber, J., Tsipras, D., Goodfellow, I., Madry, A., & Kurakin, A. (2019, February 18). On evaluating adversarial robustness. arXiv.org. https://arxiv.org/abs/1902.06705
Cath, C., Wachter, S., Mittelstadt, B., Taddeo, M., & Floridi, L. (2017). Artificial Intelligence and the ‘Good Society’: the US, EU, and UK approach. Science and Engineering Ethics. https://doi.org/10.1007/s11948-017-9901-7
Chakravorti, B., (2024). AI and machine learning AI’s Trust Problem Twelve persistent risks of AI that are driving skepticism. Available: https://hbr.org/2024/05/ais-trust-problem [2025, Mach 10]
Hassabis, D. (2022). Responsible AI development and implementation [Conference presentation].
Hemmer, P., Schemmer, M., Kühl, N., Vössing, M., & Satzger, G. (2024). Complementarity in Human-AI Collaboration: Concept, Sources, and Evidence. arXiv preprint arXiv:2404.00029.
Hoffman, J.L., Lawson-Jenkins, K., Blum, J. (2006). Trust Beyond Security: An Expanded Trust Model. Communications of the ACM 49 (7), 95-101. https://doi.org/10.1145/1139922.1139924
Hoffmann, C. P., Lutz, C., & Meckel, M. (2014). Digital Natives or Digital Immigrants? The Impact of User Characteristics on Online Trust. Journal of Management Information Systems, 31(3), 138–171. https://doi.org/10.1080/07421222.2014.995538
Holweg, M. (2022). The ethical implications of AI adoption in organizations. Berkeley Technology Law Journal, 37(2), 1-28.
Holweg, M., et al. (2022). The reputational risks of AI. California Management Review. Available: https://cmr.berkeley.edu/2022/01/the-reputational-risks-of-ai/
Holweg, M., Younger, R., & Wen, Y. (2022). The reputational risks of AI. California Management Review. https://cmr.berkeley.edu/2022/01/the-reputational-risks-of-ai/ [2025, January 31]
Horvitz, E. (1999). Principles of mixed-initiative user interfaces. In Proceedings of the SIGCHI conference on Human factors in computing systems the CHI is the limit. 59-166. New York, NY ACM Press. – References – Scientific Research Publishing. (n.d.). https://www.scirp.org/reference/referencespapers?referenceid=79734
Jacovi, A., Marasovi´c, A., Miller, T., Goldberg, Y. (2021). Formalizing trust in artificial intelligence: Prerequisites, causes and goals of human trust in ai. In Proceedings of the 2021 ACM conference on fairness, accountability, and transparency, 624–635. https://doi.org/10.1145/3442188.3445923
Jobin, A., Ienca, M., & Vayena, E. (2019). The global landscape of AI ethics guidelines. Nature Machine Intelligence, 1(9), 389-399.
Kiela, D., Firooz, H., Mohan, A., Goyal, V., Taylor, A., Ferraro, F., … & Yang, Y. (2021). The hateful memes challenge: Detecting hate speech in multimodal memes. Advances in Neural Information Processing Systems, 34, 2611-2624.
Krafft, P. M., Young, M., Katell, M., Huang, K., & Bugingo, G. (2020). Defining AI in Policy versus Practice. In Proceedings of the AAAI/ACM Conference on AI, Ethics, and Society (AIES ’20). Association for Computing Machinery, New York, NY, USA, 72–78. https://doi.org/10.1145/3375627.3375835
Lange, D., Lee, P. M., & Dai, Y. (2010). Organizational Reputation: a review. Journal of Management, 37(1), 153–184. https://doi.org/10.1177/0149206310390963
Lechterman, T. (2023). The Concept of Accountability in AI Ethics and Governance. In Bullock, J. B., Chen, Y., Himmelreich, J., Hudson, V. M., Korinek, A., Young, M. M., & Zhang, B., (eds.), The Oxford Handbook of AI Governance. Oxford University Press. Available: https://philarchive.org/rec/LECTCO-8 [2025, January 31]
Lee, J. D. & See, K. (2004). Trust in Automation: Designing for Appropriate Reliance. Human Factors, 46(1), 50–80. https://doi.org/10.1518/hfes.46.1.50_30392
Levin, M. (2024). AI: A bridge toward diverse intelligence and humanity’s future. Tufts University Working Paper. https://doi.org/10.31234/osf.io/ez263
Lewis, P., Perez, E., Piktus, A., Petroni, F., Karpukhin, V., Goyal, N., Küttler, H., Lewis, M., Yih, W., Rocktäschel, T., Riedel, S., & Kiela, D. (2020). Retrieval-Augmented Generation for Knowledge-Intensive NLP tasks. https://proceedings.neurips.cc/paper/2020/hash/6b493230205f780e1bc26945df7481e5-Abstract.html
Liu, P., Yuan, W., Fu, J., Jiang, Z., Hayashi, H., & Neubig, G. (2022). Pre-train, Prompt, and Predict: A systematic survey of prompting methods in natural language processing. ACM Computing Surveys, 55(9), 1–35. https://doi.org/10.1145/3560815
Lukyanenko, R., Maass, W. & Storey, V. C. (2022). Trust in artificial intelligence: From a Foundational Trust Framework to emerging research opportunities. Electronic Markets, 32(4), 1993–2020. https://doi.org/10.1007/s12525-022-00605-4
Makovi, K., Sargsyan, A., Li, W., Bonnefon, J., & Rahwan, T. (2023). Trust within human-machine collectives depends on the perceived consensus about cooperative norms. Nature Communications, 14(1). https://doi.org/10.1038/s41467-023-38592-5
McNamara, A., Smith, J., Murphy-Hill, E. (2018). Does ACM’s code of ethics change ethical decision making in software development? In: Proceedings of the 2018 26th ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering, 729–733. https://doi.org/10.1145/3236024.3264833
Mittelstadt, B. D., Allo, P., Taddeo, M., Wachter, S., & Floridi, L. (2016). The ethics of algorithms: Mapping the debate. Big Data & Society, 3(2).
Mitchell, M., Wu, S., Zaldivar, A., Barnes, P., Vasserman, L., Hutchinson, B., & Gebru, T. (2019). Model cards for model reporting. Proceedings of the Conference on Fairness, Accountability, and Transparency (FAT), 220-229. https://doi.org/10.1145/3287560.3287596
Mökander, J., & Floridi, L. (2022). Operationalising AI governance through ethics-based auditing: an industry case study. AI And Ethics, 3(2), 451–468. https://doi.org/10.1007/s43681-022-00171-7
Morley, J., Floridi, L., Kinsey, L., & Elhalal, A. (2019). From What to How: An Initial Review of Publicly Available AI Ethics Tools, Methods and Research to Translate Principles into Practices. Science and Engineering Ethics, 26(4), 2141–2168. https://doi.org/10.1007/s11948-019-00165-5
Munn, L. (2022). The uselessness of AI ethics. AI And Ethics, 3(3), 869–877. https://doi.org/10.1007/s43681-022-00209-w
Nasr-Azadani, M. M., & Chatelain, J. (2024, March 19). The Journey to Trustworthy AI- Part 1: Pursuit of pragmatic frameworks. https://doi.org/10.48550/arXiv.2403.15457
Nogrady, B. (2016). The real risks of artificial intelligence. Available: https://www.bbc.com/future/article/20161110-the-real-risks-of-artificial-intelligence [2025, January 31]
OECD. (2019a). Recommendation of the Council on Artificial Intelligence. Available: https://legalinstruments.oecd.org/en/instruments/OECD-LEGAL-044 [2025, January 31]
OECD. (2019b). OECD Principles on AI. Available: https://www.oecd.org/going-digital/ai/principles/ [2025, January 31]
Parasuraman, R., Sheridan, T., & Wickens, C. (2000). A model for types and levels of human interaction with automation. IEEE Transactions on Systems Man and Cybernetics – Part a Systems and Humans, 30(3), 286–297. https://doi.org/10.1109/3468.844354
Park, B., & Rogan, M. (2019). Capability reputation, character reputation, and exchange partners’ reactions to adverse events. Academy of Management Journal, 62(2), 553–578. https://doi.org/10.5465/amj.2016.0445
Pavlidis, M. (2011). Designing for Trust. CEUR Workshop Proceedings. 731. https://www.researchgate.net/publication/232286675_Designing_for_Trust
Raji, I. D., Bender, E. M., Paullada, A., Denton, E., & Hanna, A. (2021). AI and the Everything in the Whole Wide World Benchmark. arXiv (Cornell University). https://doi.org/10.48550/arxiv.2111.15366
Raji, I. D., Xu, P., Honigsberg, C., & Ho, D. E. (2022). Outsider oversight: Designing a third party audit ecosystem for AI governance. arXiv (Cornell University). https://doi.org/10.48550/arxiv.2206.04737
Rasmussen, J. (1983). Skills, rules, and knowledge; signals, signs, and symbols, and other distinctions in human performance models, in IEEE Transactions on Systems, Man, and Cybernetics, vol. SMC-13, no. 3, 257-266, https://doi.org/10.1109/TSMC.1983.6313160
Rahwan, I., Cebrian, M., Obradovich, N., Bongard, J., Bonnefon, J. F., Breazeal, C., … & Wellman, M. (2019). Machine behaviour. Nature, 568(7753), 477-486.
Reisman, D., Schultz J., Crawford K., & Whittaker M. (2018). Algorithmic Impact Assessments Report: A Practical Framework for Public Agency Accountability. AI Now Institute, Available: https://ainowinstitute.org/publication/algorithmic-impact-assessments-report-2 [2025, January 31]
Rudin, C. (2019). Stop explaining black box machine learning models for high stakes decisions and use interpretable models instead. Nature Machine Intelligence, 1(5), 206–215. https://doi.org/10.1038/s42256-019-0048-x
Russell, S. (2019). Human compatible: Artificial intelligence and the problem of control. Viking.
Russell, S., & Norvig, P. (2016). Artificial Intelligence: A Modern Approach (4th ed.). Essex, United Kingdom: Pearson.
Sattlegger, A., Nitesh, B. (2024). Beyond principles: Embedding ethical AI risks in public sector risk management practice. In Proceedings of the 25th Annual International Conference on Digital Government Research (dg.o ’24). Association for Computing Machinery, New York, NY, USA, 70–80. https://doi.org/10.1145/3657054.3657063
Shneiderman, B. (2020). Human-Centered Artificial Intelligence: Reliable, safe & Trustworthy. arXiv.org. https://arxiv.org/abs/2002.04087
Shumailov, I., et al. (2024). Training with synthetic data causes artificial neural networks to learn less. Nature. Available: https://www.nature.com/articles/s41586-024-07566-y [2025, January 31]
Singh, J., Walden, I., Crowcroft, J., & Bacon, J. (2016). Responsibility & Machine Learning: Part of a Process. SSRN Electronic Journal. https://doi.org/10.2139/ssrn.2860048
SRF. (2024). Künstliche Intelligenz – Einkaufen der Zukunft: 800 Kameras schauen zu. Schweizer Radio Und Fernsehen. Available: https://www.srf.ch/sendungen/kassensturz-espresso/kassensturz/kuenstliche-intelligenz-einkaufen-der-zukunft-800-kameras-schauen-zu [2025, January 31]
Sturm, T., Gerlacha, J., Pumplun, L., Mesbah, N., Peters, F., Tauchert, C., Nan, N., & Buxmann, P. (2021). Coordinating human and machine learning for effective organization learning. MIS Quarterly, 45(3), 1581–1602. https://doi.org/10.25300/misq/2021/16543
Te’eni, D., Yahav, I., Zagalsky, A., Schwartz, D., Silverman, G., Cohen, D., Mann, Y., & Lewinsky, D. (2023). Reciprocal Human-Machine Learning: a theory and an instantiation for the case of message classification. Management Science. https://doi.org/10.1287/mnsc.2022.03518
Thiebes, S., Lins, S. & Sunyaev, A. (2021). Trustworthy artificial intelligence. Electron Markets 31, 447–464. https://doi.org/10.1007/s12525-020-00441-4
Tschopp, M., & Ruef, M. (2020). AI & Trust -Stop asking how to increase trust in AI. Available: https://www.researchgate.net/publication/339530999_AI_Trust_-Stop_asking_how_to_increase_trust_in_AI [2025, January 31]
Vakkuri, V., Kemell, K.-K., Kultanen, J., Siponen, M., Abrahamsson, P. (2019). Ethically aligned design of autonomous systems: industry viewpoint and an empirical study. https://doi.org/10.48550/arXiv.1906.07946
Van Rooy, D., & Vaes, K. (2024). Harmonizing human-AI synergy: Behavioral science in AI-integrated design. Proceedings of the International Design Conference, 2287-2296
Veale, M., Van Kleek, M., & Binns, R. (2018). Fairness and accountability design needs for algorithmic support in high-stakes public sector decision-making. Proceedings of the 2018 CHI Conference on Human Factors in Computing Systems.
Vought, R. T. (2020). Guidance for Regulation of Artificial Intelligence Applications. Available: https://www.whitehouse.gov/wp-content/uploads/2020/01/Draft-OMB-Memo-on-Regulation-of-AI-1-7-19.pdf [2025, January 31]
Wachter, S., Mittelstadt, B., & Floridi, L. (2017). Why a right to explanation of automated Decision-Making does not exist in the General Data Protection Regulation. International Data Privacy Law, 7(2), 76–99. https://doi.org/10.1093/idpl/ipx005
Wehmeier, S., McIntosh, C., Turnbull, J. & Ashby, M. (2005). Oxford advanced learner’s dictionary of current English. 7th ed. Oxford, United Kingdom: Oxford University Press
Wei, J., Bosma, M., Zhao, V. Y., Guu, K., Yu, A. W., Lester, B., Du, N., Dai, A. M., & Le, Q. V., & Google Research. (2022). Finetuned language models are zero-shot learners. In ICLR 2022. https://arxiv.org/pdf/2109.01652.pdf
Yang, Q., Steinfeld, A., Carolyn, R., & Zimmerman, J. (2020). Re-examining Whether, Why, and How Human-AI Interaction Is Uniquely Difficult to Design. http://dx.doi.org/10.1145/3313831.3376301
Zerilli, J., Knott, A., Maclaurin, J., & Gavaghan, C. (2019). Transparency in algorithmic and human decision-making: Is there a double standard? Philosophy & Technology, 32(4), 661-683.
Zheng, Q., Xu, Z., Choudhry, A., Chen, Y., Li, Y., & Huang, Y. (2023). Synergizing Human-AI Agency: A Guide of 23 Heuristics for Service Co-Creation with LLM-Based Agents. arXiv preprint arXiv:2310.15065. https://doi.org/10.48550/arXiv.2310.15065