All these resources and more linked here: AI Testing and Assurance LinkedIn Group
AI Assurance, Evaluation & Accountability
- On the Dangers of Stochastic Parrots: Can Language Models Be Too Big? – Emily M. Bender, Timnit Gebru, Angelina McMillan-Major, Margaret Mitchell
- The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity – Parshin Shojaee, Iman Mirzadeh, Keivan Alizadeh, Maxwell Horton, Samy Bengio, Mehrdad Farajtabar
- The Pursuit of Fairness in Artificial Intelligence Models: A Survey – Tahsin Kheya, Mohamed Jenek, Sunil Aryal
- Model Cards for Model Reporting – Margaret Mitchell, Simone Wu, Andrew Zaldivar, Parker Barnes, Lucy Vasserman, Ben Hutchinson, Elena Spitzer, Inioluwa Deborah Raji, Timnit Gebru
- Datasheets for Datasets – Timnit Gebru, Jamie Morgenstern, Briana Vecchione, Jennifer Wortman Vaughan, Hanna Wallach, Hal Daumé III, Kate Crawford
- Gender Shades: Intersectional Accuracy Disparities in Commercial Gender Classification – Joy Buolamwini, Timnit Gebru
- Closing the AI Accountability Gap: Defining an End-to-End Framework for Internal Algorithmic Auditing – Inioluwa Deborah Raji, Andrew Smart, Rebecca N. White, Margaret Mitchell, Timnit Gebru, Ben Hutchinson, Jamila Smith-Loud, Daniel Theron, Parker Barnes
- Closing the AI Accountability Gap: Audits and Evaluation in the Age of Artificial Intelligence – Inioluwa Deborah Raji
- The Fallacy of AI Functionality – Inioluwa Deborah Raji, I. Elizabeth Kumar, Aaron Horowitz, Andrew D. Selbst
- The Values Encoded in Machine Learning Research – Abeba Birhane, Pratyusha Kalluri, Dallas Card, William Agnew, Ravit Dotan, Michelle Bao
- Responsible Artificial Intelligence – From Principles to Practice – Virginia Dignum
- The ROI of AI Ethics – Zalabak, Dhamodharan, Lesieur, Magnusson, Kennedy, Krishnan
Software Testing Foundations
- Developer Testing in the IDE: Patterns, Beliefs, and Behavior – Beller, Gousios, Panichella, Proksch, Amann, Zaidman
- Explore It!: Reduce Risk and Increase Confidence with Exploratory Testing – Elisabeth Hendrickson
- Lessons Learned in Software Testing: A Context-Driven Approach – James Bach, Cem Kaner, Bret Pettichord
- Perfect Software: And Other Illusions about Testing – Gerald M. Weinberg
- The Impossibility of Complete Testing – Cem Kaner
- Toward a Theory of Test Data Selection – John B. Goodenough, Susan L. Gerhart
- An Experimental Evaluation of the Assumption of Independence in Multiversion Programming – John C. Knight, Nancy G. Leveson
- Software Testing Techniques – Boris Beizer
- Testing Computer Software – Cem Kaner, Jack Falk, Hung Quoc Nguyen
Systems, Quality, Risk & Measurement
- How Complex Systems Fail – Richard I. Cook, MD
- The Secrets of Consulting: A Guide to Giving and Getting Advice Successfully – Gerald M. Weinberg
- Becoming a Technical Leader: An Organic Problem-solving Approach – Gerald M. Weinberg
- Quality Software Management, Vol. 2: First-Order Measurement – Gerald M. Weinberg
- Quality Software Management, Vol. 1: Systems Thinking – Gerald M. Weinberg
- Exploring Requirements: Quality Before Design – Donald C. Gause, Gerald M. Weinberg
- Quality Software Management, Vol. 4: Anticipating Change – Gerald M. Weinberg
- Measuring and Managing Performance in Organizations – Robert D. Austin
- Thinking, Fast and Slow – Daniel Kahneman
- Tacit and Explicit Knowledge – Harry Collins
- Rethinking Expertise – Harry Collins, Robert Evans
- The Leprechauns of Software Engineering – Laurent Bossavit
- Antifragile: Things That Gain from Disorder – Nassim Nicholas Taleb
- Fooled by Randomness: The Hidden Role of Chance in Life and in the Markets – Nassim Nicholas Taleb
- The 48 Laws of Power – Robert Greene
- Search for the Real, and Other Essays – Hans Hofmann
- Letters to a Young Contrarian – Christopher Hitchens
AI Ethics, Power & Human Consequences
- The AI Con: How to Fight Big Tech’s Hype and Create the Future We Want – Emily M. Bender, Alex Hanna
- Artifictional Intelligence: Against Humanity’s Surrender to Computers – Harry Collins
- Against the Commodification of Education – Dagmar Monett, Gilbert Paquet
- AI Now Landscape Report – Kate Brennan, Amba Kak, Sarah Myers West
- Thinking Fast, Slow, and Artificial: How AI is Reshaping Human Reasoning and the Rise of Cognitive Surrender – Steven Shaw, Gideon Nave
- The AI Mirror: How to Reclaim Our Humanity in an Age of Machine Thinking – Shannon Vallor
- Race After Technology: Abolitionist Tools for the New Jim Code – Ruha Benjamin
- Artificial Unintelligence: How Computers Misunderstand the World – Meredith Broussard
- Atlas of AI: Power, Politics, and the Planetary Costs of Artificial Intelligence – Kate Crawford
- Algorithms of Oppression: How Search Engines Reinforce Racism – Safiya Umoja Noble
- Weapons of Math Destruction – Cathy O’Neil
AI Standards, Regulation & Governance
- Hiroshima Process International Code of Conduct for Organizations Developing Advanced AI Systems
- NIST Risk Management Framework
- EU AI Act
- IEEE Standard for Algorithmic Bias Considerations – IEEE 7003-2024
- OECD AI Principles
Safety, Reliability & Software Engineering
- Margaret Hamilton – NASA Biography and Apollo Software Engineering – Margaret Hamilton
- Apollo Flight Guidance Computer Software Collection [Hamilton] – Smithsonian National Air and Space Museum
- An Experimental Evaluation of the Assumption of Independence in Multiversion Programming – John C. Knight, Nancy G. Leveson
Worth Watching / Listening To
- Mystery AI Hype Theater 3000
- Better Offline
- Test Guild Test Automation Podcast
- The Vernon Richard Show
- Engineering Quality Podcast
- EvilTester – Software Testing
- Rapid Software Testing
- Association for Software Testing
- Testing In The Pub
The BEST resource page for software testers, hosted by Huib Schoots