Skip to main content
Cornell University
Learn about arXiv becoming an independent nonprofit.
We gratefully acknowledge support from the Simons Foundation, member institutions, and all contributors. Donate
arxiv logo > cs > arXiv:2605.02640

Help | Advanced Search

Computer Science > Artificial Intelligence

(cs)
[Submitted on 4 May 2026]

Title:Trustworthy AI Suffers from Invariance Conflicts and Causality is The Solution

Authors:Ruta Binkyte, Ivaxi Sheth, Zhijing Jin, Mohammad Havaei, Bernhard Schölkopf, Mario Fritz
View a PDF of the paper titled Trustworthy AI Suffers from Invariance Conflicts and Causality is The Solution, by Ruta Binkyte and 5 other authors
View PDF HTML (experimental)
Abstract:As artificial intelligence (AI), including machine learning (ML) models and foundation models (FMs), is increasingly deployed in high-stakes domains, ensuring their trustworthiness has become a central challenge. However, the core trustworthy AI objectives, such as fairness, robustness, privacy, and explainability, are hard to achieve simultaneously, especially while preserving utility. This position paper argues that causality is necessary to understand and balance trade-offs in performance and multiple objectives of trustworthy AI. We ground our arguments in re-interpreting trustworthy AI trade-offs as incompatible invariance requirements under different changes to the data-generating process. We then illustrate that causality provides a unifying framework for understanding how trade-offs in trustworthy AI arise, and how they can be softened or resolved through selective invariance. This perspective applies to both classical ML models and large-scale FMs. Our paper discusses how causal assumptions may be applied explicitly or implicitly in modern large-scale systems. Finally, we outline open challenges and opportunities for using causality to build more trustworthy AI.
Comments: Accepted at ICML'2026
Subjects: Artificial Intelligence (cs.AI)
Cite as: arXiv:2605.02640 [cs.AI]
  (or arXiv:2605.02640v1 [cs.AI] for this version)
  https://doi.org/10.48550/arXiv.2605.02640
arXiv-issued DOI via DataCite (pending registration)

Submission history

From: Ruta Binkyte [view email]
[v1] Mon, 4 May 2026 14:26:28 UTC (2,070 KB)
Full-text links:

Access Paper:

    View a PDF of the paper titled Trustworthy AI Suffers from Invariance Conflicts and Causality is The Solution, by Ruta Binkyte and 5 other authors
  • View PDF
  • HTML (experimental)
  • TeX Source
license icon view license

Current browse context:

cs.AI
< prev   |   next >
new | recent | 2026-05
Change to browse by:
cs

References & Citations

  • NASA ADS
  • Google Scholar
  • Semantic Scholar
Loading...

Bookmark

BibSonomy Reddit

arXivLabs: experimental projects with community collaborators

arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website.

Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them.

Have an idea for a project that will add value for arXiv's community? Learn more about arXivLabs.

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)
  • About
  • Help
  • contact arXivClick here to contact arXiv Contact
  • subscribe to arXiv mailingsClick here to subscribe Subscribe
  • Copyright
  • Privacy Policy
  • Web Accessibility Assistance
  • arXiv Operational Status