Publication:

Trustworthy AI: Ensuring Reliability and Accountability from Models to Agents

Loading...
Thumbnail Image

Date

2026-05-08

Published Version

Published Version

Journal Title

Journal ISSN

Volume Title

Publisher

The Harvard community has made this article openly available. Please share how this access benefits you.

Research Projects

Organizational Units

Journal Issue

Citation

Long, Xuan. 2026. Trustworthy AI: Ensuring Reliability and Accountability from Models to Agents. Doctoral Dissertation, Harvard University Graduate School of Arts and Sciences.

Abstract

In this thesis, we develop algorithms with theoretical guarantees for ensuring reliability and accountability of Machine Learning (ML) systems. As ML systems evolve from predictive models to generative models and autonomous agents, the landscape of trustworthy AI shifted. This thesis introduces tools grounded in information theory, optimization, and statistical learning to mitigate bias, reduce arbitrary decisions, ensure content provenance, and evaluate LLM-driven agents in autonomous settings. Towards mitigating bias and arbitrariness in traditional ML models, we introduce a kernel-based method to achieve multiaccuracy across complex subpopulations that tradi- tional demographic categories may overlook. We also develop methods to address predictive multiplicity—where equally accurate models yield conflicting individual predictions. We ensure the accountability in generative AI through watermarking large language models (LLMs). We characterize the information-theoretic trade-off between watermark detection and text distortion and deriving optimal watermarking strategies by leveraging optimal transport and coding theory. Empirical evaluations show our watermarks achieve a superior detection-quality tradeoff across language generation and coding tasks. Finally, we evaluate autonomous LLM agents in multi-agent environments through the first simulator of a fully LLM-driven supply chain. While agents can outperform human experts, reducing costs by up to 67%, we identify systemic risks such as costly tail events.

Description

Other Available Sources

Research Data

Keywords

Information Theory, Large Language Models, Machine Learning, Multi-Agent Systems, Optimization, Trustworthy AI, Applied mathematics

Terms of Use

This article is made available under the terms and conditions applicable to Other Posted Material (LAA), as set forth at Terms of Service

Endorsement

Review

Supplemented By

Related Stories