Trustworthy AI: Securing Sensitive Data in Large Language Models

Feretzakis, Georgios; Verykios, Vassilios S.

Abstract:Large Language Models (LLMs) have transformed natural language processing (NLP) by enabling robust text generation and understanding. However, their deployment in sensitive domains like healthcare, finance, and legal services raises critical concerns about privacy and data security. This paper proposes a comprehensive framework for embedding trust mechanisms into LLMs to dynamically control the disclosure of sensitive information. The framework integrates three core components: User Trust Profiling, Information Sensitivity Detection, and Adaptive Output Control. By leveraging techniques such as Role-Based Access Control (RBAC), Attribute-Based Access Control (ABAC), Named Entity Recognition (NER), contextual analysis, and privacy-preserving methods like differential privacy, the system ensures that sensitive information is disclosed appropriately based on the user's trust level. By focusing on balancing data utility and privacy, the proposed solution offers a novel approach to securely deploying LLMs in high-risk environments. Future work will focus on testing this framework across various domains to evaluate its effectiveness in managing sensitive data while maintaining system efficiency.

Comments:	40 pages, 1 figure
Subjects:	Artificial Intelligence (cs.AI)
Cite as:	arXiv:2409.18222 [cs.AI]
	(or arXiv:2409.18222v1 [cs.AI] for this version)
	https://doi.org/10.48550/arXiv.2409.18222

Computer Science > Artificial Intelligence

Title:Trustworthy AI: Securing Sensitive Data in Large Language Models

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators