Trustworthy AI: Securing Sensitive Data in Large Language Models

doi:10.48550/arXiv.2409.18222

Trustworthy AI: Securing Sensitive Data in Large Language Models

Large Language Models (LLMs) have transformed natural language processing (NLP) by enabling robust text generation and understanding. However, their deployment in sensitive domains like healthcare, finance, and legal services raises critical concerns about privacy and data security. This paper proposes a comprehensive framework for embedding trust mechanisms into LLMs to dynamically control the disclosure of sensitive information. The framework integrates three core components: User Trust Profiling, Information Sensitivity Detection, and Adaptive Output Control. By leveraging techniques such as Role-Based Access Control (RBAC), Attribute-Based Access Control (ABAC), Named Entity Recognition (NER), contextual analysis, and privacy-preserving methods like differential privacy, the system ensures that sensitive information is disclosed appropriately based on the user's trust level. By focusing on balancing data utility and privacy, the proposed solution offers a novel approach to securely deploying LLMs in high-risk environments. Future work will focus on testing this framework across various domains to evaluate its effectiveness in managing sensitive data while maintaining system efficiency.

Publication:

arXiv e-prints

Pub Date:

September 2024

DOI:

10.48550/arXiv.2409.18222

arXiv:

arXiv:2409.18222

Bibcode:

2024arXiv240918222F

Keywords:

Computer Science - Artificial Intelligence

E-Print:

40 pages, 1 figure

NASA/ADS

Trustworthy AI: Securing Sensitive Data in Large Language Models

Abstract