arXiv · 2505.10315
Private Transformer Inference in MLaaS: A Survey
Abstract
Transformer models have revolutionized AI, powering applications like content generation and sentiment analysis. However, their deployment in Machine Learning as a Service (MLaaS) raises significant privacy concerns, primarily due to the centralized processing of sensitive user data. Private Transformer Inference (PTI) offers a solution by utilizing cryptographic techniques such as secure multi-party computation and homomorphic encryption, enabling inference while preserving both user data and model privacy. This paper reviews recent PTI advancements, highlighting state-of-the-art solutions and challenges. We also introduce a structured taxonomy and evaluation framework for PTI, focusing on balancing resource efficiency with privacy and bridging the gap between high-performance inference and data privacy.
Explore related subjects
Keep this discovery
Yang Li, Xinyu Zhou, Yitong Wang, Liangxin Qian, Jun Zhao. 2025-05-15. Private Transformer Inference in MLaaS: A Survey. https://arxiv.org/abs/2505.10315
Cite the original work for its findings. Save a collection to share your selection of sources.