Sunday 06 April 2025
The quest for privacy in machine learning has reached a new milestone. Researchers have developed a novel approach to fine-tune large language models while protecting sensitive user data. This breakthrough could pave the way for more widespread adoption of AI-powered applications, without compromising individual privacy.
The challenge is straightforward: as machine learning models become increasingly sophisticated, they require vast amounts of training data to learn and improve. However, this data often contains sensitive information about users, such as their personal preferences or medical history. To address this issue, researchers have been working on developing protocols that allow multiple parties to jointly train a model without revealing their individual data.
The new approach, dubbed PriFFT, uses a technique called function secret sharing (FSS) to enable secure fine-tuning of large language models. FSS is a method for encrypting and decrypting data using mathematical functions, rather than traditional encryption algorithms. This allows parties to share information about the model without revealing their individual contributions.
In the proposed system, multiple clients contribute small pieces of data to a central server, which then uses these fragments to fine-tune the language model. The twist is that each client’s contribution is encrypted using FSS, ensuring that even the server itself cannot access sensitive user information.
The researchers have tested PriFFT on several large language models, including BERT and its variants. Their results show that the system can achieve high accuracy while minimizing the amount of data shared between parties. This is a significant improvement over existing solutions, which often sacrifice performance for privacy.
One of the key advantages of PriFFT is its ability to scale up to large datasets, making it suitable for real-world applications. The system’s authors have demonstrated that PriFFT can handle datasets with millions of samples, without compromising on accuracy or efficiency.
Another notable aspect of PriFFT is its flexibility. The system can be adapted to different machine learning tasks and models, making it a versatile tool for researchers and developers. This flexibility could lead to widespread adoption across various industries, from healthcare to finance.
While there are still challenges to overcome before PriFFT becomes widely adopted, this breakthrough marks an important step towards achieving privacy-preserving machine learning. As the demand for AI-powered applications continues to grow, the need for secure and efficient data sharing methods will only intensify. With PriFFT, researchers have taken a significant leap forward in addressing this challenge, paving the way for more responsible development of AI technology.
Cite this article: “Secure Fine-Tuning of Pre-Trained Language Models via Function Secret Sharing”, The Science Archive, 2025.
Machine Learning, Privacy, Language Models, Data Sharing, Function Secret Sharing, Fss, Encryption, Bert, Accuracy, Scalability







