From the course: AI Security and Responsible AI Practices by Pearson

Unlock this course with a free trial

Join today to access over 26,400 courses taught by industry experts.

Exploring model theft attacks

Exploring model theft attacks

“

Model theft refers to unauthorized access by either copying or extracting proprietary information from large language models or any AI model, you know, out there. This type of attacks targets the valuable intellectual property contained within the model itself, like, for example, that you're seeing here in the screen, the architecture of the model, or the training data that has been used to train the model, along with other intellectual property that is valuable to the organization. And the model itself can be the actual intellectual property as well. So all the things that the attacker can obtain and get access to by performing these types of attacks is by extracting the types of architectural components like parameters and weights that you have used to either fine-tune or to train the model to perform a specific task, right? So these are several key aspects of model theft attacks, right? be launched by disgruntled employees, like insider threats or malicious insiders that can leak…

Contents