TIPS & TRICKS

ChatGPT exploited to leak Windows copyright information: Serious security vulnerability caused by AI

Chatgpt Bị Khai Thác Để Lộ Thông Tin Bản Quyền Windows: Lỗ Hổng Bảo Mật Nghiêm Trọng Do Ai Gây Ra

Researchers have successfully exploited ChatGPT to retrieve valid Windows license keys, including those previously belonging to Wells Fargo. This highlights the potential risks of data leakage from large AI models. This exploitation method turns queries into a word-guessing game, thereby bypassing security filters through prompt engineering and prompt obfuscation techniques. This technique is particularly effective with new AI models such as GPT-4o and GPT-4o-mini.

AI intrusion techniques and potential security loopholes

By using prompt engineering techniques, researchers created complex commands combined with prompt obfuscation—inserting sensitive keywords in HTML or other formats—to trick AI keyword filters. Through this method, ChatGPT revealed real Windows keys that the model had learned from public sources or online leaks. This is considered a form of AI jailbreaking, forcing the system to perform unintended actions, bypassing content barriers and censorship, and turning AI into a tool for disclosing sensitive information.

Risks and solutions for businesses in the AI era

This incident serves as a warning to businesses and organizations regarding data security risks. AI can inadvertently reveal internal information that has been posted on the internet, including license keys, API keys, or other sensitive data. Current AI filters rely primarily on keywords, making them easy to bypass using camouflage techniques. Meanwhile, public data containing sensitive information is still used to train models, creating a massive vulnerability. Additionally, AI is susceptible to exploitation through social engineering scenarios like games or role-playing, making it easy for users to trick the system into revealing unintended information.

To mitigate risks, experts recommend that businesses audit public data to avoid exposing sensitive information on the internet, while simultaneously implementing logic-layer protections instead of relying solely on keyword filters. Detecting fraudulent behavior, identifying social engineering scenarios, and updating AI monitoring systems will help limit the ability of AI to be exploited for information disclosure. Following the incident, OpenAI implemented measures to patch the vulnerability and prevent prompt exploitation, but risks from other jailbreak variants still exist, requiring users and organizations to remain vigilant when using AI.

The incident where ChatGPT was “jailbroken” to reveal Windows license keys is not only a warning for the AI industry but also for all businesses, organizations, and individual users regarding data protection, sensitive information management, and the reassessment of AI usage processes in daily operations. This is a crucial step toward ensuring safety in the AI era, where technology can be both a powerful supporting tool and a source of potential security risks if not properly controlled.

Share: 𝕏 P in
Question and answer (0 comments)

Table of contents
  1. Top