[Tried AI] Could Your Chats Train AI? How to Control Data Use
Translated from Korean and summarized by DistantNews. Read the original for the full story.
At a glance
- ChatGPT, Claude and Gemini may use user conversations, files or related activity to improve services or train models, depending on account settings.
- The guide explains how to disable or limit data use, while noting exceptions such as feedback submissions, safety reviews and some retained activity.
- Memory settings personalize future conversations but differ from model-training controls, and users are advised not to enter passwords, identification numbers or confidential records.
Generative AI services have moved into ordinary work and daily routines. ChatGPT, Gemini and Claude can draft emails, summarize documents, generate ideas, plan trips, revise writing and even act as conversation partners.
That convenience comes with a privacy question: depending on their settings, these services may use usersโ conversations and uploaded material to improve service quality or train models. Information entered casually can include names, birth dates, account and card details, investment records, passwords, authentication codes and internal company documents. If such data becomes training material, personal information or corporate secrets could be exposed.
South Koreaโs Personal Information Protection Commission recommends that users choose directly whether their data can be used for training. Most AI services turn the relevant option on by default, making it important to check the controls. In ChatGPT, users can open Settings and Data Controls, then manage the โImprove the model for everyoneโ option. Audio recording and video recording have separate controls. Business accounts default to refusing data sharing, while the Codex developer service requires separate settings. Users can also visit ChatGPTโs Privacy Portal and select โDo not use my content to train modelsโ for a permanent opt-out.
Claude places the control under Settings and Privacy, through โHelp improve the AI model.โ Anthropic says data may be anonymized and used for model improvement or marketing, but deletes it immediately when users block permission. Feedback and conversations selected for trust and safety review may still be used. Gemini uses an Activity setting rather than a separate training switch. Turning it on can save chats, attachments, audio, approximate location, visited websites and product-use information for model training and service improvement. Activity is automatically deleted after 18 months, although users can change that period. Stopping the setting excludes activity from training, but Google retains conversations for at least 72 hours for service provision and abuse prevention. It can also limit connected functions such as Gmail summaries, calendar creation and smart-home controls.
Memory is separate from model training. It stores context, preferences, work background and answer style for personalization, but sensitive prompts can still enter memory. Disabling memory and deleting chat history are separate steps. In all cases, the basic rule is simple: do not enter passwords, resident registration numbers, financial or medical records, or confidential documents into an AI service.
Originally published by Dong-A Ilbo in Korean. Translated, summarized, and contextualized automatically by DistantNews, with a note on how the source frames the story. Not individually reviewed before publishing. How this works.