
Twitch has quietly introduced a new control that allows users to stop Amazon from using their channel content to train generative AI models. The update, reflected in a revised support page, gives streamers a way to block their streams, VODs, clips, chat messages, and channel images or text from being used in future AI development by Amazon. Previously, this data was automatically included in training datasets with no simple way to opt out.
The change is significant because it marks the first time Twitch has provided a direct user-level opt-out for AI training since the practice came under public scrutiny. Amazon acquired Twitch in 2014 and has since integrated the platform's vast repository of live and recorded video content into its broader AI research and product development efforts. The new setting is located under the security section of each user's account preferences, at www.twitch.tv/settings/security, where a toggle labeled "Allow training" now appears.
Twitch's updated support documentation explains that when a user allows training, their content may be used to improve generative AI models developed by Amazon. The documentation offers a concrete example: audio from a user's stream might be used to refine speech-to-text models, which could improve captioning features on Twitch as well as across Amazon's ecosystem of products and services. The wording makes clear that the data is not limited to Twitch-specific applications but may feed into Amazon's wider AI ambitions, including models that generate or synthesize text, audio, images, or video.
A practice hidden in plain sight
Although the opt-out is new, the underlying practice is not. For years, Twitch users have had their content used to train Amazon's AI models without explicit consent or widespread awareness. The first notable confirmation came in 2024, when Twitch's then-chief monetization officer, Mike Minton, responded to a question at an event held by The Information. Asked whether Amazon uses Twitch content for AI training, Minton replied, "Yeah, for sure." He added that such use was "obviously within the bounds of user trust and within the bounds of privacy regulations, which vary all around the world," and that Twitch had "a role to play in that."
That candid acknowledgment did not trigger an immediate policy shift. Instead, the announcement now, more than a year later, seems to be a response to growing user unease and regulatory pressure around AI training on user-generated content. The timing coincides with broader industry debates about consent, compensation, and the ethics of using publicly available content to train commercial AI systems.
Twitch's support page does not explicitly state when Amazon began using Twitch content for AI training, nor does it disclose which models or products have already benefited from this data. Ars Technica has asked Twitch for details about the origin of the practice but has not yet received a response. The lack of transparency is a recurring theme in online discussions, where users have expressed uncertainty about how their content has been used and whether they have any recourse.
Why users are concerned
The automatic opt-in model means that many streamers were part of Amazon's AI training datasets without knowingly signing up. This has raised several concerns among the Twitch community. One major issue is financial: creators who produce unique content, such as commentary, gameplay, music, or educational material, often rely on the exclusivity and originality of their work. If Amazon can use that content to train AI systems that generate similar content, it could potentially undercut the creators' business models. For example, an AI trained on thousands of hours of cooking streams might someday generate cooking videos with voiceover, captions, or even visual elements that mimic a particular streamer's style, reducing the demand for that streamer's own content.
Another concern relates to broader distrust of Amazon. The company has faced criticism over its treatment of workers, its data collection practices, and its dominance in cloud computing and retail. Some users do not want to contribute, even indirectly, to a corporate giant they view as harmful. Others worry about the privacy implications of having their voice, likeness, and personal moments captured in video streams used to train AI without permission. Even though Twitch streams are public, the context and intent of a streamer's content may differ substantially from being repurposed for AI research.
There is also a less tangible but deeply felt anxiety about creativity and originality. Many streamers view their channels as personal expressions, and the idea that an AI might learn to replicate their tone, humor, or presentation style feels invasive. The support page's acknowledgment that audio can be used to improve speech-to-text models may seem benign, but the underlying principle—using individual creators' voices as training data—strikes many as a form of exploitation.
On the other hand, some users are comfortable with the arrangement. They may see AI training as a beneficial use of their content, especially if it leads to better captioning, translation, or accessibility features. Others may simply not care or may view the exposure as an unavoidable reality of broadcasting on a platform owned by Amazon. Twitch's new opt-out setting accommodates both perspectives, but it does not address the fundamental question of whether Amazon should have sought permission in the first place.
How the opt-out works
To use the new control, a user must navigate to the security settings page on Twitch's website. There, a setting labeled "Allow training" appears, and the user can toggle it off to prevent future use of their content for Amazon's generative AI models. The support page clarifies that turning off the setting will not affect any data that has already been used in previous training runs. It only applies to "future training" of models.
This limitation is important. If Amazon has already trained models on years of Twitch content, those models remain in existence and may continue to be used or refined, even if a user opts out now. The opt-out is forward-looking only, which means the vast majority of Twitch's existing content—including everything uploaded before the setting was introduced—may already be embedded in Amazon's AI systems. Twitch does not offer a way to retroactively delete data from training sets, and no such feature is planned.
The setting is available to all Twitch users, not just partners or affiliates. It covers all forms of content listed in the support page: streams, VODs, clips, stream chats, and pictures and text on a user's channel. This broad scope suggests that Amazon's AI training pipeline ingests not only video but also metadata, chat logs, and textual descriptions associated with a channel. The toggle is off by default, meaning users who do not change their settings will continue to allow Amazon to use their content.
The broader AI training landscape
Twitch's decision to offer an opt-out comes at a time when the AI industry is facing intense scrutiny over the use of public and user-generated data. Several major companies, including OpenAI, Google, and Meta, have been sued by authors, artists, and news outlets over alleged unauthorized use of copyrighted or personal material in training datasets. While Twitch's situation differs—users typically grant broad rights to the platform when they agree to its terms of service—the underlying tension is the same: to what extent can platforms monetize user content for AI purposes?
Unlike a written article or a music album, Twitch content is often ephemeral and deeply personal. Streamers broadcast their reactions, conversations, and daily lives to an audience that may range from a dozen viewers to tens of thousands. The expectation of these creators is that their content will be seen by viewers, not that it will be archived, analyzed, and used to build algorithms that mimic human creativity. The automatic opt-in model, combined with a lack of clear communication, has created an impression of stealthy data harvesting.
Industry observers note that Twitch's new opt-out, while welcome, falls short of best practices. Ideally, platforms should ask for explicit consent before using user content for AI training, and should provide clear explanations of what will be used, how it will be used, and whether the user will receive any compensation. Twitch's approach—automatically including users and then giving them an opportunity to leave—is common among technology companies, but it is increasingly being challenged by regulators and consumer advocates.
The European Union's Artificial Intelligence Act, which entered into force in 2024, imposes transparency obligations on AI developers and requires them to disclose where training data comes from. Similarly, the California Consumer Privacy Act and other privacy laws may give users certain rights over the use of their personal data for AI training. Twitch's global user base means that compliance with such laws may have been one factor in the decision to introduce the opt-out setting.
Amazon's own position on the matter is complex. The company has invested heavily in generative AI, including its Bedrock platform and its family of Titan models. Twitch provides an enormous and diverse dataset that Amazon would be reluctant to relinquish. By offering an opt-out, Twitch can maintain access to data from users who do not actively change their settings, while allowing concerned users to withdraw. This approach may reduce regulatory risk and public backlash without significantly diminishing the volume of training data.
For many Twitch streamers, the announcement is a double-edged sword. On one hand, it is reassuring to have more control over one's digital footprint. On the other, it reveals the extent to which their content has already been used without explicit consent. The timing of the change—coming years after the practice began and only after a public confirmation by an executive—suggests that Twitch and Amazon are more concerned about mitigating legal and reputational damage than about respecting user autonomy from the outset.
As the AI industry continues to evolve, the question of who owns the rights to user-generated content used in training will remain central. Twitch's opt-out is a step in the direction of user control, but it is a modest one. It does not address past use, does not offer compensation, and does not apply to data that has already been ingested. Streamers who care about this issue should review their settings today, but they should also be aware that the algorithmic genie is likely already out of the bottle.
Source:Ars Technica News
