3 Sources
[1]
Is Google Gemini trained on Google docs? One indie developer thinks so, after it told players about his unreleased game plans
* A game developer claims Gemini leaked information in his private Google Docs * Per screenshots shared online, the AI told players info the dev claims was never made public * TechRadar has tried recreating the results, without success Is Google Gemini secretly scraping our Google Docs to train its AI? That's something one game developer is wondering after a player was able to learn unreleased information about their game from the AI -- including specific details the dev claims were never shared outside of a private Google Doc. According to the developer's Reddit post, specifically Google's AI knew the name of a character 'Vantage Tripod' before it had ever been released publicly -- the most the fanbase knew was that there's a 'Tripod Fish' character planned for the game, but not this exact name. It also seemed to be able to regurgitate accurate information related to unreleased mechanics that the developer says they had only written into GDocs the day before the player's AI interaction. The developer admitted Google's AI made a few mistakes, but said that some details were scarily accurate in many ways -- accurate enough that the developer is certain Google must have scrapped his private documents. Google says Gemini can access Google Docs information, but only does so when given express permission -- such as being asked to summarise a document -- and it adds that even when it does go into your files, Gemini handles data transiently. That is, Google's AI won't retain anything. There are a couple of exceptions to this. If a Google Doc is accessible to 'Anyone with the link' and that link is posted publicly online (such as on a page or in a forum) Google could scrape it for training data. Alternatively, if you've allowed a third-party extension to scrape your Google Docs, it's possible Gemini could get access to that data -- indirectly seeing what's in your docs. If you were to type out info from your docs into Google Gemini, it would also learn about what you had written that way. Despite Google's promises, some aren't entirely convinced. The thread I shared above (along with plenty of anti-AI and anti-Google subreddits) is full of people certain that Google has drained their digital files for all the data it can find. Privacy company Proton has also published an article outlining details such as Google's privacy hub not explicitly saying it won't use your content for AI training. So many unknowns I've reached out to Google with a request for comment about what has happened here (it has yet to get back to me), and while waiting for an answer I had an attempt at recreating the responses by prompting Gemini myself with no luck. It's only reference to these details was the Reddit post information. Later screenshots shared by the developer show the friend who got the AI to divulge the details originally also failed to recreate the "fluke." The whole situation is very weird and it's impossible to tell exactly what has happened, based on the information available. While it certainly looks like Gemini has used information it shouldn't have, it's equally possible that the developer has provided the AI or the friend with access to that information without realising -- allowing Gemini to respond the way it did. The AI could also have hallucinated the information in some way, or perhaps taken inspiration from previous prompts to inform its comments. Further, it's worth noting that the only responses that would stand out are these two that get very close to the truth, the sea of incorrect guesses that the AI could have given wouldn't bat an eyelid. Regardless of what happened it's another situation that reminds us to be careful with our personal info. Digitally storing personal data has its risks beyond possible AI scraping, and sharing data with an AI often means your private info won't be private anymore. Follow TechRadar on Google News and add us as a preferred source to get our expert news, reviews, and opinion in your feeds.
[2]
Steam Developer Alleges Google Leaked Confidential Game Content via AI Search
Video game developers of any size have plenty of challenges to face, like mass refunds of their short game or two developers releasing games with the same name on the same day. A new obstacle to worry about? AI scraping private documents and leaking unannounced information about their games. That's what happened to Klub Kofta Studio, a self-described solo developer of the tower defense game Operation Octo, which launched on Steam in September 2025. In a post on Reddit, they shared Google's AI told a player the name of an upcoming Operation Octo character that Klub Kofta Studio had never publicly shared. "I had never told the name to anyone," they wrote in the post title. Polygon has reached out to Klub Kofta Studio and Google for comment, but did not hear back in time for publication. As Klub Kofta Studio explained it on Reddit, a player in their Discord server was asking the Google search AI "silly questions" about the game to get a kick out of the answers. "But the answers sometimes contained surprisingly obscure info that should not be available for a small indie game like mine." Klub Kofta Studio asked the player to throw some questions at the AI about future, unannounced content. "Somehow, the AI spitted out the exact & highly specific character name 'Vantage Tripod', which I had never mentioned to anyone," they wrote. "As far as I know, the only place where the info exists in a digital form is inside one of my own Google Docs, and this was NOT me speaking to the AI!" They shared screenshots of a snippet of their conversation on Discord, including their befuddled reactions to the AI's eerily correct answers. "My game is not big enough for there to be much noise out there to confuse Google's AI, so I guess that's why it managed to give actual, scarily real leaks on my game (without being confused by online speculations), and I have no idea how it knew all this," they said. Google's AI overview has a habit of pulling information from Reddit threads for answers, correct or otherwise. The AI overview can also be manipulated, which in turn can cause it to dole out unreliable misinformation. Last year, a YouTuber managed to convince Google's AI search that GTA 6 would include a "twerk button" through a fairly low-effort shitposting project. As people on Reddit and Bluesky have pointed out, Google automatically opts users into sharing their Google Drive data with Google to improve AI models, among other uses. (You can head to Google support to learn how to manage how your activity interacts with Google's Gemini features.) For years now, large tech companies have been quietly updating their terms of service and opting users into allowing the use of their data for AI model training. "It's what companies do. Quietly opt all your stuff into AI while trying to avoid announcing it, then bury an option to opt out. The fuckers," one person commented on the Reddit thread. Others are worried about more personal information being leaked by AI overviews based on information in users' "private" (heavy air quotes here) Google Drives. "Doesn't that mean you can dox people by asking very specific questions given that people's private personal data is included in the training sets then?......" another commentor mused. If you haven't yet, now would be a good time to look into who has access to your public and private data and what they're using it for. Living in an off-the-grid log cabin never seemed more appealing.
[3]
Google AI leaks "scarily accurate and very specific" game detail from private documents, says indie dev: "I had never told the name to anyone"
As if the rapid proliferation of artificial intelligence wasn't already scary enough, an indie developer says it managed to leak a very specific detail from their upcoming game from private documents. Over on Reddit, solo game developer Klub Kofta Studio, whose flagship game is the aquatic tower defense title Operation Octo, shared a harrowing Discord interaction they had with a player about an upcoming update. Klub Kofta attached screenshots of the conversation in which the player asked Google's Gemini AI questions about the game and received "surprisingly obscure info that should not be available for a small indie game like mine." That in and of itself isn't shocking, but Google's AI chatbot was also talking about future content leaks gleaned from conversations on the game's Discord. This was "super impressive," said Klub Kofta, who inadvertently revealed an upcoming creature casually referred to as Tripod Fish when they drew it while playing an online party game called Gartic Phone. Discord users were able to decipher that Tripod Fish was coming to the game based on the drawing, but the exact name of the creature was never revealed publicly, according to Klub Kofta. Google Gemini was able to figure it out regardless. "Out of curiosity, I chimed in and asked the player to ask the AI about future contents in my game," explained the solo dev. "Somehow, the AI spitted out the exact & highly specific character name 'Vantage Tripod,' which I had never mentioned to anyone. As far as I know, the only place where the info exists in a digital form is inside one of my own Google Docs, and this was NOT me speaking to the AI! (The player even asked the questions in incognito)." I'm hesitant to theorize on how this all could've happened considering it's just one unverified account, but Reddit users speculated that Google automatically opts in users to allow Gemini to scan Google Drive files to train its AI. I couldn't verify that claim after some digging, but if Klub Kofta's story is true, it carries some deeply disturbing implications about online privacy in the age of AI. One Reddit user who called the story "unsettling" suggested some alternate possibilities, including scanned Discord conversations between server members, but Klub Kofta insisted the creature's name was never revealed anywhere outside of their private Google Drive. "Even the discord server members have never learned this specific name. I'm certain I'd kept it a secret, and a search history on Discord (as shown in the screenshot) indeed reveals that the keyword had never been said up until this point," said the dev. "The character name has never made it into serialized data in the game yet, so the one Google Doc where I kept future character names & descriptions was the only source of information I can think of." Forget hackers, with just a few months until GTA 6, Rockstar better be real careful about who's talking to ChatGPT.
Share
Copy Link
Klub Kofta Studio claims Google Gemini revealed the exact name 'Vantage Tripod' for an unreleased Operation Octo character—information stored only in private Google Docs. The incident raises urgent questions about whether AI training on user data extends to supposedly private documents and what this means for digital privacy.
Klub Kofta Studio, the solo game developer behind tower defense title Operation Octo, experienced what they describe as a disturbing Google AI leak when a player discovered unreleased game plans through Google's AI search
1
. The incident occurred when a Discord community member asked Google Gemini "silly questions" about the game and received answers containing "surprisingly obscure info that should not be available for a small indie game like mine," according to the developer's Reddit post2
. What started as entertainment quickly turned serious when the AI revealed highly specific, confidential game content that had never been made public.The most alarming revelation came when Google Gemini accurately identified an unreleased character name as "Vantage Tripod"—a detail the game developer insists was never shared with anyone
3
. While Discord server members had previously seen a drawing of a "Tripod Fish" character during an online party game called Gartic Phone, the exact name remained undisclosed. "As far as I know, the only place where the info exists in a digital form is inside one of my own Google Docs, and this was NOT me speaking to the AI," Klub Kofta Studio explained2
. The indie developer emphasized that even Discord server members had never learned this specific name, and search history confirmed the keyword had never been mentioned until the AI revealed it3
.
Source: GamesRadar
This incident raises critical questions about whether Google Gemini trained on Google docs without explicit permission. Google states that Gemini can access Google Docs information only when given express permission—such as being asked to summarize a document—and handles data transiently without retention
1
. However, exceptions exist. If a Google Doc is set to "Anyone with the link" and that link appears publicly online, Google could scrape it for data training. Third-party extensions with access to Google Drive could also indirectly expose document contents to Gemini1
. Privacy company Proton has noted that Google's privacy hub doesn't explicitly state it won't use your content for AI training on user data1
.Reddit and Bluesky users quickly pointed out that Google automatically opts users into sharing their Google Drive data with Google to improve AI models, among other uses
2
. This practice reflects a broader industry trend where large tech companies quietly update terms of service to enable AI training on user data. "It's what companies do. Quietly opt all your stuff into AI while trying to avoid announcing it, then bury an option to opt out," one Reddit commenter observed2
. The privacy concerns around AI extend beyond game development—users worry about personal information being exposed through the AI Overview feature based on supposedly private documents. One commenter questioned whether specific queries could effectively dox people if private personal data is included in training sets2
.Related Stories
TechRadar attempted to recreate the responses by prompting Gemini with similar queries but found no success—the AI only referenced information from the Reddit post itself
1
. Later screenshots shared by Klub Kofta Studio showed that even the friend who originally obtained the AI's responses failed to recreate the "fluke"1
. This inconsistency makes the situation harder to verify, though it doesn't diminish the developer's concerns about how their unreleased game plans became accessible to AI search.While it's impossible to determine exactly what happened based on available information, several possibilities exist. The game developer may have inadvertently provided the AI or the friend with access to information without realizing it. The AI could have hallucinated the information or drawn inspiration from previous prompts. It's also worth noting that only responses matching the truth would stand out—incorrect AI guesses wouldn't raise alarms
1
. Regardless of the mechanism, this incident serves as a stark reminder about digital storage risks beyond possible AI scraping. Users should review who has access to their public and private data and understand what it's being used for2
. For developers working on Operation Octo or any unreleased project, the implications are clear: private Google Docs may not be as private as assumed, and managing data access settings has become essential in an era where AI training practices remain opaque.Summarized by
Navi
1
Technology

2
Technology

3
Science and Research
