This is an archived article last updated on Jul 5, 2023. The information may be outdated.

What is Google’s New Policy to Train AI?

Last Updated: Jul 5, 2023, 18:32 IST

Google has updated its privacy policy to allow the company to collect and use public data to train its AI models. This policy change has been met with mixed reactions, with some people praising Google for its transparency and others expressing concerns about privacy and bias.

Google updates its Privacy Policy  to use Public Data for AI Training
Google updates its Privacy Policy to use Public Data for AI Training

In its recent policy update, tech giant Google has decided to gather data from all sources available on the internet to train its AI models, including Bard.

Under the new policy, Google will be able to collect data from a variety of public sources, including social media posts, government records, and the web. This data will be used to train AI models for a variety of purposes, such as spam filtering, fraud detection, and language translation.

Google argues that using public data is necessary to train AI models that are accurate and effective. The company also says that it will take steps to protect user privacy, such as de-identifying data before it is used to train models.

The Google Policy page states: “We may share non-personally identifiable information publicly and with our partners — like publishers, advertisers, developers, or rights holders. For example, we share information publicly to show trends about the general use of our services.”

What is Google’s New Policy? 

Google's policy on collecting public data is not very transparent, so users must read the policy carefully to understand what information Google collects.

Here is what the policy update mentions, “Google uses information to improve our services and to develop new products, features, and technologies that benefit our users and the public. 

“For example, we may collect information that’s publicly available online or from other public sources to help train Google’s AI models and build products and features like Google Translate, Bard, and Cloud AI capabilities. Or, if your business’s information appears on a website, we may index and display it on Google services”, it added.

Earlier the company was using this information to update and train its language models which enhanced its already available products such as Google Translate. Now, the company has clearly mentioned that all the public data will be used to update its AI products. 

Source: Google

The above image is from Google’s policy archives in which the green colour represents newly added information. 

The Dangers of Data Scraping

This new policy update can cause severe cases of data scraping and privacy concerns. While companies typically keep user data confidential for future use and new product development, Google's new policy allows the company to use any publicly available information to train its AI models. 

This means that Google can access and process any type of data that is available on the internet, including personal information. The company has mentioned that it de-identifies the sources but it can still be a trouble. 

First, it can violate the privacy of individuals. When data is scraped without permission, individuals may not be aware that their data is being collected or how it is being used. This can lead to a number of problems, such as identity theft and financial fraud.

Second, data scraping can be used to create biased AI models. If AI models are trained on data that is scraped from the internet, the models may reflect the biases that are already present in the data. This can lead to AI models that discriminate against certain groups of people.

Finally, data scraping can disrupt the internet. The most recent example of this is the Twitter outage. When data is scraped from websites, it can slow down the websites and make them difficult to use. 

Elon Musk displayed his concerns about data scraping and he decided to limit the number of tweets that people can read per day he is also continuously working to make the platform more secure by monetising different services.  

In conclusion, the new policy can surely help Google generate powerful AI but it will be a safety hazard as well. The policies could also lead to increased data scraping and privacy violations. It is important to carefully monitor how Google implements these policies. 

Nikhil Batra
Nikhil Batra

Content Writer

Nikhil is a dedicated digital journalist and communications professional with more than five years of experience, currently working within the General Knowledge section at Jagran Josh. He has established himself as a subject matter expert in Finance, Economy, History, Technology, and Trending News, consistently delivering accurate, engaging, and easy-to-read content for a wide global audience.

Over the course of his career, Nikhil has developed deep expertise in crafting informative listicles, viral trending stories. His editorial portfolio also spans finance, historical research, and technology reporting, making him a versatile and well-rounded content professional. Every piece he produces reflects a strong balance between factual accuracy and reader engagement.

... Read More
First Published: Jul 5, 2023, 18:14 IST

FAQs

  • Did Google update their privacy policy?
    +
    Yes, Google has updated its privacy policy to allow the company to collect and use public data to train its AI models.
  • What is Google's responsible AI division?
    +
    The mission of the Responsible AI and Human Centered Technology (RAI-HCT) team does research and ensures best practices to develop responsible AI tools
  • What are the data privacy considerations with AI?
    +
    The main privacy concerns surrounding AI is the potential for data breaches and unauthorized access to personal information.

Get here current GK and GK quiz questions in English and Hindi for India, World, Sports and Competitive exam preparation. Download the Jagran Josh Current Affairs App.

Trending

Latest Education News