Also Watch: Nothing Phone (2) with Snapdragon 8+ Gen 1 chipset launched in India: Price, specs, pre-order offers, more
The complaint highlights a recent update to Google's privacy policy, which explicitly states that the company may employ publicly accessible information to train its AI models and tools, such as Bard. In response to a previous report on the policy update by The Verge, Google clarified that their policy has always been transparent about the use of publicly available information from the open web to train language models for services like Google Translate, and the update merely extends this to newer services like Bard.
This lawsuit arises at a time when a new generation of AI tools has gained significant attention for their capacity to generate written content and images in response to user prompts. However, the extensive reliance on large language models in this technology has attracted legal scrutiny regarding copyright concerns related to the incorporation of works within these data sets, as well as the potential utilisation of personal and sensitive data from ordinary users, including children, as indicated in the Google lawsuit.
Also Read Game, set and AI: Wimbledon 2023 will see AI commentary for the first time in tennis with help of IBM
Tim Giordano, one of the attorneys from Clarkson Law Firm involved in the lawsuit against Google, emphasised that "publicly available" does not imply free to use for any purpose. He stated in an interview with CNN, "Our personal information and our data are our property, and it's valuable, and nobody has the right to just take it and use it for any purpose."
The lawsuit aims to obtain injunctive relief, requesting a temporary halt on the commercial access and development of Google's generative AI tools such as Bard. In addition, it calls for unspecified damages and financial compensation to be paid to individuals, including a minor, whose data was allegedly misappropriated by Google. The firm has identified eight plaintiffs who are part of this legal action.
Giordano distinguished the benefits and alleged harms of how Google typically indexes online data for its core search engine from the new allegations of data scraping for AI training purposes. While Google's search engine directs users to attributed links that can drive engagement and potentially lead to purchases, scraping data for AI training creates an alternative version of the work, significantly altering incentives for users to purchase it.
He said that although some internet users may have grown accustomed to their digital data being collected and used for search results or targeted advertising, the same cannot be assumed for AI training. He also underlined that "people could not have imagined their information would be used this way."
Also Read
Battle of the billionaires: Elon Musk vs Mark Zuckerberg cage match could make over $1 billion
Google appeals to Supreme Court to quash antitrust directives on Android in India