Table of Contents
These companies are smart, but also very hypocritical
Reddit, Twitter and Others: Enough is Enough
The current training method of AI models "breaks" the network
Home Technology peripherals AI OpenAI and Google have a double standard: use other people's data to train large models, but never allow their own data to leak out

OpenAI and Google have a double standard: use other people's data to train large models, but never allow their own data to leak out

Jun 05, 2023 pm 03:03 PM
Google ai

In a new era of generative AI, big tech companies are pursuing a "do as I say, not as I do" strategy when it comes to using online content. To a certain extent, this strategy can be said to be a double standard and an abuse of the right to speak.

At the same time, as large language models (LLM) become the mainstream trend in AI development, both large and start-up companies are sparing no effort to develop their own large models. Among them, training data is an important prerequisite for the ability of large models.

Recently, according to Insider reports, Microsoft-backed OpenAI, Google and its backed Anthropic have been using online content from other websites or companies for training for many years. Their generative AI model . These were all done without asking for specific permission and will form part of a brewing legal battle to determine the future of the web and how copyright law is applied in this new era.

OpenAI and Google have a double standard: use other peoples data to train large models, but never allow their own data to leak out

These big tech companies may argue that they are fair use, whether that is really the case is up for debate. But they won’t let their content be used to train other AI models. So we can’t help but ask, why can these large technology companies use other companies’ online content when training large models?

These companies are smart, but also very hypocritical

Is there any solid evidence for the claim that big tech companies use other people’s online content but don’t allow others to use their own? This can be seen in the terms of service and use of some of their products.

First let’s look at Claude, which is an AI assistant similar to ChatGPT launched by Anthropic. The system can complete tasks such as summary summarization, search, assistance in creation, question and answer, and coding. It was upgraded again some time ago and the context token was expanded to 100k, which greatly accelerated the processing speed.

OpenAI and Google have a double standard: use other peoples data to train large models, but never allow their own data to leak out

Claude’s Terms of Service are as follows. You may not access or use the Service in the following ways (some of which are listed here). If any of these restrictions are inconsistent or unclear with the Acceptable Use Policy, the latter shall prevail:

  • Develop any product or service that competes with our Services, including developing or training any AI or machine learning algorithms or models
  • From our Crawl, crawl or otherwise obtain data or information from the Service

Claude Terms of Service Address: https://vault.pactsafe.io/s /9f502c93-cb5c-4571-b205-1e479da61794/legal.html#terms

Similarly, Google’s Generative AI Terms of Use states, “You may not use the Service To develop machine learning models or related technologies."

OpenAI and Google have a double standard: use other peoples data to train large models, but never allow their own data to leak out

##Google Generative AI Terms of Use Address: https: //policies.google.com/terms/generative-ai

What about OpenAI’s terms of use? Similar to Google, "You may not use the output of this service to develop models that compete with OpenAI."

OpenAI and Google have a double standard: use other peoples data to train large models, but never allow their own data to leak out

OpenAI Terms of Use Address: https://openai.com/policies/terms-of-use

These companies are smart, they know that high-quality content is critical to training new AI models, so it makes sense not to allow others to use their output in this way. But they have no scruples in using other people’s data to train their own models. How to explain this?

OpenAI, Google and Anthropic declined Insider's request for comment and did not respond.

Reddit, Twitter and Others: Enough is Enough

Actually, other companies weren't happy when they realized what was happening. In April, Reddit, which has been used for years to train AI models, plans to start charging for access to its data.

Reddit CEO Steve Huffman said, “Reddit’s data corpus is too valuable to give away that value to the largest companies in the world for free.”

Also in April this year, Musk accused Microsoft, OpenAI’s main supporter, of illegally using Twitter data to train AI models. "Time for litigation," he tweeted.

OpenAI and Google have a double standard: use other peoples data to train large models, but never allow their own data to leak out

#However, in response to Insider's comment, Microsoft said, "There are so many things wrong with this premise that I don't even know where to start. ”

OpenAI CEO Sam Altman has tried to deepen this problem by exploring new AI models that respect copyright. According to Axios, he recently said, "We are trying to develop a new model. If the AI ​​system uses your content or uses your style, you will get paid for it."

OpenAI and Google have a double standard: use other peoples data to train large models, but never allow their own data to leak out

Sam Altman

Publishers (including Insiders) will all have vested interests. Additionally, some publishers, including U.S. News Corp., are already pushing for tech companies to pay to use their content to train AI models.

The current training method of AI models "breaks" the network

A former Microsoft executive said there must be something wrong with this. Microsoft veteran and famous software developer Steven Sinofsky believes that the current training method of AI models "breaks" the network.

OpenAI and Google have a double standard: use other peoples data to train large models, but never allow their own data to leak out

Steven Sinofsky

He’s pushing The post reads, "In the past, crawled data was used in exchange for click-through rates. But now it is only used to train a model and does not bring any value to creators and copyright owners."

Perhaps, as more companies wake up, this uneven data usage in the era of generative AI will soon be changed.

The above is the detailed content of OpenAI and Google have a double standard: use other people's data to train large models, but never allow their own data to leak out. For more information, please follow other related articles on the PHP Chinese website!

Statement of this Website
The content of this article is voluntarily contributed by netizens, and the copyright belongs to the original author. This site does not assume corresponding legal responsibility. If you find any content suspected of plagiarism or infringement, please contact admin@php.cn

Hot AI Tools

Undresser.AI Undress

Undresser.AI Undress

AI-powered app for creating realistic nude photos

AI Clothes Remover

AI Clothes Remover

Online AI tool for removing clothes from photos.

Undress AI Tool

Undress AI Tool

Undress images for free

Clothoff.io

Clothoff.io

AI clothes remover

Video Face Swap

Video Face Swap

Swap faces in any video effortlessly with our completely free AI face swap tool!

Hot Tools

Notepad++7.3.1

Notepad++7.3.1

Easy-to-use and free code editor

SublimeText3 Chinese version

SublimeText3 Chinese version

Chinese version, very easy to use

Zend Studio 13.0.1

Zend Studio 13.0.1

Powerful PHP integrated development environment

Dreamweaver CS6

Dreamweaver CS6

Visual web development tools

SublimeText3 Mac version

SublimeText3 Mac version

God-level code editing software (SublimeText3)

Hot Topics

Java Tutorial
1657
14
PHP Tutorial
1257
29
C# Tutorial
1229
24
How much is Bitcoin worth How much is Bitcoin worth Apr 28, 2025 pm 07:42 PM

Bitcoin’s price ranges from $20,000 to $30,000. 1. Bitcoin’s price has fluctuated dramatically since 2009, reaching nearly $20,000 in 2017 and nearly $60,000 in 2021. 2. Prices are affected by factors such as market demand, supply, and macroeconomic environment. 3. Get real-time prices through exchanges, mobile apps and websites. 4. Bitcoin price is highly volatile, driven by market sentiment and external factors. 5. It has a certain relationship with traditional financial markets and is affected by global stock markets, the strength of the US dollar, etc. 6. The long-term trend is bullish, but risks need to be assessed with caution.

Which of the top ten currency trading platforms in the world are the latest version of the top ten currency trading platforms Which of the top ten currency trading platforms in the world are the latest version of the top ten currency trading platforms Apr 28, 2025 pm 08:09 PM

The top ten cryptocurrency trading platforms in the world include Binance, OKX, Gate.io, Coinbase, Kraken, Huobi Global, Bitfinex, Bittrex, KuCoin and Poloniex, all of which provide a variety of trading methods and powerful security measures.

What are the top ten virtual currency trading apps? The latest digital currency exchange rankings What are the top ten virtual currency trading apps? The latest digital currency exchange rankings Apr 28, 2025 pm 08:03 PM

The top ten digital currency exchanges such as Binance, OKX, gate.io have improved their systems, efficient diversified transactions and strict security measures.

Which of the top ten currency trading platforms in the world are among the top ten currency trading platforms in 2025 Which of the top ten currency trading platforms in the world are among the top ten currency trading platforms in 2025 Apr 28, 2025 pm 08:12 PM

The top ten cryptocurrency exchanges in the world in 2025 include Binance, OKX, Gate.io, Coinbase, Kraken, Huobi, Bitfinex, KuCoin, Bittrex and Poloniex, all of which are known for their high trading volume and security.

Decryption Gate.io Strategy Upgrade: How to Redefine Crypto Asset Management in MeMebox 2.0? Decryption Gate.io Strategy Upgrade: How to Redefine Crypto Asset Management in MeMebox 2.0? Apr 28, 2025 pm 03:33 PM

MeMebox 2.0 redefines crypto asset management through innovative architecture and performance breakthroughs. 1) It solves three major pain points: asset silos, income decay and paradox of security and convenience. 2) Through intelligent asset hubs, dynamic risk management and return enhancement engines, cross-chain transfer speed, average yield rate and security incident response speed are improved. 3) Provide users with asset visualization, policy automation and governance integration, realizing user value reconstruction. 4) Through ecological collaboration and compliance innovation, the overall effectiveness of the platform has been enhanced. 5) In the future, smart contract insurance pools, forecast market integration and AI-driven asset allocation will be launched to continue to lead the development of the industry.

Binance official website entrance Binance official latest entrance 2025 Binance official website entrance Binance official latest entrance 2025 Apr 28, 2025 pm 07:54 PM

Visit Binance official website and check HTTPS and green lock logos to avoid phishing websites, and official applications can also be accessed safely.

Recommended reliable digital currency trading platforms. Top 10 digital currency exchanges in the world. 2025 Recommended reliable digital currency trading platforms. Top 10 digital currency exchanges in the world. 2025 Apr 28, 2025 pm 04:30 PM

Recommended reliable digital currency trading platforms: 1. OKX, 2. Binance, 3. Coinbase, 4. Kraken, 5. Huobi, 6. KuCoin, 7. Bitfinex, 8. Gemini, 9. Bitstamp, 10. Poloniex, these platforms are known for their security, user experience and diverse functions, suitable for users at different levels of digital currency transactions

What are the top currency trading platforms? The top 10 latest virtual currency exchanges What are the top currency trading platforms? The top 10 latest virtual currency exchanges Apr 28, 2025 pm 08:06 PM

Currently ranked among the top ten virtual currency exchanges: 1. Binance, 2. OKX, 3. Gate.io, 4. Coin library, 5. Siren, 6. Huobi Global Station, 7. Bybit, 8. Kucoin, 9. Bitcoin, 10. bit stamp.

See all articles