{"id":277546,"date":"2024-04-08T07:43:06","date_gmt":"2024-04-08T14:43:06","guid":{"rendered":"https:\/\/sftarticles.wpenginepowered.com\/es\/?p=329699"},"modified":"2025-07-01T16:46:08","modified_gmt":"2025-07-01T23:46:08","slug":"openai-and-google-would-be-training-their-ai-with-youtube-videos","status":"publish","type":"post","link":"https:\/\/cms-articles.softonic.io\/en\/openai-and-google-would-be-training-their-ai-with-youtube-videos\/","title":{"rendered":"OpenAI and Google would be training their AI with YouTube videos"},"content":{"rendered":"\n<p class=\"wp-block-paragraph\">Both <strong>OpenAI<\/strong> and <strong>Google<\/strong> would have resorted to transcribing <strong>YouTube<\/strong> videos to train their AI models, something that could violate the copyright of content creators. In addition to these two companies, <strong>Meta<\/strong> itself would have taken a series of shortcuts to <strong>access as much data as possible<\/strong> to train its AI models, as reported by <strong><a href=\"https:\/\/www.nytimes.com\/2024\/04\/06\/technology\/tech-giants-harvest-data-artificial-intelligence.html\" target=\"_blank\" rel=\"noopener nofollow\" title=\"\">The New York Times<\/a><\/strong>.<\/p>\n\n\n<div class=\"sc-card-program\">\r\n  <div class=\"sc-card-program__body\">\r\n    <div class=\"sc-card-program__row clearfix\">\r\n      <div class=\"sc-card-program__col-logo\">\r\n        <img decoding=\"async\" class=\"sc-card-program__img\" alt=\"ChatGPT\" src=\"https:\/\/images.sftcdn.net\/images\/t_app-icon-s\/p\/47ef1772-2a82-4750-b97a-354b13dbd112\/3647786732\/chatgpt-ChatGPT-icon.png\" width=\"100px\" height=\"100px\">\r\n      <\/div>\r\n      <div class=\"sc-card-program__col-title\">\r\n        <span class=\"sc-card-program__title\">ChatGPT<\/span>\r\n        <a class=\"sc-card-program__button sc-card-program-internal\" href=\"https:\/\/chatgpt.en.softonic.com\/iphone\" target=\"_self\" rel=\"noopener noreferrer\">DOWNLOAD<\/a>\r\n      <\/div>\r\n      <div class=\"sc-card-program__col-rating\">\r\n        <svg class=\"rating-score__content\" xmlns=\"http:\/\/www.w3.org\/2000\/svg\" version=\"1.1\" x=\"0\" y=\"0\" viewbox=\"0 0 50 50\" enable-background=\"new 0 0 50 50\" xml:space=\"preserve\"><path class=\"rating-score__background rating-score--good\" fill=\"none\" stroke-width=\"6\" stroke-miterlimit=\"10\" d=\"M40 40c8.3-8.3 8.3-21.7 0-30s-21.7-8.3-30 0 -8.3 21.7 0 30\"><\/path><path class=\"rating-score__value rating-score__value--0\" fill=\"none\" stroke-width=\"6\" stroke-dashoffset=\"0\" stroke-miterlimit=\"10\" d=\"M40 40c8.3-8.3 8.3-21.7 0-30s-21.7-8.3-30 0 -8.3 21.7 0 30\"><\/path><text class=\"rating-score__number\" content=\"\" text-anchor=\"middle\" transform=\"matrix(1 0 0 1 25 31.0837)\" data-auto=\"app-user-score\"><\/text><\/svg>\r\n      <\/div>\r\n    <\/div>\r\n    <div class=\"sc-card-program__row\">\r\n      <span class=\"sc-card-program__description\"><\/span>\r\n    <\/div>\r\n    <div class=\"sc-card-program__row\">\r\n      <img decoding=\"async\" class=\"sc-card-program__bigpic\" src=\"\" onerror=\"this.style.display='none'\">\r\n    <\/div>\r\n    <a class=\"sc-card-program__link track-link sc-card-program-internal\" href=\"https:\/\/chatgpt.en.softonic.com\/iphone\" target=\"_self\" rel=\"noopener noreferrer\"><\/a>\r\n  <\/div>\r\n<\/div>\n\n\n\n<p class=\"wp-block-paragraph\">According to the published article, OpenAI used <strong>Whisper<\/strong>, a speech recognition tool, to <strong>transcribe over a million hours of YouTube videos<\/strong>. Then, they inputted the transcriptions into <strong><a href=\"https:\/\/en.softonic.com\/articles\/gpt-4-is-already-for-everyone-so-you-can-exploit-its-potential\" target=\"_blank\" rel=\"noopener\" title=\"\">GPT-4<\/a><\/strong>, the powerful AI system that powers the latest ChatGPT chatbot model. Google, the owner of YouTube, also transcribed YouTube videos to train their AI models.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">The transcription of videos by both companies may have infringed the copyright of content creators on their videos; previously, <a href=\"https:\/\/en.softonic.com\/articles\/nvidia-is-not-exempt-from-lawsuits-over-artificial-intelligence\" target=\"_blank\" rel=\"noopener\" title=\"\">various companies were sued for using creators&#8217; content without their permission<\/a>. In addition, the use of YouTube videos by OpenAI <strong>could also violate Google&#8217;s policies<\/strong>, which prohibit the use of their videos for &#8220;standalone&#8221; applications and &#8220;automated media (such as robots, botnets, or scrapers)&#8221; to access their videos.<\/p>\n\n\n<div class=\"wp-block-image\">\n<figure class=\"aligncenter size-large\"><img loading=\"lazy\" decoding=\"async\" width=\"1024\" height=\"575\" src=\"https:\/\/articles-img.sftcdn.net\/sft\/articles\/auto-mapping-folder\/sites\/3\/2024\/04\/youtube-2-1024x575-1-1024x575.jpg\" alt=\"\" class=\"wp-image-277550\" srcset=\"https:\/\/articles-img.sftcdn.net\/auto-mapping-folder\/sites\/3\/2024\/04\/youtube-2-1024x575-1-1024x575.jpg 1024w, https:\/\/articles-img.sftcdn.net\/auto-mapping-folder\/sites\/3\/2024\/04\/youtube-2-1024x575-1-300x169.jpg 300w, https:\/\/articles-img.sftcdn.net\/auto-mapping-folder\/sites\/3\/2024\/04\/youtube-2-1024x575-1-768x431.jpg 768w, https:\/\/articles-img.sftcdn.net\/auto-mapping-folder\/sites\/3\/2024\/04\/youtube-2-1024x575-1-800x450.jpg 800w, https:\/\/articles-img.sftcdn.net\/auto-mapping-folder\/sites\/3\/2024\/04\/youtube-2-1024x575-1-664x374.jpg 664w, https:\/\/articles-img.sftcdn.net\/auto-mapping-folder\/sites\/3\/2024\/04\/youtube-2-1024x575-1-238x134.jpg 238w, https:\/\/articles-img.sftcdn.net\/auto-mapping-folder\/sites\/3\/2024\/04\/youtube-2-1024x575-1-436x246.jpg 436w, https:\/\/articles-img.sftcdn.net\/auto-mapping-folder\/sites\/3\/2024\/04\/youtube-2-1024x575-1-370x208.jpg 370w, https:\/\/articles-img.sftcdn.net\/auto-mapping-folder\/sites\/3\/2024\/04\/youtube-2-1024x575-1-304x170.jpg 304w, https:\/\/articles-img.sftcdn.net\/auto-mapping-folder\/sites\/3\/2024\/04\/youtube-2-1024x575-1.jpg 1282w\" sizes=\"auto, (max-width: 1024px) 100vw, 1024px\" \/><\/figure>\n<\/div>\n\n\n<p class=\"wp-block-paragraph\"><strong>Matt Bryant<\/strong>, a spokesperson for Google, told The New York Times that the company was not aware of such use by OpenAI, but the article alleges that Google staff knew about OpenAI&#8217;s unauthorized use of YouTube videos and <strong>did not take action because they were doing the same thing<\/strong>. Google also told the outlet that they only train their AI with videos from creators who have agreed to have their content used in this way.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">In July 2023, Google modified its terms of service to <strong>allow the use of public online content<\/strong>, such as Google Docs and Google Maps restaurant reviews, to continue training its AI models.<\/p>\n\n\n<div class=\"sc-card-program\">\r\n  <div class=\"sc-card-program__body\">\r\n    <div class=\"sc-card-program__row clearfix\">\r\n      <div class=\"sc-card-program__col-logo\">\r\n        <img decoding=\"async\" class=\"sc-card-program__img\" alt=\"ChatGPT\" src=\"https:\/\/images.sftcdn.net\/images\/t_app-icon-s\/p\/47ef1772-2a82-4750-b97a-354b13dbd112\/3647786732\/chatgpt-ChatGPT-icon.png\" width=\"100px\" height=\"100px\">\r\n      <\/div>\r\n      <div class=\"sc-card-program__col-title\">\r\n        <span class=\"sc-card-program__title\">ChatGPT<\/span>\r\n        <a class=\"sc-card-program__button sc-card-program-internal\" href=\"https:\/\/chatgpt.en.softonic.com\/iphone\" target=\"_self\" rel=\"noopener noreferrer\">DOWNLOAD<\/a>\r\n      <\/div>\r\n      <div class=\"sc-card-program__col-rating\">\r\n        <svg class=\"rating-score__content\" xmlns=\"http:\/\/www.w3.org\/2000\/svg\" version=\"1.1\" x=\"0\" y=\"0\" viewbox=\"0 0 50 50\" enable-background=\"new 0 0 50 50\" xml:space=\"preserve\"><path class=\"rating-score__background rating-score--good\" fill=\"none\" stroke-width=\"6\" stroke-miterlimit=\"10\" d=\"M40 40c8.3-8.3 8.3-21.7 0-30s-21.7-8.3-30 0 -8.3 21.7 0 30\"><\/path><path class=\"rating-score__value rating-score__value--0\" fill=\"none\" stroke-width=\"6\" stroke-dashoffset=\"0\" stroke-miterlimit=\"10\" d=\"M40 40c8.3-8.3 8.3-21.7 0-30s-21.7-8.3-30 0 -8.3 21.7 0 30\"><\/path><text class=\"rating-score__number\" content=\"\" text-anchor=\"middle\" transform=\"matrix(1 0 0 1 25 31.0837)\" data-auto=\"app-user-score\"><\/text><\/svg>\r\n      <\/div>\r\n    <\/div>\r\n    <div class=\"sc-card-program__row\">\r\n      <span class=\"sc-card-program__description\"><\/span>\r\n    <\/div>\r\n    <div class=\"sc-card-program__row\">\r\n      <img decoding=\"async\" class=\"sc-card-program__bigpic\" src=\"\" onerror=\"this.style.display='none'\">\r\n    <\/div>\r\n    <a class=\"sc-card-program__link track-link sc-card-program-internal\" href=\"https:\/\/chatgpt.en.softonic.com\/iphone\" target=\"_self\" rel=\"noopener noreferrer\"><\/a>\r\n  <\/div>\r\n<\/div>\n","protected":false},"excerpt":{"rendered":"<p>Both OpenAI and Google would have resorted to transcribing YouTube videos to train their AI models, something that could violate the copyright of content creators. In addition to these two companies, Meta itself would have taken a series of shortcuts to access as much data as possible to train its AI models, as reported by &hellip; <a href=\"https:\/\/cms-articles.softonic.io\/en\/openai-and-google-would-be-training-their-ai-with-youtube-videos\/\" class=\"more-link\">Continue reading<span class=\"screen-reader-text\"> &#8220;OpenAI and Google would be training their AI with YouTube videos&#8221;<\/span><\/a><\/p>\n","protected":false},"author":9256,"featured_media":277548,"comment_status":"closed","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":"","wpcf-pageviews":1},"categories":[1015],"tags":[],"usertag":[],"vertical":[],"content-category":[],"class_list":["post-277546","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-news"],"aioseo_notices":[],"_links":{"self":[{"href":"https:\/\/cms-articles.softonic.io\/en\/wp-json\/wp\/v2\/posts\/277546","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/cms-articles.softonic.io\/en\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/cms-articles.softonic.io\/en\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/cms-articles.softonic.io\/en\/wp-json\/wp\/v2\/users\/9256"}],"replies":[{"embeddable":true,"href":"https:\/\/cms-articles.softonic.io\/en\/wp-json\/wp\/v2\/comments?post=277546"}],"version-history":[{"count":1,"href":"https:\/\/cms-articles.softonic.io\/en\/wp-json\/wp\/v2\/posts\/277546\/revisions"}],"predecessor-version":[{"id":313938,"href":"https:\/\/cms-articles.softonic.io\/en\/wp-json\/wp\/v2\/posts\/277546\/revisions\/313938"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/cms-articles.softonic.io\/en\/wp-json\/wp\/v2\/media\/277548"}],"wp:attachment":[{"href":"https:\/\/cms-articles.softonic.io\/en\/wp-json\/wp\/v2\/media?parent=277546"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/cms-articles.softonic.io\/en\/wp-json\/wp\/v2\/categories?post=277546"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/cms-articles.softonic.io\/en\/wp-json\/wp\/v2\/tags?post=277546"},{"taxonomy":"usertag","embeddable":true,"href":"https:\/\/cms-articles.softonic.io\/en\/wp-json\/wp\/v2\/usertag?post=277546"},{"taxonomy":"vertical","embeddable":true,"href":"https:\/\/cms-articles.softonic.io\/en\/wp-json\/wp\/v2\/vertical?post=277546"},{"taxonomy":"content-category","embeddable":true,"href":"https:\/\/cms-articles.softonic.io\/en\/wp-json\/wp\/v2\/content-category?post=277546"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}