Skip to main content

Reddit wants a better AI deal with Google: users in exchange for content

Reddit wants more users and more money from Google in exchange for an even bigger mountain of data feeding its AI machine, according to Bloomberg. The negotiation shows a new front in the struggle between Big AI and content providers as they try to harvest new revenue streams without bleeding out the very traffic and engagement that keep them alive.

A year and a half after cutting its first data-sharing deal with Google, reportedly worth $60 million a year, unnamed executives say Reddit is back at the negotiating table eyeing an even bigger role inside the company’s AI ecosystem. The platform is reportedly eager for Google to help entice users – who get an answer farmed from Reddit and leave – into posting in Reddit’s forums, which would generate more of the very content tech giants need to train their data-hungry AI models. 

Reddit also wants more money for its data. The platform is reportedly considering a dynamic pricing-style arrangement for future licensing deals with companies like Google and OpenAI, a system where pay would be determined by how useful or important content is to the answers generated by AI tools. 

Executives reportedly believe current terms do not reflect how valuable Reddit data is to AI companies. Reddit is in a stronger position than most to make this claim and it has been exceptionally useful to tech firms training AI models. In a slop-ridden internet, Reddit posts are made by real people speaking candidly, content is well-sorted by theme, and it is all ranked according to a human-run voting system, not an algorithm. Data suggests Reddit is the top cited domain for AI tools like Perplexity and Google’s AI Overviews, and adding “reddit” to Google search queries is a well-known hack to get more useful answers from search engines.  

The push for better terms — particularly terms that go beyond money and help keep content-providers alive — underscores the paradox at the center of many AI licensing deals: platforms like Reddit hold troves of data tech companies need to train their AI models, only for them to watch those same models strangle the traffic and activity that made them valuable in the first place. 



from The Verge https://ift.tt/8vicZlC

Comments

Popular posts from this blog

Pandora Stories lets artists add commentary to their own playlists

Pandora launched Stories today, a tool that lets artists and creators add voice commentary to their own playlists. The Stories feature merges podcasts with music playlists, and is meant for artists to add context to an album, or for podcasters to experiment with new storytelling formats. The feature is part of Pandora AMP, the streaming service’s free Artist Marketing Platform that helps creators promote their work. To kick off the launch, Pandora’s prepared some Stories by artists like John Legend and Daddy Yankee, who tell listeners their personal stories interspersed between their own songs. There’s also a Stories playlist called Love Songs That Aren’t Really Love Songs , which includes commentary on individual songs like a podcast... Continue reading… from The Verge - All Posts https://ift.tt/2Xz1oNc

Instagram’s updated algorithm prioritizes original content instead of rip-offs

Image: Kristen Radtke / The Verge Instagram is making significant changes to how its system recommends content, with a focus on original content and increased distribution for smaller accounts. The slew of changes were announced by the company in a blog post today. The biggest change deals with aggregators — accounts that download or screenshot other users’ videos and photos and repost them. Sometimes aggregators will credit the original poster by tagging them in the post or caption, but often, content is wholesale ripped off with no acknowledgment, and engagement is siphoned off from the person who created the content in the first place. Instagram clearly has a problem with this and will begin removing reposted content from recommendations across the platform. The... Continue reading… from The Verge - All Posts https://ift.tt/ECgcPAU

We asked camera companies why their RAW formats are all different and confusing

When you set up a new camera, or even go to take a picture on some smartphones, you’re presented with a key choice: JPG or RAW? JPGs are ready to post just about anywhere, while RAWs yield an unfinished file filled with extra data that allows for much richer post-processing. That option for a RAW file (and even the generic name, RAW) has been standardized across the camera industry — but despite that, the camera world has never actually settled on one standardized RAW format. Most cameras capture RAW files in proprietary formats, like Canon’s CR3, Nikon’s NEF, and Sony’s ARW. The result is a world of compatibility issues. Photo editing software needs to specifically support not just each manufacturer’s file type but also make changes for each new camera that shoots it. That creates pain for app developers and early camera adopters who want to know that their preferred software will just work. Adobe tried to solve this problem years ago with a universal RAW form...