Premium Only Content
LAION-5B: 5 billion image-text-pairs dataset (with the authors)
#laion #clip #dalle
LAION-5B is an open, free dataset consisting of over 5 billion image-text-pairs. Today's video is an interview with three of its creators. We dive into the mechanics and challenges of operating at such large scale, how to keep cost low, what new possibilities are enabled with open datasets like this, and how to best handle safety and legal concerns.
OUTLINE:
0:00 - Intro
1:30 - Start of Interview
2:30 - What is LAION?
11:10 - What are the effects of CLIP filtering?
16:40 - How big is this dataset?
19:05 - Does the text always come from the alt-property?
22:45 - What does it take to work at scale?
25:50 -When will we replicate DALL-E?
31:30 - The surprisingly efficient pipeline
35:20 - How do you cover the S3 costs?
40:30 - Addressing safety & legal concerns
55:15 - Where can people get started?
References:
LAION website: https://laion.ai/
LAION Discord: https://discord.com/invite/mVcgxMPD7e
LAION-5B: https://laion.ai/laion-5b-a-new-era-o...
img2dataset tool: https://github.com/rom1504/img2dataset
LAION-400M: https://paperswithcode.com/dataset/la...
Links:
TabNine Code Completion (Referral): http://bit.ly/tabnine-yannick
YouTube: https://www.youtube.com/c/yannickilcher
Twitter: https://twitter.com/ykilcher
Discord: https://discord.gg/4H8xxDF
BitChute: https://www.bitchute.com/channel/yann...
LinkedIn: https://www.linkedin.com/in/ykilcher
BiliBili: https://space.bilibili.com/2017636191
If you want to support me, the best thing to do is to share out the content :)
If you want to support me financially (completely optional and voluntary, but a lot of people have asked for this):
SubscribeStar: https://www.subscribestar.com/yannick...
Patreon: https://www.patreon.com/yannickilcher
Bitcoin (BTC): bc1q49lsw3q325tr58ygf8sudx2dqfguclvngvy2cq
Ethereum (ETH): 0x7ad3513E3B8f66799f507Aa7874b1B0eBC7F85e2
Litecoin (LTC): LQW2TRyKYetVC8WjFkhpPhtpbDM4Vw7r9m
Monero (XMR): 4ACL8AGrEo5hAir8A9CeVrW8pEauWvnp1WnSDZxW7tziCDLhZAGsgzhRQABDnFy8yuM9fWJDviJPHKRjV4FWt19CJZN9D4n
-
3:28:55
Price of Reason
11 hours agoTrump Means Business! Disney's F4 Hail Mary Pass! Assassin's Creed Shadows Art Book SUCKS?
44.7K4 -
8:00:07
SpartakusLIVE
9 hours ago#1 Shadow BANNED Hero
18.9K -
2:17:46
Kim Iversen
9 hours agoTrump To SMUG Netanyahu: Let's Clear “All” Palestinians From Gaza! | RFK Jr, Tulsi Move On To Round Two
69.9K422 -
30:25
Standpoint with Gabe Groisman
1 day agoDemocrats Are Stalling Trump Appointments with Senator Rick Scott
90.8K24 -
1:00:24
The StoneZONE with Roger Stone
10 hours agoAnthony Fauci’s Brutal History Of Animal Torture Exposed! | The StoneZONE w/ Roger Stone
62.6K18 -
1:03:38
Man in America
11 hours agoUSAID Corruption, $21 TRILLION Missing, & the End of the US Global Empire?
66.8K44 -
56:38
Flyover Conservatives
11 hours ago6 Steps to Take Advantage of Trump’s New Golden Age! - Clay Clark | FOC Show
42.8K2 -
1:15:25
Glenn Greenwald
10 hours agoTulsi and RFK Jr. Approved by Key Senate Committees; Trump Meets Netanyahu: Wants to Cleanse Gaza; Pro-Palestinian Group Suspended at UMich | SYSTEM UPDATE #402
107K106 -
1:43:57
Danny Polishchuk
10 hours agoThe Funniest Call In Show On Earth - Live From New York City's Best Comedy Club
59.8K1 -
1:41:13
megimu32
10 hours agoON THE SUBJECT: Will the Super Bowl Be WOKE??!
44.3K10