The Transformer paper was published around 2016–2017, and it still took approximately 5 years until the ChatGPT moment in late 2022.
Snapshot · Huberman Lab
The Transformer paper was published around 2016–2017, and it still took approximately 5 years until the ChatGPT moment in late 2022.
Where this was said
At 23:05 · chapter starts 21:19
Andrew Huberman raises the question of whether the data-plus-algorithm recipe that worked for vision was similarly applied to sound and speech recognition. Fei-Fei Li confirms that ImageNet's success did indeed open the floodgates: every sub-area of AI — speech recognition, sound classification, natural language processing — received a significant boost. She offers the charming example of Stanford colleagues using machine learning to analyze whale songs. But the more transformative moment, she argues, came with the Transformer architecture around 2016–2017, which proved dramatically more powerful than the ImageNet-era AlexNet and was ideally suited for the vast abundance of text available on the internet. Companies like OpenAI and Google quickly rallied around this architecture, though it still took approximately five years from the paper's publication to the ChatGPT moment in late 2022.
Quickly forming opinions on how the Twitter algorithm and platform worked allowed the speaker to grow rapidly on the platform.
There are more than 90,000 Flock surveillance cameras currently in use around the United States.
A 2023 report estimated that 10 million Americans own Ring cameras, roughly 1 in 5 households having a video-enabled doorbell.
The ImageNet dataset collected 15 million images to drive machine learning, becoming a cornerstone of the modern AI revolution.
A Stanford graduate student benchmarked human performance on the ImageNet 1,000-category challenge at roughly 4% error rate, a figure AI surpassed by 2016.
From the 2012 ImageNet breakthrough, it took only about 3–4 more years for AI algorithms to surpass human performance in naming 1,000 object categories.
Flock's surveillance network scans more than 20 billion license plates per month across the United States.
OpenAI's Sora, released in January 2024, demonstrated AI's ability to generate realistic video from text prompts, marking a key milestone in video generation.
AlphaGo's Move 37 against Lee Sedol was a move that human Go masters had never considered, illustrating a unique form of AI creativity within constrained mathematical rules.
We use essential and analytics cookies to run Vuci. To understand how the site is used: Privacy Policy.
Install Vuci on your phone
Add it to your home screen for a faster, app-like experience.
Install Vuci on your phone
Tap the Share button, then “Add to Home Screen”.
A new version is available
Reload to get the latest Vuci.