Computer vision pioneer Jitendra Malik on why machines still underestimate vision and must learn like children.

Jitendra Malik: A professor at UC Berkeley and one of the seminal figures in computer vision, both before and after the deep learning revolution. Cited over 180,000 times, he has mentored many world-class AI researchers.
Lex Fridman interviews Jitendra Malik, a foundational computer vision researcher, about why vision is far harder than it appears since most human visual processing is subconscious. Malik argues that perception is fundamentally tied to action and that current systems rely too heavily on supervised, feed-forward learning. He advocates for systems that learn like children: multimodally, incrementally, physically, and through exploration. The conversation spans autonomous driving skepticism, segmentation, 3D understanding, the limits of the Turing test, and the present-day risks of AI in recommender systems.
Canonical
“shout out to my favorite flavor of linux ubuntu mate 2004 once again get it at expressvpn.comlexpod”— Lex Fridman
Netflix
“if you watch netflix if you enjoy watching movies you're using your perception system to interpret the movie ultimately your enjoyment of that movie means you'll subscribe to netflix”— Lex Fridman
Tesla
“i mean i i've seen examples of this in uh actually i mean i own a tesla and it has various safety features built in”— guest
Alison Gopnik
“my colleague alison gopnik has together with a couple of authors co-authors has this book called the scientist in the crib referring to children”— guest
Peter Medawar
“there is a british scientist uh in fact he won a nobel prize peter medover who has a book on on this and uh basically he calls it the research is the art of the soluble”— guest
Fyodor Dostoyevsky
“now let me leave you with some words from prince mishkin and the idiot by dostoyevsky beauty will save the world”— Lex Fridman