It does not reason through fixed rules
Instead of saying 'this is a car' with an endless checklist, we show it millions of images. The network learns recurring shapes, proportions, volumes, and reflections that usually appear when a vehicle is in front of the camera.







