They are also overly US centric.
One of the questions asks you to click on only the school buses. I had to Google how you tell the difference between a school bus and not a school bus.
Also is it a crosswalk if it's at an intersection or is it only a crosswalk if it's in the middle of a road somewhere?
The questions either need to be not cultural or they need to be adapted for where they detect the user is coming from, the first option seems easier.
I think the reason AI are better than humans is that the AI is just as stupid as the image classifier.