People don't know who Dylan is. This guy has a history of seeking out negative spins on Chinese issues and using his position as mod of r/hardware to keep them up.
I fail to see novelty here. What's the size difference between the model and and all of the 64x32 image training data? If the difference is not significant, you're basically almost just scrubbing a video, right?
Modeless, can you please answer this "trick" question?
If a flat piece of cardboard that is 10 inches wide occupies 10 degrees of my horizontal FOV, how many degrees of my horizontal FOV will 20 inches wide flat piece of cardboard occupy? Both cardboards are positioned at the same distance from my eyes, of course.