Five takeaways from Bright Data’s VLA panel on why web-scale pretraining, curation speed, and data provenance now define the race to build real-world robots. Vision-Language-Action models need clean successful demonstrations for imitation learning, while world models need diverse outcomes including failures; panelists estimate one hour of web video provides roughly the learning value of five minutes of robot teleoperation data.