I'm skeptical of any Engineering loop that doesn't include reality (as in touch grass) feedback. Pure logic and reasoning is the domain of Maths and Science (philosophy). Surely it will work, but it will not "be able to solve any learning loop".
I'm almost certain the goal of this startup is to make physical automated research labs guided by RL
How is that different than video input?
I'm almost certain the goal of this startup is to make physical automated research labs guided by RL