Reading title and this ^ comment before loading the page made my brain assume that this is about combining front and rear image sensors, and additionally using eye/facial tracking to infer user intention + satisfaction to help with stabilization/framing/post-processing. I'm glad(?) we aren't there yet.