I like it as I'm vibecoding an (airgappable) browser AI workspace myself, but in terms of putting the models to use, just exposing the chat interface feels a bit limiting to me.
My take on this concept: https://github.com/willaaam/gemma-4-E2B-webgpu-vision