Two local models in a browser video editor, measured: Gemma 4 E2B understands requests but can't make editing decisions; EmbeddingGemma 2 made video search actually work
TL;DR: I built a browser video editor whose AI runs entirely on WebGPU, with no server. Gemma 4 E2B as an agent: it understands requests well (54 right of 76 rephrasings of known edits, none wrong in a way that looked right) but makes bad decisions (a fixed rule beat it 32:0 at choosing edits from…
Read the full story at r/LocalLLM ↗