I dont this is top most priority. We do this with codex or claude in chrome and validate the UX.<p>I think BDD is going to be back in style for agent generate code.<p>When you generate at scale discovering, maintaining and scaling test is the major problem, thats why were katana [1] a behavior driven testinf utility thay discovers behavior and maintain, generates test cases.<p>[1] <a href="https://github.com/adaptive-scale/katana" rel="nofollow">https://github.com/adaptive-scale/katana</a>
"It reads the screen, not the DOM", but is built completely on Playwright? At first it made me think you have some visual model at play, but doesn't actually seem that way.
We built something like this haha
CI with regular e2e tests usually gets pretty expensive and slow. How does cost compare to regular playwright tests?