Testing AI-written code is one thing, testing it against real customer data raises the…
This is a dev post classified by Jev as Testing & observability (a tool drop), kept by the Dev Radar because it carries real work, not commentary.
Testing AI-written code is one thing, testing it against real customer data raises the stakes. How do you give testing agents the access it needs while keeping that data isolated? Lark built an AI test platform that maps a customer's product, writes end-to-end tests, and keeps running them as the code changes. That means testing environments with real user data and credentials in play. Lark also needed to run Docker inside the sandbox itself to spin up a customer's own dev environment for testing. @e2b's MicroVM isolation gave Lark strict security boundaries plus Docker-in-sandbox suppor
Posted by Vasek Mlejnsky (9k followers) 2 days ago · 27 likes · 1.3k views · view the original post on X. Kept by the Dev Radar as Testing & observability.
More dev work like this
- What does observability look like when a judgment costs nothing? — @hugorcd
- When agent traces turn into a debugging pile, this repo gives you a place to inspect them. — @DanKornas
- I’m glad @strawgate's standards for coffee are lower than his standards for code review… — @kayvz
- Gitar finds bugs, fixes them under your team's rules, and verifies every fix by running… — @SonarSource
- One AI agent is easy to monitor. 1,000 agents running across your org? Different problem. — @AlphaSignalAI
- OPENAI 🔥: Chrome extensions are now supported in the ChatGPT desktop browser! — @testingcatalog
- Regressions happen. But you don't want to find out from app store reviews or spicy… — @expo
- e2e + jev from @typesafeai ⚡ — @o_kwasniewski
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 13.5k posts from 4.7k X accounts over the last 21 days, 1.5k tools, 12 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-19 19:24 UTC. Full method.