This is Part 2 of a two-part series. Part 1 covered the problem, the CLIP baseline, and BLIP. This...
When the Picture Doesn't Match the Label — Part 2: Feature-Based Detection with OCR and VQA
This is Part 2 of a two-part series. Part 1 covered the problem, the CLIP baseline, and BLIP. This...
If you're building an app that generates the same character across many scenes, you've probably hit...
Prompt injection is the #1 attack against AI agents. Nobody solves it well. I built L1.9 — a prompt...
I use Codex to implement frontend pages from Figma designs. When the Figma file is well structured,...