devwithjev
โ€” reading nowโ€” views
Submit a build

Jev Chooses Caption Positions

Lahfir uses Jev to decide where each caption goes. MediaPipe detects face boxes and a mouth keypoint, while a Face Mesh check rejects the back of a head; deterministic code scores how well each mouth moves with the audio.

Lahfir@mdlahfir๐•
Jev experiment: Caption positioning Captioning is solved. Deciding where each caption goes is still manual work. MediaPipe gives 1. face boxes, 2. a mouth keypoint, and 3. a Face Mesh check that rejects the back of a head. The deterministic code turns that into one number: how well each mouth moves in time with the audio. Jev turns it into a decision. Two-choice questions per line: 1. which face s
Sep 21, 2026X postsView on X

Also filed under Tools & apps