PodcastIntel
Sign in Get Started Free
AI Engineer

Computer Use at the Edge of the Statistical Precipice — Pierluca D'Oro, Programma Labs

Aug 14, 2026 · 0:17:28
AI Summary
  • Small, screen-agnostic scripts can match frontier model performance.
  • Replaying recorded actions blindly is a valid agent strategy on deterministic benchmarks.
  • Pass@k metric on deterministic environments reflects replay script success rate.

Guests on This Episode

PD
Pierluca D'Oro
1 podcast appearance

More from AI Engineer

View all episodes →

Get AI Summaries for Every New Episode

Subscribe to AI Engineer and get AI summaries, guest tracking, and email digests delivered automatically.

Sign Up Free →