Background agents (the tasks your AI delegates work to) can now look at images, capture page screenshots, and understand videos, just like the main agent. Visual work delegates reliably: captioning a photo gallery, writing honest alt text, matching a reference design, picking a video poster frame, or reviewing a page's layout no longer fails partway because the helper agent could not see what it was working with.
Available on vision-capable models across all task roles - explorer, coder, auditor, and researcher.