VISTA
AI overview
Sign in and the AI will write an overview from our coverage.
Headlines · 2
- VISTA boosts multimodal agents with visual memory and active recall
A team led by Kaiming He introduced VISTA, a framework that gives multimodal agents direct visual input, lossless visual memory and tools to inspect past frames. According to the reported paper results, Claude Opus 5 improved from 40.68 to 100 on 25 public ARC-AGI-3 games, while GPT-5.6 Sol improved from 13.33 to 99 without changing the underlying models.
MIT科技评论中文 · 🔥 0 - VISTA Gives Frontier Models Visual Memory for ARC-AGI-3
A paper from Kaiming He’s team presents VISTA, a harness that gives vision-language models direct access to game images, persistent frame-by-frame visual memory, and model-controlled inspection tools. The source reports that Claude Opus 5 completed all 25 public ARC-AGI-3 games with a perfect score, while GPT-5.6 Sol achieved 99, attributing the gains to improved visual access and memory rather than additional model training.
创业邦 科技 · 🔥 0
Experience and discussion from the community
Share my VISTA experienceAsk about VISTA
Nobody has shared their experience with VISTA yet.