Bijan Bowen tests GPT-6.1 Sol through a browser desktop, games and interactive 3D scenes. He varies reasoning effort and speed settings, then reruns a pool-party task without fast mode because the initial comparison was not equivalent.
The browser desktop and skating game are functional but less elaborate than his recent Sonnet 5.5 results. A follow-up improves the skating scene and adds a successful replay camera, while the subway shooter follows the revised instructions and produces a more convincing result.
A detailed recreation of a television apartment, a RuneScape-style interface and an interactive car model show stronger object and interface detail. Important inaccuracies remain in spatial layout, game mechanics and the car's shape; these demonstrations are qualitative tests rather than a comprehensive benchmark.
The robot-arm experiment navigates more cautiously but does not complete its task. A legacy-laptop experiment also fails and partly switches to a different model, so it cannot establish Sol's capability there. Bijan Bowen reports modest subscription usage, while explicitly noting uncertainty about whether plan allowances had changed.
Watch on YouTube




