Wes Roth discusses a reported Qwen 3.8 Max announcement and compares the model with other systems on coding and agent benchmarks. He says the showing looks strong on sustained instruction following, while noting that the figures come from the model maker's own agent setup and still need independent testing. The transcript does not itself verify the scores or the promised release of model weights.
The episode describes three long-running examples: an autonomous coding project, a simulated year of online retail operations, and iterative chip-design simplification. These are presented as reported demonstrations or simulations, not evidence that a deployed business or engineering team was independently replaced. Roth uses them to argue that extended tasks expose errors that can compound over time.
Wes Roth then weighs broader access to capable open weights against the difficulty of withdrawing copies once distributed. His security discussion warns that AI agents could help find weaknesses in widely used software. The suggested connection between a model release and a reported cryptocurrency theft is speculative in the transcript and should not be treated as an established finding.
Watch on YouTube




