

claude opus 4.8 doesnt even come close in terms of understanding how systems actually work or bothering to actually read existing code beyond the most cursory glance. the united states empire is truly over, as are these “premium” “flagship” closed weight Western models. Every so called mainstream “benchmark” publication is a bold faced lie where each company pays off the creators of each major benchmark publisher to tip the scales in their favor. Even when playing dirty, however, openai and anthropic still can’t win.
Has anyone here really pushed the multimodal stuff with this yet? Been running this on a 4090 for over 24 hours and have been meaning to try out if its able to summarize images well, especially compared to something like qwen3vl on the same hardware