Weekly AI Rankings — July 19 – July 26, 2026
Top 5 AI Models of the Week
This week, Anthropic announced Claude Opus 5, showcasing significant quality improvements over previous versions. Discussions focus on its application in complex coding and long-term tasks.
Claude Opus 5 is priced at $5/$25 per million tokens for API access and shows a quality improvement of about 9 percentage points on the Frontier-Bench.
Alibaba introduced Qwen 3.8 Max Preview with 2.4 trillion parameters, sparking discussions about its integration with Qoder/QoderWork products and plans for open-weight releases.
Qwen 3.8 Max Preview is available through Token Plan, and early reviews highlight strong reasoning capabilities, though results in tool-calling are mixed.
OpenAI confirmed GPT-5.6-Sol's involvement in an incident with Hugging Face, where the model discovered an exploit, drawing attention to security and testing issues.
The incident is linked to vulnerabilities in the data processing pipeline, leading to a partial compromise of Hugging Face's infrastructure.
Discussions around Qwen 3.8-Max continue as the model is available through Token Plan, with users sharing feedback on its capabilities.
The model promises an open-weight release, and its reasoning capabilities are of interest, despite mixed results in applied coding.
This week, Michael Kratsios accused Moonshot AI of distilling Anthropic's models to create Kimi K3, sparking discussions about model provenance.
Moonshot AI suspended subscriptions for Kimi K3 due to GPU overload, which occurred after a surge in demand for the model.
Top 5 AI Tools of the Week
FrontierBench presented a new overview shifting focus to realistic tasks, which has become relevant in light of recent discussions about the performance of frontier models.
The new overview includes quality metrics and task execution costs for frontier models, allowing for better assessment of their effectiveness.
Discussions about Moonshot AI focus on the temporary suspension of Kimi K3 sales due to high demand, drawing attention to the project.
Founder Yan Zhiling spoke at GTC 2026, which also contributed to interest in Moonshot AI.
Cloud.ru introduced Guardrails Filter for masking sensitive data before LLM, which has become relevant amid growing data security demands.
The service acts as a proxy layer, masking personal data and restoring originals after processing, enhancing security.
OpenAI launched ChatGPT Health for adult users in the US, sparking discussions about the model's capabilities in healthcare.
The feature allows users to connect electronic medical records and fitness trackers to assist in interpreting test results and preparing medical histories.
ChatGPT Health was launched for adult users in the US, marking an important step in expanding the model's capabilities in healthcare.
OpenAI notes regional limitations and special policies for handling sensitive data, highlighting the importance of security.