SecuritylabSeptember 4, 2026🇷🇺Translated from Russian

September 2026 AI Model Rankings: Fable 5.1 Tops Intelligence Index as Competition Tightens Across GPT-5.6 Sol, Grok 4.6 and Muse Spark 1.3

The start of September 2026 proved to be a rare moment when the roster of the strongest language models had to be almost completely rewritten. Anthropic introduced Fable 5.1 and the restricted Mythos 5.1, Meta updated Muse Spark to version 1.3, Google launched Gemini 3.8 Flash, and Alibaba refreshed Qwen3.8-Max. Summer releases such as GPT-5.6 Sol, Grok 4.6, Kimi K3, GLM-5.3 and DeepSeek V4 Pro stayed in the race and continue to compete with the newcomers.

Compiling a conventional ranking from smartest to least capable has become difficult. Contemporary models run in several reasoning-depth modes, and the gap between low, high and max settings can exceed ten points on a single test. Speed, token consumption, long-context pricing, image support, tool-use quality and the ability to execute long autonomous action sequences now vary widely.

It is therefore more useful to treat the market as several overlapping races. Fable 5.1 currently claims the highest quality for complex reasoning. GPT-5.6 Sol and Grok 4.6 deliver comparable performance at lower cost. Muse Spark 1.3 unexpectedly entered the top tier on capability-to-cost ratio. Gemini 3.8 Flash bets on speed and multimodality. Kimi, GLM, Qwen and DeepSeek have nearly erased the former boundary between Chinese and American models.

How to Compare Models Without Misleading Yourself

The Artificial Analysis Intelligence Index serves as a convenient common scale. This independent metric combines tests on knowledge, programming, scientific reasoning, tool use and long-horizon agent tasks. The methodology is more reliable than vendor tables because models are evaluated under comparable conditions.

Scores cannot be read as an intelligence quotient. A result of 66 versus 61 does not mean the first model is eight percent smarter. The index reflects average performance on a specific test suite. On Russian-language article writing, accounting-document analysis, website layout or debugging enterprise applications the ordering may differ.

Reasoning modes alter the picture even more. GPT-5.6 Sol scores around 61 in max mode, 59 in xhigh and 57 in high. Fable 5.1 reaches 66 only at maximum depth. Gemini 3.8 Flash scores 59 in high, 57 in medium and 52 in low. Comparing Gemini against Claude without specifying the mode is about as useful as comparing cars without stating engine power.

Market Leaders at the Start of September 2026

The current top tier is unusually dense. After Fable 5.1 the gap between several flagships fits within a few points. Expensive models do not always complete a given task more cheaply; some systems generate far more intermediate text, reason longer or invoke external tools more frequently.

  • Claude Fable 5.1, max – 66 points, 1 M context, $10/$50 per million tokens, available
  • Claude Opus 5, max – 63 points, 1 M context, $5/$25, available
  • Muse Spark 1.3, max – 62 points, 1 M context, price not announced, limited access
  • GPT-5.6 Sol, max – 61 points, ~1.05 M context, $4/$20, available
  • Grok 4.6, high – 61 points, 500 k context, $2/$6, available
  • Muse Spark 1.3, xhigh – 61 points, 1 M context, $1.25/$4.25, available
  • Kimi K3, max – 60 points, ~1 M context, $3/$15, open weights
  • GLM-5.3, max – 60 points, up to 1 M context, $1.40/$4.40, open weights
  • Gemini 3.8 Flash, high – 59 points, ~1.05 M context, $0.75/$3.75, available
  • Qwen3.8-Max – 58 points (previous snapshot), 1 M context, ~$2/$6, updated to 0902
  • DeepSeek V4 Pro 0813, max – 53 points, 1 M context, $0.66–1.32/$1.98–3.96, open weights

Additional caveats apply. Independent tests of the updated Qwen3.8-Max-0902 have not yet produced a comparable final score, so the 58-point figure refers to the prior version. Muse Spark 1.3 maximum mode remains in limited testing; the public xhigh mode scores 61. For Grok 4.6 the best independent result occurs in high mode, while xhigh unexpectedly scores slightly lower.

Pricing also requires notes. The Gemini 3.8 Flash tariff is temporarily reduced until the end of 2026. GPT-5.6 Sol is sold at a temporarily lowered price. DeepSeek pricing changes between peak and off-peak hours. Grok 4.6 doubles in price for requests longer than 200 k tokens. A simple API-price column therefore hides half of the real economics.

Related articles

SecuritylabOther

How Social Media Recommendation Algorithms Shape Content, Behavior, and Regulation Worldwide

Social media platforms such as TikTok, Instagram, and YouTube rely on sophisticated recommendation systems that analyze user behavior to predict engagement rather than simply ordering posts chronologically. These algorithms form candidate sets and rank content using signals including watch time, likes, comments, and device data, creating personalized feeds that can diverge significantly between users. Research shows mixed effects on political views and mental health, with personalization capable of both amplifying polarization and reducing it depending on optimization goals. Regulatory responses vary: China mandates transparency and youth mode limits on Douyin, the EU's Digital Services Act requires non-profiled alternatives for large platforms, and Russia enforces disclosure rules under Federal Law No. 408-FZ since 2023. Practical steps allow users to reset recommendations and provide negative feedback to retrain feeds. The article debunks myths of dopamine addiction and inevitable radicalization while highlighting both risks and benefits of algorithmic curation.

AntiMalwareOther

Russia Proposes 22% VAT and 100-Ruble Customs Fee on All Foreign Online Purchases

The Russian Ministry of Finance has submitted draft amendments to the Tax Code and the federal budget for 2027-2029 that would impose a standard 22% VAT on goods purchased from foreign online stores and marketplaces. In addition, every postal shipment valued at up to 200 euros would incur a separate customs fee of 100 rubles. Electronic trading platforms are designated as tax agents and are expected to pass the new costs directly to buyers through higher final prices. The proposals eliminate the current duty-free threshold for low-value cross-border orders under EAEU rules and replace an earlier plan for gradual VAT increases with an immediate full-rate application. Officials state the measures are intended to create equal competitive conditions for Russian sellers and to further formalize the economy. The changes have not yet been adopted and must still undergo government and legislative review.

AntiMalwareOther

VK Tech Rolls Out Unified AI Assistant Across VK WorkSpace Corporate Tools

VK Tech is preparing a unified AI Assistant for its VK WorkSpace platform, along with semantic search and an MCP server to connect autonomous agents. The features target small and medium businesses as well as large corporations by automating routine tasks in messaging, email, and calendar services. Semantic search will allow users to locate documents, emails, and messages by meaning rather than exact titles or phrasing. Meeting recordings will be automatically transcribed with extraction of key topics, decisions, and action items. The AI Assistant launches first in the Messenger module of the On-Premise version, with cloud SaaS availability scheduled for October and later expansion to Mail, Disk, and Calendar. Access remains strictly limited by user permissions and company policies. An MCP server will also be introduced, enabling AI agents to query data, prepare meeting materials, and perform actions inside the platform services, initially supporting Messenger and extending to other modules by year-end.

AntiMalwareOther

Russia Weighs Passenger Fees of Up to 1000 Rubles to Fund Domestic Aircraft Leasing

Russian authorities are discussing a new surcharge on air tickets to cover leasing and operating costs of locally produced planes. The proposed fees range from 500 rubles on domestic flights to 1000 rubles on international routes, with some sources mentioning higher figures of 700 and 2000 rubles respectively. The initiative involves the Ministry of Transport, the Ministry of Industry and Trade, and Rostec, though no final decisions on amounts or collection methods have been made. Revenue estimates suggest the measure could generate 68 to 84 billion rubles annually based on 2025 passenger volumes, with total collections targeted at 200-250 billion rubles between 2027 and 2030. Funds would subsidize the gap between high production costs of models such as the MC-21 and SJ-100 and the discounted prices offered to airlines, while also addressing elevated maintenance and fuel expenses. Officials expect serial production after 2030 to lower costs and eliminate the need for the surcharge, yet experts caution that added fees risk raising ticket prices and reducing overall passenger traffic.