The mystery lasts a little over a day: Union Alpha, the stealth model that appeared on OpenRouter, OpenCode and Cloudflare, which had already sparked a buzz for its coding level, is in fact Pareto 26.9, the system developed by The Unbiased.
To reveal it, it was the same company, after the account @aitrackerbot managed to trace its identity within twenty-four hours. Before the official confirmation, the community had already started suspecting it: an analysis of the tokenizer had ruled out the initial hypothesis of a Zhipu AI/GLM model, bringing Union Alpha closer to an Llama 3-type architecture.
The focal point of the reveal is precisely this: Pareto does not correspond to a single set of weights, like a traditional model, but to a blended system that runs multiple LLMs in parallel on each request, dynamically synthesizing the results based on the task, with a verification mechanism that calls upon a more powerful model only when the task truly requires it.
The Unbiased does not publicly disclose the specific models that compose the system at any given moment, because the combination would change too frequently to remain an accurate information, while stating that it will announce every change before bringing it into production.
So no training runs are planned: according to The Unbiased, earlier Pareto versions taught the company that benchmarks can be misleading, while the private beta phase helped identify which models were particularly well suited for specific tasks.
Union Alpha, and thus Pareto, handles multimodal inputs of text and images, supports tool calling and works with a very wide context window: 262,144 tokens, with a maximum output of 131,072 tokens.
Based on the most recent public model card available (relating to version 26.8, not yet updated to 26.9 at the time of writing this article), Pareto achieves a composite score of 73.5 across seven public benchmarks at a cost of $0.10 per task, with results of 86.0 on Terminal-Bench 2.1, 86.0 on SWE-Bench Verified, 90.4 on GPQA-Diamond, 47.0 on Humanity’s Last Exam, 69.2 on arXivMath, 58.0 on DRACO and 77.9 on MMMU-Pro (the multimodal benchmark).
The comparison is made with models like Claude Opus 5, Claude Fable 5, GPT-5.6 Sol, Kimi K3 and DeepSeek V4F-0731. In terms of latency, The Unbiased reports no additional delay on agentive tasks, but up to three times slower on particularly complex reasoning tasks.
The Unbiased chose to launch Pareto 26.9 anonymously on OpenRouter, OpenCode and Cloudflare to gather real-user feedback, understand what difficulties would emerge at scale, and use these learnings ahead of the public launch of Pareto 26.10, scheduled for next month (according to a source, the date indicated would be October 10, 2026, not confirmed directly by The Unbiased).
The company specifies that it does not store prompts nor outputs from users, and believes that the absence of data retention should become the standard also for other models.
The initial communication spoke of free and unlimited access for a week, but the reality was different: Union Alpha was launched on September 16, 2026, and already the next day, September 17, 2026, at about 32 hours distance, free access was halted to make way for the paid version, precisely due to the outsized demand generated in the first hours.
During the short free trial phase, Pareto 26.9 offered coding parameters comparable to Astra’s, but at a clearly lower cost.
Now that the company is coordinating with platforms to enable paid access, Pareto 26.9’s price list stands at $2.50 (about €2.18) per million input tokens, $0.25 (about €0.22) per million cached tokens, and $7.50 (about €6.53) per million output tokens.
In the early hours of Wednesday, the demand reached one billion tokens per minute, rendering the model practically unusable due to its slowness. The Unbiased describes itself as a small-scale entity, and admits it did not start with sufficient computing capacity to sustain the interest generated. AWS helped the company triple capacity overnight, but even that was not enough to fully satisfy the demand.
For anyone who still wants to try the system through Cloudflare AI Gateway, access requires a Cloudflare account with a domain you own, a credit card and a prepaid minimum deposit of $10 (plus a fee of about $0.50); the integration is via API endpoints compatible with the OpenAI standard, easily usable in existing workflows.
The Unbiased considers systems like Pareto, based on the coordinated work of multiple models, the path to achieving more intelligence at a lower cost, and defines them as a new field of research.
From here on, the company states it intends to carry forward this research openly, publishing details on its official model card. The invitation to users is to continue using Pareto 26.9 and report where the system fails, ahead of the public launch of Pareto 26.10 next month.
During the HUAWEI CONNECT 2026 event, in Shanghai, Huawei unveiled Atlas 960E, a new AI…
Honor is preparing to present to the public a new smartphone that combines an extreme…
The renowned Asian brand HTC expands its business horizons by officially bringing its smart glasses…
After the announcement from last week and the pre-order phase, iPhone 18 Pro and iPhone…
If the value remains equal to or higher than this threshold, the component is considered…
Xiaomi has launched in China the new Mijia Smart Air Purifier 6C, an air purifier…