Nebius to Acquire Eigen AI, Deepening Its Token Factory Inference Bet

Nebius acquires Eigen AI to boost Token Factory AI inference platform, shown as GPU cloud infrastructure

Nebius, the Amsterdam-headquartered AI infrastructure company, announced on April 30, 2026 that it has agreed to acquire Eigen AI, a deal the company says will strengthen Nebius Token Factory — its managed platform for running AI models in production — as a “frontier inference platform.” Financial terms were not disclosed in the announcement.

Executive Summary

The announcement is short on detail but clear in direction: Nebius is buying its way further up the stack. Token Factory is the company’s inference service — inference being the work of actually running a trained AI model to answer queries, as opposed to the one-time job of training it. By acquiring Eigen AI, Nebius signals that it wants to compete on the software and efficiency of serving models, not only on the raw GPU capacity underneath.

That matters because inference is where the AI infrastructure market’s recurring revenue increasingly lives. Training runs are lumpy, contract-driven, and dominated by a handful of frontier labs; inference demand grows with every application that puts a model in front of end users. A GPU cloud that can serve tokens more efficiently than rivals can either undercut them on price or keep the margin — and an in-house optimization team is one of the few durable ways to get that edge.

Inference Is Becoming the Real Battleground

For the past several years, the headline numbers in AI infrastructure have come from training: giant clusters, multi-year capacity contracts, gigawatt campuses. But training is a capital-intensive land grab with a small set of customers. Inference — serving billions of model queries a day — is the volume business, and its economics are decided by software as much as hardware. Techniques like smart request batching, caching, and model-serving optimizations can multiply how many tokens a given GPU produces per second, which translates directly into cost per query.

Nebius framing the deal around making Token Factory a “frontier inference platform” tells you where it thinks the fight is heading. Frontier-scale models are expensive to serve, and the providers who serve them cheapest — without sacrificing latency or reliability — will win the workloads of AI application companies that live and die on unit economics.

Vertical Integration in the AI Cloud Race

Nebius belongs to the cohort often called neoclouds — specialist GPU cloud providers that grew up renting accelerator capacity, distinct from hyperscalers like AWS, Microsoft Azure, and Google Cloud. The strategic risk for any neocloud is commoditization: if all you sell is access to the same Nvidia hardware everyone else buys, price competition eventually erodes margins. The escape route is moving up the stack into managed platforms, and inference services are the most natural rung.

Acquiring an inference-focused company rather than building everything internally is a classic vertical-integration play: own the layer that differentiates your commodity input. Hyperscalers and inference-API specialists are pursuing the same layer, so the competitive logic is straightforward — Nebius needs Token Factory to be more than a thin wrapper around GPUs, and buying specialized talent and technology is faster than growing it.

Buy Versus Build, and What a Thin Release Does and Does Not Establish

It is worth being precise about what the announcement substantiates. It establishes that Nebius has agreed to acquire Eigen AI and that Nebius intends the deal to bolster Token Factory’s inference capabilities. It does not disclose a purchase price, Eigen AI’s size, its customers, or the specific technology being acquired — so any claim about how much this improves Token Factory’s performance or economics is, for now, unverifiable from the source material. “Strengthening” language in an acquisition release is aspiration until integration results show up in benchmarks, pricing, or customer wins.

Still, the pattern is credible. Across the industry, inference-optimization teams — often small groups with deep expertise in GPU kernels, serving engines, and scheduling — have become prized acquisition targets, because a handful of engineers can move serving costs by double-digit percentages. If Eigen AI fits that profile, the deal is less about revenue than about capability: the acqui-hire economics of the AI era, where talent density in a narrow specialty commands strategic premiums.

Background

Nebius Group emerged in 2024 from the restructuring of Yandex N.V., the Dutch holding company that divested its Russian assets and refocused on AI infrastructure, resuming trading on Nasdaq that year. Since then, Nebius has expanded aggressively — building GPU data-center capacity in Europe and the United States and signing large capacity agreements, including a multibillion-dollar GPU deal with Microsoft announced in September 2025. Token Factory, launched in late 2025, is its managed inference platform and a centerpiece of its push beyond raw compute rental into higher-margin platform services, of which the Eigen AI acquisition is the latest step.

Source: Nebius agrees to acquire Eigen AI, strengthening Nebius Token Factory as a frontier inference platform — company announcement dated April 30, 2026, distributed via Google News.