{"id":168773,"date":"2026-05-01T12:45:07","date_gmt":"2026-05-01T11:45:07","guid":{"rendered":"https:\/\/www.intelligentcio.com\/eu\/?p=168773"},"modified":"2026-05-01T12:45:08","modified_gmt":"2026-05-01T11:45:08","slug":"nebius-agrees-to-acquire-eigen-ai-strengthening-nebius-token-factory-as-a-frontier-inference-platform","status":"publish","type":"post","link":"https:\/\/www.intelligentcio.com\/eu\/2026\/05\/01\/nebius-agrees-to-acquire-eigen-ai-strengthening-nebius-token-factory-as-a-frontier-inference-platform\/","title":{"rendered":"Nebius agrees to acquire Eigen AI, strengthening Nebius Token Factory as a frontier inference platform"},"content":{"rendered":"\n<p><em>The deal combines Eigen AI\u2019s advanced inference stack with Nebius\u2019s global AI infrastructure to enhance performance, scalability and production efficiency.<\/em><\/p>\n\n\n\n<p>The AI cloud company Nebius has announced an agreement to acquire Eigen AI, a leading inference and model optimisation company.<\/p>\n\n\n\n<p>The acquisition will strengthen Nebius Token Factory as a frontier managed inference platform for production AI, combining a battle-tested optimisation stack with Nebius\u2019s global compute capacity and AI cloud platform, and will add elite inference research talent to the company\u2019s established in-house AI R&amp;D capabilities.<\/p>\n\n\n\n<p>Following close, Eigen AI\u2019s inference and post-training optimisation layers will be integrated directly into Nebius Token Factory, which provides enterprise-grade autoscaling endpoints and fine-tuning pipelines across all major open-source models. The two companies have already delivered jointly optimised implementations of leading open source models that ranked among the fastest on Artificial Analysis.<\/p>\n\n\n\n<p>The acquisition also accelerates Amsterdam-based Nebius\u2019s expansion in the US. Eigen AI\u2019s founding team \u2013 researchers who have developed optimisation techniques and tools the industry runs on \u2013 will join Nebius to establish a Nebius engineering and research presence in the San Francisco Bay Area.<\/p>\n\n\n\n<p>Roman Chernin, co-founder and Chief Business Officer, Nebius, said: &#8220;We are operating in a capacity-scarcity world where AI builders need optimised inference and infrastructure scale. The integration of Eigen AI\u2019s optimisation capabilities and founding team will establish Nebius Token Factory at the frontier of inference, offering customers market-leading model performance and unit economics with massive compute capacity to back it at scale.&#8221;<\/p>\n\n\n\n<p>New York based Eigen AI\u2019s founding team brings deep expertise from research that shapes how the industry deploys inference today. Co-founders Ryan Hanrui Wang and Wei-Chen Wang are alumni of MIT\u2019s HAN Lab, led by Professor Song Han, a pioneering researcher in AI computing and model efficiency.<\/p>\n\n\n\n<p>Ryan\u2019s pioneering Sparse Attention (SpAtten) work is the most-cited HPCA paper since 2020, while Wei-Chen received the MLSys 2024 Best Paper Award for Activation-aware Weight Quantisation (AWQ) quantisation \u2013 now the standard for 4-bit model serving in production deployments. Co-founder Di Jin, an MIT CSAIL PhD, brings deep expertise in post-training and large-scale model alignment, having contributed to Meta&#8217;s Llama 3 and Llama 4 post-training and co-authored the CGPO RLHF framework.<\/p>\n\n\n\n<p>Ryan Hanrui Wang, co-founder and CEO, Eigen AI, said: &#8220;We\u2019re proud to join Nebius and work alongside the Token Factory team to push the boundaries of inference performance. Nebius has built a world-class AI cloud with a deep engineering culture that perfectly aligns with our own. Together, we are removing the friction of AI model customisation and deployment so developers can run models reliably in production without managing the underlying infrastructure.&#8221;<\/p>\n\n\n\n<p>Inference is now the fastest-growing segment of AI, forecast to account for about two-thirds of compute demand this year. Open-source model usage is rising alongside it. With more workloads moving into production, the system optimisation layer is becoming critical infrastructure.<\/p>\n\n\n\n<p>Running inference efficiently in production is inherently complex and requires deep expertise across the entire execution stack, from how models are represented, to how GPU kernels execute them, to how workloads are scheduled in real time.<\/p>\n\n\n\n<p>Open-source models typically ship unoptimised, and newer architectures such as Mixture-of-Experts (MoE), Compressed Sparse Attention (CSA), reasoning and long-context models introduce additional challenges around memory, routing and compute efficiency. Most teams do not have the capacity to solve these problems in-house.<\/p>\n\n\n\n<p>Eigen AI addresses this challenge with a full-stack optimisation approach that spans the entire model lifecycle. From post-training and fine-tuning to production inference optimisation, across all major open-source models in production demand, including GPT-OSS, Gemma, Qwen, Llama, Nemotron, DeepSeek, GLM, Kimi and MiniMax.<\/p>\n\n\n\n<p>By integrating Eigen AI\u2019s optimisation layer directly into Nebius Token Factory, Nebius removes this bottleneck across the lifecycle. The system, model- and kernel-level techniques developed by the Eigen team are designed to extract materially better performance from hardware out of the box, delivering higher throughput and lower cost per inference without additional engineering overhead.<\/p>\n\n\n\n<p>As a result, Nebius Token Factory customers will benefit from faster time to production, significantly better unit economics and the ability to adopt new models more quickly. Existing Eigen AI customers will gain access to Nebius\u2019s global AI infrastructure and platform capabilities.<\/p>\n\n\n\n<p>The deal consideration will be paid in a combination of cash and Nebius\u2019s Class A shares with aggregate value as of signing, based on Nebius\u2019s 30-day weighted average stock price, of approximately US$643 million, subject to adjustments. The transaction is expected to close in the coming weeks &#8211; subject to certain customary conditions, including antitrust clearance.<\/p>\n","protected":false},"excerpt":{"rendered":"<p>The deal combines Eigen AI\u2019s advanced inference stack with Nebius\u2019s global AI infrastructure to enhance performance, scalability and production efficiency. The AI cloud company Nebius has announced an agreement to acquire Eigen AI, a leading inference and model optimisation company. The acquisition will strengthen Nebius Token Factory as a frontier managed inference platform for production [&hellip;]<\/p>\n","protected":false},"author":58,"featured_media":168774,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"_acf_changed":false,"footnotes":""},"categories":[21213,21064,93],"tags":[23594,19517,23725,76,25273,25271,3881,21279,25272,25274],"class_list":["post-168773","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-acquisitions","category-ai","category-top-stories","tag-ai-cloud","tag-ai-infrastructure","tag-ai-optimisation","tag-digital-transformation","tag-eigen-ai","tag-inference-platform","tag-machine-learning","tag-nebius","tag-open-source-models-2","tag-san-francisco-bay-area"],"acf":[],"publishpress_future_workflow_manual_trigger":{"enabledWorkflows":[]},"_links":{"self":[{"href":"https:\/\/www.intelligentcio.com\/eu\/wp-json\/wp\/v2\/posts\/168773","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.intelligentcio.com\/eu\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.intelligentcio.com\/eu\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.intelligentcio.com\/eu\/wp-json\/wp\/v2\/users\/58"}],"replies":[{"embeddable":true,"href":"https:\/\/www.intelligentcio.com\/eu\/wp-json\/wp\/v2\/comments?post=168773"}],"version-history":[{"count":1,"href":"https:\/\/www.intelligentcio.com\/eu\/wp-json\/wp\/v2\/posts\/168773\/revisions"}],"predecessor-version":[{"id":168775,"href":"https:\/\/www.intelligentcio.com\/eu\/wp-json\/wp\/v2\/posts\/168773\/revisions\/168775"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.intelligentcio.com\/eu\/wp-json\/wp\/v2\/media\/168774"}],"wp:attachment":[{"href":"https:\/\/www.intelligentcio.com\/eu\/wp-json\/wp\/v2\/media?parent=168773"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.intelligentcio.com\/eu\/wp-json\/wp\/v2\/categories?post=168773"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.intelligentcio.com\/eu\/wp-json\/wp\/v2\/tags?post=168773"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}