Lenovo unveils purpose-built AI inferencing servers

News
Jan 7, 20263 mins

New Lenovo servers cover high-performance tasks and edge computing.

Medium shot of female technician working on a tablet in a data center full of rack servers running diagnostics and maintenance on the system
Credit: Frame Stock Footage / Shutterstock

Lenovo Group Ltd. has introduced a range of new enterprise-level servers designed specifically for AI inference tasks. The servers are part of Lenovo’s Hybrid AI Advantage lineup, a family of inferencing devices.

Nvidia has captured the training space, where large language models (LLMs) are generated, but the inferencing space, where the LLMs are put to work doing things like answering questions and making decisions, is wide open with no clear leader.

But it is growing fast. Futurum Group estimates the global AI inference infrastructure market will grow from $5.0 billion in 2024 to $48.8 billion by 2030, for a six-year CAGR of 46.3%.

Lenovo says that moving from training to action turns the significant capital committed to AI into tangible business return, and invaluable competitive gain. Its new AI Inferencing suite executes AI workloads across an organization’s cloud, data center, and edge to wherever they deliver the greatest value.

“Enterprises today need AI that can turn massive amounts of data into insight the moment it’s created,” said Ashley Gorakhpurwalla, executive vice president at Lenovo and president of Lenovo Infrastructure Solutions Group in a statement. “With Lenovo’s new inferencing-optimized infrastructure, we are giving customers that real-time advantage—transforming massive amount of data into instant, actionable intelligence that fuels stronger decisions, greater security, and faster innovation.”

The first server is the Lenovo ThinkSystem SR675i, a high-end server featuring AMD Eypc server CPUs and Nvidia Blackwell GPUs and is built to handle large language models at scale and speed up simulations in sectors like healthcare, manufacturing, and finance.

There is also the Lenovo ThinkSystem SR650i, which offers high-density GPU computing power for faster AI inference and is intended for easy installation in existing data centers to work with existing systems.

Finally, there is the Lenovo ThinkEdge SE455i for smaller, edge locations such as retail outlets, telecom sites, and industrial facilities. Its compact design allows for low-latency AI inference close to where data is generated and is rugged enough to operate in temperatures ranging from -5°C to 55°C.

All of the servers include Lenovo’s Neptune air- and liquid-cooling technology and are available through the TruScale pay-as-you-go pricing model.

In addition to the new hardware, Lenovo introduced new AI Advisory Services with AI Factory Integration. This service gives access to professionals for identifying, deploying, and managing best-fit AI Inferencing servers. It also launched Premier Support Plus, a service that gives professional assistance in data center management, freeing up IT resources for more important projects.

Andy Patrizio is a freelance journalist based in southern California who has covered the computer industry for 20 years and has built every x86 PC he’s ever owned, laptops not included.

Andy writes the Data Center Explorer blog for Network World. His work has appeared in a variety of publications, including Tom's Guide, Wired, Dr. Dobbs Journal, Tech Target, Business Insider, and Data Center Knowledge. Earlier in his career, he held editorial positions at IT publications like InternetNews, PC Week and InformationWeek.

Andy holds a BA in Journalism from the University of Rhode Island.

More from this author