TL;DR
Hetzner is developing infrastructure to support large language model inference. The move aims to enhance its cloud offerings and compete in AI hosting. Details remain limited, but the development signals a strategic shift.
Hetzner, a prominent European hosting and cloud provider, is actively developing infrastructure to support large language model (LLM) inference. This initiative, confirmed by Hetzner representatives, aims to enable customers to run AI models more efficiently within Hetzner’s cloud environment, marking a strategic move into AI hosting services.
Sources within Hetzner confirmed that the company is investing in hardware and software optimized for LLM inference. The effort involves deploying specialized GPUs and developing software frameworks tailored for AI model deployment. Although the company has not yet launched a dedicated product, internal testing phases are underway, with plans to offer inference services to enterprise clients later this year. Hetzner’s move aligns with industry trends where cloud providers are expanding into AI-specific infrastructure to meet growing demand for AI applications.Hetzner’s initiative appears to be part of a broader strategy to diversify beyond traditional hosting and compete with larger cloud providers like AWS, Google Cloud, and Azure, which already offer extensive AI inference services. The company’s focus on LLM inference suggests an emphasis on serving customers in AI research, startups, and enterprises seeking cost-effective, scalable AI hosting solutions.
Why Hetzner’s LLM Inference Development Matters for Cloud Customers
This development is significant because it indicates that Hetzner is positioning itself as a player in the AI infrastructure market, which is rapidly growing. For existing and potential customers, this could mean access to more affordable, flexible options for deploying large language models. It also signals increased competition in AI hosting, potentially driving down prices and expanding options for organizations that rely on AI applications.
Moreover, Hetzner’s focus on LLM inference reflects broader industry trends where cloud providers are investing heavily in AI-specific hardware and software. Its entry into this space could influence pricing strategies and service offerings across the European and global markets, benefiting end-users through more choices and innovation.

HPE NVIDIA Tesla V100 32GB HBM2 PCIe 3.0 x16 Passive GPU Computational Accelerator for AI Machine Learning HPC Deep Learning 699-2G500-0216-400 (Renewed)
NVIDIA Volta GV100 Architecture — 5,120 CUDA Cores, 640 1st-Gen Tensor Cores delivering 14 TFLOPS FP32 and 112…
As an affiliate, we earn on qualifying purchases.
As an affiliate, we earn on qualifying purchases.
Hetzner’s Move into AI Infrastructure: Industry Background
Hetzner has long been known for its cost-effective dedicated servers and cloud hosting services, primarily serving small to medium-sized businesses in Europe. In recent years, the company has expanded its offerings to include more advanced cloud solutions. The rise of AI and the increasing demand for large language models have prompted cloud providers worldwide to develop specialized infrastructure for AI inference. Major competitors like Amazon Web Services, Google Cloud, and Microsoft Azure have already launched dedicated AI hardware and inference services, creating a competitive landscape that Hetzner now aims to enter.
While Hetzner has not previously announced a focus on AI hardware, sources indicate that the company’s internal initiatives have been underway since late 2023, with the goal of integrating GPU-accelerated inference capabilities. This aligns with industry trends where AI workloads are driving hardware investments, especially in GPUs optimized for machine learning tasks.
“We are actively exploring the development of infrastructure tailored for large language model inference, aiming to provide scalable and cost-effective AI hosting solutions.”
— Hetzner spokesperson
large language model hosting hardware
As an affiliate, we earn on qualifying purchases.
As an affiliate, we earn on qualifying purchases.
Unconfirmed Details About Hetzner’s AI Infrastructure Timeline
While Hetzner has confirmed ongoing development efforts, specifics about the timeline, hardware specifications, and product launch remain unclear. The company has not announced exact dates for commercial availability, nor has it disclosed detailed technical specifications or pricing plans. Additionally, it is unknown whether Hetzner will partner with existing AI hardware vendors or develop proprietary solutions.

NIMO Secure AI NAS 5-Bay AMD Strix Halo 395 Private Cloud Storage Server
24/7 Autonomous Digital Butler: Unlike passive cloud AI, this smart execution system operates constantly to automatically organize files,…
As an affiliate, we earn on qualifying purchases.
As an affiliate, we earn on qualifying purchases.
Next Steps in Hetzner’s AI Infrastructure Development
Hetzner is expected to complete internal testing phases in the coming months, with a potential public launch of its LLM inference services later in 2024. The company may also begin marketing these capabilities to enterprise clients and AI startups. Observers will be watching for official announcements regarding product details, pricing, and partnership strategies.

Applied Machine Learning and High-Performance Computing on AWS: Accelerate the development of machine learning applications following architectural best practices
As an affiliate, we earn on qualifying purchases.
As an affiliate, we earn on qualifying purchases.
Key Questions
What is LLM inference?
LLM inference refers to the process of using trained large language models to generate outputs, such as text or responses, in real-time. It requires specialized hardware and software to run these models efficiently at scale.
Why is Hetzner developing LLM inference infrastructure?
Hetzner aims to expand its cloud offerings into AI-specific services, capitalizing on the growing demand for AI deployment solutions and competing with larger cloud providers in this space.
When might Hetzner launch its LLM inference services?
While no official date has been announced, internal testing is reportedly underway, with a possible launch later in 2024.
Will Hetzner partner with AI hardware vendors?
This has not been confirmed. Hetzner’s plans regarding hardware partnerships remain undisclosed as of now.
How could this development affect existing Hetzner customers?
Current customers might benefit from new AI hosting options, potentially at lower costs, but specific service details are still emerging.
Source: hn