2026 AI Infra Summit
Sponsor Spotlight · 2026
Power Efficiency: 10-30% More Performance Per Watt
Key takeaways- With 160 neoclouds starting up globally, the pitch is to optimize the power a data center already has rather than build more generation capacity.
- The management layer boots first on the platform, collects forensic telemetry on every component, and load-balances the idle periods in XPU and CPU operation.
- Treating racks and pods as single units with no human intervention is claimed to yield 10–30% more tokens per dollar or per watt than cooling-only approaches.
Gopi Reddy Sirineni, President and CEO of Axiado, presents his company's intelligent platform management solution that addresses power efficiency challenges in AI data centers by monitoring system components, identifying idle periods in XPU and CPU operations, and performing automated load balancing across infrastructure. Axiado operates as a foundational management layer that boots first on platforms and enables racks and pods to function as single units, delivering 10 to 30% more tokens per dollar or per watt compared to traditional optimization methods.
More from this showcase
Other executive interviews in 2026 AI Infra Summit.
AI Cooling Tech Evolution
- Every watt a GPU draws creates a matching thermal obligation, so power and cooling have to be solved as one problem as AI compute density climbs.
- Cooling is moving from air to single-phase and two-phase direct-to-chip, which works only when the system is engineered as a whole rather than assembled from standalone components.
- Power compute effectiveness is the metric that matters — the share of each watt that reaches the GPU rather than the supporting infrastructure.
Abstract
Jeff Moore, VP of Strategic Partnerships at Aegis, participates on a panel at AI Infra in Santa Clara discussing thermal and power challenges created by high-wattage AI compute GPUs, which are driving the evolution from air cooling to advanced direct-to-chip cooling solutions. He emphasizes power compute effectiveness as a critical metric for maximizing GPU watt allocation and highlights the industry's long-term commitment to solving these infrastructure challenges.
Performance at Scale, Energy Efficiency & Cybersecurity
- Performance at scale dominates the summit conversation, spanning compute, data centers, edge, and physical AI, with chiplet architectures drawing the most excitement.
- Energy efficiency has to be architected from the silicon up through racks to the whole data center, not bolted on afterwards.
- The attack surface has widened since Spectre and Meltdown to CPUs, NPUs, XPUs and whole semiconductor subsystems, so security must run from silicon through firmware to software.
Abstract
Michal Siwinski, Executive Vice President, Chief Product and Marketing Officer, and GM of Security Solutions at Arteris, identifies three key themes at the AI Infrastructure Summit: performance at scale across compute and data centers with emphasis on chiplet architectures, energy efficiency from silicon to data center level, and comprehensive cybersecurity integration. Siwinski notes that the attack surface has expanded significantly since the Intel Spectre and Meltdown vulnerabilities, requiring Arteris to partner with clients on security solutions spanning semiconductors through entire systems.
Astera Labs Taurus 200G Smart Signal Conditioners
- Taurus 200G-per-lane signal conditioners cover Ethernet, UALink, and PCIe in the industry's first OCP-compatible footprint spanning both retimers and redrivers.
- Smart swap lets customers switch between retimer and redriver at any point in the design cycle, with one Cosmos software stack across both.
- The Taurus 4 3.2T retimer and redriver were shown driving PRBS patterns across a backplane mockup, trading reach against power for scale-up links.
Abstract
Aanchal Sharma, Senior Director of Product Management at Astera Labs, introduces the company's Taurus 200 gig per lane smart signal conditioners featuring an innovative "smart swap" capability that enables customers to interchange between retimers and redrivers at any design cycle stage using a single Cosmos software stack. The demonstration showcases the Taurus 4 3.2T retimer and redriver in a backplane connection mockup for scale-up applications across Ethernet, UALink, and PCIe, with both products offering integrated telemetry and diagnostics in an OCP-compatible footprint.