Tokens Per Watt Over Peak Performance: Retrofitting Heterogeneous Silicon for Production Inference

Tokens Per Watt Over Peak Performance: Retrofitting Heterogeneous Silicon for Production Inference
by Kalar Rajendiran on 09-24-2026 at 2:00 pm

Image for Article 2 SemiWiki Tokens per Watt or per Dollar

👉 This is Part 2 of an editorial series on the evolving economics of AI inference. If you missed Part 1, where I broke down the shift toward full-system integration, the lessons from the AI Infra Summit, and the “no forks” open-source philosophy, you can read it HERE: Beyond the Accelerator: Why Silicon Challengers … Read More


WEBINAR: Porting, Analysing & Optimising AI Models onto Edge GPUs

WEBINAR: Porting, Analysing & Optimising AI Models onto Edge GPUs
by Admin on 09-21-2026 at 9:59 pm

Running AI models efficiently at the edge is about more than hardware TOPS. In this session, we’ll walk through the end-to-end process of porting, analysing and optimising modern AI models for edge GPUs, using Qwen 3.5 as a practical example. Learn how to take a model from framework to deployment, identify performance bottlenecks,… Read More


Webinar: LPO: Inside the First Standards-Compliant 112G Optical PHY

Webinar: LPO: Inside the First Standards-Compliant 112G Optical PHY
by Admin on 09-21-2026 at 9:27 pm

Featured Speaker:

  • Kant Deshpande, Architect, Technical Product Management, Synopsys

As AI/ML clusters scale toward 800G, 1.6T, and beyond, the optical DSP/retimer functions inside conventional transceivers (FRO) are becoming a significant bottleneck — adding power, latency, and cost to every link at a time when interconnects… Read More


Broadcom’s AI Engine Shifts Into Overdrive

Broadcom’s AI Engine Shifts Into Overdrive
by Daniel Nenni on 09-20-2026 at 8:00 am

The XPU Effect Broadcom AI

Broadcom’s fiscal third-quarter 2026 results show how artificial-intelligence infrastructure is shifting from general-purpose acceleration toward customized compute and large-scale networking. The company reported AI semiconductor revenue of $16.7 billion, up 221% year over year and 54% sequentially. That expansion… Read More


Webinar: The Convergence of Edge Accelerators

Webinar: The Convergence of Edge Accelerators
by Admin on 09-16-2026 at 3:36 pm

*Company Email Required for Registration*

ABSTRACT

Neural rendering. Offloading overloaded NPUs. Protection against evolving AI models. Developer-friendly programming. Reduced system complexity. There are many reasons why edge AI systems benefit from a GPU that delivers fast, efficient and fully programmable AI acceleration.… Read More


Podcast EP365: How Agentrys is Revolutionizing Chip Design with Mark Ren

Podcast EP365: How Agentrys is Revolutionizing Chip Design with Mark Ren
by Daniel Nenni on 09-11-2026 at 10:00 am

Daniel is joined by Mark Ren, founder and CEO of Agentrys. Mark has 26 years of EDA and AI R&D experience spanning IBM Research and NVIDIA Research, driving design automation innovations that power modern chip design. He received the IBM Corporate Award for contributions to the design closure for high-performance microprocessors.… Read More


Webinar: The Ghost in the GDSII: AI’s New Role in Physical Design

Webinar: The Ghost in the GDSII: AI’s New Role in Physical Design
by Admin on 09-08-2026 at 11:27 pm

Every chip ends its journey as a GDSII file — millions of polygons handed to the foundry. But hidden inside today’s layouts is a new kind of intelligence. AI is quietly reshaping physical design: guiding floorplanning and placement, predicting congestion and timing before a single wire is routed, and autonomously tuning… Read More


Webinar: From Waveforms to AI: Simulating 6G Before It Exists (America’s Session)

Webinar: From Waveforms to AI: Simulating 6G Before It Exists (America’s Session)
by Admin on 09-03-2026 at 1:59 am

September 24, 2026 | 10:00 AM PDT

6G standardization is accelerating: 3GPP RAN1 is actively evaluating candidate waveforms, new duplexing schemes, and advanced channel coding proposals, with each organization running proprietary simulators under subtly different assumptions. This makes direct comparison of results nearly

… Read More

Webinar: The future of systems engineering: How SysML v2 and generative AI transform engineering

Webinar: The future of systems engineering: How SysML v2 and generative AI transform engineering
by Admin on 09-03-2026 at 1:51 am

Engineering teams are dealing with increasing system complexity, stricter integration requirements and the shift toward software-defined products. Generative AI is emerging as a valuable capability, but it is not sufficient on its own.

To manage complexity and convergence across software and hardware, organizations need… Read More


PDF Solutions CONNECT 2026 Conference

PDF Solutions CONNECT 2026 Conference
by Admin on 09-03-2026 at 12:38 am

ABOUT THE CONFERENCE

Where Analytics Meets Innovation

Introducing Exensio® Aurora — AI-first, highly scalable architecture for Exensio Analytics.

At PDF Solutions CONNECT, we’ll introduce Exensio Aurora, a purpose-built solution architecture designed to handle semiconductor manufacturing data at petabyte scale

… Read More