Loading
From Micro-Kernels to Macro-Impact: How yasp.compile Sped Up IBM Granite 4.0 by up to 3x End-to-End Over torch.compile
AI-written GPU kernels have been winning on isolated operations for over a year, but those wins rarely survive the trip to a real model — Google’s AlphaEvolve moved Gemini training by roughly 1%. Benc …
yasp, Microsoft, and AMD: Achieving 2.9x Speedup for MiniGPT on Azure’s AMD Radeon™ PRO V710 GPU
Benchmarking AI inference on Azure, yasp’s agentic compiler delivered a 2.91× speedup for the MiniGPT Block on AMD’s Radeon PRO V710 GPU over torch.compile. This technical deep-dive breaks down exactl …
Real Intent releases Riven AI agent for chip sign-off
https://vikshy.com/stories/real-intent-releases-riven-ai-agent-for-chip-sign-off/
As AI clusters scale to tens of thousands of GPUs, bandwidth alone is no longer enough. Infrastructure teams need deep visibility into the health and performance of the fabric itself. Discover how Int …
The next bottleneck in AI infrastructure isn't bandwidth, it's knowing why a link failed at 3am. ZeroFlap (ZF) Optics and PILOT give operators a way to actually see the optical layer instead of guessi …
FROM AI COMPUTING TO 800V DC WHY SST VALIDATION MUST COME FIRST
AI is transforming power systems from supporting infrastructure into a critical foundation that directly impacts computing deployment efficiency. SSTs, connecting medium-voltage grids, 800V DC systems …
Why SoC Interconnects Have Outgrown the Bus: The NoC-Centric, CHI-Native Path Forward
If your team is architecting an AI accelerator, automotive ADAS, or hyperscaler SoC — the bus-to-NoC migration is no longer a question of if. It is a question of how and when.
Cache Coherency 101: Why It Matters in Multi-Core SoCs.
This one is a guide for SoC architects who are dealing with the coherency problem at scale. Not a light read — but if you're building compute, AI, or automotive SoCs today, these are the decisions tha …
Harnessing the Power of AMBA CHI: The Protocol Powering Modern Compute Infrastructure
CHI is how coherent CPU clusters participate in a shared memory system. This paper covers what it actually does — from the channel model to the direct-transfer features that define modern high-perform …
Coherent vs Non-Coherent Fabrics — Choosing the Right NoC for Your SoC.
Coherent vs Non-Coherent Fabrics — Choosing the Right NoC for Your SoC.
157 Results