
Hammerspace Breaks IO500 Barriers: How They Built the Fastest NFS-Based Benchmark Ever w/ Jon Flynn
About this episode
In this landmark 100th episode of Data Unchained, host Molly Presley sits down with Jonathan Flynn, Director of Applied Systems at Hammerspace, live from Supercomputing 2025. Together they explore the performance engineering breakthroughs that enabled Hammerspace and Samsung to deliver a historic IO500 10 Node Production result using only standard Linux, the upstream NFSv4.2 client, and off the shelf NVMe hardware.
This episode breaks down how the Hammerspace Data Platform delivered more than a 33 percent gain over earlier submissions, doubled overall bandwidth, and achieved an unprecedented 809 percent improvement in the IO Hard Read test using Samsung PM1753 Gen 5 NVMe SSDs. Jonathan explains the Linux kernel innovations, metadata advancements, IO path optimization, parallel file system breakthroughs, and multi instance file placement strategies that allowed Hammerspace to reach genuine HPC class performance without proprietary clients or custom networking.
Listeners get a detailed walkthrough of the architectural differences between Research and Production IO500 submissions, the impact of metadata redundancy, the performance benefits of NFSd direct and NFS direct, the role of ZFS locking improvements, and how upstream Linux contributions directly advanced the state of HPC and AI data infrastructure. Jonathan also highlights the evolution of MLPerf benchmarking, the benefits of tier zero storage, and how Hammerspace performance engineering is unlocking new levels of efficiency and scalability for AI training, scientific workloads, and large scale analytics.
This episode is essential for AI architects, HPC engineers, kernel developers, data scientists, and infrastructure leaders building the next generation of high performance data platforms.
Cyberpunk by jiglr | https://soundcloud.com/jiglrmusic
Music promoted by https://www.free-stock-music.com
Creative Commons Attribution 3.0 Unported License
https://creativecommons.org/licenses/by/3.0/deed.en_US
Hosted on Acast. See acast.com/privacy for more information.
Get every episode summarized
Each time Data Unchained publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.
Email me new episodesFree for 3 shows. No card needed.
Hosts & guests
No transcript yet
This episode has not been transcribed. Request it and it moves to the front of the queue.
More episodes
More from Data Unchained

Global Namespace for Precision Medicine: The Data Breakthrough w/ William Baird
Data Unchained

Understanding AI Readiness Before Implementation w/ Tim Gasper & Juan Sequeda
Data Unchained

Data Storage Shortage: What Options Do You Have? w/ Chris Mellor
Data Unchained

Inside the SSD Shortage Crisis w/ Tom Coughlin
Data Unchained