
technologyJul 6, 202559:41pending
TornadoVM: The Need for GPU Speed
About this episode
An airhacks.fm conversation with Michalis Papadimitriou (@mikepapadim) about:
starting with Java 8, first computer experiences with Pentium 2, doom 2 and Microsoft Paint, university introduction to Object-oriented programming using Objects First and bluej IDE, Monte Carlo simulations for financial portfolio optimization in Java, porting Java applications to OpenCL for GPU acceleration achieving 20x speedup, working at Huawei on GPU hardware, writing unit tests as introduction to TornadoVM, working on FPGA integration and Graal compiler optimizations, experience at OctoAI startup doing AI compiler optimizations for TensorFlow and PyTorch models, understanding model formats evolution from ONNX to GGUF, standardization of LLM inference through Llama models, implementing GPU-accelerated Llama 3 inference in pure Java using TornadoVM, achieving 3-6x speedup over CPU implementations, supporting multiple models including Mistral and working on qwen 3 and deepseek, differences between models mainly in normalization layers, GGUF becoming quasi-standard for LLM model distribution, TornadoVM's Consume and Persist API for optimizing GPU data transfers, challenges with OpenCL deprecation on macOS and plans for Metal backend, importance of developer experience and avoiding python dependencies for Java projects, runtime and compiler optimizations for GPU inference, kernel fusion techniques, upcoming integration with langchain4j, potential of Java ecosystem with Graal VM and Project Panama FFM for high-performance inference, advantages of Java's multi-threading capabilities for inference workloads
Michalis Papadimitriou on twitter: @mikepapadim
Get every episode summarized
Each time airhacks.fm podcast with adam bien publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.
Email me new episodesFree for 3 shows. No card needed.
Hosts & guests
No transcript yet
This episode has not been transcribed. Request it and it moves to the front of the queue.
More episodes
More from airhacks.fm podcast with adam bien

From Manchester to Mountain View: Binary Translators, JVMs, and Android
airhacks.fm podcast with adam bien
May 8, 20261:05:26pending

Migrating Ruby Monoliths to Java, Agentic AI Foundation and MCP
airhacks.fm podcast with adam bien
Apr 28, 202659:22pending

Apache PLC4X, Industrial Protocol Drivers, and the JDBC of Industrial Automation
airhacks.fm podcast with adam bien
Apr 22, 202656:50failed

Green Java with Quarkus: Performance Benchmarks, SBOM, and Serverless Architectu...
airhacks.fm podcast with adam bien
Apr 11, 20261:08:34failed