The SC #125

Issue #125 of The SC. Weekly supercomputing news. Microsoft goes all in on its AI chips, Nvidia gives us an LLM router, Google extends AlphaFold as an HPC product and AWS gives us K8 inspired EC2 extensions.

The SC #125

You’d be forgiven for thinking that if you exclude all the mainstream AI news, especially everything that’s more about financing and power grids that nothing much of note happened last week. Even my AI research bots turned up little to write about. A little old fashioned manual digging in release notes did turn up a couple of interesting stories though.

We have word from Microsoft that that they will not only be releasing an update to their own AI chips (Maia) but also rolling out a surprisingly aggressive number of them. Treat that statement with a pinch of salt but it does rather confirm the very heterogenous nature of compute. Yes, present tense. Now.

Nvidia gave us an open source LLM router which I suppose could be interesting in some domains of HPC to try and control cost where we have converged LLM inference and classical compute. I’m not sure what the workflow for that looks like though. I guess you need the routing decision before the scheduling decision else you can’t allocate resources correctly. I wonder how long it is before we see an LLM router integrated into a HPC scheduler. For LLM inference only workloads this seems like a bit of a no brainer surely? 

How is it possible that even with AI genies writing code at a rate that eclipses my fat fingers on a keyboard that I still have more possible ideas than engineering capacity. 

Riddle me that batman. 

Talking of converged AI and HPC compute, Google gives us AlphaFold for HPC as an addition to their catalogue of offerings. Looks like an out of the box way to run AlphaFold on Google Batch to reduce your compute wall time.

AWS also gave us Kubernetes style health checks for the application layer in EC2 (see the release notes in Noteworthy). Please let this be another nail in the coffin of HPC on K8.

Lastly, we welcomed Eduardo to the team last week and I had a bit of a rant about synthetic and natural intelligence.


All the News in Depth

Cloud and vendor releases for Supercomputing & HPC (and AI too if you change the filters)

Microsoft is still investing in its own Nvidia alternatives for AI focused compute with the next generation of Maia processor due for release imminently.

Nvidia provides an open source LLM router

Google adds to their converged AI/HPC capability with AlphaEvolve for HPC


HMx Labs Updates

Not sure if this is a rant, man shouting at clouds or just incoherent mutterings of a mad man. I can say none of it was written, edited or illustrated by AI. I think I had a few interesting things to say but do feel free to disagree

AI: Context, Taste & Inspiration
Somewhere between a rant and some random thoughts on AI.

Also, welcome to Eduardo!

Eduardo Joins HMx Labs
Welcome to Eduardo Visinheski who joins us at HMx Labs as our newest HPC Engineer.

Off Topic

AI progress is accelerating on an exponential curve right? Right? Maybe not?

Is the Dunning Kruger effect just nonsense? I don’t know, but I can say I agree with Sabine’s last point, that data cloud doesn’t look like it has any kind of correlation at all!


Know someone else who might like to read this newsletter? Forward this on to them or even better, ask them to sign up here: https://cloudhpc.news