Last night, in a basement in Florida, one conversation ran across two kinds of silicon…
This is a dev post classified by Jev as Hosting & infra (a tool drop), kept by the Dev Radar because it carries real work, not commentary.
Last night, in a basement in Florida, one conversation ran across two kinds of silicon at once. Two NVIDIA DGX Sparks read the prompt. The finished cache crossed an RDMA link straight into a Mac Studio's memory. Apple silicon wrote the answer. No file in the middle, no re-reading, one model, one thought — GLM-5.3-Flash with ~380,000 tokens of context in view. That hop should not exist. CUDA and Metal were never meant to touch each other's memory. @ashhart built the bridge that makes them: MCDMA — Metal CUDA Direct Memory Access. Kernel-level RDMA on macOS, written by one person, open source.
Posted by Volatile Markets (1.5k followers) 2 h ago · 16 likes · 424 views · view the original post on X. Kept by the Dev Radar as Hosting & infra. Tools mentioned: mcdma.
More dev work like this
- I have to retract what I posted as the fastest speed of Qwen 3.8 Flash on a single DGX… — @yume_arasaki
- How does a global bank navigate regulations and rapid AI cycles without vendor lock-in?… — @RedHat
- Level up your skills while shaping open source priorities! — @CloudNativeFdn
- 🤙 DeepSeek V4.1 Flash. > 100 t/s reached. max 178 t/s stream decode at b1. — @qubitium
- Azure AI Engineer Roadmap Explained! — @AiswaryaVenkit1
- 文科生有救了,妈妈再也不怕我的localhost:8888别人访问不了了! — @mylifcc
- Kubernetes Node Failure & Recovery in action 👇 — @twtayaan
- 一个 3.3 万 Star 的 Rust 项目,刚刚发布 1.0。 — @mylifcc
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 17k posts from 4.9k X accounts over the last 21 days, 1.9k tools, 12 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-20 22:16 UTC. Full method.