- 🚀 Make LLMs run on spyre accelerator and CPUs (yes CPUs)
- 🔩 Bend vLLM into places it wasn't designed for
- 🏗️ Fix build systems across architectures
- 📦 Make multi-arch containers actually behave
- ⚡ CPU-only LLM inference pipelines
- 🏛️ s390x(cpu) and spyre support for modern ML stacks
- ☸️ Infra that works beyond a single machine
🔥 squeezing every drop of performance.
⚠️ making PyTorch do questionable things
☸️ running clean infra on Kubernetes



