for anyone digging into what it takes to run M3 with blazing-fast inference, the kernel work from our team and @FireworksAI_HQ is now open:
MiniMax MSA: https://github.com/MiniMax-AI/MSA/tree/fireworks-msa
Fireworks kernels: https://github.com/fw-ai/minimax-kernels