Taalas buried the lede for the amazing demo of their first tape out.
15k tokens per second with a llama 8b model, ability to scale that up etched onto silicon
As models satisfice etching makes sense, particularly ternary..
Try it out https://chatjimmy.ai
Bullish for $AMD