Why so dense?
The dense models code well on my M1 mac but they are horrendously slow. The Mixture of Experts I find pretty close for coding but have much more acceptable speed on Mac hardware. One big advantage of PI is the context by default is pretty small - in the hundreds of tokens. Even Opencode is ten thousand+ tokens before you've even typed anything and that pre-fill speed although better on M4/M5 is still going to be painful once you get a decent sized context.