Frontier open-weight AI models now run on a 64GB Mac by compressing only the per-token expert slice to 2 bits — type0 | type0