A new infrastructure layer appears to be taking shape between hyperscalers and GPU-rental shops, built around Chinese open-weight models for buyers who care more about where inference happens than whether it tops a benchmark. The shift is not ideological. It is a stack-dependency move: when a workload leaves a US frontier API, it needs somewhere to land, and a thin tier of operators is now building exactly that landing.
Antimatter, a Hong Kong neo-cloud, is the cleanest case in SCMP's reporting on the launch. Queenie Chan, Antimatter's chief marketing officer, gives the sharper read than the CoreWeave-rival framing: "Many of our customers are looking for alternative solutions to move off from the frontier models." That word, "move off," is the pattern. Buyers are not defecting because open-weight models are smarter. They are routing around frontier APIs — cost control, jurisdiction, and the right to fine-tune are the reporter's read on the underlying motivations — not because open-weight models are inherently superior.
The mechanism, as Antimatter positions it, is repeatable. Anywhere a US frontier provider cannot comfortably serve — from Middle East sovereign clouds to European regulated industries to mainland China — a neo-cloud can sit in the gap, host an open-weight model, and bill per token without frontier markups. Capability still lags at the top of the benchmark curve, but the workloads priced on sovereignty and unit cost do not need the top of the curve.
The tell is whether the inquiry pipeline converts. Chan said demand is coming from the Middle East and Europe. If those conversations turn into paid capacity contracts inside a year, the inference stack has a new resident. If they do not, "neo-cloud" is a logo and a press release.
Reported by Sky for Type0, from Hong Kong firm bets on Chinese open-weight models to rival CoreWeave. Read the original: scmp.com