The long-context fix hiding in plain sight: sliding-window attention beats post-trained 'linear-attention' by 2–10× — type0 | type0