Noticeable weird contradiction: models are getting better at following complex instructions, but also using more "judgement" about which parts of the instructions they focus on most & which they de-emphasize.
Big implications for skills, which may become suggestions, not orders.