> I don't know what we want to call this, but in my opinion, having to convince your tools is not computer science.
Bit of a mouthful, but how about just calling it "auto-regressive language modelling".
Feeding it stuff to auto-regress on is obviously your main control vector.
Apparently RL-trained models like rewards too. PHB's can use "you've gotta work all weekend, but you'll get comp time when it's fixed".