Regulation, the Harness, and RL Steering of LLMs: A Research View
Why Amodei-style AI regulation could handicap American AI, why the software harness is the new layer above LLMs, and the formal landscape of RL steering and fine-tuning, from RLHF and PPO to DPO and GRPO.