#reasoning-models
Wiki 2
- Controllable Thinking Style Training-injected markers that toggle the style of a reasoning model's chain-of-thought without changing the answer pipeline
- Stealing Reasoning Traces from Proprietary LLM APIs Encrypted chain-of-thought blocks are replayable across sessions, users and models, so a cheap sibling model will decode a frontier model's hidden reasoning