Use Jailbreaking to reverse the CoT process of ChatGPT o1-preview

Background Recently, OpenAI announced gpt-o1-preview and there are some interesting updates, including safety improvements regarding “CoT Safety Alignment”. The o1 model family represents a transition from fast, intuitive thinking to now also using slower, more deliberate reasoning. What’s Chain-of-Thought Safety Similar to how a human may think for a long time before responding to a […]