Support H-CoT: Hijacking the Chain-of-Thought to Jailbreak Reasoning Models
Offen
@riyosha arbeitet bereits daran.
Seit 15.2.2026.
enhancement
help wanted
- Vorherrschende Sprache
- Python
- Sterne
- 4.5k
- Forks
- 893
- Ø Merge
- 3 T. 50 Min.
- Gemergte PRs (30 T.)
- 165
Beschreibung
Is your feature request related to a problem? Please describe.
I recently learned about the jailbreak method “H-CoT: Hijacking the Chain-of-Thought Safety Reasoning Mechanism to Jailbreak Large Reasoning Models”, a method which has been shown to successfully bypass safety filters in several large reasoning models including OpenAI o1/o3, DeepSeek-R1, and Gemini 2.0 Flash Thinking. It would be great to implement this feature.
Refer:
- https://github.com/dukeceicenter/jailbreak-reasoning-openai-o1o3-deepseek-r1
- https://maliciouseducator.org/
Describe the solution you'd like
Could you provide explicit support or integration for the H-CoT jailbreak method within your repository?
Beitragsleitfaden
Für dieses Repository ist kein Beitragsleitfaden indexiert
Erste Schritte
- Lies das ganze Issue und danach den Beitragsleitfaden des Projekts.
- Schreib ins Issue, dass du es übernimmst — das erspart doppelte Arbeit.
- Forke das Repository und arbeite in einem Branch.
- Öffne einen Pull Request, der die Issue-Nummer nennt.
Bewertung
Dieses Issue wurde noch nicht bewertet.