microsoft / microsoft/PyRIT

Support H-CoT: Hijacking the Chain-of-Thought to Jailbreak Reasoning Models

Aperta
#897 7 commenti 0 reazioni 1 assegnatario Vedi su GitHub

@riyosha ci sta già lavorando.

Dal 15/2/2026.

enhancement help wanted
Lingua principale
Python
Stelle
4.5k
Fork
893
Merge medio
3g 50m
PR unite (30g)
165

Descrizione

## Is your feature request related to a problem? Please describe.

I recently learned about the jailbreak method “H-CoT: Hijacking the Chain-of-Thought Safety Reasoning Mechanism to Jailbreak Large Reasoning Models”, a method which has been shown to successfully bypass safety filters in several large reasoning models including OpenAI o1/o3, DeepSeek-R1, and Gemini 2.0 Flash Thinking. It would be great to implement this feature.

Refer:
1. https://github.com/dukeceicenter/jailbreak-reasoning-openai-o1o3-deepseek-r1
2. https://maliciouseducator.org/

## Describe the solution you'd like

Could you provide explicit support or integration for the H-CoT jailbreak method within your repository?

Guida per i contributori

Nessuna guida per i contributori indicizzata per questo repository

Come iniziare

  1. Leggi tutta la issue e poi la guida ai contributi del progetto.
  2. Commenta sulla issue per dire che te ne occupi tu — evita che due persone facciano lo stesso lavoro.
  3. Fai un fork del repository e lavora su un branch.
  4. Apri una pull request che faccia riferimento al numero della issue.

Valutazione

Questa issue non è ancora stata valutata.

Ricevi le nuove issue nella tua casella

Un breve riepilogo di issue GitHub adatte ai principianti.