Skip to main content
buildradar
Sign in

RobustNLP/CipherChat

@RobustNLP

A framework to evaluate the generalization capability of safety alignment for LLMs

Stars
629
Forks
67
Language
Python
License
MIT
Last push
11 months ago
Pythonsecurityllmlarge-language-modelschatgptalignmentjailbreakgpt-4-0613

No related intel yet

This repo has not appeared in any of the sources the radar tracks. The collector runs on a schedule — check back once it covers this repo.