How to use from the
Use from the
Transformers library
# Use a pipeline as a high-level helper
from transformers import pipeline

pipe = pipeline("text-classification", model="jackhhao/jailbreak-classifier")
# Load model directly
from transformers import AutoTokenizer, AutoModelForSequenceClassification

tokenizer = AutoTokenizer.from_pretrained("jackhhao/jailbreak-classifier")
model = AutoModelForSequenceClassification.from_pretrained("jackhhao/jailbreak-classifier")
Quick Links

Jailbreak Classifier

Classifies prompts as jailbreaks or benign. This is a fine-tune checkpoint of bert-base-uncased on the jailbreak-classification dataset.

Training Details

Training Data

Fine-tuned on the jailbreak-classification dataset.

Training Procedure

Training Hyperparameters

Fine-tuning hyper-parameters:

  • learning_rate = 5e-5
  • train_batch_size = 8
  • eval_batch_size = 8
  • lr_scheduler_type = linear
  • num_train_epochs = 5.0
Downloads last month
4,694
Inference Providers NEW

Datasets used to train jackhhao/jailbreak-classifier

Space using jackhhao/jailbreak-classifier 1