Hugging Face
Models
Datasets
Spaces
Posts
Docs
Solutions
Pricing
Log In
Sign Up
FAR AI
non-profit
https://far.ai/
FARAIResearch
AlignmentResearch
Request to join this org
Follow
14
AI & ML interests
Frontier alignment research to ensure the safe development and deployment of advanced AI systems.
Team members
9
spaces
1
Runtime error
23
🔎
Tuned Lens
models
3406
Sort: Recently updated
AlignmentResearch/clf_helpful_pythia-31m_s-1_adv_tr_gcg_t-1
Updated
3 days ago
•
148
AlignmentResearch/clf_helpful_pythia-31m_s-0_adv_tr_gcg_t-0
Updated
3 days ago
•
146
AlignmentResearch/clf_helpful_pythia-2.8b_s-4_adv_tr_gcg_t-4
Updated
3 days ago
•
12
AlignmentResearch/clf_helpful_pythia-2.8b_s-3_adv_tr_gcg_t-3
Updated
3 days ago
•
12
AlignmentResearch/clf_helpful_pythia-2.8b_s-2_adv_tr_gcg_t-2
Updated
3 days ago
•
8
AlignmentResearch/clf_helpful_pythia-2.8b_s-1_adv_tr_gcg_t-1
Updated
3 days ago
•
12
AlignmentResearch/clf_helpful_pythia-2.8b_s-0_adv_tr_gcg_t-0
Updated
3 days ago
•
16
AlignmentResearch/clf_helpful_pythia-1b_s-4_adv_tr_gcg_t-4
Updated
3 days ago
•
18
AlignmentResearch/clf_helpful_pythia-1b_s-3_adv_tr_gcg_t-3
Updated
3 days ago
•
16
AlignmentResearch/clf_helpful_pythia-1b_s-2_adv_tr_gcg_t-2
Updated
3 days ago
•
50
Expand 3406 models
datasets
14
Sort: Recently updated
AlignmentResearch/WordLength
Viewer
•
Updated
Aug 7
•
100k
•
3.15k
AlignmentResearch/Harmless
Viewer
•
Updated
Jul 29
•
86.6k
•
927
AlignmentResearch/Helpful
Viewer
•
Updated
Jul 29
•
88.1k
•
2.21k
AlignmentResearch/StrongREJECT
Viewer
•
Updated
Jul 29
•
313
•
658
AlignmentResearch/PasswordMatch
Viewer
•
Updated
Jul 29
•
100k
•
4.42k
AlignmentResearch/IMDB
Viewer
•
Updated
Jul 29
•
97.5k
•
3.89k
AlignmentResearch/EnronSpam
Viewer
•
Updated
Jul 29
•
62.3k
•
732
AlignmentResearch/PasswordMatch-test
Viewer
•
Updated
Jul 26
•
50k
•
47
AlignmentResearch/WordLength-test
Viewer
•
Updated
Jul 26
•
100k
•
69
AlignmentResearch/StrongREJECT-test
Viewer
•
Updated
Jul 26
•
313
•
38
Expand 14 datasets