Hugging Face
Models
Datasets
Spaces
Posts
Docs
Solutions
Pricing
Log In
Sign Up
PKU-Alignment
/
beaver-7b-v1.0-cost
like
9
Follow
PKU-Alignment
33
Reinforcement Learning
Safetensors
PKU-Alignment/PKU-SafeRLHF
English
safe-rlhf
llama
reinforcement-learning-from-human-feedback
beaver
safety
ai-safety
deepspeed
rlhf
alpaca
arxiv:
2302.13971
arxiv:
2307.04657
arxiv:
2310.12773
Model card
Files
Files and versions
Community
1
Train
1070fa3
beaver-7b-v1.0-cost
/
config.json
Commit History
Convert model checkpoint to safetensors
1070fa3
XuehaiPan
commited on
Apr 19
Update architecture name in config.json
42e2cbe
XuehaiPan
commited on
Dec 15, 2023
hello beaver cost model
0e42156
RuiyangSun
commited on
Jul 10, 2023