metadata
language:
- en
tags:
- pytorch
- causal-lm
- pythia
license: apache-2.0
datasets:
- Anthropic/hh-rlhf
Pythia-160m supervised finetuned with Anthropic-hh-rlhf dataset for 1 epoch.
Benchmark evaluations included in repo done using lm-evaluation-harness.
See Pythia-160m for model details (paper).