Model save
Browse files
README.md
CHANGED
@@ -34,10 +34,10 @@ This model was trained with SFT.
|
|
34 |
|
35 |
### Framework versions
|
36 |
|
37 |
-
- TRL: 0.
|
38 |
-
- Transformers: 4.
|
39 |
- Pytorch: 2.6.0+cu124
|
40 |
-
- Datasets: 3.
|
41 |
- Tokenizers: 0.21.1
|
42 |
|
43 |
## Citations
|
@@ -49,7 +49,7 @@ Cite TRL as:
|
|
49 |
```bibtex
|
50 |
@misc{vonwerra2022trl,
|
51 |
title = {{TRL: Transformer Reinforcement Learning}},
|
52 |
-
author = {Leandro von Werra and Younes Belkada and Lewis Tunstall and Edward Beeching and Tristan Thrush and Nathan Lambert and Shengyi Huang and Kashif Rasul and Quentin
|
53 |
year = 2020,
|
54 |
journal = {GitHub repository},
|
55 |
publisher = {GitHub},
|
|
|
34 |
|
35 |
### Framework versions
|
36 |
|
37 |
+
- TRL: 0.18.1
|
38 |
+
- Transformers: 4.50.0.dev0
|
39 |
- Pytorch: 2.6.0+cu124
|
40 |
+
- Datasets: 3.6.0
|
41 |
- Tokenizers: 0.21.1
|
42 |
|
43 |
## Citations
|
|
|
49 |
```bibtex
|
50 |
@misc{vonwerra2022trl,
|
51 |
title = {{TRL: Transformer Reinforcement Learning}},
|
52 |
+
author = {Leandro von Werra and Younes Belkada and Lewis Tunstall and Edward Beeching and Tristan Thrush and Nathan Lambert and Shengyi Huang and Kashif Rasul and Quentin Gallou{\'e}dec},
|
53 |
year = 2020,
|
54 |
journal = {GitHub repository},
|
55 |
publisher = {GitHub},
|