bryanchrist
/

MATHWELL

Model card Files Files and versions Community

bryanchrist commited on Nov 10, 2024

Commit

a4b46ae

·

verified ·

1 Parent(s): 7243ed0

Update README.md

Files changed (1) hide show

README.md +16 -6

README.md CHANGED Viewed

@@ -28,11 +28,21 @@ The following `bitsandbytes` quantization config was used during training:
 ## Citation
 ```bash
-@inproceedings{christ_mathwell_2024,
-	title = {{MATHWELL}: {Generating} {Educational} {Math} {Word} {Problems} {Using} {Teacher} {Annotations}},
-	url = {https://openreview.net/forum?id=jNsjlRfpk0},
-	booktitle = {The 2024 {Conference} on {Empirical} {Methods} in {Natural} {Language} {Processing}},
-	author = {Christ, Bryan R. and Kropko, Jonathan and Hartvigsen, Thomas},
-	year = {2024},
 }
 ```

 ## Citation
 ```bash
+@inproceedings{christ-etal-2024-mathwell,
+    title = "{MATHWELL}: Generating Educational Math Word Problems Using Teacher Annotations",
+    author = "Christ, Bryan R  and
+      Kropko, Jonathan  and
+      Hartvigsen, Thomas",
+    editor = "Al-Onaizan, Yaser  and
+      Bansal, Mohit  and
+      Chen, Yun-Nung",
+    booktitle = "Findings of the Association for Computational Linguistics: EMNLP 2024",
+    month = nov,
+    year = "2024",
+    address = "Miami, Florida, USA",
+    publisher = "Association for Computational Linguistics",
+    url = "https://aclanthology.org/2024.findings-emnlp.696",
+    pages = "11914--11938",
+    abstract = "Math word problems are critical K-8 educational tools, but writing them is time consuming and requires extensive expertise. To be educational, problems must be solvable, have accurate answers, and, most importantly, be educationally appropriate. We propose that language models have potential to support K-8 math education by automatically generating word problems. However, evaluating educational appropriateness is hard to quantify. We fill this gap by having teachers evaluate problems generated by LLMs, who find existing models and data often fail to be educationally appropriate. We then explore automatically generating *educational* word problems, ultimately using our expert annotations to finetune a 70B language model. Our model, MATHWELL, is the first K-8 word problem generator targeted at educational appropriateness. Further expert studies find MATHWELL generates problems far more solvable, accurate, and appropriate than public models. MATHWELL also matches GPT-4{'}s problem quality while attaining more appropriate reading levels for K-8 students and avoiding generating harmful questions.",
 }
 ```