Bart: update example for #3140 compatibility (#3233)

* Update bart example docs
This commit is contained in:
Sam Shleifer
2020-03-12 10:36:37 -04:00
committed by GitHub
parent 72768b6b9c
commit 2e81b9d8d7
3 changed files with 22 additions and 4 deletions

View File

@@ -1,4 +1,4 @@
### Get the CNN/Daily Mail Data
### Get the CNN Data
To be able to reproduce the authors' results on the CNN/Daily Mail dataset you first need to download both CNN and Daily Mail datasets [from Kyunghyun Cho's website](https://cs.nyu.edu/~kcho/DMQA/) (the links next to "Stories") in the same folder. Then uncompress the archives by running:
```bash
@@ -32,6 +32,7 @@ unzip stanford-corenlp-full-2018-10-05.zip
cd stanford-corenlp-full-2018-10-05
export CLASSPATH=stanford-corenlp-3.9.2.jar:stanford-corenlp-3.9.2-models.jar
```
Then run `ptb_tokenize` on `test.target` and your generated hypotheses.
### Rouge Setup
Install `files2rouge` following the instructions at [here](https://github.com/pltrdy/files2rouge).
I also needed to run `sudo apt-get install libxml-parser-perl`