Matt Crane
b4dc87d46f
Explicitly open files with utf-8 encoding ( #35 )
...
Same issue as castorini/data#19
2017-08-29 14:20:18 -04:00
Michael Tu
a6fc10818f
Fix SM Model Internal Reproducibility Bug ( #34 )
...
* Fix SM Model reproducibility bug
vocab is in different order every time, causing unseen words to use
different random states
* Make requirements.txt usable from conda and pip
The existing torch requirement does not work with conda or pip.
Also upgrade pytorch version while we are at it.
2017-07-20 06:52:26 +08:00
Michael Tu
a0755e6aa3
SM Model Jupyter Notebook Tutorial ( #32 )
...
Notebook tutorial for SM CNN.
2017-07-04 18:28:17 -04:00
Salman Mohammed
4c081645a7
clean up the relation prediction model for Simple QA - Ferhan's paper ( #28 )
...
+ cleaned up the model code for the simple qa directory
+ created vocab objects for pre-loading word embeddings easily
2017-06-19 14:24:27 -04:00
Matt Crane
945b1fa6c0
Fix vocab caching issue ( #29 )
...
With the line as-was the vocab cache was stored as b'the' rather than the, meaning that word2vec wasn't found for terms causing massive performance loss (AP 0.71 cf 0.77).
2017-06-19 13:02:36 -04:00
Michael Tu
43dd6fdb1f
GPU Support for SM Model ( #26 )
...
Add code to using GPU for the SM Model (#25 ). To use the GPU if one is available, add the --cuda optional parameter when calling main.py.
2017-06-02 08:47:13 -04:00
Gaurav Baruah
92789cb9f5
E2e sweep ( #24 )
...
Now ensuring that the bridge process raw candidate sentences fetched from the index, exactly as was done for the best performing SM model.
2017-05-29 19:47:42 -04:00
gauravbaruah
9bd5b7bb2a
Corpus idf ( #23 )
...
as part of sourcing-IDF-from-index and e2e experiments.
2017-05-07 22:28:35 -04:00
rosequ
212aa6fb93
renaming sm_model; faster bridge ( #18 ) ( #22 )
...
+ renamed sm_model
+ faster bridge by obtaining the IDF scores of a term directly from the Java server
2017-05-03 17:24:47 -04:00