mirror of https://github.com/deepset-ai/haystack.git synced 2025-08-28 18:36:36 +00:00

History

* initial commit

* Add latest docstring and tutorial changes

* added comments and fixed bug

* fixed bugs, added benchmark and added documentation

* Add latest docstring and tutorial changes

* fix type: ignore comment

* fix logging in benchmark

* fixed distillation config

* Add latest docstring and tutorial changes

* added type annotations

* fixed distillation loss calculation

* added type annotations

* fixed distillation mse loss

* improved model distillation benchmark config loading

* added temperature for model distillation

* removed uncessary imports, added comments, added named parameter calls

* Add latest docstring and tutorial changes

* added some more comments

* added distillation test

* fixed distillation test

* removed unnecessary import

* fix softmax dimension

* add grid search

* improved model distillation benchmark config

* fixed model distillation hyperparameter search

* added doc strings and type hints for model distillation

* Add latest docstring and tutorial changes

* fixed type hints

* fixed type hints

* fixed type hints

* wrote out params instead of kwargs in DistillationDataSilo initializer

* fixed type hints

* fixed typo

* fixed typo

Co-authored-by: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>

2021-11-26 18:49:30 +01:00

data_scripts

Create time and performance benchmarks for all readers and retrievers (#339 )

2020-10-12 13:34:42 +02:00

config.json

Implement OpenSearch ANN (#1225 )

2021-07-26 10:52:52 +02:00

distillation_config.json

Model Distillation (#1758 )

2021-11-26 18:49:30 +01:00

model_distillation.py

Model Distillation (#1758 )

2021-11-26 18:49:30 +01:00

nq_to_squad.py

Create time and performance benchmarks for all readers and retrievers (#339 )

2020-10-12 13:34:42 +02:00

reader_results.csv

Choose correct similarity fns during benchmark runs & re-run benchmarks (#773 )

2021-02-03 11:45:18 +01:00

reader.py

Refactoring of the haystack package (#1624 )

2021-10-25 15:50:23 +02:00

README.md

add readme

2020-10-22 15:32:56 +02:00

results_to_json.py

Implement OpenSearch ANN (#1225 )

2021-07-26 10:52:52 +02:00

retriever_index_results.csv

Implement OpenSearch ANN (#1225 )

2021-07-26 10:52:52 +02:00

retriever_query_results.csv

Implement OpenSearch ANN (#1225 )

2021-07-26 10:52:52 +02:00

retriever_query_results.md

Choose correct similarity fns during benchmark runs & re-run benchmarks (#773 )

2021-02-03 11:45:18 +01:00

retriever_simplified.py

Refactoring of the haystack package (#1624 )

2021-10-25 15:50:23 +02:00

retriever.py

Refactoring of the haystack package (#1624 )

2021-10-25 15:50:23 +02:00

run.py

Integrate sentence transformers into benchmarks (#843 )

2021-04-09 17:24:16 +02:00

templates.py

Update Milvus benchmarks (#1128 )

2021-06-02 13:09:45 +02:00

utils.py

Rename every occurrence of 'embed_passages' with 'embed_documents' (#1667 )

2021-10-28 12:17:56 +02:00

README.md

Benchmarks

Run the benchmarks with the following command:

python run.py [--reader] [--retriever_index] [--retriever_query] [--ci] [--update-json]

You can specify which components and processes to benchmark with the following flags.

--reader will trigger the speed and accuracy benchmarks for the reader. Here we simply use the SQuAD dev set.

--retriever_index will trigger indexing benchmarks

--retriever_query will trigger querying benchmarks (embeddings will be loaded from file instead of being computed on the fly)

--ci will cause the the benchmarks to run on a smaller slice of each dataset and a smaller subset of Retriever / Reader / DocStores.

--update-json will cause the script to update the json files in docs/_src/benchmarks so that the website benchmarks will be updated.