QontoFAQ: A better Information Retrieval Benchmark [R]
Mirrored from r/MachineLearning for archival readability. Support the source by reading on the original site.
Retrieval benchmarks sometimes feel benchmaxxed by models, so we wanted to find a way to tie it as close as possible to my objective: finding the article that answers a product question right.
We worked on a new metric which seems more proportional to document relevance, and built up a benchmarking dataset to measure embedding models.
Here is an article on the approach: https://medium.com/qonto-way/qontofaq-benchmarking-information-retrieval-acd89600ebe1
and the associated code: https://github.com/qonto/qonto-faq-benchmark
[link] [comments]
More from r/MachineLearning
-
Simulating fault tolerance with stage skipping in pipeline-parallel training [R]
Sep 22
-
LinearSolveBench: new benchmark for linear solvers [P]
Sep 22
-
Understanding and Enhancing Kimi Delta Attention [R]
Sep 22
-
Xiaomi releases MiMo-V2.6: "Frontier intelligence, all the modalities, built in public." [N]
Sep 22
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.