Published April 24, 2026 | Version v1

Evaluation of LLM-geneated boolean search queries

Authors/Creators

Description

prompts.csv contains the prompts and respective keyword for embedding.
The first row is the name of the corresponding file in ordered_items_per_prompt/

ordered_items_per_prompt/ contains one file per prompt.
Each of those files has the IDs of all items in the observatory, ordered by relevance.
The second column is the distance of the embedding of the keyword and the embedding of the item.
The third column indicates whether the item is relevant ('y') or irrelevant ('n' or blank).

For each (prompt,model) combination, a file is created in results_rewrites/
It contains the generated query for that (prompt,model).

For each (prompt,model) combination, a file is created in results_items/
It contains the items that match the generated query in the observatory.

results/metrics contains the macro-metrics for each prompt.
The columns are: model, precision, recall, latency.

Files

test_set.zip

Files (338.9 kB)

Name Size Download all
md5:1bd295aba3ea68903c2ee50e55b620c9
338.9 kB Preview Download