haiku.rag/evaluations/evaluations/datasets
2026-01-24 13:48:09 +02:00
..
__init__.py Add evaluation dataset for multi-modal q&a 2026-01-22 14:17:09 +02:00
hotpotqa.py replace pyright with ty type checker 2026-01-19 15:50:14 +02:00
open_rag_bench.py Customize orb QA prompt to not use LaTeX as gpt-oss Ollama implementation fails to parse it properly 2026-01-24 13:48:09 +02:00
repliqa.py replace pyright with ty type checker 2026-01-19 15:50:14 +02:00
wix.py Customize orb QA prompt to not use LaTeX as gpt-oss Ollama implementation fails to parse it properly 2026-01-24 13:48:09 +02:00