haiku.rag/evaluations
2025-09-30 11:31:50 +03:00
..
datasets Refactor evaluations so that we can perform with multiple datasets.. 2025-09-30 11:31:50 +03:00
__init__.py Refactor evaluations so that we can perform with multiple datasets.. 2025-09-30 11:31:50 +03:00
benchmark.py When populating the eval db, check if chunks for the document have been created. Protects against interrupts 2025-09-30 11:31:50 +03:00
config.py Refactor evaluations so that we can perform with multiple datasets.. 2025-09-30 11:31:50 +03:00
llm_judge.py Refactor evaluations so that we can perform with multiple datasets.. 2025-09-30 11:31:50 +03:00