haiku.rag/evaluations/evaluations
2025-11-21 13:26:39 +02:00
..
datasets Restructure into uv workspace to support minimal and full installations 2025-11-04 17:59:12 +02:00
__init__.py Restructure into uv workspace to support minimal and full installations 2025-11-04 17:59:12 +02:00
benchmark.py Name evaluation runs. Use only LLMJudge evaluator. 2025-11-21 13:26:39 +02:00
config.py Restructure into uv workspace to support minimal and full installations 2025-11-04 17:59:12 +02:00
llm_judge.py Restructure into uv workspace to support minimal and full installations 2025-11-04 17:59:12 +02:00
prompts.py Restructure into uv workspace to support minimal and full installations 2025-11-04 17:59:12 +02:00