haiku.rag/evaluations/evaluations
2026-04-29 12:24:06 +03:00
..
datasets remove dataset-specific system prompts 2026-04-28 14:33:25 +03:00
evaluators citation retrieval scoring 2026-04-28 12:46:16 +03:00
__init__.py Restructure into uv workspace to support minimal and full installations 2025-11-04 17:59:12 +02:00
benchmark.py pin judge model to ollama:qwen3.6 2026-04-29 12:24:06 +03:00
config.py remove dataset-specific system prompts 2026-04-28 14:33:25 +03:00
optimization.py remove dataset-specific system prompts 2026-04-28 14:33:25 +03:00
skill_runner.py benchmark RAG and analysis skills via --target 2026-04-28 12:09:19 +03:00