haiku.rag/evaluations/evaluations
2026-05-18 16:49:43 +03:00
..
datasets split open_rag_bench dataset into orb_text and orb_multimodal variants 2026-05-06 12:49:03 +03:00
evaluators Bump pydantic-ai, prepare for 2.* 2026-05-18 15:14:29 +03:00
__init__.py Restructure into uv workspace to support minimal and full installations 2025-11-04 17:59:12 +02:00
benchmark.py analysis.model inherits qa.model when unset; per-skill vision gate 2026-05-18 16:49:43 +03:00
config.py remove dataset-specific system prompts 2026-04-28 14:33:25 +03:00
optimization.py remove dataset-specific system prompts 2026-04-28 14:33:25 +03:00
skill_runner.py benchmark RAG and analysis skills via --target 2026-04-28 12:09:19 +03:00