This website requires JavaScript.
Explore
Help
Sign in
danvics
/
haiku.rag
Watch
1
Star
0
Fork
You've already forked haiku.rag
0
Code
Issues
Pull requests
Projects
Releases
Packages
Wiki
Activity
Actions
f843fb8140
haiku.rag
/
evaluations
/
evaluations
History
Yiorgis Gozadinos
f843fb8140
Use non-thinking judge in evals
2025-11-25 13:00:25 +02:00
..
datasets
Use Mean Reciprocal Rank for single document evaluation as metric. Use Mean Average Precision for variable document evaluation as metric
2025-11-24 14:49:30 +02:00
evaluators
Add support for per-model configuration settings including thinking, temperature and max_tokens
2025-11-25 12:14:24 +02:00
__init__.py
Restructure into uv workspace to support minimal and full installations
2025-11-04 17:59:12 +02:00
benchmark.py
Use non-thinking judge in evals
2025-11-25 13:00:25 +02:00
config.py
Default eval db location, evaluation script
2025-11-25 11:07:20 +02:00
prompts.py
Restructure into uv workspace to support minimal and full installations
2025-11-04 17:59:12 +02:00