Yiorgis Gozadinos
|
00f2a40a60
|
Refresh benchmarks doc and remove unused eval datasets
|
2026-06-01 10:40:51 +03:00 |
|
Yiorgis Gozadinos
|
ce7201271f
|
split open_rag_bench dataset into orb_text and orb_multimodal variants
|
2026-05-06 12:49:03 +03:00 |
|
Yiorgis Gozadinos
|
5c3a864a0b
|
rename open_rag_bench eval DBs to text/multimodal variants
|
2026-05-05 12:55:06 +03:00 |
|
Yiorgis Gozadinos
|
7bfb818d77
|
remove dataset-specific system prompts
|
2026-04-28 14:33:25 +03:00 |
|
Yiorgis Gozadinos
|
6812435ba0
|
cleanup
|
2026-03-12 12:05:50 +02:00 |
|
Yiorgis Gozadinos
|
c7a9ad8583
|
Customize orb QA prompt to not use LaTeX as gpt-oss Ollama implementation fails to parse it properly
|
2026-01-24 13:48:09 +02:00 |
|
Yiorgis Gozadinos
|
0d3ac5142d
|
Fix answer in orb
|
2026-01-22 15:08:33 +02:00 |
|
Yiorgis Gozadinos
|
e6b239bd6c
|
Add evaluation dataset for multi-modal q&a
|
2026-01-22 14:17:09 +02:00 |
|
Yiorgis Gozadinos
|
5e55c0df54
|
replace pyright with ty type checker
|
2026-01-19 15:50:14 +02:00 |
|
Yiorgis Gozadinos
|
7526735059
|
hotpotqa adapter
|
2025-12-12 10:56:37 +02:00 |
|
Yiorgis Gozadinos
|
06b4f6a224
|
Add format parameter for text-to-DoclingDocument conversion
|
2025-12-08 15:56:21 +02:00 |
|
Yiorgis Gozadinos
|
b5892699a0
|
Use Mean Reciprocal Rank for single document evaluation as metric. Use Mean Average Precision for variable document evaluation as metric
|
2025-11-24 14:49:30 +02:00 |
|
Yiorgis Gozadinos
|
2f9c907031
|
Restructure into uv workspace to support minimal and full installations
|
2025-11-04 17:59:12 +02:00 |
|