Process Reward Models (PRMs) trained on step-level error labels automatically annotated by formal verification tools.
Ryo Kamoi PRO
ryokamoi
AI & ML interests
NLP
Organizations
VisOnlyQA
Dataset for evaluating the visual perception capabilities of LVLMs.
-
VisOnlyQA: Large Vision Language Models Still Struggle with Visual Perception of Geometric Information
Paper • 2412.00947 • Published • 8 -
ryokamoi/VisOnlyQA_Eval_Real_v1.1
Viewer • Updated • 900 • 207 -
ryokamoi/VisOnlyQA_Eval_Synthetic
Viewer • Updated • 700 • 126 • 2 -
ryokamoi/VisOnlyQA_Train
Viewer • Updated • 70k • 590 • 2
FoVer
Process Reward Models (PRMs) trained on step-level error labels automatically annotated by formal verification tools.
VisOnlyQA
Dataset for evaluating the visual perception capabilities of LVLMs.
-
VisOnlyQA: Large Vision Language Models Still Struggle with Visual Perception of Geometric Information
Paper • 2412.00947 • Published • 8 -
ryokamoi/VisOnlyQA_Eval_Real_v1.1
Viewer • Updated • 900 • 207 -
ryokamoi/VisOnlyQA_Eval_Synthetic
Viewer • Updated • 700 • 126 • 2 -
ryokamoi/VisOnlyQA_Train
Viewer • Updated • 70k • 590 • 2
datasets 19
ryokamoi/FoVer-misc
Updated • 116
ryokamoi/FoVer-FormalLogic-Llama-3.1-8B
Viewer • Updated • 10.7k • 39
ryokamoi/FoVer-FormalLogic-Qwen-2.5-7B
Viewer • Updated • 10.7k • 45
ryokamoi/FoVer-FormalProof-Llama-3.1-8B
Viewer • Updated • 10.7k • 44
ryokamoi/FoVer-FormalProof-Qwen-2.5-7B
Viewer • Updated • 10.7k • 48
ryokamoi/FoVer-FormalLogic-FormalProof-Llama-3.1-8B-LastStepBalanced-40k
Viewer • Updated • 40k • 47
ryokamoi/FoVer-FormalLogic-FormalProof-Qwen-2.5-7B-LastStepBalanced-40k
Viewer • Updated • 40k • 93
ryokamoi/VisOnlyQA_Eval_Real_v1.1
Viewer • Updated • 900 • 207
ryokamoi/VisOnlyQA_Eval_Synthetic
Viewer • Updated • 700 • 126 • 2
ryokamoi/VisOnlyQA_metadata
Viewer • Updated • 3 • 85