AI & ML interests
None defined yet.
Recent Activity
View all activity
Papers
Alternating Reinforcement Learning for Rubric-Based Reward Modeling in Non-Verifiable LLM Post-Training
OpenRubrics: Towards Scalable Synthetic Rubric Generation for Reward Modeling and LLM Alignment
models 10
OpenRubrics/RubricRM-4B-Rubric
196k • Updated • 7
OpenRubrics/RubricRM-4B-Judge
196k • Updated • 11
OpenRubrics/RubricRM-4B-Rubric-v2
196k • Updated • 11
OpenRubrics/RubricRM-8B-Judge
308k • Updated • 16
OpenRubrics/RubricRM-8B-Rubric
308k • Updated • 19
OpenRubrics/RubricRM-8B-Judge-v2
308k • Updated • 230
OpenRubrics/RubricRM-8B-Rubric-v2
308k • Updated • 102
OpenRubrics/RubricRM-4B-Judge-v2
196k • Updated • 67 • 1
OpenRubrics/RubricARM-8B-Rubric
308k • Updated • 150 • 3
OpenRubrics/RubricARM-8B-Judge
308k • Updated • 215 • 3