Text Generation
Transformers
PyTorch
English
qwen3
reward_model
nvidia
conversational
text-generation-inference
yizhujiao commited on
Commit
dcf5e57
·
verified ·
1 Parent(s): 116ce2d

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +3 -0
README.md CHANGED
@@ -31,6 +31,9 @@ license_link: >-
31
 
32
  This model achieves **state-of-the-art performance** on the average score on three major reward modeling benchmarks (RewardBench, RM-Bench, and RMB) by addressing the "judgment diffusion" problem where models spread attention too thinly across evaluation criteria.
33
 
 
 
 
34
  ### Key Features
35
 
36
  - 🎯 **Adaptive Focus**: Dynamically selects 1-3 critical evaluation dimensions per instance
 
31
 
32
  This model achieves **state-of-the-art performance** on the average score on three major reward modeling benchmarks (RewardBench, RM-Bench, and RMB) by addressing the "judgment diffusion" problem where models spread attention too thinly across evaluation criteria.
33
 
34
+ Code Link: https://github.com/yzjiao/BR-RM
35
+
36
+
37
  ### Key Features
38
 
39
  - 🎯 **Adaptive Focus**: Dynamically selects 1-3 critical evaluation dimensions per instance