Scale AI
company
Verified
AI & ML interests
None defined yet.
Recent Activity
Papers
ResearchRubrics: A Benchmark of Prompts and Rubrics For Evaluating Deep Research Agents
Chasing the Tail: Effective Rubric-based Reward Modeling for Large Language Model Post-Training
datasets
19
ScaleAI/SA2_bowlstack0
Viewer
•
Updated
•
200
•
173
ScaleAI/dummy_mcp
Viewer
•
Updated
•
16
•
47
ScaleAI/PRBench
Viewer
•
Updated
•
1.65k
•
928
•
5
ScaleAI/researchrubrics
Viewer
•
Updated
•
101
•
139
•
9
ScaleAI/swe-oec-claude-expert
Viewer
•
Updated
•
1.27k
•
62
•
1
ScaleAI/VisualToolBench
Viewer
•
Updated
•
1.19k
•
119
•
1
ScaleAI/TutorBench
Viewer
•
Updated
•
1.47k
•
235
ScaleAI/SWE-bench_Pro
Viewer
•
Updated
•
731
•
14.4k
•
39
ScaleAI/BioRiskEval
Viewer
•
Updated
•
156k
•
59
ScaleAI/TutorBench_sample
Viewer
•
Updated
•
30
•
26