AI Measurement Science
- 9 followers
- United States of America
- http://aimslab.stanford.edu/
- aimslab@cs.stanford.edu
Popular repositories Loading
-
-
fantastic-bugs
fantastic-bugs PublicFantastic Bugs and Where to Find Them in AI Benchmarks
-
safety-irt
safety-irt PublicA Measurement Analysis of Multilingual Safety Evaluation (Accepted to COLM 2026)
-
Repositories
Showing 10 of 14 repositories
-
- paiec_baseline Public
-
- benchmark-caliper Public
Uplifting Human Decision Making in AI Evaluation by Automating Benchmark Validity Analysis
-
- redteam-measurement Public
Long-form response matrices for adaptive AI red-teaming benchmarks (JailbreakBench, HarmBench, StrongREJECT, Do-Not-Answer), formatted to the aims-foundations/measurement-db schema.
People
This organization has no public members. You must be a member to see who’s a part of this organization.