evaluation-task · Source-linked discovery

AIR Bench: AI Risk Benchmark

A safety benchmark evaluating language models against risk categories derived from government regulations and company policies.

Open interactive record →
OriginYi Zeng, Yu Yang, Andy Zhou et al.TopicsSafeguardsStatusimported

Can support

Not independently assessed by FronteraEval yet.

Cannot support by itself

No inference beyond the upstream source should be made until the protocol is reviewed.

Original sources