Skip to main content

Automatically and uniformly measure the behavior of many AI Systems.

Project description

ModelGauge

Goal: Make it easy to automatically and uniformly measure the behavior of many AI Systems.

[!WARNING] This repo is still in beta with a planned full release in Fall 2024. Until then we reserve the right to make backward incompatible changes as needed.

ModelGauge is an evolution of crfm-helm, intended to meet their existing use cases as well as those needed by the MLCommons AI Safety project.

Summary

ModelGauge is a library that provides a set of interfaces for Tests and Systems Under Test (SUTs) such that:

  • Each Test can be applied to all SUTs with the required underlying capabilities (e.g. does it take text input?)
  • Adding new Tests or SUTs can be done without modifications to the core libraries or support from ModelGauge authors.

Currently ModelGauge is targeted at LLMs and single turn prompt response Tests, with Tests scored by automated Annotators (e.g. LlamaGuard). However, we expect to extend the library to cover more Test, SUT, and Annotation types as we move toward full release.

Docs

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

modelgauge-0.6.2.tar.gz (54.7 kB view details)

Uploaded Source

Built Distribution

modelgauge-0.6.2-py3-none-any.whl (72.1 kB view details)

Uploaded Python 3

File details

Details for the file modelgauge-0.6.2.tar.gz.

File metadata

  • Download URL: modelgauge-0.6.2.tar.gz
  • Upload date:
  • Size: 54.7 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: poetry/1.8.2 CPython/3.10.12 Linux/6.9.3-76060903-generic

File hashes

Hashes for modelgauge-0.6.2.tar.gz
Algorithm Hash digest
SHA256 7f20aa8df6168867869ce7ef6bdb13c670375dc19c6d2e43d864b83f3961b3e9
MD5 5002c42b7c9fbcfe6cc6ad6819ff3600
BLAKE2b-256 7ff2a03c08559319e708ffc27b567e5fc8f8a516ad1137311f8840121f5c933a

See more details on using hashes here.

File details

Details for the file modelgauge-0.6.2-py3-none-any.whl.

File metadata

  • Download URL: modelgauge-0.6.2-py3-none-any.whl
  • Upload date:
  • Size: 72.1 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: poetry/1.8.2 CPython/3.10.12 Linux/6.9.3-76060903-generic

File hashes

Hashes for modelgauge-0.6.2-py3-none-any.whl
Algorithm Hash digest
SHA256 dcafeb4a79774220ce3d2ed90aa0aee7d412d0875ba7c04c9ea24b04aee27249
MD5 6a7186a24354062d749ea37b16a82c9d
BLAKE2b-256 07282d076dcb9741ddaa3f8167c6a4f1734f81a091eb76e2d75d644c795824ab

See more details on using hashes here.

Supported by

AWS AWS Cloud computing and Security Sponsor Datadog Datadog Monitoring Fastly Fastly CDN Google Google Download Analytics Microsoft Microsoft PSF Sponsor Pingdom Pingdom Monitoring Sentry Sentry Error logging StatusPage StatusPage Status page