Skip to content
#

inverse-propensity-scoring

Here are 2 public repositories matching this topic...

Off policy evaluation audit for logged bandits: measures how much of a target policy's probability mass sits outside what the log could have produced, shows that a 95 percent interval around the standard estimator covers 0.26 of the time, and returns exit 2 rather than a number when the estimate would be about a different quantity.

  • Updated Aug 26, 2026
  • Python

Add this topic to your repo

To associate your repository with the inverse-propensity-scoring topic, visit your repo's landing page and select "manage topics."

Learn more