2007, Data Processing, Measurement and Uncertainty

Measuring Validity and Reliability of Human Ratings

The Unofficial Google Data Science Blog

JULY 18, 2023

E ven after we account for disagreement, human ratings may not measure exactly what we want to measure. Researchers and practitioners have been using human-labeled data for many years, trying to understand all sorts of abstract concepts that we could not measure otherwise. That’s the focus of this blog post.

Measurement

Measurement Metrics Uncertainty Slice and Dice

Changing assignment weights with time-based confounders

The Unofficial Google Data Science Blog

JULY 22, 2020

For example, consider a smaller website that is considering adding a video hosting feature to increase engagement on the site. The fantasy football and video hosting examples, which we will discuss in more detail later, highlight situations where this design might be considered, despite potential complexity in the analysis.

Experimentation

Experimentation Statistics Testing Strategy

Data Leaders Brief

Measuring Validity and Reliability of Human Ratings

Changing assignment weights with time-based confounders

Webinars

Stay Connected