SYNTOLOGY HomeExplorerAtlasCodeMethodologyAboutDevelopersFeedPricing
Paper · 2306.06342 · 2023

Distribution-free inference with hierarchical data

arXiv · PDF · Open in the Atlas

Code that ran

We lifted 3 functions out of this paper's own repositories and ran 0 of them in a sandbox. "Ran" means the function executed on a synthesized input and returned a value. It is not a reproduction of the paper's results.

FunctionStatusWhere it lives
generate_l96_data Not yet run rebeccawillett/distribution-free-inference-with-hierarchical-data/lorenz96-experiment/l96_data.py
code served (permissive licence) · get_code("4c483237935d9c99")
lorenz96 Not yet run rebeccawillett/distribution-free-inference-with-hierarchical-data/lorenz96-experiment/l96_data.py
code served (permissive licence) · get_code("a4866eeb6683a7bc")
save_l96_data Not yet run rebeccawillett/distribution-free-inference-with-hierarchical-data/lorenz96-experiment/l96_data.py
code served (permissive licence) · get_code("59992655a0d0d369")

Repositories linked to this paper

Some links come from the archived Papers with Code dataset (CC BY-SA 4.0): attribution and licence.

Abstract

This paper studies distribution-free inference in settings where the data set has a hierarchical structure -- for example, groups of observations, or repeated measurements. In such settings, standard notions of exchangeability may not hold. To address this challenge, a hierarchical form of exchangeability is derived, facilitating extensions of distribution-free methods, including conformal prediction and jackknife+. While the standard theoretical guarantee obtained by the conformal prediction framework is a marginal predictive coverage guarantee, in the special case of independent repeated measurements, it is possible to achieve a stronger form of coverage -- the "second-moment coverage" property -- to provide better control of conditional miscoverage rates, and distribution-free prediction sets that achieve this property are constructed. Simulations illustrate that this guarantee indeed leads to uniformly small conditional miscoverage rates. Empirically, this stronger guarantee comes at the cost of a larger width of the prediction set in scenarios where the fitted model is poorly calibrated, but this cost is very mild in cases where the fitted model is accurate.

For agents

The same record, over MCP at https://syntology.ai/mcp:

get_harvested_code_for_paper("2306.06342")
get_code_for_paper("2306.06342")
have("2306.06342")

Connect an agent — have() is free.