3-element Vector{Float64}:
0.5226655090702952
-0.13343797382338
-0.3513147652124974
2026-09-23
\[ \def\Er{{\mathrm{E}}} \def\cov{{\mathrm{Cov}}} \def\var{{\mathrm{Var}}} \def\R{{\mathbb{R}}} \]
The Question of Identification
Can we learn the true value of the parameter \(\theta_0\) if we knew the true distribution of the data \(P_X\)?
Definition
Let \(X\) be an observed random vector with distribution \(P_X\). Let \(\mathcal{P}\) be aprobability model — a collection of probabilities such that \(P_X \in \mathcal{P}\). Then \(\theta_0 \in \R^k\) is identified in \(\mathcal{P}\) if there exists a known \(\psi: \mathcal{P} \to \R^k\) s.t.
\[ \theta_0 = \psi(P_X) \]
\(\theta_0 =\) mean of \(X\), then \(\theta_0\) is identified by \[ \psi_\mu(P) = \int x dP(x) \] in \(\mathcal{P} = \{P : \int x dP(x) < \infty \}\)
Generally, descriptive statistics identified in a broad probability model with just regularity restrictions to ensure the statistics exist
\[ Y = \alpha + \beta X + \epsilon \]
\(\mathcal{P} = \{P_{X,Y}:\) \(Y=\alpha + \beta X + \epsilon\),
\(| \mathrm{Cov}(X,Y) | < \infty\), \(0 < \mathrm{Var}(X) < \infty\)
\(\mathrm{Cov}(X, \epsilon) = 0\) \(\}\)
\(\beta\) identified as
\[ \beta = \frac{\int (x - \Er X) (y - \Er Y ) dP_{X,Y}(x,y)} {\int (x - \Er X)^2 dP_{X}(x)} = \frac{ \cov(X,Y) }{ \var(X) } \]
\[ Y = X'\beta + \epsilon \]
\[ Y = 1\{ \beta_0 + \beta_1 X > u \} \]
Is \(u \sim N(0,1)\) innocuous?
Data:
Parameter: \(\theta_0 = \Er[Y_{i,1} - Y_{i,0}] =\) average treatment effect
Assume:
What if production function is \(Y_i = A_i F(L_i, K_i, M_i)\), can the elasticity of output with respect to \(M\) be identified?
Regression being a linear approximation to \(\Er[Y|X]\) does not mean \(\beta = \Er[X X']^{-1} \Er[X Y]\) necessarily has the sign you want
In example below, \(\Er[Y|x_1=1, x_2] > \Er[Y|x_1=0,x_2]\), but \(\beta_1 < 0\)
Definition: Observationally Equivalent
Definition: Observationally Equivalent Parameters
Definition: (Non-Constructive) Identification
\(s_0 \in S\) is identified if there is no \(s\) that is observationally equivalent to \(s_0\)
\(\theta_0\) is identified (in \(S\)) if there is no observationally equivalent \(\theta \neq \theta_0\)
\[ Y = X'\beta + \epsilon \]
If \(\beta\) and \(\tilde{\beta}\) observationally equivalent, then \(0 = \Er[X(Y - X'\beta)] = \Er[X(Y-X'\tilde{\beta})]\), so \(0 = \Er[XX'](\beta - \tilde{\beta})\)
\(X = [X_1\, X_2]'\), if rank \(\Er X X' = 1\), then \(\beta_1, \beta_2\) is observationally equivalent to any \(\tilde{\beta}_1, \tilde{\beta}_2\) s.t. \[ \tilde{\beta}_1 + \tilde{\beta}_2 \frac{\cov(X_1, X_2)}{\var(X_1)} = \beta_1 + \beta_2 \frac{\cov(X_1, X_2)}{\var(X_1)} \]
\(\theta_0 = \lambda( \beta ) = \beta_1 + \beta_2\frac{\cov(X_1, X_2)}{\var(X_1)}\) is identified if rank \(\Er [X X'] \geq 1\)
\(Y_i = 1\{\beta_0 + \beta_i X_i \geq U_i \}\)
\(\Er[Y|X] = \int \frac{e^{\beta_0 + \beta X_i}} {1+e^{\beta_0 + \beta X_i}} dF_\beta(\beta)\)
Non-constructive and constructive identification of \(F_\beta\) in Fox et al. (2012)