Sign in

stephenekhansen.bsky.social

@stephenekhansen.bsky.social
48 followers 23 following 9 posts
PostsRepliesMedia
stephenekhansen.bsky.social @stephenekhansen.bsky.social · 22/03/2025
Highly recommend @yabramuvdi.bsky.social new Substack on Large Language Models (in Spanish) substack.com/@yabramuvdi. I have learned so much working with Yabra over the years, and I think you will too!
substack.com
Yabra Muvdi | Substack
Desarrollo, escribo y enseño sobre modelos de lenguaje.
010
stephenekhansen.bsky.social @stephenekhansen.bsky.social · 12/12/2024
More generally, establishing procedures for valid inference in the growing world of AI-generated indicators is a major future challenge. arxiv.org/abs/2402.15585
arxiv.org
Inference for Regression with Variables Generated by AI or Machine Learning
It has become common practice for researchers to use AI-powered information retrieval algorithms or other machine learning methods to estimate variables of economic interest, then use these estimates ...
000
stephenekhansen.bsky.social @stephenekhansen.bsky.social · 12/12/2024
We also show that an IV strategy that uses a human-labeled sample to purge the measurement error in generated variables works poorly when the number of labels is small relative to the unlabeled data.
100
stephenekhansen.bsky.social @stephenekhansen.bsky.social · 12/12/2024
We provide an illustration of how bias correction increases the estimated impact of remote work on wages across occupations.
100
stephenekhansen.bsky.social @stephenekhansen.bsky.social · 12/12/2024
We provide a simple bias correction formula that applied researchers can easily use. This restores valid inference and has quantitatively important effects even when AI/LLM are extremely accurate.
100
stephenekhansen.bsky.social @stephenekhansen.bsky.social · 12/12/2024
We consider the realistic case where algorithms become more precise as the sample size increases. In this setting, **point estimates are biased** but **standard errors are correct**. This is the opposite of the typical generated regressor problem.
100
stephenekhansen.bsky.social @stephenekhansen.bsky.social · 12/12/2024
Suppose we treat an AI-generated variable as "data" in a regression model. One intuition is that measurement error biases coefficient estimates. Another is that ignoring uncertainty biases standard errors. Which is it?
100
stephenekhansen.bsky.social @stephenekhansen.bsky.social · 12/12/2024
📢 **new results** LLMs and AI can be used to extract measures from text like sentiment, beliefs, and uncertainty. What can go wrong when plugging these measures into regressions and how to fix the problem? Read more below and check out arxiv.org/abs/2402.15585 for details #EconSky
arxiv.org
Inference for Regression with Variables Generated by AI or Machine Learning
It has become common practice for researchers to use AI-powered information retrieval algorithms or other machine learning methods to estimate variables of economic interest, then use these estimates ...
2204
stephenekhansen.bsky.social @stephenekhansen.bsky.social · 07/12/2024
Oh no, my cover is blown! j/k, thank you Fatih. Hoping this new place has more econ/ML content and fewer Joe Rogan clips.
011