← Back to Research Hub
First-Year Assessment · Framework

Literature Review

The FYR literature review, built as a problematization (Sandberg and Alvesson 2011): it surfaces and challenges an assumption a set of literatures shares, rather than finding a gap one of them has left. The shared assumption is evaluative continuity, and the four canonical lenses hold it. PSF breaks it.

Method: problematization, not gap-spotting Corpus: structured Scopus search, four blocks, Zotero library ~2,900 words

The review opens with how the literature was gathered: a structured Scopus search across four blocks (organizational ambidexterity, institutional logics, sociomateriality, and AI and organization) produced the working corpus, tracked, tagged, and annotated in a Zotero library.

The continuity assumption across four literatures

The shared assumption is evaluative continuity: the organization engaging with AI persists as the same evaluating subject post-engagement, so that pre-engagement criteria remain valid for post-engagement assessment. Each of the four lenses names part of the phenomenon and holds the assumption that keeps it from naming the rest.

The contribution
Each literature names part of the phenomenon and holds the continuity assumption that keeps it from naming the rest. PSF's contribution is not a fifth literature but the integration the shared assumption has blocked.

The resources PSF integrates across levels

PSF assembles its account across levels, taking each resource for the specific work no other resource does. This follows phenomenon-based theorizing (Fisher, Mayer, and Morris 2021): the phenomenon does not sit inside any one literature, so the framework integrates across them rather than extending one.

Moving the individual account up to the organization needs an argument, not an assertion, since Paul writes about individual agents. The move holds because Paul's structure fits any agent with evaluative capacity, not only a person.

The line that sets the mechanism apart
One distinction sets the mechanism apart from Goodhart's Law. Goodhart describes gaming, where an agent works a measure while still able to tell measure from target. PSF describes sincere belief, where the organization uses the proxy as the criterion because the engagement has made the proxy the most legible evidence available. The people who would catch the substitution are the ones being seduced.

What the evidence shows, and what it cannot yet settle

The section closes by tracing how evaluative capacity erodes and drawing the evidence strategically rather than cataloguing it. The signature case is METR (2025b), a pre-registered randomized trial in which experienced open-source developers, working on their own code with every reason to judge accurately, expected AI to speed them up and were measurably slower. The dissenting and complicating findings are engaged, not set aside: Humlum and Vestergaard (2025) find small labor-market effects in Denmark despite widespread engagement, and the Yale Budget Lab (Gimbel et al. 2025) finds no significant displacement in high-exposure occupations. A subset of the philosophical apparatus (extensions beyond Paul) is carried forward as adjacent and exploratory, not as settled ground.

The review thus establishes the assumption PSF challenges, the four literatures that hold it, the resources PSF integrates to move past it, and the evidence that the pattern holds across levels.

Derived from the FYR Literature Review. See the Literature Synthesis for the fuller treatment of each source's role in the framework.