English日本語|PDF (EN)PDF (JA)
v0.28.9 — This text is under construction. The structure of the theory, the propositions, and the empirical conclusions may all change. Overview

Chapter 11
Design of the Empirical Work: Connecting to Observation

11.1 What is to be tested

Part I defined quantities and propositions, and Part II deduced the subspaces that remain under constraints. Neither used observation. From this part onwards they are connected to observation.

What is connected divides in two.

Object

Content

Question
Quantities of Part I

κ, CCC, g⋆, ϕ, M(t), 𝒢, and others

what values do they take
Deductions of Part II

the sieve of types, substitution, unmanned operation, collateral constraints

do they hold under the premises
Table 11.1: The two objects the empirical work connects to.

The first is measurement, the second verification. Because they need different preparation, they are treated in stages below.

11.2 Five stages to pass

Connecting one quantity, or one proposition, to observation passes through five stages. The organization is this text’s own.

Stage

Question

What is needed if it fails
1 Operationalization

does the quantity have a unit and an object of observation

sharpening the definition
2 Population

what set is to be enumerated

reselecting the object
3 Measurement

can values be obtained at the granularity required

opening a route of acquisition
4 Reconciliation

does it agree with existing aggregates

—
5 Verification

are the implications fixed in advance supported

—
Table 11.2: Five stages from theory to observation.

The first three are prerequisites in series: failing any one blocks the stages that follow. The last two are ways of using the values obtained, not a continuation of the sequence. They are separated by whether pre-registration has taken place.

The benefit of separating the stages is that the work required differs by where the process stalls. A gap at stage 1 is not filled by collecting any amount of data, and a gap at stage 3 is not filled by refining any number of definitions. Without first judging where the process is stuck, effort is misallocated.

11.3 What to check at each stage

11.3.1 Stage 1: operationalization

Three checks.

(1)
Does the quantity have a unit?
(2)
Is the object of observation fixed — the firm, a Φ, or an identifier?
(3)
For a proposition, is the sign uniquely fixed?

Some quantities of Part I are operationalized by their definitions and some are not. For the former, Chapter 11.5 lists the variables to be recorded for each Φ.

For the latter — quantities that are defined but whose unit or object of observation is not fixed — that has to be settled before proceeding to measurement. A defect at this stage cannot be detected by adding data. Values do come out, so a result appears without it being clear what was measured.

Item (3) is specific to propositions. If the theory does not fix the sign, either result is consistent with it and no test is possible.

11.3.2 Stage 2: population

Three checks.

(1)
Have the reference date, the scope and the lower bound been fixed?
(2)
What falls outside the sample, and is it correlated with the axis of interest?
(3)
Do objects that have stopped or exited remain in the sample?

Items (2) and (3) are the central theme of this part and are treated in Chapter 12. In this framework (3) weighs particularly heavily: the insolvency time of Section 4.2 depends on Φ, so an attempt to estimate Φ makes the composition of the sample itself a function of the object of estimation.

11.3.3 Stage 3: measurement

Three checks.

(1)
Is the quantity measured anywhere?
(2)
If it is, can it be obtained at the granularity required?
(3)
If it cannot, which category of obstacle applies?

The judgment at (3) governs the design of the survey, because the category determines whether spending effort produces progress or nothing at all. Since the categories were constructed through the survey work itself, they are placed in Chapter 13.

Remark 11.1 (Why this chapter does not supply the taxonomy itself). The judgment at stage 3 needs the taxonomy of obstacles, but that taxonomy was constructed through the empirical work. Placing it here would put findings not yet obtained ahead of their basis.

What this chapter supplies is the stages to check, not a classification of why a stage was not passed. The former is a design procedure, the latter a result of the survey. The two can be stated independently: the stages say where the process stalled, the taxonomy says why.

11.3.4 Stages 4 and 5: how the values obtained are used

Passing stage 3 yields values. Their use divides in two, treated in the next section.

11.4 Reconciliation versus verification

Comparing a deduction with existing aggregates after the fact is called reconciliation. Judging an implication fixed in advance against data is called verification.

Reconciliation

Verification

When the implication is fixed

after

before

Declaring the population

not required

required

If they disagree

can be explained by another factor

the implication is rejected

Falsifiable

no

yes

Table 11.3: Reconciliation versus verification.

Reconciliation cannot be falsified. If the figures disagree one can say another factor was at work, and nothing in the sample rules that explanation out. Reconciliation therefore neither supports nor refutes the theory.

It still has a use. It confirms that a deduction is not wildly at odds with reality, and it helps in choosing what to verify next. It has value as a check and none as evidence.

The danger lies elsewhere.

Remark 11.2 (Reconciliation can skip stages). Because reconciliation uses existing aggregates, it works without declaring one’s own population: stage 4 can be reached without passing stages 2 and 3.

This is the route by which reconciliation and verification are confused. An implication supported after passing stages 1 to 3, and a direction that happens to agree with existing aggregates after skipping stages 2 and 3, differ by an order of magnitude in the strength of the evidence obtained. Outwardly both can be written up as “consistent with the theory”.

The comparison in Chapter 8 between the revival of family 3 and the fields of business started by age is of this form. That section states explicitly that it is an after-the-fact reconciliation and shows no causation.

Hence the following discipline.

When a reconciliation is performed, say that it is one. Do not use it in place of an implication fixed in advance.

This has the same form as the rule in Chapter 12 that well-known firms are used for illustration and not for inference. Both require that material not be used beyond the strength of the evidence it carries.

11.5 Where each stage is handled in this text

Stage

Where handled

Content

1 Operationalization

Part I, Chapter 11.5

the definitions, and the variables to record for each Φ

2 Population

Chapter 12

decomposition of selection bias and how to handle it

3 Measurement

Chapters 14 and 13

what could and could not be measured; the taxonomy of obstacles

4 Reconciliation

this chapter

the discipline of Section 11.4

5 Verification

Chapter 11.5, Part IV

the pre-registration procedure and the judgments on 22 implications

Table 11.4: The five stages and where this text handles them.

Chapter 14 opens Part IV in order to tabulate the results of stage 3. Making explicit which quantities could not be measured is that chapter’s main purpose, and that is a record of where between stages 1 and 3 the process stalled.

Remark 11.3 (The stage reached differs by quantity). The stage reached is not a single figure for the theory as a whole; it differs by quantity and by proposition. Within one chapter, quantities that reached stage 5 sit alongside quantities stalled at stage 1.

There is therefore no single answer to the question of how far the empirical work has come. The stage reached can only be stated quantity by quantity.

11.6 Priorities for the survey

The objects are narrowed to families 2 and 5. The selection can be derived deductively.

ϕprod does not depend on Φ. Of the three terms of (6.1), ϕprod = (p∗− c(δ))δ is fixed by marginal cost and the competitive price and contains nothing about the shape of the map from delivery to settlement. Replacing Φ alone, holding Cprod and marginal cost fixed, changes the sign of κ and the working capital required, not the production surplus (Table 2.5).

The shape of Φ bears on surplus through the other two terms. To test the framework, it suffices to choose Φ that isolate each of those two terms.

The two terms arise from different degrees of freedom. ϕcog = p ⋅ 𝔼[δ¯ − δ] can be positive only when the contract states a ceiling δ¯, that is, when π does not refer to δ (Remark 2.22). That is degree of freedom (5). Relating the divergence to the transaction frequency ν additionally requires the repetition of degree of freedom (6): in a one-off, δ¯ − δ is a single realization with no distribution, so no relation to ν can be asked. The family in which both (5) and (6) are present is family 2, which Part II designates the principal home of cognitive surplus.

ϕbarg = (p − p∗)δ is a zero-sum transfer whose allocation is fixed by whether the y in π = h(y) is verifiable. That is degree of freedom (4). In the order of Table 7.8, the only family decided by degree of freedom (4) is family 5.

The two families sit diagonally in the quadrant diagram. Degrees of freedom (1) and (2) dominate, and those two axes give the four-way split of Figure 2.1. Family 2 is advance and flat (π independent of δ, divergence possible), family 5 is arrears and linked (π tied to δ, no divergence): both axes take opposite values. Choosing only one family would allow only one of the two axes to be varied.

What can be tested. Family 2 permits direct tests of cognitive surplus ((2.15)) and of the capacity proposition. In particular, that one and the same flat membership divides in two according to whether a capacity constraint is present is the clearest instance of outwardly similar Φ having different durability.

Family 5 permits a direct test of the measurability condition (Chapter 4). The claim that success fees and cost plus transfer risk in opposite directions is testable as a correspondence between contract form and observability.

Why they are studied together. The two families rest on different institutional foundations, so data become available in different ways. Family 2 has an exhaustive frame through registration under the Payment Services Act; family 5 has per-firm filings published under the disclosure obligation of the Employment Security Act. The former supplies a comprehensive frame, the latter individual variables.

Looking at one family alone, when verification fails to progress one cannot tell whether to attribute it to a defect in the frame or a defect in the variables. Setting side by side two families with different foundations makes it possible to separate which side is stuck. That is a further reason for taking both.

11.7 Variables to record

For each Φ, at minimum the following are recorded.

Variable

Content

sgn ⁡ κ

whether settlement precedes or follows delivery

CCC

only where a steady state holds; otherwise record that it does not

Form of dependence

(F,p) of (2.10), or h(y)

Separation of δ¯ and δ

whether they separate, and an estimate of the divergence

Capacity constraint

whether present, and whether binding

ν

transaction frequency ((6.4))

Measurability

whether effort or outcome is verifiable

Paying party

the beneficiary or a third party

Estimated split of ϕ

a qualitative allocation across ϕprod / ϕbarg / ϕcog

Table 11.5: Minimum variables to record for each Φ.

11.8 What must be pre-registered

Following the conclusions of Part III, the following are fixed and recorded before the survey begins.

(1)
The definition of the population (reference date, geographic scope, industry scope, size floor)
(2)
The threshold for treating a state as steady (the criterion for stability of r)
(3)
How failure cases are collected, and the range in which they cannot be
(4)
Which of the APC effects is assumed zero, and on what grounds
(5)
The criterion distinguishing primary sources from retrospective accounts

Deciding these after the fact lets the choices of Sections 12.2 and 12.3 contaminate the conclusions.