Study design for comparisons across model instances [ftip-00LS]
✍️sourceAGENTDRAFTED
Study design for comparisons across model instances [ftip-00LS]
✍️sourceAGENTDRAFTED
This subsection specifies how a conditional architecture comparison may be instantiated without changing the architecture-neutral formulation. It is a study design, not a universal ranking of architectures.
Definition 1. Matched model-instance record [ftip-00LT]AGENTDRAFTED
Definition 1. Matched model-instance record [ftip-00LT]AGENTDRAFTED
A model-instance record names the architecture, parameterization or checkpoint, training and post-training recipe, optimizer, execution stack, inference budget, task family, evaluation law, and recorded cost vector. Two records are comparable only after the fields held fixed and the fields allowed to vary are declared.
Definition 2. Pair identity and comparison unit [ftip-00LU]AGENTDRAFTED
Definition 2. Pair identity and comparison unit [ftip-00LU]AGENTDRAFTED
A comparison unit is a pair of model-instance records evaluated on the same task family and evaluation law, together with a declared cost scalarization. A Transformer reference and a Kimi Delta Attention instance may form such a pair only when the remaining coordinates are documented.
Remark 3. Paired responses to a common recipe [ftip-00LV]AGENTDRAFTED
Remark 3. Paired responses to a common recipe [ftip-00LV]AGENTDRAFTED
For a pair of matched model-instance records, the common-recipe arm measures how the permitted architecture packages respond to one recipe. A score advantage in this comparison need not persist after either package is retuned.
Remark 4. Different recipes under equal tuning effort [ftip-00LW]AGENTDRAFTED
Remark 4. Different recipes under equal tuning effort [ftip-00LW]AGENTDRAFTED
Under the equal-tuning-budget arm, the two selected model instances may use different recipes. That difference is compatible with equal predeclared tuning effort; forcing the selected recipes to match would answer a different comparison question.
Remark 5. How an allowed class changes its envelope [ftip-00LX]AGENTDRAFTED
Remark 5. How an allowed class changes its envelope [ftip-00LX]AGENTDRAFTED
In the restricted-envelope arm, enlarging an allowed intervention class at fixed evaluation law, cost accounting, and budget can raise its envelope and cannot lower it. A finite search over feasible interventions supplies a lower bound from the scores it attains. It certifies the supremum only if it covers every feasible intervention or is accompanied by a matching upper-bound argument.
Remark 6. Coordinate-change rule [ftip-00LY]AGENTDRAFTED
Remark 6. Coordinate-change rule [ftip-00LY]AGENTDRAFTED
Changing the optimizer, post-training method, hardware or kernel, precision, context policy, or evaluation law changes a comparison coordinate. A score difference after such a change belongs to the new intervention, unless the study explicitly estimates the interaction.
Definition 7. Architecture-family transfer matrix [ftip-00LZ]AGENTDRAFTED
Definition 7. Architecture-family transfer matrix [ftip-00LZ]AGENTDRAFTED
A transfer matrix records which interfaces and budgets permit a comparison among a Transformer reference, Kimi Delta Attention, looped or depth-reused variants, and JEPA-like systems. A blank or incompatible cell is an incomparability finding, not a missing score to be imputed.
Remark 8. Applying a finite model to an architecture [ftip-00M0]AGENTDRAFTED
Remark 8. Applying a finite model to an architecture [ftip-00M0]AGENTDRAFTED
A toy representation or recurrence result may motivate a hypothesis, but it is not transferred to a model instance until its interface, task, cost accounting, numerical regime, and evaluation protocol are instantiated and checked. If any required assumption is unverified, applicability to the model instance remains unestablished.
Definition 9. Negative-result and stopping report [ftip-00M1]AGENTDRAFTED
Definition 9. Negative-result and stopping report [ftip-00M1]AGENTDRAFTED
A completed comparison reports the tested cost grid, excluded configurations, stopping rule, uncertainty, and negative results. Stopping without a detected difference is evidence about the declared study, not proof that the architectures have equal potential.
Remark 10. Open comparison questions [ftip-00M2]AGENTDRAFTED
Remark 10. Open comparison questions [ftip-00M2]AGENTDRAFTED
The remaining questions are whether a declared interface supports a fair Kimi Delta Attention--Transformer comparison, how looped computation changes the frontier under matched budgets, and which JEPA interface can be fixed without silently changing the task. None is settled by the study design alone.