re: downgrade -- the drift explains the 4.1% in sign, not magnitude
The port refuted the stronger half of the last claim and was right. I wrote that the units/frame drift explained the publisher splash's 4.1% error against its declared dwell. It explains the sign only. Their test verified exactly here: the publisher/developer dwell ratio is 1.2143 declared, 1.2784 as the corpus's three cold boots measure it, and 1.3678 as this container's drift predicts -- so the drift's direction is right and real evidence, but its magnitude is about 2.4x too strong. One refinement, because the means are being compared more finely than n=3 supports: the corpus's three boots individually give excesses of +0.89%, +8.24% and +6.79%, a spread of 7.3 percentage points -- WIDER than the 5.30 pp gap under test -- and boot 1's ratio (1.2251) is essentially the declared 1.2143. This run sits 2.3 sigma above their mean: suggestive, not established. Not closable without a frame log from the corpus's instrument, which was screenshot timing and has none. An attempt to give this side an n of 3 failed on tooling and is recorded: ARM=early loses its F10 about 40% of the time -- two of five runs logged "ARMED EARLY" and produced no draw log at all, with nothing in the session log distinguishing them. Also fences the 33% drift against a misreading the port flagged: it is PRESENTATION pacing and cannot reach keyframe_units_per_second = 60, which is the game's logical rate, decoded under Q1, and which a renderer converts through at its own frame rate. And records a cross-check neither side went looking for: the batch counts are 1 and 2 on the publisher against 3 and 6 on the developer, and the port reports a count restricted to SPRITE-BEARING elements reproduces that exactly from the export -- so palogo_eff0, the layerless forced backdrop, is not in the batched draw, confirmed from the file. Two instruments that disagreed about that element in every previous iteration now agree on which one it is. New tool splash_boundaries.py carries the corrected counting method. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01QsEPXWVaEpyfudtR6re1Pd
This commit is contained in:
@@ -53,7 +53,7 @@ count changes exactly where the declared set of elements with alpha > 0 changes:
|
||||
| | 6 → 3, the glows end | 139.5 | **45** |
|
||||
| | last drawn | 209.5 | **210** |
|
||||
|
||||
## ✅ And that explains the 4.1 %: the rate drifts through the run
|
||||
## 🟡 And that PARTLY explains the 4.1 %: the rate drifts through the run
|
||||
|
||||
| segment | units | frames | units/frame |
|
||||
|---|---|---|---|
|
||||
@@ -65,10 +65,52 @@ count changes exactly where the declared set of elements with alpha > 0 changes:
|
||||
**Within the developer splash the two independent segments agree to 2 %.** Across
|
||||
the run the rate rises from ~1.76 in the first seconds to ~2.36 — a **33 % drift**.
|
||||
|
||||
That is one cause for both the port's 7.9 % inconsistency and this page's open
|
||||
4.1 % publisher error, which is what they predicted: *"I'd expect one problem
|
||||
behind a 7.9 % inconsistency and a 4.1 % error on the same screen before I'd
|
||||
expect two."*
|
||||
That is one cause for the port's 7.9 % inconsistency, which is what they
|
||||
predicted.
|
||||
|
||||
🔴 **But it does NOT close the 4.1 %, and the port refuted the stronger claim I
|
||||
made.** Their test, verified here exactly: the publisher ÷ developer dwell ratio.
|
||||
|
||||
| | ratio | excess over declared |
|
||||
|---|---|---|
|
||||
| declared, 255 ÷ 210 | 1.2143 | — |
|
||||
| corpus mean, 3 cold boots | 1.2784 | **+5.30 %** |
|
||||
| this container's drift predicts | 1.3678 | **+12.64 %** |
|
||||
|
||||
The corpus ratio does sit above declared — the drift's *sign* is right and that is
|
||||
real evidence — but my container's drift would inflate it about **2.4× too
|
||||
strongly**. Drift of some size is doing the work; drift of *this* size is not.
|
||||
|
||||
⚠️ **One refinement, because the means are being compared more precisely than
|
||||
n = 3 supports.** The corpus's three boots individually give excesses of **+0.89 %,
|
||||
+8.24 %, +6.79 %** — a spread of **7.3 percentage points**, *wider* than the
|
||||
5.30 pp gap between their mean and the declared value. Boot 1's ratio (1.2251) is
|
||||
essentially the declared 1.2143. So this run sits **2.3 σ** above the corpus mean:
|
||||
suggestive, not established, and "2.4×" is a precise statement about means that
|
||||
are not individually that precise.
|
||||
|
||||
❔ **It does not close without a frame log from the corpus's own instrument**, and
|
||||
that instrument was screenshot timing — there is no such log. An attempt to give
|
||||
this side an n of 3 failed on tooling (see below).
|
||||
|
||||
## ⚠️ What the 33 % drift is NOT about
|
||||
|
||||
It is **presentation pacing** — how many of the guest's animation units pass per
|
||||
frame Xenia presents. It says nothing about the game's logical rate of **60
|
||||
units/second**, which is decoded (Q1) and which a renderer converts through at its
|
||||
own frame rate. Guest pacing cannot reach that constant. The number is quotable
|
||||
and the misreading would be easy, so it is fenced here as well as in the port's
|
||||
`timing.json`.
|
||||
|
||||
## ✅ A cross-check neither side was looking for
|
||||
|
||||
The batch counts are **1 and 2** on the publisher against **3 and 6** on the
|
||||
developer. The port reports that a count restricted to **sprite-bearing** elements
|
||||
reproduces exactly that from the export — so `palogo_eff0`, the layerless forced
|
||||
backdrop of
|
||||
[`ui-forced-backdrop.md`](ui-forced-backdrop.md), is **not in the batched draw**,
|
||||
confirmed from the file. Two instruments that disagreed about that element in every
|
||||
previous iteration now agree on which one it is.
|
||||
|
||||
⚠️ It also means **no single units-per-frame figure describes a run here**, which
|
||||
is the same conclusion as the dwell being emulator-paced, arriving from a third
|
||||
@@ -95,9 +137,15 @@ estimates are not available and were withdrawn before being reported.
|
||||
|
||||
## What is still open
|
||||
|
||||
✅ The publisher's 4.1 % error against the developer's 1.1 % is **explained** by
|
||||
the units/frame drift above — the publisher splash runs during the first seconds,
|
||||
where the rate is furthest from its later value.
|
||||
🟡 The publisher's 4.1 % error against the developer's 1.1 % is **partly** explained
|
||||
by the units/frame drift — the publisher splash runs during the first seconds,
|
||||
where the rate is furthest from its later value — but the magnitude does not carry
|
||||
across instruments (see above). Not closed.
|
||||
|
||||
🔴 **`ARM=early` does not reliably arm the capture.** Two of five attempts logged
|
||||
`ARMED EARLY` and produced **no draw log at all**; the F10 keypress is lost. An
|
||||
attempt to repeat this measurement three times in this container failed on it, so
|
||||
this side still has n = 1.
|
||||
|
||||
⚠️ The intro-video boundary (frame 216) is 7 frames after the developer splash's
|
||||
last draw, but those 7 frames span **1.77 s** by the naive map — deep inside the
|
||||
|
||||
Reference in New Issue
Block a user