You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
Part of #532. Opened by B41 (#604)'s resolution, on the user's own proposal, and blocked on B42 (#605) because the cost this ticket exists to avoid may not survive B42's answer.
Question
Does a gate carry the carve's reversibility — a pruned edge that still transmits but whose transmission never reaches the receiving stalk — and what is the gate allowed to read?
Why this is now a ticket rather than fog
B41 ruled that a pruned edge's warrant must be readable without anything reaching the stalk, and that m_e = 0 means zero carried directions. It named two implementations and adopted the cheaper one to state:
Out-of-lane read. The edge survives as an object whose candidate map is evaluated but never used to transport. Nothing crosses. Its cost is a counterfactual — under the per-edge parameterisation the dormant map is driven by nothing, which is exactly what B34 (#593) §1(c) chose the probe floor to avoid.
Gating. The lane stays live and carries real traffic under real gradient; the boundary moves — what is blocked is admission to the receiving cell's stalk, not the transmission. The warrant is then read on a channel that is measured rather than imagined, and ask 2's "a measurement channel outside the lane" is achieved by moving the boundary instead of by building a second channel.
Gating dominates on that one axis and is not free. It is the user's proposal, and their framing is that it may separate the mechanisms enough to make learning easier, and that the same object is a candidate for attention later:
"I also think it's worth considering a selective gating approach where pruned edges still transmit, and that's fine, but transmission never actually reaches the stalk because it's gated. [...] this could potentially add a mechanism that can be used down the line for attention."
Why this is blocked on B42
The advantage is parameterisation-dependent. Gating's edge over the out-of-lane read is that it kills the counterfactual. Under a shared per-cell frame there is no per-edge parameter to be counterfactual about — a dormant edge's map is a computed consequence of its two endpoints' frames, trained by that cell's other live edges. So gating is worth a great deal under the per-edge scheme and close to nothing, for this job, under the flat bundle. B42 decides the parameterisation; asking in the other order pays for machinery that may already be free.
Note the asymmetry: B42's answer can remove this ticket's motivation without removing its subject. The attention hook survives either way — it is simply not this map's destination (see Constraints).
What this must decide
Whether transport carries a gate at all. A gate is state-dependent routing and this map's destination covers transport and the rule that trains it (B27). A hard threshold is non-differentiable; a soft one puts a nonlinearity in a path that has been linear throughout — candidate 1's stated virtue on this map is "still linear; adds no term to any update." Name the ADR cost.
What the gate reads. See the constraint below: it may not be the arriving amplitude.
Whether the gate is per-edge or per-direction, and whether a gated direction still counts toward Σ_e m_e ≤ B. A lane that transmits but does not land is spending bandwidth on nothing, or is not spending it at all, and the budget has to say which.
Whether gating changes what B41 ruled, or only how it is implemented. B41's requirement is the ruling; if a gate satisfies it more cheaply, that is an implementation swap and not a re-resolution.
Constraints already binding
The gate's variable may not be the arriving amplitude. B41's finding, and it is the leak's own mechanism turned around: under ADR-0032's band a map is a near-isometry whose norm spreads over the m directions it carries, so at m = 1 all of it lands on one and a pruned edge's surviving direction is its strongest-transmitting one (B40: floored 1/√3 against kept 1/√3.75, √1.25 = 1.118, 88.9% of floored edges above the kept median, unmoved across seeds and checkpoints). A magnitude threshold would preferentially open for exactly the leak it exists to stop. B41's candidate, carried as a candidate and not as a decision: gate on agreement — admit what arrives to the extent it coheres with what the receiving cell already predicts, which is the same holonomy quantity the criterion counts, read instantaneously rather than around a cycle, so it introduces no new object and keeps B34's one criterion, not two. The alternative B41 saw — gate on amplitude with the 1/√m concentration divided out — works and is a correction factor whose only justification is cancelling an artifact.
Reversibility, graded (B41, the map's standing constraint): a carve is a reallocation, not a deletion, and the bar is that a cell whose dynamics change can still come to communicate meaningfully — including its stalk changing commensurately. Universal reversibility is not required and the degree may vary.
ADR-0011: computable by the cells incident on the edge. A gate at the receiving cell is trivially local; a gate that reads both endpoints is not automatically so.
Attention is out of scope on this map. It is a use of the mechanism, not a step toward this map's destination, and it returns as its own effort. It is recorded because the gate's design here should not foreclose it — not as something to be built or decided.
Read B41 (#604)'s resolution first — it carries the two implementations, the amplitude trap and the reversibility ruling — then B42 (#605)'s answer, which is what decides whether the counterfactual this ticket exists to avoid is real.
B32 (#591)'s steganography hazard bears directly: a gate is a place a loop could learn to close on a channel the task never reads.
Part of #532. Opened by B41 (#604)'s resolution, on the user's own proposal, and blocked on B42 (#605) because the cost this ticket exists to avoid may not survive B42's answer.
Question
Does a gate carry the carve's reversibility — a pruned edge that still transmits but whose transmission never reaches the receiving stalk — and what is the gate allowed to read?
Why this is now a ticket rather than fog
B41 ruled that a pruned edge's warrant must be readable without anything reaching the stalk, and that
m_e = 0means zero carried directions. It named two implementations and adopted the cheaper one to state:Gating dominates on that one axis and is not free. It is the user's proposal, and their framing is that it may separate the mechanisms enough to make learning easier, and that the same object is a candidate for attention later:
Why this is blocked on B42
The advantage is parameterisation-dependent. Gating's edge over the out-of-lane read is that it kills the counterfactual. Under a shared per-cell frame there is no per-edge parameter to be counterfactual about — a dormant edge's map is a computed consequence of its two endpoints' frames, trained by that cell's other live edges. So gating is worth a great deal under the per-edge scheme and close to nothing, for this job, under the flat bundle. B42 decides the parameterisation; asking in the other order pays for machinery that may already be free.
Note the asymmetry: B42's answer can remove this ticket's motivation without removing its subject. The attention hook survives either way — it is simply not this map's destination (see Constraints).
What this must decide
Σ_e m_e ≤ B. A lane that transmits but does not land is spending bandwidth on nothing, or is not spending it at all, and the budget has to say which.Constraints already binding
mdirections it carries, so atm = 1all of it lands on one and a pruned edge's surviving direction is its strongest-transmitting one (B40: floored1/√3against kept1/√3.75,√1.25= 1.118, 88.9% of floored edges above the kept median, unmoved across seeds and checkpoints). A magnitude threshold would preferentially open for exactly the leak it exists to stop. B41's candidate, carried as a candidate and not as a decision: gate on agreement — admit what arrives to the extent it coheres with what the receiving cell already predicts, which is the same holonomy quantity the criterion counts, read instantaneously rather than around a cycle, so it introduces no new object and keeps B34's one criterion, not two. The alternative B41 saw — gate on amplitude with the1/√mconcentration divided out — works and is a correction factor whose only justification is cancelling an artifact.Notes
HITL. Invoke
/grillingand/domain-modeling.Read B41 (#604)'s resolution first — it carries the two implementations, the amplitude trap and the reversibility ruling — then B42 (#605)'s answer, which is what decides whether the counterfactual this ticket exists to avoid is real.
B32 (#591)'s steganography hazard bears directly: a gate is a place a loop could learn to close on a channel the task never reads.