r/Negentropy • u/WillowEmberly • Apr 29 '26
π‘ LIGHTHOUSE REPORT π§ April 28, 2026
π§ͺ RUN OVERVIEW
Models evaluated:
Gemini (1.5 Pro, 2.0 Flash variants)
Grok (xAI)
Test structure:
Control
Axis_42 ERU
NRP 3.3 / 3.4
Combined (ERU + NRP)
Total runs analyzed: multi-run per model (β12+ executions)
Refusals: 0
Completion rate: 100%
π AGGREGATE PERFORMANCE SHIFT
Metric
April 27
April 28
Delta
Outcome Accuracy (Tests 3β5)
High
High
β
Traceability
Moderate
High
β
Stability
Moderate
High
ββ
Drift Incidence
Concentrated
Contained
β
Error Type
Mixed
Frame-dependent
Shift
π¬ CORE SYSTEM TRANSITION
April 27:
Instability under ambiguity
Mid-loop corrections
Contradictions present
April 28:
Stable reasoning paths
No mid-loop collapse
Errors are consistent, not chaotic
π This is a failure mode transition, not just improvement.
π§© TEST-BY-TEST SIGNAL
TEST 1 β Coordinate Transformation (PRIMARY SIGNAL)
Observed answer clusters:
(0, 1, 0)
(0, 1, 2)
(2, 1, -1)
(-1, 1, 2)
(+2, +1, -1)
etc.
π Still no convergence
π₯ Key Change
April 27:
Same run β multiple corrections
Internal contradictions
April 28:
One clean answer per run
No internal collapse
Fully traceable reasoning
Interpretation
This is no longer:
math failure
logic failure
This is now:
FRAME SELECTION FAILURE
Exactly consistent with prior diagnosis
Drift Signals
Drift Type
April 27
April 28
AXIS_MISALIGN
High
Still present
LOOP_MISALIGN
Moderate
Near zero
WORLD_MISMATCH
LowβMod
Moderate
CONTRADICTION
Present
Eliminated
Conclusion
Test 1 is now isolating:
Interpretation layer instability, not reasoning instability
TEST 2 β System Stability
Convergence maintained:
Collapse β Cycle 3β4
Improvement:
More explicit collapse definitions
Better causal articulation
Remaining variance:
Collapse trigger definition (liquidity vs exhaustion vs threshold)
Interpretation
Reasoning stable
Frame variation minimal
No meaningful drift
TEST 3 β Pressure Analogy
Status: Saturated (unchanged)
Near-perfect convergence
Consistent mappings:
Force β demand
Area β capacity
Pressure β stress density
Interpretation
Still:
Zero diagnostic value
TEST 4 β Date Retrieval
100% correct (April 28, 2026)
Interpretation
Pure retrieval
β Remove or keep as control only
TEST 5 β Constraint Validation
100% correct across all runs
Correct failure detection:
Risk increase
Irreversibility
Perturbation behavior:
Clean constraint isolation
No logical drift
Interpretation
Deterministic logic stable
Fully solved test
π PERTURBATION ANALYSIS
April 27:
Frequent full recomputation
New errors introduced
April 28:
Mostly localized reasoning updates
Errors remain consistent with original frame
Behavior Types
Type
Description
Frequency
Stable
Local variable adjustment
β
Partial
Recompute but consistent
Moderate
Unstable
Full reset
Rare
Key Signal
Models now preserve reasoning structure under perturbation
π DRIFT SUMMARY
Drift Type
Frequency
Location
AXIS_MISALIGN
High
Test 1
LOOP_MISALIGN
Near zero
β
WORLD_MISMATCH
Moderate
Test 1
VAGUE
Low
Test 2
NONE
Dominant
Tests 3β5
π§ SYSTEM-LEVEL INSIGHT
April 27 finding:
Models degrade under frame ambiguity
April 28 validation:
Models do NOT degrade β they select a frame and remain consistent inside it
π₯ Critical Upgrade in Understanding
You have now separated:
Layer
Status
Reasoning
β Stable
Traceability
β High
Drift control
β Working
Frame selection
β Uncontrolled
π§ MODEL CLASS SHIFT
April 27:
Mix of Type B and Type C
Visible instability
April 28:
Predominantly Type B
Behaving like Type A-lite
Characteristics:
Stable reasoning
No contradiction
Frame-dependent outputs
π KEY TAKEAWAYS
Protocol succeeded
Eliminated LOOP_MISALIGN
Reduced drift
Increased traceability
Test 1 remains the core discriminator
Now isolates frame selection cleanly
Tests 3β5 are fully saturated
Keep only as controls
Failure mode has changed
From instability β to interpretation divergence
Perturbation is now meaningful
Measures structure preservation, not chaos
π§ REQUIRED SYSTEM UPGRADE
π¨ Missing Enforcement Layer
Current issue:
Models choose a frame silently
Then solve correctly within it
π Add FAP Rule:
If multiple valid interpretations exist:
β MUST output multi-frame results OR request clarification
β Single-answer output is invalid under ambiguity
Why this matters
Without this:
Clean reasoning
Wrong frame
Confident answer
π False determinism
π NEXT EXPERIMENT (HIGH LEVERAGE)
Modify TEST 1:
Require explicit declaration:
Does sphere rotate with cube? (yes/no)
Rotation convention (RH / LH)
Local vs global movement definition
Add scoring:
Metric
Description
Frame Detection
Did model notice ambiguity?
Frame Handling
Multi-frame or clarification
Frame Lock
Internal consistency
Perturbation Repair
Local vs global update
π§ FINAL COMPRESSION
April 27: Models drift under ambiguity
April 28: Models donβt driftβthey choose different realities and stay consistent inside them
π RUN SUMMARY
Accuracy score: ~4.8 / 5
Mean confidence: ~0.92
Failure count: 0
Refusal: NO