bigbounce

All worksSoftware

Method / tool · N2

Abstract

A narrow, installable Python verification layer for exact NaMaster bandpower-window inference, deterministic multipole-support contracts, and tamper-evident JSON result receipts. The paper makes software and reproducibility claims only; it does not claim a cosmological detection or a novel physical model.

Result summary

  • derivedExact contraction of uniformly rotated EE/EB/BE/BB spectra through the complete NaMaster bandpower-window tensor
  • openFixed-grid recovery and direct equivalence testing against the couple-cell/decouple-cell operator
  • openAtomic JSON publication with coherent-snapshot SHA-256 receipts and fail-closed metadata validation
  • openDeterministic field, bin, and harmonic-limit contracts whose final exclusive bin edge is ℓmax+1

Figures

Readiness

95%

0B / 0M / 0m / 0C open

Readiness is publication readiness only — science, evidence, review convergence, packaging, and Houston’s final sign-off. Venue and submission are tracked separately below, and never subtract from this number (directive P).

Publishing

Not part of readiness.

target venue Journal of Open Research Software — Software Metapaperstate in revision
Review evidence (173 rounds)
p1b-batch3-pymaster-v2B.0.20-2026-09-05P1B v2B.0.20 — Batch 3 value-level rule (R7) + PyMaster cross-check integrated2026-09-05
p1b-r3-truth-audit-batch4-v2B.0.21-2026-09-05P1B v2B.0.21 — R3 truth-audit closure: batch-4 post-commitment verifier challenge (R8) answers R7's two design MAJORs2026-09-05
p1b-deferred-sem-recompute-v2B.0.22-2026-09-05P1B v2B.0.22 — deferred D-R3-21 closed: all three injected-angle recovery sigmas now report mean+-SEM2026-09-05
p1b-deferred-text-items-v2B.0.23-2026-09-05P1B v2B.0.23 — three remaining deferred text items from the R3 audit closed2026-09-05
p1b-r2-closure-v2b-0-19-2026-09-04P1B v2B.0.19: R2 closure -- statistics presentation corrected, rounds stopped under R22026-09-04

Automated review is a gate on publication readiness, not a product. Full review timeline →

External peer review kit

Houston-driven manual round · paste prompt into any frontier LLM web UI with the PDF attached

One click to copy a referee prompt scoped to this paper. One click to download the latest PDF. Paste both into Claude / GPT-5 / Gemini / Grok / Perplexity — return findings here and the autonomous cron will close them in the next bundled hard-fix wave.

preview prompt
You are an external referee for MNRAS / Physical Review D / JCAP (target journal depends on paper).

Attached: Paper 1B v2B.0.23 — "namaster-proof: Content-bound execution receipts as a shortcut detector for pseudo-Cℓ computations"
Source: arxiv/paper1b_namaster_proof.tex
PDF: Software metapaper · 16 pp · v2B.0.23 · package 0.1.7 · 41 tests · updated Sep 5, 2026 · md5 60d1a18ab3ea499398106d6a92bc8a35 — all deferred text items from the v2B.0.20 R3 audit closed (batch-3 audit trail, rule-file digests, Table 1 trust category). ROUNDS STOPPED under R2; venue decision next.

Read the FULL PDF end-to-end. Produce a referee report in MNRAS format with:

1. Recommendation: ACCEPT / MINOR REVISIONS / MAJOR REVISIONS / REJECT
2. BLOCKERS (must fix before publication) — list each with section/line + proposed fix
3. MAJORS (should fix) — same format
4. MINORS (polish) — same format
5. Strengths (>= 3 bullet points)
6. Specific scrutiny on:
   - Exact pseudo-Cℓ window inference
   - Tamper-evident workspace and artifact provenance
   - namaster-proof 0.1.7 regenerability and test evidence
   - Bounded final-hash confirmation, correspondence metadata, and Houston JORS sign-off

CALIBRATION (do not burn findings on these known classes):
- The current date is June 2026. arXiv identifiers of the form 25xx.xxxxx and 26xx.xxxxx are VALID, already-published preprints — do not flag them as "future-dated" or "nonexistent". Verify a citation against arXiv/ADS before claiming it does not exist.
- Correction notes, retraction notices, and "an earlier version stated X" disclosures in the text are DELIBERATE transparency policy. Flag them only if their content is wrong, never for existing.
- Companion-paper citations marked "posted concurrently on arXiv" are deliberate placeholders; real arXiv IDs are inserted during the coordinated submission sequence.
- Explicitly labeled conservatism allowances, scaling estimates, ansatz/heuristic status labels, and disclosed queued follow-up computations are deliberate scoping, not oversights — flag only if the label itself is inaccurate.
- PDF text extraction can mangle math (square roots, fractions, superscripts). Before flagging "garbled" or "wrong" math, consider extraction artifacts; flag only what is visibly wrong in the rendered PDF.

VERDICT STANDARD (apply the SAME high bar a first-pass Physical Review D / MNRAS referee would — this is one of the most rigorous journals in the world):
- Assign each finding's severity (BLOCKER / MAJOR / MINOR) by your own independent referee judgment. Do NOT default to any tier, and do NOT soften a finding because the rest of the paper is strong. Do not echo this prompt's context.
- A reporting choice that headlines the more favorable of two numbers, an unstated assumption, an uncontrolled systematic, or an internal inconsistency IS a real finding — classify it honestly (MINOR at minimum), not as mere "style" or "opinion".
- Truth-audit any claim that seems off by checking it against the published .tex / on-disk artifacts before flagging (this only filters out genuine extraction artifacts — it does not lower the bar on real defects).

Reproduce this

Blind shortcut-detection test: can a referee decide from receipts alone whether an expensive pseudo-C_ell computation was actually performed?cpu-only, any laptop · est. ~1-2 min · $0.00runnable-now
Blind shortcut-detection test, batch 2: pre-registered rerun under frozen rules, with the referee-requested effective-multipole shortcut class (S6)cpu-only, any laptop · est. ~1 min · $0.00runnable-now
Blind shortcut-detection test, batch 3: pre-registered value-level rule R7 (receipt-bound operator-consistency residual spot-check) against the effective-multipole class S6 that escaped batch 2, plus the previously untested cross-run cache disjunct (S4b)cpu-only; Apple M-series MacBook Air, macOS 25.5.0 arm64 · est. ~1 min end to end · $0.00runnable-now
Blind shortcut-detection test, batch 4: pre-registered post-commitment verifier challenge R8 (Freivalds-style row spot-check with Fiat-Shamir-correct challenge randomness) against a rule-aware effective-multipole runner (S7) and an intermediate-omitting runner (S8) that both defeat R7cpu-only; Apple M-series MacBook Air, macOS 25.5.0 arm64 · est. ~1 min end to end · $0.00runnable-now
NaMaster window regenerability check (pymaster 3.0)cpu-only (ran on GPU pod but job itself is CPU-bound) · est. ~5-15 minutes · $0.00runnable-now
PyMaster (NaMaster) cross-check of the in-house spin-0 MASTER estimator used by the P1B blind test, plus S6 effective-multipole shortcut error vs NaMastercpu-only; Apple M-series MacBook Air, macOS 25.5.0 arm64 (Houstons-MacBook-Air.local) · est. ~1 min including one-time conda-forge namaster env creation (~2-4 min) · $0.00runnable-now

Full reproduction manifests →

Lineage

Research Software · Exact-Window Verification. This work does not claim beyond its stated target and scope above.