CORTEXA
← Browse
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-24Cited by 0

Honest, Leakage-Free Operational Earthquake Forecasting: A Multi-Region CSEP Testbed with Pre-Registered Negatives and the Horizon-Dependent Value of Geodetic Context

Felipe Santibañez-Leal

Version 2.1 (revised). Operational earthquake forecasting (OEF) issues calibrated conditional probabilities of future seismicity; the Epidemic-Type Aftershock Sequence (ETAS) model is its de-facto benchmark, and under fair, prospective, CSEP-style testing no machine-learning temporal point process has been shown to beat a well-fit ETAS. This preprint describes a complete, honesty-first OEF system and, on it, a leakage-free multi-region CSEP testbed built to answer, rather than assert, four questions. The system fits a regime-tiled space-time ETAS with a full hygiene pipeline (rolling magnitude of completeness, moment-magnitude homogenization, dual-catalog declustering, propagated uncertainty), against a mandatory adaptive smoothed-seismicity Poisson null and a transparent Reasenberg-Jones fallback, calibrated by isotonic regression with a genuine epistemic-plus-aleatory uncertainty triad, and it emits both gridded and catalog-based forecast representations so over-dispersion is scored honestly. Skill is established only by winning CSEP comparison tests (information gain per earthquake, IGPE, in nats) against both baselines, under a strict forecast clock that engineers against five leakage modes, with pre-registered ship rules and shuffled-label negative controls. Four artifact-backed findings result, each with its honest negative kept on the record. (i) ETAS adds skill over the stationary null only where there is triggering to exploit (Japan +0.072 nats at one day; low-seismicity interiors exactly 0.0), quantifying the high-versus-low-seismicity bias. (ii) A convex log-score-optimal stack of ETAS-family variants earns a real, significant global gain (+0.011 over 751 events) that does not generalize to the canonical active margins; under the pre-registered rule it is a no-ship, and a temporally adaptive variant is refuted as a multiple-comparison artifact. (iii) The binding one-to-seven-day consistency failure is count over-dispersion, not spatial shape: a catalog-based number test passes a window the Poisson number test rejects, and the frozen-intensity forecast systematically under-counts by about 28 percent because it omits within-window secondary triggering. (iv) The value of a GNSS-strain geodetic context covariate, added to a Hawkes-structured neural point process, is horizon-dependent: it does not beat ETAS at the one-to-seven-day operational horizon (mean IGPE -0.053 over eight weekly windows) but beats it robustly at thirty days (global +0.115 over 2166 events, positive in 9/10 windows and in every high-seismicity region), because the calibrated model deploys a time-flat geodetic background rather than triggering. We conclude that base tiled ETAS is at or near the practical ceiling for mean-rate one-to-seven-day IGPE over global M>=5, and that the honest place for a geodetic covariate is a longer, background-dominated outlook. The manuscript includes the full methodology and equations, the leakage-free evaluation protocol, complete per-region and per-experiment results, five purpose-driven figures generated deterministically from the committed artifacts, and appendices. This is an independent research and education tool; it is not an operational alarm system and must not be used for life-safety decisions. Code, configurations, provenance manifests, and all committed artifacts (MIT): https://github.com/fsantibanezleal/CAOS_SEISMIC . Static forecast viewer: https://seismic.fasl-work.com .

View free PDFSource page

Related papers

openalexZenodo (CERN European Organization for Nuclear Research)2026-07-24

Code and data for: Leakage-audited machine learning versus ETAS for earthquake forecasting in the Sea of Marmara

Basri Kerem Alhan, Kenessary Khabat

Code, processed data products, configuration, and results artifacts for "Machine learning versus ETAS for earthquake forecasting in the Sea of Marmara: a leakage-audited negative result and a closed-form scoring artifact" (Alhan & Khabat, submitted to Seismica). Version 1.2.0 acc…

View free PDFSource page
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-26

Supplementary Materials for "Beyond grades: multi-target deep learning for early academic risk detection"

Miguel Angel Rodríguez Ortiz, Luis Anido-Rifón, Pedro C. Santana-Mancilla

This repository contains the supplementary materials associated with the article: “Beyond Grades: Multi-Target Deep Learning for Early Academic Risk Detection” The materials support the transparency, reproducibility, interpretability, and pedagogical analysis of the leakage-free…

View free PDFSource page
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-26

The Value of Data in the Pre-AI Era | 前AI时代的数据价值

WU, JEFFI CHAO HUI

《前AI时代的数据价值》简介 本文作者巫朝晖(Jeffi Chao Hui Wu)基于跨越四十年的个人实证记录与多领域系统构建实践,系统性地提出了“前AI时代数据”这一核心学术概念,并将其严格界定为:2022年底生成式人工智能(Generative AI)以低成本、高仿真度大规模介入公共互联网内容生产之前,由真实人类大脑、真实的物理环境与真实的社会交互所产出的原始数字记录。作者认为,在当今海量AI生成文本、影像与逻辑推演泛滥的“数字噪音膨胀”时代,此类数据正从传统档案升格为兼具唯一性与不可复制性的稀缺基础资源,其价值遵循严格的“数据年龄”准则——即形成时…

View free PDFSource page
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-26

A co-measurement gap: no standard paradigm jointly tracks a felt valence signal and an independent viability endpoint under within-individual perturbation (position note)

Hiroaki Aizawa

POSITION / RESEARCH-GAP NOTE. Version 1.1 (2026-07-26); first published 2026-07-18. This note establishes no empirical result and must not be cited as one. It argues a structural gap and gives the design specification of the study that would close it. — The gap — A family of acco…

View free PDFSource page
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-24

Controllable Generative AI: A Technical Review of Model Parameters for Hallucination Mitigation, Fine-Tuning, and Replicable Model Construction

Prateek Dutta

Generative artificial intelligence (GenAI) systems, particularly large language models (LLMs), expose a dense and interacting set of architectural, optimization, and inference-time parameters that jointly determine factual reliability, adaptation quality, and reproducibility. Thi…

View free PDFSource page