Skip to content
José Luis Delgado
Papers

Papers by research line.

Preprints and manuscripts grouped by research line.

Line 01

ML / AI / Formal Systems

Evaluation records, tool-call traces, proof obligations, model boundaries, state invariants, and audit-ready run histories.

Under review 2026
Score is Not Circuit: Evaluating Mechanistic Localization as a Score-to-Circuit Stack

A study of mechanistic localization as a score-to-circuit pipeline, separating scorer quality from circuit quality across calibration, selection, and evaluation.

Under review 2026
SUBGOALCALC: Certified Decomposition for Lean Theorem-Proving Agents

A certified Lean 4 decomposition layer for theorem-proving agents, with proof obligations, witnesses, dependencies, replay semantics, and schedule selection.

Under review 2026
From Circuits to Claims: Post-Selection Certificates for Mechanistic Interpretability

A claim-centered audit framework for selected circuits in mechanistic interpretability, reporting preservation, localization, minimality, stability, robustness, power, and structural recovery.

Under review 2026
CAGE: Causal Authorization Graph Enforcement for Tool-Using AI Agents

A runtime reference monitor for field-level authorization in tool-using agents, enforcing source-field-sink authority over typed dependency graphs before protected actions execute.

Under review 2026
Auditing Loose Scoring in IFEval: A Human-Adjudicated Study of Verifiable Instruction Following

A human-adjudicated audit of IFEval strict-fail/loose-pass disagreements, separating valid corrections, loose false positives, and ambiguous cases.

Line 02

Crypto / PQC / Security

Measurement artifacts, certificate-chain behavior, parser and import validation, decapsulation-test harnesses, threat models, and operational regression cases.

Manuscript 2026
What Can Verifiable Decapsulation Tests Certify? Pass Bounds and Fault-Recognition Limits for FO-Based KEMs

A black-box analysis of what confirmation-code-augmented Fujisaki-Okamoto decapsulation tests can certify, with pass bounds, list-hit obstructions, and dependency-cone limits.

Preprint 2026
Observability for Post-Quantum TLS Readiness: A Multi-Surface Evidence Framework

A multi-surface framework for post-quantum TLS observability across passive evidence, active probing, certificate chains, registry knowledge, and explicit uncertainty.

Preprint 2026
From Public-Key Linting to Operational Post-Quantum X.509 Assurance for ML-KEM and ML-DSA: Registry-Driven Policy, Mutation-Based Evaluation, and Import Validation

A workflow-centric assurance framework for ML-KEM and ML-DSA in X.509, covering certificate profiles, SPKI representation, private-key import, policy registries, and mutation-based evaluation.

Manuscript 2026
The State-Aware Enrollment Plane of Post-Quantum ACME: Measurement, Regime Shifts, and Hardened Workflows

A state-aware study of post-quantum ACME enrollment, comparing standard issuance, reusable-authorization renewal, certificate-based fast re-enrollment, and hardened workflow variants.

Manuscript 2026
Recovering the Right WebAuthn Account: Intent-Bound Recovery Under Realistic Topologies

A topology-aware WebAuthn recovery model for multi-account and multi-authenticator settings, with intent-bound client behavior, provenance checks, and symbolic corroboration.