arXiv:2606.24768v2 Announce Type: replace-cross
Abstract: This paper presents general strong duality results when testing hypotheses by betting against them. A bet is an e-variable for a composite null hypothesis $\Pcal$: a nonnegative random variable $X$ whose expected value is at most one under every $P \in \mathcal P$. Following Kelly, Breiman, Cover, Shafer, and Grunwald et al. (2024), we study a natural minimax \emph{log-optimality} criterion: given a composite alternative $\Qcal$, we characterize the ``GROW value'' $\sup_{X} \inf_{Q} \E_{Q}[\log X]$. This paper generalizes the results of Larsson et al. (2025) from (arbitrary $\mathcal P$ and) simple $\mathcal Q$ to arbitrary $\mathcal Q$. We prove that there always exists a minimizing information-projection pair between the weak-$*$ closures of the convex hulls of arbitrary $\mathcal P$ and $\mathcal Q$, and show that the GROW value for \emph{bounded} e-variables always equals their relative entropy. We also prove a similarly general strong duality for the REGROW criterion with bounded e-variables and arbitrary bounded offsets. Under various assumptions our results extend to unbounded e-variables, and examples show that without any assumptions such extensions fail. Our results are analogous to those in Larsson et al. (2026), swapping tests for bounded e-variables, minimax risk for the GROW criterion, and total variation for relative entropy.
Science Journals
arXiv:2606.25824v2 Announce Type: replace-cross
Abstract: A method for testing the hypothesis of an ultrametric organization of the energy landscape of RNA secondary structures is proposed, based on the analysis of the kinetics of transitions between basins of attraction. The method is based on a kinetic metric constructed from the spectral decomposition of the symmetrized Kramers transition rate matrix and the Mahalanobis distance, and it is not an ultrametric by construction. A computational scheme has been developed that includes automatic filtering of noise eigenmodes and a procedure for analyzing disconnected structure graphs. The performance of the method is demonstrated on a sample of reference and random RNAs.
arXiv:2606.26910v3 Announce Type: replace-cross
Abstract: Photonic integrated circuits offer a scalable and robust route toward quantum information technologies by consolidating photon sources and linear optical networks onto compact, wafer-manufacturable chips. Although silicon photonics has enabled diverse discrete-variable quantum breakthroughs -- spanning multiphoton entanglement, quantum networking, and photonic qubit fusion for quantum computing -- scaling these platforms beyond proof-of-principle demonstrations remains severely constrained by a critical system-level bottleneck. Optical loss compounds rapidly across photon generation, routing, and state analysis, causing multiphoton generation probabilities to plummet exponentially as circuit depth and complexity grow. Here we overcome this rate-loss barrier by demonstrating a monolithic, ultralow-loss silicon nitride (Si$_3$N$_4$) integrated photonic platform engineered for high-performance discrete-variable quantum information processing. Our architecture seamlessly integrates narrowband photon-pair sources with low-loss qubit-fusion circuits and reconfigurable state-analysis interferometers. The on-chip sources prepare Einstein-Podolsky-Rosen (EPR) states with a fidelity of 0.9875(3) and exhibit near-unity photon indistinguishability, yielding a heralded Hong-Ou-Mandel interference visibility of 0.990(6). By executing on-chip fusion of two EPR states, we synthesize and characterize four-photon Greenberger-Horne-Zeilinger states with a record fidelity of 0.943(8) and a fourfold count rate of 27 Hz -- more than two orders of magnitude higher than previous silicon-photonic implementations. Combined with standard CMOS-compatible fabrication on 150-mm-diameter wafers, these results establish ultralow-loss Si$_3$N$_4$ integrated photonics as a definitive, manufacturable platform for deployable, large-scale quantum information processors.
arXiv:2607.01741v2 Announce Type: replace-cross
Abstract: Reinforcement Learning (RL) is a sequential decision-making framework in which an agent learns optimal policies through interaction with an environment by maximizing cumulative rewards. Among RL methods, Bayesian Reinforcement Learning (BRL) addresses common practical challenges related to data scarcity by leveraging prior knowledge about the environment and sequential belief updates. However, most BRL approaches require an explicit likelihood function, which is frequently inaccessible or intractable in real-world settings.
We propose Likelihood-Free Iterated Batch Importance Sampling (LF-IBIS), a novel algorithm for BRL that updates the agent's beliefs online as new interactions become available. By combining Approximate Bayesian Computation with Iterated Batch Importance Sampling, LF-IBIS enables full Bayesian inference in settings where the environment dynamics are not described by an explicit or tractable likelihood. The method yields approximate posterior distributions over both environment parameters and optimal policies, providing a quantification of policy uncertainty useful for a Bayesian treatment of the exploration-exploitation trade-off. We test the method on a simulation study in response-adaptive randomization in clinical trials, where closed-form posteriors enable validation. Additional experiments address settings where the posterior has no closed form and illustrate online policy updating based on the posterior distribution of the optimal policy.
arXiv:2607.02947v2 Announce Type: replace-cross
Abstract: Public official-information request records contain process signals. They can support research, workflow review, and analyst-led assessment. Yet they also mix observed correspondence, platform states, inferred events, and legal outcomes. FOI-O is a reusable process-modelling method and verification infrastructure for Freedom of Information administration. It is a global model that began with the New Zealand Official Information Act and has since iterated through the Australian Commonwealth and New South Wales settings. The NZ package remains the mature reference implementation; the Australian work remains provisional pending empirical evaluation and jurisdiction-specific legal validation. FOI-O models request records, observed correspondence, controlled vocabularies, provenance, review queues, release metadata, and bounded analysis rules. Legally meaningful outcomes require certification by an authorised decision-maker. Its typed operational and semantic contracts are supported by deterministic examples, process models, fixture-only process-mining exports, quality gates, and tests. This article describes the motivation, architecture, ontology-development method, its versioned extraction and review protocol, how related repositories share data and evidence, and how the method may be adapted for Australia, validation evidence, and implementation boundaries. The project is not legal advice, is not an official government publication, and does not certify release, refusal, redaction, charging, extension, transfer, complaint, or publication outcomes.
arXiv:2607.05642v2 Announce Type: replace-cross
Abstract: As quantum networks move toward practical deployment, standardized performance monitoring becomes essential. This article proposes a structured monitoring framework for quantum networks with performance metrics, including quality (e.g., entanglement fidelity, QBER, loss, dark count rate), throughput and latency (e.g., entanglement rate, waiting time), timing (e.g., coincidence window, production and coincidence jitter), and exogenous factors (e.g., temperature, humidity, vibrations). These measurements enable real-time observability, benchmarking, and control, supporting use cases such as fault diagnosis, adaptive timing, and entanglement routing. Additionally, we implement a non-invasive prototype environmental monitoring system integrated with the quantum network infrastructure at Oak Ridge National Laboratory, demonstrating practical feasibility of live data collection and alert generation. Furthermore, we discuss the challenges of real-time monitoring and the trade-offs between observability and system performance. This work establishes a foundation for developing advanced quantum network monitoring systems and lays the groundwork for future autonomous control and quantum software-defined networking.
arXiv:2607.08793v3 Announce Type: replace-cross
Abstract: Sepsis is a leading cause of mortality, yet optimal treatment policies remain contested. Existing reinforcement learning (RL) approaches learn fixed strategies for sepsis treatment, limiting adaptability to changing clinical objectives during inference. We propose EHRMPC, a framework that decouples learning patient dynamics from optimizing treatment by training a patient digital twin in the form of a generative electronic health record (EHR) model. The digital twin predicts clinical trajectories under interventions and enables model predictive control (MPC) to optimize treatments via inference-time planning over simulations. We evaluate EHR-MPC on a multicenter ICU sepsis cohort spanning 8 hospitals in the Mass General Brigham health system using both off-policy importance sampling and on-policy simulation-based evaluation. Relative to RL baselines, EHR-MPC achieves comparable off-policy performance and improved simulation performance. Unlike RL, this work frames sepsis treatment optimization as inference-time control over learned patient dynamics, establishing a general framework for decision making with generative clinical models.
arXiv:2607.09151v2 Announce Type: replace-cross
Abstract: High Performance Computing (HPC) centers are expanding to encompass resources that extend beyond traditional computing. By extending resources to quantum computing, hybrid quantum-classical workflows tackle complex optimization problems that have never before been possible. However, integrating quantum processing units (QPUs) into cloud-native and scientific workload managers presents a unique orchestration challenge: remote quantum devices introduce a second, external queue -- a two-queue problem -- alongside the queue owned by the traditional scheduler. In this work we present Fluence, a Kubernetes scheduler plugin backed by the Fluxion graph-based scheduler, that enables informed, gang-scheduled placement for quantum-classical workloads and custom resources. We evaluate Fluence across three scenarios using AWS Braket simulators and real QPUs. First, under node contention, Fluence's atomic gang placement all but eliminates the wasted node-time that a default scheduler accrues by partially placing gangs. Second, we introduce a synchronization primitive for the two-queue problem in which a single producer submits a shared quantum task while consumers remain scheduling-gated, reducing worker idle time by roughly 5x under short device queues and by orders of magnitude when a real device queue stretched to hours. Third, cost- and queue-aware backend selection pins the cheapest or shortest-queue device satisfying a workload, cutting mean per-run cost by roughly 70x and time-to-result from hours to under a minute. Together, these results show that quantum-awareness can be added to a cloud-native scheduler without modifying user containers.
arXiv:2607.11135v3 Announce Type: replace-cross
Abstract: The tryptophan residues in tubulin $\alpha\beta$-dimers form an ordered aromatic network that has been proposed to support quantum exciton transport even under physiological environmental noise. Existing studies of this system mostly assume white-noise dephasing, but the statistical properties of the protein-solvent bath coupled to tryptophan sites remain uncharacterised under physiological conditions. Here we characterise this fluctuation bath via all-atom molecular dynamics simulations of a solvated tubulin dimer at 310 K, combining high-frequency and long-time trajectories with 10 fs and 10 ps sampling intervals. The resulting autocorrelation of the site-energy fluctuations is tri-exponential, with three well-separated decay modes: sub-100-fs and picosecond fluctuations driven by water dynamics, and a nanosecond mode originating from protein conformational rearrangements. All three modes fall deep within the non-Markovian regime. We further demonstrate that the slow protein mode introduces strong quasi-static disorder, which results in Anderson localisation, while the two fast water modes frequently tune chromophore pairs through resonance, enabling environment-assisted quantum transport (ENAQT). On the full eight-site network, the coloured-noise bath confines excitons predominantly to strongly coupled proximal tryptophan pairs, in marked contrast to the more uniform delocalisation predicted by the standard white-noise Haken-Strobl model. Our workflow generalises to other pigment-protein systems with solvent-exposed chromophores.
arXiv:2607.13165v2 Announce Type: replace-cross
Abstract: For integer vectors R,S let A(R,S) denote the class of (0,1)-matrices with row sum vector R and column sum vector S. Its interchange graph G(R,S) has A(R,S) as its vertex set, two matrices being adjacent when they differ by a single 2 x 2 interchange. Brualdi conjectured that G(R,S) is Hamiltonian for every R,S. We prove the stronger statement that G(R,S) is maximally Hamiltonian: Hamilton-laceable when bipartite, and Hamilton-connected when not. The proof is a structural induction on the number of matrices in the class, organized by the structure theory of interchange graphs. Deleting inactive lines and splitting invariant positions expresses any class as a Cartesian product, reducing the argument to the prime factors. The bipartite classes are products of complete transposition graphs; we settle them together, without induction, by proving they are paired 2-disjoint-path-coverable and hence Hamilton-laceable, using a recent theorem of Coleman, Fischberg, Gong, Harrington and Wong on paired disjoint path covers. The non-bipartite classes divide into three cases: products assembled from smaller factors, a base of Johnson graphs and small classes, and the large prime classes, treated by a pivot-and-fiber construction whose line quotients are matroid base-exchange graphs. The complete argument has been machine-checked in the Lean 4 proof assistant from first principles together with seven cited results of the literature; the disjoint-path-cover results it imports are themselves proved within the formalization.
arXiv:2607.13291v2 Announce Type: replace-cross
Abstract: Quantum networking testbeds lack a distinct plane for coordinating distributed measurements and collecting experimental data across heterogeneous devices. To address this gap, we present the Measurement Plane, a dedicated plane that complements the data, control, and management planes rather than replacing or extending their pipelines. The contribution is presented as a distributed framework that organizes measurement functions into four layers: application, experiment coordination, capability, and resource agents. Our design separates user workflows from device-specific control.
We implemented the framework as containerized microservices connected through publish--subscribe messaging, and validated it on a two-node quantum networking setup connected by an optical network. The framework successfully coordinated remote nodes to execute coincidence measurement and polarization entanglement distribution experiments with visibility interference of up to 98 percent. This evaluation demonstrated the effectiveness of the framework for supporting complex, distributed quantum experiments, enabling online measurement and feedback, and significantly reducing manual configuration and execution effort.
arXiv:2607.14087v2 Announce Type: replace-cross
Abstract: We prove a stochastic comparison for Gaussian maxima. Let $R$ be an $m\times m$ correlation matrix satisfying $R-\mathbf{1} \mathbf{1}^{\mathsf T}/m\succeq0$, let $X\sim\mathcal{N}(0,R)$, and let $Z_1,\ldots,Z_m$ be independent standard Gaussian random variables. Then $\max_{1\leq i\leq m}X_i \leq_{\mathrm{st}} \max_{1\leq i\leq m}Z_i$, or equivalently, $\mathbb{P}\{X_i\leq c\text{ for every }i\}\geq\Phi(c)^m$ for every $c\in\mathbb{R}$. This comparison resolves the Weak Simplex Conjecture: among $d+1$ equiprobable equal-energy signals in $\mathbb{R}^d$ transmitted over an additive white Gaussian noise channel, the regular simplex maximizes the probability of correct maximum-likelihood decoding at every signal-to-noise ratio. It also proves the inequality asserted by the Simplex Mean Width Conjecture and gives an exact formula for the largest number of equiprobable messages that can be sent at prescribed energy and error probability by a deterministic no-feedback AWGN code under a per-codeword energy constraint. The proof combines a Gaussian product inequality for log-concave functions with an adaptive tilting argument that makes the inequality applicable to the one-sided threshold events defining the maximum. A lean formalization of this argument is available at https://github.com/abhmul/weak-simplex-conjecture-lean.
arXiv:2607.14089v2 Announce Type: replace-cross
Abstract: Strained membrane resonators have emerged as a promising platform for optomechanical accelerometry; however, the desired combination of low frequency and high $Q$-mass product requires a rethinking of their dissipation dilution engineering. Applying Bayesian optimization to a Si$_3$N$_4$ membrane, we discover a class of sail-like trampoline resonators in which the frequency is decreased by an order of magnitude while preserving the $Q$-mass product. We demonstrate centimeter-scale sails with kHz frequencies, $Q\sim10^7$ and $Q\times\text{mass}\sim$ 10 g. Vertically integrating a 7 kHz device with a nanoribbon, we realize a monolithic cavity optomechanical accelerometer with a room temperature thermal noise of $40\;\text{n}g_0/\sqrt{\text{Hz}}$, sufficient to resolve $\mu g_0/\sqrt{\text{Hz}}$ ambient vibration over a bandwidth of 4 kHz with a displacement imprecision of $10^{-14}\;\text{m}/\sqrt{\text{Hz}}$. Cryogenic arrays of sail membranes may be attractive for new physics searches and distributed quantum sensing experiments.
arXiv:2607.14637v2 Announce Type: replace-cross
Abstract: For close binaries and star-planet systems, tidal interactions mediate the energy transfer between the orbital motion and the internal flows of the bodies involved, thus playing a central role in their evolution. For equilibrium tides, the associated energy transfer is commonly modeled through an effective viscosity acting on the tidal flow. However, the scaling of viscous dissipation efficiency with tidal frequency $\omega_\text{T}$ remains debated, particularly when $\omega_\text{T}$ greatly exceeds the convective eddy turnover frequency $\omega_\text{c}$. Previous numerical studies have addressed this issue by subjecting a turbulent convective flow to an oscillating background shear mimicking equilibrium tides. In this work, we adopt a novel three-layered convective box -- designed to represent a stellar convection zone sandwiched between two stable layers -- driven by an external periodic forcing. We quantify tidal dissipation efficiency by the forcing power on the flow in steady state. Our results yield a shallower scaling of tidal power per unit mass with $\omega_\text{T}$ than reported in earlier shear-flow simulations. This scaling is consistent with the prediction by \cite{Terquem2021}, suggesting that the effective turbulent viscosity depends only weakly on $\omega_\text{T}$, although our simulations are restricted to $\omega_\text{T}\lesssim 10\omega_\text{c}$. Moreover, we find no evidence of inverse energy transfer (or ``negative viscosity''), a phenomenon observed in some prior shear-flow simulations. We further investigate the influence of rotation within the same local framework. Slow rotation ($\Omega\lesssim \omega_\text{T}$) tends to enhance the tidal power, whereas fast rotation ($\Omega\gtrsim\omega_\text{T}$) significantly suppresses it. We discuss the limitations of our approach and the broader implications of our findings.
arXiv:2607.15159v2 Announce Type: replace-cross
Abstract: We construct a new family of Calderbank-Shor-Steane (CSS) codes using the generator and parity-check matrices of Low-Density Generator Matrix (LDGM) codes, with row operations applied to both matrices in order to achieve the desired quantum rate. Decoding is performed in an iterative manner, by applying message passing over the associated graph, and discrete Density Evolution (DDE) is used to optimize performance in the depolarizing channel. The proposed construction offers high flexibility and easiness in the design, producing quantum codes that possess excellent error correction capabilities. By properly designing the structure of the code, we are able to control and bound the weight of the stabilizer generators to a small value, which results in codes particularly well suited for fault-tolerant quantum computation. At the same time, these codes achieve very good performance in terms of error correction capability.
arXiv:2607.15700v2 Announce Type: replace-cross
Abstract: Quartet correlations in neutron-rich Te isotopes are investigated within the quartet Bardeen-Cooper-Schrieffer (BCS) framework. Taking $^{100}$Sn as an inert core, we consider two valence protons and valence neutrons occupying the $2d_{5/2} \oplus 1g_{7/2}$ model space, and solve the quartet BCS variational equations with a charge-independent isovector pairing interaction. The effective pairing strength is constrained from empirical neutron pairing gaps in the Te isotopic chain. We find that the valence quartet number increases as the valence neutron number is enlarged from $N_{\rm val}=2$ to $14$. The same increasing behavior is also found for the condensed quartet component. The proton occupation of the $1g_{7/2}$ orbit is strongly enhanced relative to the conventional like-particle BCS reference and is driven close to the degeneracy-weighted limit. These results suggest that additional valence neutrons enhance the quartet admixture in the correlated quartet BCS state, while redistributing the fixed proton weight from pair-like configurations to quartet configurations.
arXiv:2607.16114v2 Announce Type: replace-cross
Abstract: It is well-known that combinatorial circuits are modeled mathematically by string diagrams in monoidal categories. Given a gate set $\Sigma$, the circuits over $\Sigma$ can be thought of as string diagrams in the free monoidal category generated by $\Sigma$. In this model, circuit semantics are then given by monoidal functors out of this free category. For quantum circuits, this functor is often valued in the category of unitary matrices. This model suffices for concrete quantum circuits, but fails to describe parameterized families of quantum circuits, such as those which arise in the analysis of ansatz circuits. In this paper, we introduce an approach to parameterized circuit semantics, which is based on enriched category theory. We first introduce an abstract categorical construction, and use this to gain new insights on controlled operations and quantum communication. We then study the special cases of Cartesian monoidal parameters and monoidal closed parameters, both endowing the parameterized semantics with useful constructions. We conclude by showing that the monoidal closed case can be used to unify two perspectives on quantum control.
arXiv:2512.08100v2 Announce Type: replace-cross
Abstract: We construct Locally Recoverable Codes (LRCs) with availability $2$ from a family of fibered surfaces. To obtain the locality and availability properties, and to estimate the minimum distance of the codes, we combine techniques coming from the theory of one-variable function fields and from the theory of fibrations on surfaces. When the locality parameter is $r=3$, we obtain a sharp bound on the minimum distance of the codes. In that case, we give a geometric interpretation of our codes in terms of doubly elliptic surfaces. In particular, this provides the first instance of an error correcting code constructed using a (doubly elliptic) K3 surface.
arXiv:2509.00078v2 Announce Type: replace-cross
Abstract: The emergence of large language models (LLMs) has transformed spoken dialog systems, yet the optimal architecture for real-time on-device voice agents remains an open question. While end-to-end approaches promise theoretical advantages, cascaded systems (CSs) continue to outperform them in language understanding tasks, despite being constrained by sequential processing latency. In this work, we introduce ChipChat, a novel low-latency CS that overcomes traditional bottlenecks through architectural innovations and streaming optimizations. Our system integrates streaming (a) conversational speech recognition with mixture-of-experts, (b) state-action augmented LLM, (c) text-to-speech synthesis, (d) neural vocoder, and (e) speaker modeling. Implemented using MLX, ChipChat achieves sub-second response latency on a Mac Studio without dedicated GPUs, while preserving user privacy through complete on-device processing. Our work shows that strategically redesigned CSs can overcome their historical latency limitations, offering a promising path forward for practical voice-based AI agents.
arXiv:2607.11969v2 Announce Type: replace-cross
Abstract: Point-adjustment (PA), for years the default scoring protocol in time-series anomaly detection (TSAD), was shown by Kim et al. (2022) to award near-perfect F1 to random anomaly scores. The field adopted a suite of replacement metrics (PA%K, range-based precision/recall, affiliation precision/recall, and Volume-Under-the-Surface, VUS, ROC/PR). We ask, independently and adversarially, whether these resist no-skill detectors on real benchmarks, and find the answer turns entirely on one overlooked variable: N, the number of random attempts an adversary reports the best of. Under a single honest run (N=1), not one replacement metric is gameable on any of six benchmarks (UCR, SMD, SMAP, MSL, NAB, PSM): a random detector reaches 90% of the best real detector's score on at most 11% of series for affiliation-F1, 5% for the ROC family, and 2% for the PR-based metrics and PA%K. But under best-of-N reporting, the seed-shopping endemic to ML, the metrics split sharply. affiliation-F1 and every ROC-based metric inflate steeply, affiliation crossing gameable (25% of series) by N=3 and reaching 0.98 at the full pool (N=41), the ROC family crossing by N=9-11; the PR-based metrics and PA%K stay near-flat at every N, floored near the anomaly prevalence (the lone exception is NAB at large N). A paired test finds VUS-ROC inflated on 131 series where its sibling VUS-PR is not, and never the reverse. The ROC-vs-PR split follows from the order-statistic behaviour of AUC under extreme class imbalance (a random PR-AUC is floored at prevalence); affiliation inflates by a second route, its extreme single-run leniency (already fragile at N=1). We release a pip-installable stress-test harness, and recommend reporting single-run scores or disclosing N and preferring PR-based metrics, which resist best-of-N inflation on nearly every benchmark.
arXiv:2607.13105v2 Announce Type: replace-cross
Abstract: Dissipative cat qubits exponentially suppress one Pauli error channel with the mean photon number, leaving the conjugate bit-flip error as the dominant failure mode. This strong noise bias makes the full machinery of general quantum error correction unnecessary: a code need only protect against a single error type, and any classical linear code can be promoted to a Clifford stabilizer code that does exactly this. We use this observation to build a \emph{bit-flip-only} quantum Reed--Solomon (RS) code. Starting from the maximum-distance-separable RS code $[7,3,5]$ over $\GF(2^3)$, we binary-expand it to the linear code $[21,9,6]$ over $\GF(2)$ and realize it as a $[[21,9,d_X=6,d_Z=1]]$ bit-flip code whose stabilizers are products of $Z$ operators. Because no phase-flip correction is attempted, the construction discards the redundancy that standard quantum RS codes spend on correcting $Z$ errors -- which a strongly biased cat qubit renders unnecessary -- and yields a shallow Clifford circuit that samples directly in Stim. Errors are decoded by an optimal bounded-distance syndrome-lookup table. We then introduce a \emph{Tornado} architecture: a two-layer concatenation that wraps every position of the outer RS code in an inner distance-three repetition code, yielding a $[[63,9,18]]$ code decoded by a two-stage inner majority vote and outer lookup decoder. Monte-Carlo simulation shows that at a physical bit-flip rate $p=0.1$ the Tornado code reaches a logical error rate $\pL \approx 5.3\times10^{-3}$, below both parent codes, and that its logical error rate scales as $\pL \propto p^{6}$ at low $p$, in contrast to $p^{2}$ for the repetition code and $p^{3}$ for the standalone RS code. We give the exact construction, the error and circuit model, an asymptotic scaling analysis, and an honest account of the overhead cost and single-shot assumptions.
arXiv:2606.15950v2 Announce Type: replace-cross
Abstract: Conformal prediction gives prediction intervals with finite-sample coverage when the data are exchangeable. Many time-indexed datasets are not exchangeable: they have seasons, recurring regimes, changing frequencies, or other forms of structured dependence. This paper studies a simple way to use that structure. We propose spectral adaptive conformal prediction, a method that forms weighted conformal quantiles using local spectral similarity and then updates the target miscoverage level online. The spectral weights choose calibration residuals that look relevant to the current test point. The adaptive update corrects the long-run miss rate when uncertainty changes over time. The theory makes both parts controllable. We give an approximate coverage bound that splits the error into a spectral mismatch term and an effective-sample-size term, prove that kernel spectral weighting never increases the mismatch term relative to uniform weighting, show that a bandwidth of order N^(-1/(d+2)) balances the two terms, and establish an unconditional long-run calibration bound for the adaptive update that holds for every sample path without independence or stationarity. Simulations with recurring regimes and slowly changing frequencies, together with four real-data examples spanning monthly, weekly, and daily U.S. and European series, show when the hybrid method improves on strong adaptive baselines and when it does not, and an effective-sample-size safeguard, computable at prediction time without outcomes, detects and repairs the one observed failure.
arXiv:2605.20807v2 Announce Type: replace
Abstract: Subject-driven text-to-image generation still struggles to preserve high-frequency identity details such as logos, patterns, and text. Existing methods typically operate directly in RGB space, which often leads to detail degradation under substantial edits. We propose a two-stage framework that decouples structure from appearance by first predicting a Canny map and then rendering the final image conditioned on both the source appearance and the predicted structure. To improve text handling, we further introduce a fully automatic pipeline that constructs a 100k-pair text-aware dataset with cross-view textual consistency. Experiments, including GPT-4.1-based evaluation and a knowledge distillation study, show clear gains over selected baselines and suggest that intermediate structural prediction is an effective route for high-fidelity subject-driven generation. Our dataset and code will be made publicly available.
arXiv:2407.16288v4 Announce Type: replace
Abstract: Uncrewed Aerial Vehicles (UAVs) offer agile, cost-effective, and efficient solutions for communication relay networks. However, their modeling and control are challenging, and the mismatch between simulations and actual conditions limits real-world deployment, while maintaining adequate situational awareness remains essential for safe operation. Several studies have proposed integrating UAV operations with immersive digital technologies, such as Digital Twin (DT) and Extended Reality (XR), to overcome these challenges. This paper provides a comprehensive overview of the latest research and developments involving immersive digital technologies for UAVs. We explore the use of Machine Learning (ML) techniques, particularly Deep Reinforcement Learning (DRL), to improve the capabilities of DT for UAV systems, and present a case study of a DT-driven DRL pipeline that couples bidirectional physical-digital synchronization with online recursive least-squares channel calibration for UAV resource allocation. We further present a second case study in which a diffusion-augmented digital twin, kept statistically faithful to the physical swarm by the same online calibration loop, drives multi-UAV velocity coordination. We identify and discuss key research gaps, and propose countermeasures based on Generative AI (GAI), emphasizing the significant role of AI in advancing DT technology for UAVs. Furthermore, we review and discuss how the XR technology can transform UAV operations with the support of GAI, and examine its practical challenges. Finally, we propose future research directions to further develop the application of immersive digital technologies for UAV operation.
arXiv:2607.15758v1 Announce Type: new
Abstract: Vision-Language Model (VLM) agents have advanced zero-shot object-goal navigation, yet single-frame reasoning leaves them without the cross-step behavioral awareness an embodied navigator requires, producing recurring failures such as dead-end stalls, in-room loops, and circuitous approaches to detected targets. Prompt-based remedies inflate token budgets across multi-submodule episodes and still struggle to encode inherently spatial signals such as angles, map cells, and viewpoint coordinates. In this paper, we propose SkillNav, an extensible behavioral skill framework for VLM-based navigation that treats the curiosity value map already maintained by modern VLM navigators as a writable substrate on which composable skills inscribe behavioral memory at zero token cost. Skills are stratified into three tiers by their level of behavioral authority, namely soft scaling for proportional reweighting, lower-bound boost for region-level guarantees, and hard override for threshold-triggered forced actions, and cooperate across tiers under a fixed composition order that establishes a predictable, declared priority among skills. This design turns capability improvement into skill registration: new behaviors plug in without retraining the VLM or disturbing existing skills, opening a path for continual refinement. A minimal prompt channel complements the score-level skills with category-level semantic hints, yielding a dual-representation design in which spatial memory lives on the map and semantic memory in short prompts. Training-free, SkillNav establishes new state-of-the-art SPL across MP3D (25.5), HM3D v0.1 (39.3), and HM3D v0.2 (43.2), improving SPL by up to 6.0 absolute over the strongest prior method, and achieves the highest Success Rate on HM3D v0.1 (69.7) and v0.2 (75.9).