Forskningsradar

Science Journals

Peer-reviewade publikationer — 54780 artiklar

Co-design approach to aperture masking for imaging through atmospheric turbulence
arXiv:2607.09265v1 Announce Type: new Abstract: Aperture masking interferometry is a technique originally designed to alleviate the influence of atmospheric turbulence on images recorded on ground-based telescopes. In this communication, we explore the optimization of the aperture mask by an optical/digital co-design approach in order to obtain diffraction-limited images of relatively bright objects imaged through turbulence. We show that, with a few simplifying assumptions, it is possible to express the Mean Square Error of the restored image as a function of the chosen mask, of the spatial Power Spectral Density of the observed object and of the noise level, without actually computing any image. This allows us to optimize the aperture mask with a reduced computing cost. We also implement a multi-frame myopic algorithm to estimate jointly the observed object, the piston and the tip-tilt in front of each sub-aperture, and check by simulations that the aperture masks obtained indeed allow a satisfactory image reconstruction.
Risk-Aware General-Utility Markov Decision Processes
arXiv:2607.09298v1 Announce Type: new Abstract: We study general-utility Markov decision processes (GUMDPs) with risk-aware objectives. In this framework, an agent aims to optimize a risk measure of the distribution of objective values, where the objective function depends on the frequency of visitation of states induced by the agent's policy. First, we motivate, propose, and formalize risk-aware GUMDPs, which enable agents and decision makers to trade off expected performance by risk aversion while benefiting from the rich set of objectives that can be cast under the framework of GUMDPs. We focus our attention on the entropic risk measure (ERM). Second, we show how we can solve risk-aware GUMDPs with ERM objectives by resorting to online planning techniques. In particular, we propose an approach based on Monte Carlo Tree Search (MCTS) to provably solve risk-aware GUMDPs up to any desired accuracy. Third, we provide a set of experimental results showcasing that our approach is successful when optimizing for a spectrum of risk-aware behaviors in the context of GUMDPs under diverse tasks (standard MDPs, maximum state entropy exploration, imitation learning, and multi-objective MDPs).
TextileNet: Towards Zero-shot Text-style Segmentation of Manuscripts
arXiv:2607.09299v1 Announce Type: new Abstract: Automatic writer identification systems have progressed remarkably in recent years, yet their deployment in archival paleography remains limited by the scarcity of labeled training data, open scribe sets, and degraded image quality. We present TextileNet, a fully convolutional multi-task network trained exclusively on synthetic data to produce dense pixel-level texture embeddings, which we transfer zeroshot to historical manuscript analysis. As an original contribution to evaluation methodology, we designed a paleographic visual quiz of 80 pair and triplet questions and administered it to a range from lay participants to senior paleographers under strict anonymity, establishing to our knowledge for the first time a human baseline for script-style discrimination on late medieval text. We employ TextileNet embeddings to perform zero-shot retrieval on sub-word granularity for hand and gender identification. Our experimental results help in building the credibility of TextileNet in the paleographic domain, but more than that demonstrate in experimental terms that the question of gender in handwriting needs to be treated with caution.
SHARE: Optimizing Secure Hub Allocation and Routing Efficiency in Payment Channel Networks
arXiv:2501.04236v3 Announce Type: replace Abstract: Payment channel hub (PCH), by leveraging a powerful hub to reliably provide off-chain payment services, offers an effective enhancement to payment channel networks (PCNs). However, existing approaches typically rely on a single hub to relay transactions and provide relationship anonymity between participants. This design lacks flexibility under high-frequency transaction scenarios and fail to adequately balance the security of off-chain payments with PCH efficiency. Moreover, current PCNs often adopt source routing, where each transaction path is predetermined without considering the dynamic distribution of large-scale payment requests, leading to load imbalance and even transaction deadlocks. To address these issues, we propose SHARE, a multi-PCH distributed routing scheme based on trusted execution environments (TEE), designed to optimize secure hub allocation and routing efficiency in PCNs. For the multi-hub allocation problem, SHARE balances the management and synchronization costs among participants, and employs mixed-integer linear programming along with supermodular optimization techniques to transform the NP-hard problem into a solvable form, enabling optimal or approximate solutions across various PCN scales. At the routing layer, SHARE integrates global network state with local sender requests to design a TEE-assisted, privacy-preserving distributed routing protocol that dynamically adjusts multipath flow rates, achieving high-throughput and deadlock-free transaction forwarding. We formally prove the security of the SHARE protocol under the universally composable framework. Experimental results demonstrate that SHARE achieves a 43.6% improvement in transaction success ratio and an over 181.5% enhancement in system throughput compared to state-of-the-art PCN solutions, effectively realizing a secure extension of PCNs.
AV-Master: Dual-Path Comprehensive Perception Makes Better Audio-Visual Question Answering
arXiv:2510.18346v3 Announce Type: replace Abstract: Audio-Visual Question Answering (AVQA) requires models to effectively utilize both visual and auditory modalities to answer complex and diverse questions about audio-visual scenes. However, existing methods lack sufficient flexibility and dynamic adaptability in temporal sampling and modality preference awareness, making it difficult to focus on key information based on the question. This limits their reasoning capability in complex scenarios. To address these challenges, we propose a novel framework named AV-Master. It enhances the model's ability to extract key information from complex audio-visual scenes with substantial redundant content by dynamically modeling both temporal and modality dimensions. In the temporal dimension, we introduce a dynamic adaptive focus sampling mechanism that progressively focuses on audio-visual segments most relevant to the question, effectively mitigating redundancy and segment fragmentation in traditional sampling methods. In the modality dimension, we propose a preference-aware strategy that models each modality's contribution independently, enabling selective activation of critical features. Furthermore, we introduce a dual-path contrastive loss to reinforce consistency and complementarity across temporal and modality dimensions, guiding the model to learn question-specific cross-modal collaborative representations. Experiments on four large-scale benchmarks show that AV-Master significantly outperforms existing methods, especially in complex reasoning tasks.
SpikeATac: A Multimodal Tactile Finger with Taxelized Dynamic Sensing for Dexterous Manipulation
arXiv:2510.27048v3 Announce Type: replace Abstract: In this work, we introduce SpikeATac, a multimodal tactile finger combining a taxelized and highly sensitive dynamic response (PVDF) with a static transduction method (capacitive) for multimodal touch sensing. Named for its `spiky' response, SpikeATac's 16-taxel PVDF film sampled at 4 kHz provides fast, sensitive dynamic signals to the very onset and breaking of contact. We characterize the sensitivity of the different modalities, and show that SpikeATac provides the ability to stop quickly and delicately when grasping fragile, deformable objects. Beyond parallel grasping, we show that SpikeATac can be used in a learning-based framework to achieve new capabilities on a dexterous multifingered robot hand. We use reinforcement learning from human feedback to fine-tune the behavior of a policy to modulate force. Our hardware platform and learning pipeline together enable a difficult dexterous and contact-rich task that has not previously been achieved: in-hand manipulation of fragile objects. Videos are available at https://roamlab.github.io/spikeatac/ .
Computing Isomorphisms between Products of Supersingular Elliptic Curves
arXiv:2503.21535v2 Announce Type: replace-cross Abstract: The Deligne-Ogus-Shioda theorem guarantees the existence of isomorphisms between products of supersingular elliptic curves over finite fields. In this paper, we present methods for explicitly computing these isomorphisms in polynomial time, given the endomorphism rings of the curves involved. Our approach leverages the Deuring correspondence, enabling us to reformulate computational isogeny problems into algebraic problems in quaternions. Specifically, we reduce the computation of isomorphisms to solving systems of quadratic and linear equations over the integers derived from norm equations. We develop $\ell$-adic techniques for solving these equations when we have access to a low discriminant subring. Combining these results leads to the description of an efficient probabilistic Las Vegas algorithm for computing the desired isomorphisms. Under GRH, it is proved to run in expected polynomial time.
GatedLinear: Adaptive Routing of Complementary Linear Bases for Time Series Forecasting
arXiv:2607.09537v1 Announce Type: new Abstract: Time series forecasting requires models to capture diverse, often mutually exclusive, temporal dynamics, from smooth trend continuation to nonstationary drift and strict phase-aligned recurrence. While recent deep learning models have improved accuracy, they typically force these diverse patterns through a single computational backbone governed by fixed algorithmic inductive biases (e.g., self-attention or spectral filtering). This single-mechanism approach often struggles with the profound heterogeneity of real-world series, where different variables and forecast horizons necessitate fundamentally different predictive treatments. To address this, we propose GatedLinear: a lightweight framework that frames forecasting as the adaptive routing of complementary linear bases. GatedLinear leverages a pool of three specialized mechanisms: a global trend-seasonal basis for smooth projection, a difference-based incremental basis for nonstationary drift, and a phase-aligned recurrence basis for explicit cyclic reuse. To dynamically orchestrate these distinct behaviors, we introduce a Tri-Factorized Fusion Gate that disentangles routing decisions into channel-specific preferences, horizon-aware offsets, and phase-indexed biases derived from known future time marks. This design allows the model to perform highly granular, point-wise soft routing across different predictive regimes without stacking computationally heavy neural modules. Experiments on standard benchmarks show that our method achieves state-of-the-art or highly competitive accuracy against recent complex foundational models, while offering explicitly interpretable routing patterns and operating with a substantially smaller parameter footprint.
CUPID: Reconstructing UV Texture Maps for Interpretable Person-of-Interest Deepfake Detection
arXiv:2606.20302v2 Announce Type: replace Abstract: Deepfakes targeting a high-profile individual, known as Person-of-Interest (POI), are a threat to modern democracies and societies. Current POI deepfake detection methods still struggle to combine robustness to post-processing, efficiency and interpretability, key aspects of modern deepfake detectors. In this paper we propose CUPID, a POI video deepfake detector that combines UV texture maps, a facial appearance representation derived from 3D face reconstructions, with the representation learning capabilities of the Masked Autoencoder (MAE). Our method does not require any deepfake videos in its training phase. Moreover, it does not even require including a specific POI in the training set: the combination of UV texture maps extracted from real video frames and the MAE context-guided reconstruction yields a latent space that captures rich and discriminative facial features even for identities unseen during training. In the testing phase, the embeddings extracted from a query video depicting the POI can be matched against pristine reference videos to assess the video authenticity. Furthermore, operating in the UV space naturally provides an additional layer of interpretability. Specifically, we can extract decoded residual maps that highlight which facial regions of a test video deviate most from the identity representation of the corresponding POI. Experiments on four deepfake datasets show that CUPID outperforms the current state of the art on most datasets and achieves the best overall robustness against strong downscaling and compression, while also providing substantially faster inference. Our experimental code will be released at https://github.com/polimi-ispl/CUPID.
The Recurrent Nova TCrB: A Method for Predicting the Next Eruptive Event in Nova Cycles
arXiv:2607.05200v2 Announce Type: replace-cross Abstract: The symbiotic recurrent nova (SyRNe) TCrB (T-Coronae Borealis) is perhaps the most famous example of the group of known four symbiotic nova systems, for which at least two previous nova eruptions are known and accurately recorded: in 1866 and 1946. B.E. Schaefer (2023) has identified the dates of two other previous eruptive events: in 1787 and 1217. Its peak magnitude V was found to be 2.50+-0.10, making it the brightest of its class. In its quiescent phase, TCrB is the brightest of all known novae, with a mean magnitude of 9.8. Careful studies, especially photometric ones, have led to different predictions for the next nova eruption, taking into account the recurrence times extrapolated from previous eruptions, which an average value about 80 years. Schaefer, in particular, has produced various forecasts, including one made in 2023 based on B and V light curves for the period: 1842-2022, which predicts the next nova eruption should occur in 2025.5+-1.3 and is therefore still valid today. Using the Schaefer's remarkable work in accurately determining the key physical parameters that drive the dynamics of the TCrB symbiotic system, we propose here a new semi-empirical method to derive the variations in the nova recurrence time, Trec, and thus obtain a forecast estimate for the next eruption for the date: 26-Feb-2027, which is currently compatible and consistent with the observed behavior and would also justify the supposed "delay" for the next event of this nova as commented by various authors.
On the Complexity of Low-Rank Matrix Signing and Entrywise Power Matrix Factorization
arXiv:2607.04875v2 Announce Type: replace Abstract: Given a nonnegative matrix $X$, a factorization rank $r$ and {a positive integer $p$}, entrywise power matrix factorization (EPMF) looks for a low-rank matrix $X_r$ such that $X = |X_r|^{\circ p}$ (exact case) or $X \approx |X_r|^{\circ p}$ (approximate case), where $(\cdot)^{\circ p}$ denotes the componentwise exponent. EPMF includes the modulus model ($p=1$) and componentwise square factorization ($p=2$) as special cases, the latter being closely related to the square root rank. We analyze the computational complexity of the exact decision problem and the Frobenius-norm approximation problem, and establish a complete complexity landscape. In the exact case, we show that EPMF is equivalent to the combinatorial problem of flipping the signs of the entries of a given matrix $X$ to obtain a rank-$r$ matrix, which we refer to as the low-rank matrix signing (LRMS) problem. We first show that LRMS, and hence exact EPMF, is strongly NP-hard, improving a weak NP-hardness result for the square-root-rank (Math. Prog., 2015). We then show that LRMS can be solved in polynomial time when $r$ is fixed. Moreover, when the rank $r$ is part of the input, we show that for generic matrices the algorithm is fixed-parameter tractable (FPT) in the parameter $r$; in fact, the running time is fixed-parameter linear in the number of entries of the input matrix. In the approximate case using the Frobenius norm as an error measure, we show that EPMF is NP-hard, already when $r=2$, the smallest nontrivial case.
Sequential Multi-Step Nanoimprint Lithography Fabrication of Zero-Mode Waveguide Nanoaperture Arrays to Enhance Single Molecule Fluorescence Detection
arXiv:2607.09141v1 Announce Type: new Abstract: Zero-mode waveguides (ZMWs) enable single-molecule fluorescence detection at micromolar concentrations by confining light to nanoscale volumes, overcoming the diffraction limit of confocal microscopy. However, their widespread adoption is hindered by high fabrication costs and limited throughput of traditional methods like focused ion beam or electron-beam lithography. Here, we introduce a scalable cost-effective approach using sequential nanoimprint lithography (NIL) combined with hydrofluoric acid etching to fabricate ZMW arrays with tunable diameters from a single initial master. In our sequential nanoimprint approach, each stamped NIL output serves as a master for the next nanoimprint generation. By leveraging the shrinkage of sol-gel nanopatterns during annealing, we achieve a cumulative diameter reduction from 230 nm to 115 nm over four successive imprints, all based on the same initial master. The resulting ZMWs exhibit detection volumes reduced by up to 1000-fold and fluorescence enhancement exceeding 16x, achieving performance comparable to state-of-the-art focused ion beam-fabricated devices. Eliminating the need for multiple master structures significantly expands the scalability of nanoimprint lithography approaches. By lowering the nanofabrication barriers and making ZMW arrays more accessible, the sequential NIL method paves the way towards broader adoption of nanophotonic devices in single-molecule biophysics, biosensing, and surface patterning applications.
Programming over Thinking: Efficient and Robust Multi-Constraint Planning
arXiv:2601.09097v4 Announce Type: replace Abstract: Multi-constraint planning involves identifying, evaluating, and refining candidate plans while satisfying multiple, potentially conflicting constraints. Existing large language model (LLM) approaches face fundamental limitations in this domain. Pure reasoning paradigms, which rely on long natural language chains, are prone to inconsistency, error accumulation, and prohibitive cost as constraints compound. Conversely, LLMs combined with coding- or solver-based strategies lack flexibility: they often generate problem-specific code from scratch or depend on fixed solvers, failing to capture generalizable logic across diverse problems. To address these challenges, we introduce the Scalable COde Planning Engine (SCOPE), a framework that disentangles query-specific reasoning from generic code execution. By separating reasoning from execution, SCOPE produces solver functions that are consistent, deterministic, and reusable across queries while requiring only minimal changes to input parameters. SCOPE achieves state-of-the-art performance while lowering cost and latency. For example, with GPT-4o, it reaches 93.1% success on TravelPlanner, a 61.6% gain over the best baseline (CoT) while cutting inference cost by 1.4x and time by ~4.67x. Code is available at https://github.com/DerrickGXD/SCOPE.
Matched Generators for the Karhunen--Lo\`eve Transform: A Double-Commutator Eigenvalue Theory
arXiv:2607.08788v1 Announce Type: cross Abstract: The Karhunen--Lo\`eve transform (KLT) diagonalizes the covariance of a second-order process and is optimal for mean-square truncation. Which classical transform it reduces to is governed by the symmetry commutant of the covariance: when the kernel commutes with a group action, the KLT eigenfunctions are the irreducible representation functions of that group, recovering the Fourier, cosine, Mellin, and spherical-harmonic systems. We study the inverse question. Given a covariance $R$ and a finite-dimensional space of candidate generators, the generator nearest to commuting with $R$, the minimizer of $\delta(A,R)=\|[R,A]\|_F/(\|R\|_F\|A\|_F)$, is the smallest-eigenvalue solution of a double-commutator eigenvalue problem $\mathrm{ad}_R^2(A^\ast)=\lambda A^\ast$, a Hermitian generalized eigenvalue problem of size the number of generators, independent of dimension. The framework recovers hidden transforms as well as classical ones: a variational characterization turns the existence of a commuting generator into a spectral condition, and a tridiagonal commutant-uniqueness result yields the prolate spheroidal, cosine, and discrete orthogonal-polynomial bases as exact recoveries, with matrix-valued extensions, and produces a continuum of transforms interpolating between and beyond the classical families. When symmetry is approximate, the coding penalty of the symmetry-adapted blockwise transform equals the multi-information among the sectors, an exact threshold between the fixed and data-driven transforms. We further give a graph-automorphism characterization of permutation structure, a sequential deflation for non-Abelian symmetry, and stability bounds under estimation error. As an application, the KLT of a two-paradigm covariance is synthesized from its two known generators, without forming the mixed covariance, reaching the full-data transform's compaction from few observations.
Geometric planted matchings in high dimensions: The power of multiple views
arXiv:2607.09026v1 Announce Type: cross Abstract: We study the problem of recovering the correspondence between a collection of $n$ points in $\mathbb{R}^d$ and a noisy, permuted version of those points. In the high-dimensional regime $d=\omega(\log n)$, under a Gaussian model with noise variance $\sigma^2=d/(b\log n)$, prior work identifies $b=2$ as the threshold for almost exact recovery. We prove that this threshold is all-or-nothing: for every fixed $b<2$, no estimator recovers a positive fraction of the matching, and even estimating the matched point cloud in Euclidean distance is asymptotically no better than ignoring the correspondence. On the other hand, we consider a multi-view generalization of the problem where $K$ noisy, independently permuted copies of the same latent point cloud are observed. Here we show that a simple polynomial-time procedure recovers all relative matchings up to $o(n)$ errors whenever $b>K/(K-1)$. Thus multiple views can break the impossibility barrier $b=2$ for the original matching problem: in particular, for $3/2 < b < 2$, the two-view model has no nontrivial recovery, but a third view makes all latent correspondences efficiently recoverable.
Transition Matching Distillation for Fast Video Generation
arXiv:2601.09881v2 Announce Type: replace Abstract: Large video diffusion and flow models have achieved remarkable success in high-quality video generation, but their use in real-time interactive applications remains limited due to their inefficient multi-step sampling process. In this work, we present Transition Matching Distillation (TMD), a novel framework for distilling video diffusion models into efficient few-step generators. The central idea of TMD is to match the multi-step denoising trajectory of a diffusion model with a few-step probability transition process, where each transition is modeled as a lightweight conditional flow. To enable efficient distillation, we decompose the original diffusion backbone into two components: (1) a main backbone, comprising the majority of early layers, that extracts semantic representations at each outer transition step; and (2) a flow head, consisting of the last few layers, that leverages these representations to perform multiple inner flow updates. Given a pretrained video flow model, we first introduce a flow head to the model, and adapt it into a conditional flow map. We then apply distribution matching distillation to the student model with flow head rollout in each transition step. Extensive experiments on distilling Wan2.1 1.3B and 14B text-to-video models demonstrate that TMD provides a flexible and strong trade-off between generation speed and visual quality. In particular, TMD outperforms existing distilled models under comparable inference costs in terms of visual fidelity and prompt adherence. Project page: https://research.nvidia.com/labs/genair/tmd
Preference Conditioned Multi-Objective Reinforcement Learning: Decomposed, Diversity-Driven Policy Optimization
arXiv:2602.07764v2 Announce Type: replace Abstract: Multi-objective reinforcement learning (MORL) seeks to train agents capable of balancing conflicting objectives. While single preference-conditioned policies offer a highly scalable solution, existing approaches remain brittle in practice, frequently failing to recover dense Pareto fronts. We demonstrate that this failure stems from two structural pathologies: destructive advantage cancellation caused by premature Early Scalarization (ES), and representational mode collapse across the preference space. To overcome these bottlenecks, we introduce $D^3PO$, a PPO-based framework that fundamentally reorganizes multi-objective optimization. By preserving per-objective learning signals through a decomposed pipeline and integrating preferences only after trust-region stabilization (Late-Stage Weighting), $D^3PO$ improves credit assignment under conflicting objectives. Concurrently, a scaled diversity regularizer encourages behavioral divergence proportional to preference distance. $D^3PO$ operates entirely within the efficient linear scalarization regime shared by standard deep MORL baselines. By reducing information loss caused due to linear scalarization rather than relying on expensive non-linear utility functions, it suggests that optimization bottlenecks play a significant role. Across available standard benchmarks, including high-dimensional and many-objective environments, $D^3PO$ consistently discovers broader, higher-quality Pareto fronts than prior methods, exceeding state-of-the-art hypervolume and expected utility using a single deployable policy.
Universal Neural Network Based Calibration and Control of Programmable Classical and Quantum Photonic Integrated Processors
arXiv:2607.09301v1 Announce Type: new Abstract: Efficient calibration and control of programmable photonic integrated circuits are fundamental for scaling quantum and classical optical computing processors. While neural network-based models offer an architecture-agnostic solution, existing approaches suffer from limited learning and generalization capabilities due to the many-to-one mapping problem between sets of control signals and optical responses, and biased training datasets derived from uniform current sampling. In this work, we propose a universal calibration and control framework employing tandem neural networks combined with two novel data generation strategies: architecture-aware sampling based on Haar measure principles, and optimized sampling, a physics-agnostic approach utilizing differential evolution. We experimentally validate these methods on 3x3 and 4x4 coherent MZI meshes, demonstrating that our approach addresses the sampling bias inherent in previous works. When evaluated using random unitary matrices, our solution outperforms standard uniform sampling baselines by ~2 bits of precision. Furthermore, we experimentally extend the application of this framework to coherent detection, achieving precise control over both amplitude and phase, and validate its impact on photonic neural network tasks.
Quantifying nanoparticle size effect on the photoacoustic generation efficiency
arXiv:2607.09359v1 Announce Type: cross Abstract: Photoacoustic (PA) signal generation in colloidal suspensions of optically absorbing nanoparticles is dominated by the thermal expansion of water for gold nanoparticles, but remains mostly unexplored for organic nanoparticles. Here, we derive a model where the PA generation efficiency scales with particle size and thermoelastic contrast with water. The model is validated using solid lipid nanoparticles labeled with several BODIPY dyes. This experimental validation paves the way for quantitative PA characterization of nanomaterials and rational design of PA contrast agents.
Sticky Routing: Training MoE Models for Memory-Efficient Inference
arXiv:2607.08780v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models activate only a sparse subset of experts per token, yet consecutive tokens frequently activate different experts -- causing constant weight swapping between slow storage and fast memory on edge devices. Existing remedies are either system-level (caching heuristics) or post-hoc (router fine-tuning), leaving the root cause unchanged during pretraining. We propose StickyMoE, a differentiable routing consistency loss that penalises abrupt expert switches between adjacent tokens, encouraging the router to maintain the same expert assignment across semantically coherent spans. StickyMoE requires no architectural changes, adds a single hyperparameter lambda, and unlike post-hoc methods, allows expert representations and routing decisions to co-adapt from the first training step. Experiments on small-scale MoE language models show that StickyMoE reduces the expert switch rate by up to 60% with less than 4% perplexity degradation, Pareto-dominating post-hoc fine-tuning on the quality-locality frontier. Routing temporal locality is most efficiently instilled at training time.
Performance and User Response of Android's Smartphone-Based Alerts in the 2025 Marmara Ereglisi Earthquake
arXiv:2607.08975v1 Announce Type: new Abstract: This study presents a comprehensive evaluation of Googles Android Earthquake Alert (AEA) system during the Mw 6.2 Marmara Ereglisi, Turkiye earthquake. AEA detected the event 5.31 seconds after its initiation, alerting over 16 million users. Warning times for weak shaking (MMI III) reached up to 150 seconds, with a median of 56 seconds. While near-source warning windows were shorter, the system achieved 90% true positives and 99% precision overall. The high density of the phone network enabled faster detection than traditional stations, even for this offshore epicenter. Feedback data shows AEA recipients were highly likely to take protective actions, such as drop, cover, and hold on, or warn others. Timely alerts substantially increased user engagement, perceived usefulness, and future trust. These results highlight how crowd-sourced technology and behavioral insights can effectively enhance seismic resilience on a massive scale.
Potential Functions as Types
arXiv:2607.08547v2 Announce Type: replace Abstract: Amortized analysis can be framed from the physicist's view, amenable to manual verification in dependent type theory using potential functions, and the banker's view, amenable to automated inference in substructural type theory using type-level credit annotations. In this work, we synthesize these perspectives in Calf, a dependent type theory cost verification. From the physicist's view, we present a fracture and gluing theorem that renders every type as containing a fusion of an abstraction function and a potential function. By construction, every program between two such types must preserve abstraction, to facilitate modularity of behavior, and conserve potential, to facilitate modularity of cost. Incorporating the banker's view, we synthetically construct type operators for credits and debits. We then define Giralf, a graded substructural dependent type theory for programming with credits and debits, which is semantically interpreted as a sub-language of Calf. Finally, we adapt an inference algorithm to transform a limited class of Calf programs into Giralf counterparts, automating the cost analysis of common algorithms in Calf.
An Incremental Sampling and Segmentation-Based Approach for Motion Planning Infeasibility
arXiv:2501.11434v3 Announce Type: replace Abstract: We present a simple and easy-to-implement algorithm to detect plan infeasibility in kinematic motion planning. Our method involves approximating the robot's configuration space to a discrete space, where each degree of freedom has a finite set of values. The obstacle region separates the free configuration space into different connected regions. For a path to exist between the start and goal configurations, they must lie in the same connected region of the free space. Thus, to ascertain plan infeasibility, we merely need to sample adequate points from the obstacle region that isolate start and goal. Accordingly, we progressively construct the configuration space (initially assumed to be entirely free) by sampling from the discretized space and updating the bitmap cells representing obstacle regions. Subsequently, we partition this partially built configuration space to identify different connected components within it and assess the connectivity of the start and goal cells. We illustrate this methodology on five different scenarios with configuration spaces having up to 5 degrees-of-freedom (DOF). Additionally, we discuss further optimizations designed to significantly accelerate the proposed algorithm. The scalability of our approach to higher-dimensional configuration spaces is also examined, with experimental demonstrations involving 6-DOF and 7-DOF robots.
RIS-Assisted Downlink Pinching-Antenna Systems: GNN-Enabled Optimization Approaches
arXiv:2511.20305v2 Announce Type: replace Abstract: This paper investigates a reconfigurable intelligent surface (RIS)-assisted multi-waveguide pinching-antenna (PA) system (PASS) for multi-user downlink information transmission, motivated by the unknown impact of the integration of emerging PASS and RIS on wireless communications. First, we formulate sum rate (SR) and energy efficiency (EE) maximization problems in a unified framework, subject to constraints on the movable region of PAs, total power budget, and tunable phase of RIS elements. Then, by leveraging a graph-structured topology of the RIS-assisted PASS, a novel three-stage graph neural network (GNN) is proposed, which learns PA positions based on user locations, and RIS phase shifts according to composite channel conditions at the first two stages, respectively, and finally determines beamforming vectors. Specifically, the proposed GNN is achieved through unsupervised training, together with three implementation strategies for its integration with convex optimization, thus offering trade-offs between inference time and solution optimality. Extensive numerical results are provided to validate the effectiveness of the proposed GNN, and to support its unique attributes of viable generalization capability, good performance reliability, and real-time applicability. Moreover, the impact of key parameters on RIS-assisted PASS is illustrated and analyzed.
ReinforceGen: Hybrid Skill Policies with Automated Data Generation and Reinforcement Learning
arXiv:2512.16861v2 Announce Type: replace Abstract: Long-horizon manipulation has been a long-standing challenge in the robotics community. We propose ReinforceGen, a system that combines task decomposition, data generation, imitation learning, and motion planning to form an initial solution, and improves each component through reinforcement-learning-based fine-tuning. ReinforceGen first segments the task into multiple localized skills, which are connected through motion planning. The skills and motion planning targets are trained with imitation learning on a dataset generated from 10 human demonstrations, and then fine-tuned through online adaptation and reinforcement learning. When benchmarked on the Robosuite dataset, ReinforceGen reaches 80% success rate on all tasks with visuomotor controls in the highest reset range setting. Additional ablation studies show that our fine-tuning approaches contribute to an 89% average performance increase. Finally, ReinforceGen demonstrates significant improvement through fine-tuning in our real-world evaluations. More results and videos are available at https://reinforcegen.github.io.