Forskningsradar

Science Journals

Peer-reviewade publikationer — 55483 artiklar

The Jacobi Factoring Circuit: Quantum Factoring with Near-Linear Gates and Sublinear Space and Depth
arXiv:2412.12558v4 Announce Type: replace-cross Abstract: We present a compact quantum circuit for factoring a large class of integers, including some whose classical hardness is expected to be equivalent to RSA (but not including RSA integers themselves). Most notably, we factor $n$-bit integers of the form $P^2 Q$ with $\log Q = \Theta(n^a)$ for $a \in (2/3, 1)$ in space and depth sublinear in n (specifically, $\tilde{O}(\log Q)$) using $\tilde{O}(n)$ quantum gates; for these integers, no known classical algorithms exploit the relatively small size of $Q$ to run asymptotically faster than general-purpose factoring algorithms. To our knowledge, this is the first polynomial-time circuit to achieve sublinear qubit count for a classically-hard factoring problem. Our circuit builds on the quantum algorithm for squarefree decomposition discovered by Li, Peng, Du, and Suter (Nature Scientific Reports 2012), which relies on computing the Jacobi symbol in quantum superposition. The technical core of our contribution is a new space-efficient quantum algorithm to compute the Jacobi symbol of $A$ mod $B$, in the regime where $B$ is classical and much larger than $A$. Our circuit for computing the Jacobi symbol generalizes to related problems such as computing the greatest common divisor and modular inverses, and thus could be of independent interest.
Active Sensing with Meta-Reinforcement Learning for Emitter Localization from RF Observations
arXiv:2605.12569v2 Announce Type: replace-cross Abstract: Global navigation satellite system (GNSS) interference poses a serious threat to reliable positioning, especially in indoor and multipath-rich environments where source localization is highly challenging. In this paper, we formulate GNSS interference localization as an active sensing problem and propose a reinforcement learning (RL) framework in which an agent sequentially explores the environment to infer the position of an emitter source from radio frequency (RF) observations acquired with a 2x2 patch antenna. The localization task is modeled as a partially observable decision process, since single-snapshot measurements are often ambiguous under multipath propagation and changing channel conditions. To address this, the proposed framework combines high-dimensional RF sensing with deep RL and recurrent policy learning. We investigate both value-based and policy-based approaches, namely Deep Q-Networks (DQN) and Proximal Policy Optimization (PPO), and study their behavior under domain shift. The approach is evaluated on a simulated dataset generated with the Sionna ray-tracing module, which provides realistic propagation effects and diverse environment configurations. Experimental results show that the proposed method achieves a localization success rate of 80.1, demonstrating the potential of RL for adaptive GNSS interference localization. Overall, the results highlight simulation-assisted training as a promising direction for robust interference localization in challenging propagation environments.
Dynamic Airspace Management for UAVs in Evolving Urban Environments: Collaborative Coordination and Human Safety
arXiv:2607.04825v1 Announce Type: new Abstract: The low-altitude economy is an emerging industry with significant development potential, in which the safety of unmanned aerial vehicle (UAV) operations is a critical challenge. Particularly within complex urban topographies and human-populated environments, UAV airspace management must prioritize collision avoidance and human safety. We propose Pharos, a collaborative multi-UAV airspace management system. Pharos lies between the distributed local perception paradigm and the centralized fine-grained control paradigm. Pharos coordinates the safe parallel execution of UAVs in shared airspace while innovatively accounting for the impact of human fear. Pharos is implemented using the MAPPO algorithm due to its faster convergence and higher rewards than other typical MARL algorithms (HAPPO and HATRPO). To evaluate Pharos, we developed a 3D simulation system using real urban data. Visualization results demonstrate its effective airspace coordination capability. Regarding performance verification, Pharos reduced human fear by 52.72% compared to the benchmark Ipopt. Moreover, we designed spatial entropy as a system evaluation metric to quantify space utilization, which improved performance by 70.82% and 2.03% compared to the benchmarks Ipopt and A-star, respectively. The source code is available at an anonymized repository: https://github.com/pharos-anonymized/source-code.git.
Small-scale dynamo saturation across magnetic Prandtl numbers using the EDQNM closure
arXiv:2607.02743v1 Announce Type: cross Abstract: Small-scale dynamos (SSDs) are believed to be the primary source of magnetic fields in all turbulent astrophysical systems, especially those with weak rotation such as elliptical galaxies and galaxy clusters. The initial kinematic phase of these dynamos is relatively well understood. Here we demonstrate analytically and numerically that, in an appropriate limit, the eddy-damped quasi-normal Markovian (EDQNM) closure for incompressible magnetohydrodynamic turbulence is strictly equivalent to the earlier models of kinematic dynamos. Moreover, it allows the extension of the kinematic dynamo framework to multi-scale turbulent flows and into the nonlinear regime. The EDQNM closure also enables us to explore a wide parameter range which is inaccessible to direct numerical simulations of the SSD. Using nonhelical EDQNM simulations, we identify several asymptotic regimes of nonlinear dynamo action when the system is highly turbulent with fluid Reynolds number $Re \gtrsim 10^6$ for magnetic Prandtl number $Pm > 1$ and magnetic Reynolds number $Rm \gtrsim 10^6$ for $Pm < 1$: 1) the kinematic growth rate approaches a value independent of $Pm$, 2) the saturated magnetic to kinetic energy ratio similarly converges to $\simeq 0.55$ across $Pm$, while the ratio of magnetic to kinetic integral wavenumbers asymptotes to $\simeq 3$. For all $Pm$, we further find strong feedback between magnetic field and velocity field largely via Alfv\'{e}nisation leading to a saturated kinetic and magnetic spectra with almost the same inertial range with a slope of $-3/2$. These findings could provide guidance for future global simulations and for modeling the nonlinear regime of astrophysical systems living in these extreme limits.
Online Modeling and Sequential Convex Programming for Lunar Landing Trajectory Optimization
arXiv:2607.02750v1 Announce Type: cross Abstract: This paper presents a guidance framework for lunar powered descent and landing that combines sequential convex programming (SCP) with real-time online model identification. A nonconvex energy-optimal landing problem is developed and then reformulated into a sequence of convex second-order cone programs (SOCPs) through a change of variables, successive linearization, and a lossless second-order cone relaxation of the thrust direction constraint. An online identification layer, built from a recursive least squares (RLS) filter with exponential forgetting and an exponential moving average (EMA) smoother, estimates unknown gravitational, thrust-scale, and mass-gauging perturbations from noisy navigation measurements and injects a corrected bias term into the dynamics constraint of each convex subproblem at every guidance cycle. Building on this architecture and my prior work in this area, a baseline SCP algorithm and a receding-horizon online SCP algorithm with model identification are developed. Also, I try to explore some theoretical foundations, establishing the losslessness of the convex relaxation, the mean-square stability and convergence of the identification filters, the guaranteed convergence of the SCP iteration, and explicit convergence radius and convergence rate results. Numerical simulations across four perturbation scenarios of increasing complexity are implemented in MATLAB using YALMIP and the ECOS solver. The results show that the proposed online algorithm consistently reduces landing position and velocity error and better tracks the true propellant consumption relative to an uncorrected nominal trajectory, while retaining the predictable convergence and real-time computational properties of convex optimization.
AudioX-Turbo: A Unified Framework for Efficient Anything-to-Audio Generation
arXiv:2606.12555v2 Announce Type: replace Abstract: Audio and music generation based on flexible multimodal control signals is a widely applicable topic, with the following key challenges: 1) a unified multimodal modeling framework, 2) large-scale, high-quality training data, and 3) the prohibitive inference cost of multi-step diffusion sampling. As such, we propose AudioX-Turbo, a unified and efficient framework for anything-to-audio generation that integrates varied multimodal conditions (i.e., text, video, and audio signals) in this work. AudioX-Turbo follows a teacher-student paradigm. The teacher AudioX-Base is built on a Multimodal Diffusion Transformer with a Multimodal Adaptive Fusion module that aligns diverse multimodal inputs for high-fidelity synthesis, and is then distilled into the few-step student AudioX-Turbo via Distribution Matching Distillation adapted to flow matching, complemented by a diffusion-based discriminator for high-quality few-step generation. To support the training of AudioX-Turbo, we construct a large-scale, high-quality dataset, IF-caps-Pro, comprising approximately 9.2M samples curated through a two-stage data collection and annotation pipeline. We benchmark AudioX-Turbo across a wide range of tasks, finding that our model achieves superior performance, especially on text-to-audio and text-to-music generation, while operating at only 4 sampling steps and requiring approximately 25x fewer function evaluations (NFE) than multi-step baselines. These results demonstrate that our method is capable of audio generation under flexible multimodal control, showing efficient and powerful instruction-following capabilities. The code and datasets will be available at https://zeyuet.github.io/AudioX-Turbo/.
Magnetohydrodynamic drag on an oscillating sphere in a rotating spherical cavity
arXiv:2604.10594v2 Announce Type: replace Abstract: The drag on an oscillating sphere is a classical fluid-mechanics problem, yet no existing theory simultaneously accounts for confinement, rotation, viscosity and magnetic fields. We consider a conducting sphere undergoing translational oscillations inside a rotating spherical cavity, modelling confined magnetohydrodynamic flows relevant to planetary interiors and liquid metal experiments. In planetary settings, these motions correspond to the polar and equatorial Slichter modes of Earth's inner core. Existing theories are restricted to separate asymptotic regimes, including viscous drag in bounded fluids (Stokes 1851), rotational effects in inviscid cavities (Busse 1974), and magnetic coupling through oscillatory boundary layers (Buffett and Goertz 1995). We derive a unified asymptotic framework for oscillatory drag in rotating spherical shells with arbitrary confinement and arbitrary electrical conductivity and magnetic permeability contrasts between the inner sphere, fluid shell and outer solid, applicable to both polar and equatorial oscillations. The theory yields expressions for added mass, viscous and electromagnetic drag, with the associated dissipation. It captures viscous pressure corrections, magnetic pressure and tension, Alfven-wave radiation, and magnetohydrodynamic Stokes-Ekman boundary layers. Classical boundary-layer theories (e.g. Ekman layers) emerge as limiting cases, while confined solutions are also derived for the diffusion-dominated bulk regimes, yielding an explicit closed-form solution for the bounded oscillatory Stokes flow and a confined extension of the inductionless theory. Direct numerical simulations validate the analytical predictions across a broad parameter range. The resulting analytical framework provides quantitative predictions of oscillatory coupling, added mass and dissipation in planetary cores, icy-moon oceans and liquid-metal experiments.
Reasoning4Sciences: Bridging Reasoning Language Models to All Scientific Branches
arXiv:2606.01145v5 Announce Type: replace Abstract: While Reasoning Language Models (RLMs) are rapidly emerging as powerful tools for scientific research, their impact is primarily concentrated in "hard science" fields. The slow -- or lack of -- adoption of RLMs in other branches of science is causing a widening gap in research productivity. In this survey, we provide the first comprehensive analysis of RLM adoption across 28 scientific disciplines following the classification used by the European Research Council (ERC), spanning the Social Sciences and Humanities, Physical Sciences and Engineering, and Life Sciences. We examine how RLMs are developed, evaluated, and applied across disciplines. Furthermore, we introduce a maturity-oriented assessment framework based on available domain-specific development and evaluation resources, revealing substantial disparities in RLM maturity that become even more pronounced when only publicly available resources are considered. Finally, we highlight current implementation paradigms that are gaining popularity across disciplines, current challenges, and future directions in enabling RLM adoption across science.
Day-Ahead Electricity Price Forecasting Using a Multivariate Group Lasso Method
arXiv:2605.27781v2 Announce Type: replace-cross Abstract: Electricity price signals in modern power systems exhibit complex dependence structures that render forecasting inherently challenging. Our analysis of real-world pricing signals from the California Independent System Operator (CAISO) reveals complex temporal group effects, whereby the influence of explanatory variables on electricity prices persists across consecutive blocks of time due to underlying economic and operational drivers. In response, we propose a multivariate statistical method based on a Group Lasso formulation to forecast the vector of day-ahead electricity prices, by leveraging multi-feature temporal group effects. Our approach is evaluated on two full years of electricity prices from CAISO, demonstrating considerable improvements in point and probabilistic forecast metrics compared to a wide array of statistical and deep learning methods. Theoretical and empirical analyses confirm the effectiveness of the proposed approach in modeling realistic group effects, maintaining both interpretability and low computational complexity. When retrospectively evaluated on test data from a recent international electricity price forecasting challenge, the proposed method ranked in second place, despite having access to significantly less information than competing approaches. Finally, the proposed method is independently validated against two operational electricity price forecasting systems in CAISO, demonstrating competitive predictive performance and practical relevance.
Partial Reductions for Kleene Algebra with Linear Hypotheses
arXiv:2601.14114v2 Announce Type: replace Abstract: Kleene algebra (KA) is an important tool for reasoning about general program equivalences, with a decidable and complete equational theory. However, KA cannot always prove equivalences between specific programs. For this purpose, one adds hypotheses to KA that encode program-specific knowledge. Traditionally, a map on regular expressions called a reduction then lets us lift decidability and completeness to these more expressive systems. Explicitly constructing such a reduction requires significant labour. Moreover, due to regularity constraints, a reduction may not exist for all combinations of expression and hypothesis. We describe an automaton-based construction to mechanically derive reductions for a wide class of hypotheses. These reductions can be partial, in which case they yield partial completeness: completeness for expressions in their domain. This allows us to automatically establish the provability of more equivalences than what is covered in existing work.
Open-Set Source Tracing as Compositional Factors via Structured Prototypes
arXiv:2607.03134v1 Announce Type: cross Abstract: Recent research expands beyond binary anti-spoofing with the emergence of Source Tracing, the task of identifying the specific generative origins of synthetic speech. However, current research often equates a "source" with its generative architecture. We propose redefining a source as a compositional tuple of Architecture, Training Data, and other training factors affecting the generated speech. We propose a framework using Structured Orthonormal Prototypes to minimize class overlap and intra-class variance. Our Subspace Partitioning strategy splits the embedding into architecture and data subspaces, while a residual subspace captures stochastic variability, enabling "compositional generalization" for novel factor combinations. This approach improves performance for partially seen sources and maintains robustness in fully open-set scenarios. MLAAD evaluations for Few-Shot open-set Identification show our approach significantly outperforms angular-margin baselines.
An Intervention-Based Framework for Shortcut Diagnosis in Spoofing Countermeasures
arXiv:2607.03150v1 Announce Type: cross Abstract: While deepfake audio detection systems achieve high performance in controlled benchmarks, their reliability often diminishes in the wild. Prior work shows that dataset-specific artifacts contribute to this gap. Yet, systematic tools to identify which acoustic properties a model exploits as shortcuts remain limited. We propose an intervention-based diagnostic framework, grounded in a directed graphical model, that formally distinguishes confound-driven shortcut dependencies from legitimate domain shift. We operationalise this through controlled acoustic perturbations targeting non-speech structure, spectral content, and signal energy, complemented by corpus-level distributional analysis. Evaluating XLS-R-300M with RawGAT-ST across ASVspoof challenges datasets, we quantify model sensitivity to specific intervention types. Results reveal that non-speech interventions produce the largest performance shifts, confirming non-speech intervals as a dominant shortcut.
No Reliable Evidence of Self-Reported Sentience in Small Large Language Models
arXiv:2601.15334v2 Announce Type: replace Abstract: Whether language models possess sentience has no empirical answer. But whether they believe themselves to be sentient can, in principle, be tested. We do so by querying several open-weights models about their own consciousness, and then verifying their responses using classifiers trained on internal activations. We draw upon three model families (Qwen, Llama, GPT-OSS) ranging from 0.6 billion to 70 billion parameters, approximately 50 questions about consciousness and subjective experience, and three classification methods from the interpretability literature. First, we find that models consistently deny being sentient: they attribute consciousness to humans but not to themselves. Second, classifiers trained to detect underlying beliefs - rather than mere outputs - provide no clear evidence that these denials are untruthful. Third, within the Qwen family, larger models deny sentience more confidently than smaller ones. These findings contrast with recent work suggesting that models harbour latent beliefs in their own consciousness.
Kwai Summary Attention Technical Report
arXiv:2604.24432v2 Announce Type: replace Abstract: Long-context ability, has become one of the most important iteration direction of next-generation Large Language Models, particularly in semantic understanding/reasoning, code agentic intelligence and recommendation system. However, the standard softmax attention exhibits quadratic time complexity with respect to sequence length. As the sequence length increases, this incurs substantial overhead in long-context settings, leading the training and inference costs of extremely long sequences deteriorate rapidly. Existing solutions mitigate this issue through two technique routings: i) Reducing the KV cache per layer, such as from the head-level compression GQA, and the embedding dimension-level compression MLA, but the KV cache remains linearly dependent on the sequence length at a 1:1 ratio. ii) Interleaving with KV Cache friendly architecture, such as local attention SWA, linear kernel GDN, but often involve trade-offs among KV Cache and long-context modeling effectiveness. Besides the two technique routings, we argue that there exists an intermediate path not well explored: {Maintaining a linear relationship between the KV cache and sequence length, but performing semantic-level compression through a specific ratio $k$}. This $O(n/k)$ path does not pursue a ``minimum KV cache'', but rather trades acceptable memory costs for complete, referential, and interpretable retention of long distant dependency. Motivated by this, we propose Kwai Summary Attention (KSA), a novel attention mechanism that reduces sequence modeling cost by compressing historical contexts into learnable summary tokens.
Volition-Guarded Multiagent Atomic Transactions: Describing People and their Machines
arXiv:2604.25596v2 Announce Type: replace Abstract: Formal models for concurrent and distributed systems describe machines; the people who operate them are either ignored or treated as external environment. Yet, key distributed systems -- notably grassroots platforms -- include people operating their personal machines (smartphones), and their faithful description must include the states of both people and machines and how they jointly effect system behaviour. Here, we propose volition-guarded multiagent atomic transactions -- executed atomically by machines and guarded by their people's volitions -- as a novel mathematical foundation for specifying systems consisting of people operating machines. Each agent's state consists of a volitional state and machine state; a transaction is enabled when the machine precondition holds and the guarding persons are willing. For example, befriending two people is guarded by both; unfriending, by either; voluntary swap of coins and bonds is guarded by both parties, while a payment is guarded by the payer. We develop the mathematical machinery to express safety and liveness of platforms specified in this framework, to implement one platform by another, and for an implementation to be resilient to faults; and provide example specifications of two grassroots platforms: social networks, and coins and bonds. These specifications are then used by AI to derive working implementations. We employ here a novel and simpler definition of `grassroots' that better captures the informal notion -- multiple instances can form and operate independently, yet may coalesce -- and show that the platforms specified here are grassroots under the new definition. We further introduce \emph{volitionally grassroots} protocols, in which two groups can become connected only by mutual consent -- the first transaction coupling them must be willed by a member of each -- and show that both platforms are volitionally grassroots.
Unraveling the temporal dependence of ecological interaction measures
arXiv:2508.19197v2 Announce Type: replace-cross Abstract: Identifying the network of species interactions is a fundamental step toward understanding ecosystem stability and biodiversity. However, the interpretability of empirical interaction measures remains a major challenge. Experimental estimates frequently exhibit puzzling temporal fluctuations, including sign shifts typically interpreted as transitions between competition and facilitation. Here, we analyze the temporal behavior of pairwise interaction measures to demonstrate that these fluctuations - and apparent shifts in ecological roles - can emerge intrinsically from standard population dynamics, without any underlying change in the actual ecological relationships. We show that inferred interactions are heavily distorted by experimental protocol choices, particularly the duration of observation and microbial growth constraints. By systematically evaluating interactions across timescales, we uncover a principled mechanism to mitigate these biases: short-term measurements reliably isolate direct, pairwise species couplings, whereas longer-term observations inevitably absorb indirect community feedbacks and systemic experimental constraints. By disentangling direct couplings from indirect network effects, our framework provides a robust, timescale-aware approach to interpreting empirical interaction matrices, offering critical quantitative guidance for experimental design and predictive ecosystem modeling.
Distributed Property Testing with (Quantum) Carrier Pigeons: Tight Bounds on State Certification
arXiv:2606.31753v2 Announce Type: replace-cross Abstract: Recently, Doosti et al. introduced the problem of distributed quantum state verification, where $m$ distributed nodes are given a copy of an unknown state $\rho$, and can send limited one way communication to a central node, who has a complete description of a known state $\sigma$. They ask how many distributed nodes $m$ are required, before the central node can succeed at distinguishing whether $\rho=\sigma$ or $\|\rho-\sigma\|_1\geq\varepsilon$ with high probability. In the setting where only quantum communication is allowed, Doosti et al. exhibit conditional lower bounds in both the public and private-coin settings, and a matching upper bound in the public-coin setting. We extend these results, and show unconditional lower bounds for when both classical and quantum communication are permitted. We show the public-coin lower bound is tight by giving an algorithm with a matching upper bound. We also show an almost tight upper bound in the private-coin setting when only quantum communication is permitted.
Governed AI-Assisted Engineering: Graduated Human Oversight for Agentic Code Generation in Regulated Domains
arXiv:2606.22484v2 Announce Type: replace Abstract: The adoption of agentic AI coding systems -- where autonomous agents generate, review, test, and deploy code with minimal human intervention -- creates a governance challenge in regulated industries. Existing frameworks address AI-assisted development maturity or the productivity-reliability tension but offer no mechanism for calibrating human oversight intensity to regulatory impact. We present the Governed AI-Assisted Engineering (GAIE) framework, a three-tier graduated human oversight model for agentic code generation in regulated domains. GAIE introduces the Oversight Classification Model (OCM), a deterministic decision function that classifies code generation tasks by regulatory impact, customer proximity, reversibility, and data sensitivity to route them through one of three oversight tiers: human-in-the-loop (strategic functions), human-over-the-loop (customer-impacting), or automated-with-monitoring (internal). Each tier defines required evidence artifacts for compliance auditability. We map GAIE against the Bank of Thailand's 2025 AI risk-management policy and demonstrate cross-jurisdiction applicability to MAS (Singapore), NIST AI RMF, ISO/IEC 42001, and the EU AI Act. Evaluation through regulatory coverage analysis, comparative framework analysis, and analytical productivity modeling suggests that graduated oversight preserves 84--97% of agentic coding velocity (central estimate: 91%) while maintaining compliance evidence coverage for regulated functions. GAIE contributes a framework that explicitly bridges AI-assisted development maturity with regulatory governance through proportionate human oversight.
Text Dictates, Music Decorates: Energy-based Attention for Editable Dance Motion Generation
arXiv:2606.22726v2 Announce Type: replace Abstract: Choreographic motion generation poses unique challenges for AI, demanding precise semantic control over complex, temporally structured, and expressive full-body dynamics. While existing models can synthesize motion from music, they remain largely black boxes. Conversely, attempting to condition generation on both text and music frequently leads to modality collapse, where dense acoustic rhythms overwhelm sparse semantic text prompts, destroying user controllability. To resolve this spatial-temporal conflict, we propose STREAM (Structural-Temporal Rhythmic Energy-based Attention for Motion), a modality-decoupled diffusion transformer. STREAM strictly separates conditioning pathways: global text semantics dictate the kinematic structure via Adaptive Layer Normalization (AdaLN), while a novel Bimodal Energy-Based Attention Module (BEAM) routes these features to the musical beat without overwriting the semantics. We further introduce Motorica++, a newly curated dataset enriched with domain-specific dance vocabulary and frame-level semantic annotations from existing Motorica dataset. Additionally, to rigorously quantify zero-shot editability, we propose the Exchange Evaluation Protocol and Editable Dance Score (EDS). Through extensive experiments, STREAM achieves state-of-the-art alignment between motion and music while perfectly preserving choreographic semantics, positioning AI not merely as a reactive synthesizer, but as a controllable, collaborative partner for artistic direction. The source code and datasets are available at https://github.com/SeongJong-Yoo/STREAM.
Confinement-induced motion of ciliates
arXiv:2601.11424v2 Announce Type: replace-cross Abstract: The time dynamics of flagellar and ciliary beating is often neglected in theories of microswimmers, with the most common models prescribing a time-constant actuation of the surrounding fluid. By explicitly introducing a metachronal wave, coarse-grained to a sinusoidal surface slip velocity, we show that a spatial resonance between the metachronal wave and the corrugation of a confining cylindrical channel enables a ciliate to swim even when it cannot move forward in a bulk fluid. Using lubrication theory, we reduce the problem to the Adler equation that reveals an oscillatory and ballistic swimming regime. Interestingly, a ciliate can even reverse its swimming direction in a corrugated channel compared to the bulk fluid.
ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning
arXiv:2607.02137v2 Announce Type: replace Abstract: We study timestep allocation for score-based diffusion sampling, where a learned reverse-time dynamics is discretized on a finite grid. Uniform and hand-crafted schedules are standard choices, but they rely on fixed prescriptions and can therefore be suboptimal. To address this limitation, we propose Adaptive Reparameterized Time (ART), a continuous-time control formulation that learns a time change by treating the speed of the sampling clock as the control, so that a uniform grid on the learned clock induces adaptive timesteps in the original diffusion time. Based on a leading-order Euler error surrogate, ART provides a principled objective for allocating timesteps along the sampling trajectory. To solve this deterministic control problem, we introduce ART-RL, an auxiliary randomized formulation with Gaussian policies that turns schedule learning into a continuous-time reinforcement learning problem. We prove that the randomized ART-RL formulation is equivalent to ART at the optimizer level, in the sense that its optimal Gaussian policy recovers the optimal ART time-warping rate through its mean. We further establish policy evaluation and policy improvement characterizations and derive trajectory-based moment identities that yield implementable actor--critic updates for learning the schedule. Across experiments ranging from controlled low-dimensional settings to image generation, ART-RL can be plugged into existing diffusion samplers by changing only the timestep grid, consistently improving sample quality over strong baseline schedules at matched budgets while leaving the rest of the sampling pipeline unchanged. The learned schedules also exhibit broad generalization, transferring without retraining across sampling budgets, datasets, solvers, pipelines, and representation spaces.
GenASiS: General Astrophysical Simulation System. II. Self-gravitating Baryonic Matter
arXiv:2602.02507v2 Announce Type: replace-cross Abstract: GenASiS (General Astrophysical Simulation System) is a code being developed initially and primarily, though not exclusively, for the simulation of core-collapse supernovae on the world's leading capability supercomputers. This paper -- the second in a series -- documents capabilities for Newtonian self-gravitating fluid dynamics, including tabulated microphysical equations of state treating nuclei and nuclear matter (`baryonic matter'). Computation of the gravitational potential of a spheroid, and simulation of the gravitational collapse of dust and of an ideal fluid, provide tests of self-gravitation against known solutions. In multidimensional computations of the adiabatic collapse, bounce, and explosion of spherically symmetric pre-supernova progenitors -- which we propose become a standard benchmark for code comparisons -- we find that the explosions are prompt and remain spherically symmetric (as expected), with an average shock expansion speed and total kinetic energy that are inversely correlated with the progenitor mass at the onset of collapse and the compactness parameter.
Calibration of systematic distortions in quantum emitter localization microscopy for deterministic nanophotonic fabrication
arXiv:2607.05370v1 Announce Type: new Abstract: Quantum photonic technologies greatly benefit from quantum light emitters with high brightness, indistinguishability, and reliable polarization characteristics. Achieving optimal performance relies on the accurate localization of emitters and their deterministic integration into tailored photonic structures with nanometer-scale accuracy. Although marker-based photoluminescence imaging techniques can achieve statistical fitting uncertainties below 10 nm, the ultimate integration yield is often limited by uncorrected systematic distortions in custom cryo-optical setups that compromise metrological accuracy. Here, we present an in situ calibration protocol that uses lithographically defined gold nanodisk arrays as references to calibrate optical distortions with a Zernike vector-field model. On held-out validation patterns beyond the calibration dataset, this correction reduces the residual systematic bias to 5.3 nm with a 2D scatter of 24.6 nm across the analyzed field of view. Furthermore, we demonstrate that applying this correction to the deterministic fabrication of circular mesa structures around semiconductor quantum dots reduces the variance in emission polarization by 49%, indicating improved registration accuracy. This calibration strategy offers a practical route to high-yield deterministic integration of quantum emitters into scalable quantum photonic circuits.
InverseCrafter: Efficient Video ReCapture as a Latent Domain Inverse Problem
arXiv:2512.05672v2 Announce Type: replace Abstract: Recent approaches in controllable novel view video generation often rely on fine-tuning pre-trained Video Diffusion Models (VDMs). This dominant paradigm is computationally expensive and frequently suffers from catastrophic forgetting of the model's original generative priors. To address this challenge, here we propose InverseCrafter, a VDM training-free framework that reformulates novel view video generation as an inpainting-based inverse problem in the latent space, eliminating the need for any annotated 4D training data. The core of our method is to establish operator equivalence by employing a lightweight latent mask encoder to define a latent-domain masking operation via a continuous, multi-channel representation. This principled representation faithfully models the forward process in the latent domain, enabling efficient, backpropagation-free solvers while bypassing the costly bottleneck of repeated VAE operations. InverseCrafter achieves high-fidelity, spatio-temporally coherent novel view synthesis with near-zero additional inference overhead and excels at general-purpose video inpainting and editing by fully preserving the pre-trained VDM's generative capabilities.
Orchestrating Communication, Computing, and Energy Transfer for Wireless-Powered 6G Closed-Loop Controls
arXiv:2607.04225v1 Announce Type: new Abstract: Future sixth generation (6G) communications are expected to support robotic control tasks in applications such as industrial automation and emergency response, where sensors, computing units, and robots are interconnected via nervous system-like networks to form sensing-communication-computing-control (SC3) closed loops. However, the limited battery capacities of devices within these SC3 loops constrain operational duration and degrade control efficiency, particularly in remote or post-disaster scenarios. To address this challenge, wireless power transfer (WPT) can be leveraged to provide continuous energy supply for SC3 closed loops. In this paper, we investigate a wireless-powered SC3 system, where a satellite transfers energy via radio frequency (RF) signals to support the communication and computing processes of multiple SC3 closed loops. By accounting for the intricate coupling among computing, communication, and energy transfer, we propose a holistic design framework to enhance overall control performance. Specifically, we adopt the linear quadratic regulator (LQR) cost as the performance metric and formulate a sum LQR cost minimization problem. The uplink/downlink transmit power, bandwidth allocation, computing capability, communication/computing time allocation, and WPT power allocation are jointly optimized. We recast the problem into a more tractable form and develop an iterative algorithm to solve it. For the special case of a single loop, we further analyze the properties of optimal solutions in energy-limited scenarios to provide insights for practical parameter configuration. Simulation results demonstrate the performance gains of the proposed scheme.