Forskningsradar

Science Journals

Peer-reviewade publikationer — 54780 artiklar

RoboTTT: Context Scaling for Robot Policies
arXiv:2607.15275v1 Announce Type: new Abstract: Recent robot foundation models operate with single-step or short-history visuomotor context. We introduce Test-Time-Training Robot Policies (RoboTTT), a robot model and training recipe that scale visuomotor context to 8K timesteps, three orders of magnitude beyond state-of-the-art policies, without growing inference latency. At this context length, we unlock new robot capabilities: one-shot in-context imitation from human video demonstrations, on-the-fly policy improvement, robustness to perturbations, and stronger performance on multi-stage, long-horizon tasks. We also observe, for the first time, steady gains in closed-loop performance as pretraining context length scales. At its core, RoboTTT integrates Test-Time Training into robot foundation models such as Vision-Language-Action policies, yielding a sequence model whose recurrent state consists of fast weights, parameters updated by gradient descent during both training and inference, compressing histories into weight space and retrieving contextual information for long-context conditioning. To scale training context length, the recipe combines sequence action forcing with truncated backpropagation through time. On challenging real-robot manipulation tasks, RoboTTT improves overall performance by 87% over the single-step context baseline and fully completes a five-minute, ten-stage assembly task, which no baseline ever does. RoboTTT trained with 8K-timestep context outperforms the same model pretrained with 1K timesteps by 62%, suggesting context length as a new scaling axis for robot foundation models. Videos are available at https://research.nvidia.com/labs/gear/robottt/
Upwind embedded boundary SBP operators: New high order numerical schemes for arbitrarily shaped domains with Cartesian grids
arXiv:2607.14403v1 Announce Type: cross Abstract: Embedded boundary summation by parts (SBP) methods define finite differencing based derivative operators with the added feature that the boundary need not coincide with a grid cell, allowing a boundary to be embedded on a regular Cartesian grid. This is achieved by the introduction of interpolation/extrapolation operators that match the accuracy of the boundary closure. These methods have been used to perform black hole excision simulations on a domain with a spherical boundary embedded in a regular Cartesian grid, demonstrating their usefulness for nonlinear problems. In this work, new operators are derived using this embedded boundary framework to increase the order of accuracy of the interior and boundary closure while minimizing the boundary error. Additionally, these novel operators improve the spectral properties on the grid by generalizing to an upwind scheme that has better dispersion relation preserving properties compared to traditional SBP schemes for wave equations. These operators are tested with the curvilinear scalar wave equation on a 3D multiblock grid with an excision sphere embedded in the center block to demonstrate the robustness and accuracy of these novel embedded operators.
Precise sample covariance spectral norm error -- an RDT view
arXiv:2607.14460v1 Announce Type: cross Abstract: We study the sample covariance error of centered Gaussians. A remarkable breakthrough [66] established the correct error scaling order and explicitly revealed the critical role of both the effective rank and the true covariance spectrum. In this work, we move beyond scaling characterizations and determine the precise limiting value of the error's spectral norm. To do so, we develop a generic framework based on Random Duality Theory (RDT). Within this framework, we first determine closed-form, explicit RDT-based upper bounds. We then establish complementary lower bounds by introducing a novel bilinear-quadratic RDT lower-bounding mechanism. By combining this mechanism with a two-replica systems bounding strategy, we show that our lower and upper bounds match in large-dimensional contexts. Our theoretical results are supplemented with numerical evaluations and simulations, demonstrating an excellent agreement already for problem sizes on the order of thousands.
Automatically Evolving Prompt Guidelines for Task-Specific Optimization
arXiv:2607.14105v1 Announce Type: new Abstract: For Large Language Models to reliably answer user queries, users must clearly specify requirements, context, and constraints. In practice, however, user queries are often underspecified, forcing models to infer unstated assumptions that may misalign with the actual user intent. Existing prompt engineering guidelines aim to mitigate this issue, they are typically generic and task-agnostic, limiting their practical utility. Additionally, existing guidelines are formed manually and in a non-systematic way. To this end, we study prompt guideline optimization: the problem of automatically generating task-specific guidelines that help write better-specified prompts for a given task and model. Our key observation is that existing (completed) task examples (aka reference answers) often implicitly encode the missing information required to complete underspecified queries, including behavioral constraints, contextual assumptions, and evaluation criteria. We therefore propose AGOPS, an automatic approach that evolves task-specific guidelines via an optimization scheme that involves a prompt LLM writer, a solver LLM and prompt evolution, which maximize downstream effectiveness on a set of examples (user queries with reference answers). At inference time, our guidelines help users write well-specified prompts, boosting the effectiveness of LLMs. We show across mathematical reasoning, medical question answering, and coding tasks, that prompt underspecification leads to major drops (up to 95.3%) in downstream task performance (compared to well-specified prompts) and, perhaps more importantly, that this drop can hardly be recovered by existing prompt optimization techniques. Users following AGOPS guidelines can regain this loss (increasing performance between 15.5 to 81.7% on average) consistently across all benchmarks.
ReBind: Multi-Reference Video Editing via Structured Instructions with Explicit Reference Relationships
arXiv:2607.14681v1 Announce Type: new Abstract: Recent diffusion-based video generation models have made significant progress in multi-reference image-conditioned video editing. However, existing methods still struggle to coordinate information from multiple visual sources accurately. We identify a critical deficiency in existing approaches. Existing editing instructions lack explicit reference relationships, and most multimodal large language models (MLLMs) cannot generate them reliably. To address this problem, we propose ReBind, a systematic framework that introduces semantic instructions with embedded reference tokens as the intermediate representation for multi-reference image-conditioned video editing. Our key insight is embedding reference tokens at semantic positions to eliminate ambiguity and establish precise bindings between visual attributes and their sources. We develop ReBind-Instruct, a specialized MLLM that learns to establish explicit bindings between visual attributes and their reference sources through a two-stage progressive scheme for precise reference relationships. We further develop ReBind-Edit, which enables lightweight adaptation of text-to-video models to coordinate multiple references by binding visual attributes to their designated sources. Extensive experiments demonstrate that ReBind substantially outperforms general-purpose MLLMs in instruction quality and achieves state-of-the-art performance among open-source methods on reference image conditioned video editing. Our project webpage: https://rebind-mrv2v.github.io/.
The orientation of the Amazonian geoglyphs as a clue for their interpretation
arXiv:2607.14201v1 Announce Type: new Abstract: Amazonian earthworks, also called geoglyphs, are thousands of man-made earthen structures, mostly of geometrical shape, which progressively emerged from the tropical forest due to progressive deforestation. They were probably built between the fifth century BC and the end of the first millennium AD, but archaeological investigation on the culture of their builders is yet at the beginning. Nevertheless, a ceremonial rather than practical function seems likely, at least for those having a very regular shape. In the present paper, simple remote-sensing technique, combined with the methodological approach of modern Archaeoastronomy, are applied to study for the first time their orientation. The analysis takes in consideration virtually all known squared and rectangular structures for a total of 326 earthworks. The results show without doubts a non-random choice for their orientation and a clear interest of their builders for the annual cycle of the Sun.
A low-rank hierarchical framework for the non-Markovian stochastic Schr\"odinger equation with convergence analysis
arXiv:2607.14689v1 Announce Type: new Abstract: We propose and analyze a novel numerical framework for the non-Markovian stochastic Schr\"odinger equation (NMSSE) based on a low-rank approximation of the bath correlation functions. By decomposing the memory kernel into a finite-dimensional representation, we derive a truncated system of hierarchical equations that effectively balances computational tractability with physical fidelity. A rigorous convergence analysis is established for the hierarchical framework under mild assumptions. We demonstrate that our formulation serves as a mathematical generalization of the Hierarchy of Pure States (HOPS), encompassing it as a special case while offering a more flexible representation of non-Markovian effects. Numerical experiments across several benchmark models are presented to illustrate the validity and efficacy of the proposed method.
Perturbation Analysis of Maximal Quantum Leakage
arXiv:2607.14469v1 Announce Type: cross Abstract: Maximal quantum leakage (MQL) is a worst-case information leakage measure that quantifies an adversary's inference advantage gained from accessing quantum encoding of classical data with arbitrary measurements. While MQL admits an exact characterization for a given ensemble of quantum states, its robustness to implementation imperfections has not been systematically studied. In this paper, we analyze the sensitivity of maximal quantum leakage under perturbations of the quantum encoding. We establish a continuity bound in terms of the trace distance between ideal and perturbed quantum states, and show, via an example, that this bound is attainable. We further derive fidelity-based and relative-entropy-based sufficient conditions for bounding the variation of maximal quantum leakage, and illustrate numerically that these conditions can be loose.
BridgeFlow: Fast and Robust SE(2)-Equivariant Motion Planning with Flow Matching
arXiv:2607.14725v1 Announce Type: new Abstract: In robotic motion planning, equivariance to rigid body transformations is crucial for robust spatial generalization. However, current learning-based planners face a critical dilemma: they either lack inherent equivariance, treating transformed tasks as novel scenarios, or enforce it via computationally expensive specialized architectures that bottleneck real-time inference. To break this trade-off, we propose BridgeFlow, a fast and strictly SE(2)-equivariant generative motion planning framework. Rather than relying on heavy equivariant networks, BridgeFlow achieves exact spatial equivariance via a lightweight task-centric canonicalization module, enabling generalization using standard architectures. To further accelerate inference, we pair a Brownian bridge informative prior with context-aware mini-batch optimal transport. This constructs a straightened vector field that minimizes transport costs and stabilizes training. Furthermore, environmental awareness is explicitly embedded via Classifier-Free Guidance. Evaluations in dense 2D environments and on a 7-DoF Franka manipulator demonstrate that BridgeFlow achieves up to a 15x inference speedup and a 2x higher valid trajectory rate over state-of-the-art diffusion baselines, alongside robust generalization to entirely unseen environments and arbitrary spatial transformations.
WorkDrive: Roadwork Chain of Causation for Autonomous Driving
arXiv:2607.14727v1 Announce Type: new Abstract: Autonomous driving vision-language models (VLMs) struggle in roadwork zones, where familiar visual cues such as lane markings and permanent signs are altered or absent, and temporary devices such as cones and barriers redefine the drivable corridor. VLMs can detect these objects, but without explicit guidance they anchor their reasoning on familiar elements from pre-training and fail to connect work-zone observations to correct planning decisions. We propose WorkDrive, a framework that constructs perception-grounded causal reasoning for work zones and aligns it with trajectory prediction. An automated multitask perception pipeline extracts structured scene facts and injects them into a Chain-of-Causation (CoC) annotation pipeline, redirecting the annotator's attention to domain-specific elements. The resulting reasoning labels are used for supervised fine-tuning, followed by reinforcement learning with a single reward: consistency between lateral meta-actions and the predicted trajectory. On ROADWork, the largest public work-zone dataset, the proposed roadwork CoC reduces trajectory average displacement error (ADE) by 9.0\%, and consistency-based GRPO yields a further 3.0\%, achieving progressive improvement over the trajectory-only baseline. Code and data will be publicly released.
VQ-Touch: A Data-Efficient Tactile Generation Framework Across Sensors and Scenarios
arXiv:2607.14728v1 Announce Type: new Abstract: Tactile image generation significantly reduces the dependency on expensive and wear-prone sensors by synthesizing high-fidelity tactile data, offering an efficient solution for tactile information acquisition in robotic perception and human-machine interaction systems. However, existing methods depend on large-scale, diverse datasets from specific sensors and lack efficient data utilization and robust generalization capabilities, struggling in vision-limited environments. To address this, we introduce VQ-Touch, a tactile generation framework that supports both cross-sensor and multi-scenario applications. Specifically, to efficiently extract complex deformation and texture features from the data, we propose DM-VQGAN, an effective tactile representation learner. Furthermore, we introduce a discrete diffusion decoder with a unified conditioning interface, supporting multimodal generation tasks such as images and labels, and enhances the model's generalization capability through few-shot mixed training, thus achieving compatibility with current mainstream sensors and their variants. Experiments show that VQ-Touch surpasses state-of-the-art methods in multiple tasks.
The Misclassification of Autistic Writing as AI-Generated
arXiv:2607.14729v1 Announce Type: new Abstract: Recent findings suggest that detection models for artificial intelligence (AI) cannot accurately identify AI-generated text and may exhibit bias against certain minority groups. In the present study, anecdotal claims that autistic writers more often have their work flagged as AI-generated are examined empirically. A corpus of approximately 60,000 Reddit posts split into "likely-autistic" and "general-Reddit" subcorpora is used to compare the distribution of probabilities output by the OpenAI GPT-2 detection model. Differences in textual features between subcorpora are observed and compared to reported features of AI-generated text. Results showed that while less than two-percent of either subcorpus was flagged as AI-generated by the model, significantly more texts from the likely-autistic subcorpus were flagged. Connections between features of text with likely-autistic authors and AI-generated text were not straightforward. The widespread use of AI-detection models with a potential bias against autistic writers in their output prompts ethical scrutiny, and the authors recommend further critical examination of the models themselves as well as their use in academic contexts.
Track fitting at the full LHC collision rate
arXiv:2607.14793v1 Announce Type: cross Abstract: The LHCb experiment at the Large Hadron Collider underwent a major upgrade before the LHC Run 3 data taking period, employing an all-software approach in its trigger system. Here we present a fast implementation of a Kalman filter, used in the first trigger stage since the 2025 data taking period, allowing to determine parameter estimates of charged-particle trajectories at the full LHCb collision frequency of 30 MHz. This approach replaces computationally expensive magnetic field map lookups and numerical integration methods with fast analytical parameterisations while maintaining the mathematical framework of Kalman filtering. Implemented on approximately 500 GPUs within the first-level trigger, the algorithm has replaced the previous partial track fitting algorithm in the real-time trigger environment at the cost of a 2% increase in processing time. Compared to the previous fitter this parameterised Kalman filter shows a significantly improved momentum resolution, resulting in a factor of two improvement in the invariant mass resolutions for reconstructed D0 and J/{\psi} hadrons. It additionally demonstrates greater robustness against detector misalignment effects and substantially sharpens the discrimination between genuine particle trajectories and accidental background, more than doubling the rejection of the latter at no cost to genuine-track efficiency, for a standard selection working point.
Goal-Oriented Semantic Communication for Distributed ISAC-Enabled Vehicle Coordination
arXiv:2607.15111v1 Announce Type: new Abstract: Vehicle coordination at unsignalized intersections relies on accurate real-time vehicle state acquisition and reliable command-and-control (C&C) signal delivery. However, existing studies typically treat sensing, communication, and control separately, which may lead to redundant transmissions, outdated state information, and unreliable vehicle coordination. In this paper, we investigate a new scenario of distributed integrated sensing and communication (ISAC)-enabled vehicle coordination at intersections, where multiple roadside units (RSUs) collaboratively transmit sensing signals for vehicle state acquisition and C&C signals for vehicle movement control under the management of a central base station (BS). To improve signaling efficiency, we propose a unified goal-oriented semantic communication (GSC) framework, which transmits sensing and C&C signals only when they are semantically important for improving intersection traffic throughput. Specifically, an extended Kalman filter (EKF) is adopted to predict vehicle states and fuse distributed sensing measurements. A masked hybrid proximal policy optimization (MHPPO) framework is then developed to jointly determine sensing transmission decisions, C&C transmission decisions, and C&C signal contents based on a value-of-information (VoI) reward. Furthermore, we propose an uncertainty-aware transmission design (UTD), including robust beamforming and VoI-based time-division power allocation, to improve sensing and communication reliability under vehicle state uncertainty and inter-RSU interference. Simulation results show that our proposed framework achieves 100% collision-free vehicle coordination with significantly reduced signaling overhead compared with predictive ISAC baselines adapted from state-of-the-art related studies and several ablation baselines.
Transverse Optomechanical Interaction Mediated by Mechanically Induced Symmetry Breaking: Hamiltonian Dynamics
arXiv:2607.14502v1 Announce Type: new Abstract: In cavity optomechanics, the interaction between light and motion is usually introduced via the shift of cavity resonances in response to mechanical displacement. Here we present an analysis of Hamiltonian dynamics of an optomechanical system with a different form of optomechanical coupling, in which mechanical motion dynamically couples otherwise independent optical modes. In the language of Schwinger pseudospin operators, the dispersive coupling can be interpreted as "longitudinal" while the mode-coupling mechanism corresponds to a transverse interaction. The latter is well known in cavity and circuit QED but was given only scarce attention in cavity optomechanics. Unlike the traditional dispersive/dissipative coupling, the mode-coupling optomechanical interaction generates rich Hamiltonian dynamics even in the absence of external drive or dissipation. For instance, under certain initial conditions this dynamics is characterized by a Hamiltonian Hopf bifurcation controlled by the total photon power injected into the system. Below the bifurcation threshold and for large enough non-linearity, mechanical modulation of optical amplitudes generates a broad spectrum of multiple sidebands covering a frequency interval larger than ten mechanical frequencies. Above the threshold, the frequency of optical oscillations becomes dependent on the mechanical amplitude, while mechanical degrees of freedom return to oscillating at their bare frequency. The scope of this work is limited to the study of purely Hamiltonian dynamics to demonstrate that the mechanically mediated mode-coupling optomechanical interaction provides an alternative method of coherent control of energy exchange between light and mechanical motion.
Optimization dynamics of Transformer backflow neural quantum states for the two-dimensional Hubbard model
arXiv:2607.14875v1 Announce Type: cross Abstract: Building on the multi-determinant Transformer backflow neural quantum state (NQS) ansatz and the associated multi-stage training workflow for the doped two-dimensional Hubbard model, we investigate how the optimization dynamics of the NQS depend on several key optimization and architectural hyperparameters. The workflow consists of neural-network backflow (NNB) initialization, supervised Transformer pre-training, and main energy optimization using the Moment-Adaptive ReConfiguration Heuristic (MARCH) within variational Monte Carlo. Using the doped $4\times4$ periodic Hubbard model at $U=8$ as a baseline, we examine how the update-norm threshold, Transformer width, number of determinant channels, and Monte Carlo batch size affect convergence. We find that a moderate update constraint improves the efficiency of MARCH optimization, larger Transformer width and more determinant channels improve the expressive capacity of the ansatz, and larger Monte Carlo batches reduce sampling noise in the update direction. We further test the same workflow at half filling, weaker interaction strength, open boundary conditions, and on a larger $8\times8$ doped lattice. These results identify practical optimization trends for Transformer backflow NQSs and highlight the balance between ansatz expressivity, MARCH update stability, and Monte Carlo sampling quality.
Measuring Spatial Clustering via Metropolis-Hastings Diffusion Distance
arXiv:2607.14880v1 Announce Type: cross Abstract: We propose a novel measure of the discrepancy between two probability distributions $f$ and $g$ on a graph - which we call the diffusion distance - that measures the rate of convergence of $f$ to $g$ under a graph-constrained Markov chain with stationary distribution $g$. As a default choice for this Markov chain, we use the Metropolis-Hastings transition matrix targeting $g$ with proposals given by a random walk on the graph. Our primary case of interest is when the second distribution $g$ is uniform, in which case the diffusion distance becomes a measure of spatial clustering in $f$. Used in this way, (Metropolis-Hastings) diffusion distance to uniformity extends Moran's $I$-type measures of spatial autocorrelation by incorporating global graph geometry rather than just local patterns. Indeed, Moran's $I$, the most well-known measure of spatial autocorrelation, can be viewed as a one-step heuristic for diffusion distance, so long as specific spatial weights are used. We establish theoretical bounds and a stability result for our measure, connecting it to graph spectra and optimal transport. We then turn our attention to outlining a statistical test for spatial clustering using diffusion distance. Under permutation null models, we derive high-probability bounds on diffusion distance underpinned by exact spectral formulas for convergence of distributions, enabling an efficient statistical test for spatial clustering on large datasets. We empirically compare diffusion distance to Moran's $I$ both as a numerical measure and as a statistical test. We show that diffusion distance exhibits higher power on synthetic data using a stochastic block model. Empirical analysis of Black population distributions for 100 U.S. cities shows that diffusion distance detects subtle differences in urban segregation patterns that Moran's $I$ does not.
Finite-Dimensional Feedback Stabilization of Nonautonomous Stochastic Parabolic Equations
arXiv:2607.14911v1 Announce Type: cross Abstract: We investigate finite-dimensional feedback stabilization for nonlinear nonautonomous stochastic parabolic equations driven by $Q$-Wiener, covering both additive and multiplicative perturbations. The control is given by a finite linear combination of localized indicator-type actuators whose supports are selected as part of the construction and may have arbitrarily small total measure. The feedback law is constructed by means of oblique projections onto suitable finite-dimensional subspaces. Within the variational Gelfand triple framework, we prove well-posedness of the closed-loop system under standard coercivity, growth, and global Lipschitz assumptions. By appropriately choosing the actuator configuration and feedback strength, we establish exponential mean-square stabilization of the stochastic dynamics and, for pure multiplicative noise, almost-sure stabilization. A fully discrete three-layer implementation complements the theoretical results. Numerical experiments illustrate the influence of number of actuators, noise intensity, and nonlinear effects on the closed-loop stabilization behavior.
Stochastic ultimatum game: Spite-driven resource feedback fosters fairness
arXiv:2607.14914v1 Announce Type: cross Abstract: Resource scarcity can fundamentally encourage antisocial behaviour, whereas resource abundance can promote fair behaviour. Experimental evidence indeed suggests that scarcity induces spiteful behaviour, while repeated interactions enhance fairness. However, existing studies of game--environment feedback systems are largely confined to the evolution of cooperation and they overlook the interplay between resources, spite, and fairness. To address this lacuna, we develop a stochastic ultimatum game framework in which an offerer and an accepter repeatedly interact to negotiate exploitation of a self-renewable resource under the ownership of the offerer. Successful agreements deplete the resource, whereas unsuccessful agreements inhibit exploitation and facilitate replenishment. The mutation--selection driven two-species stochastic evolutionary dynamics reveal that the emergence of spite and fairness strongly depends on the resource growth rate. Fairness predominantly prevails for resources with high growth rates. Intriguingly, low resource growth rates give rise to a resource feedback loop driven by spite: spiteful behaviour dominates in the depleted state, facilitating transition of the resource state to replete state which, in turn, promotes fairness through repeated interactions.
Lossy compression of weighted graph adjacency matrices by transform coding
arXiv:2607.14834v1 Announce Type: new Abstract: In this paper, we propose a compression framework for weighted graphs in which the graph topology is transmitted losslessly and edge weights are compressed lossily. A challenge in the lossy compression of edge weights is that the underlying relationships between edges are ambiguous. To address this issue, we first transform the unweighted graph into the corresponding line graph, whose nodes represent the edges of the original graph and whose edges encode the relationships between them. The line graph transform allows us to regard edge weights as a graph signal defined on the line graph. Instead of transmitting the edge-weight vector, we first transform it with a graph filter bank on the line graph. Then, quantization and entropy coding are performed on the transformed coefficients of the edge weight vector. In addition to the lossy compression method, we formalize edge smoothness on the line graph and show that it serves as a measure of the difficulty of compression. The proposed smoothness measure can be easily calculated without converting to a line graph. This provides insight into the expected compression performance of a given weighted graph. Experiments on synthetic and real-world data validate the effectiveness of the proposed method by comparing it with existing matrix preprocessing methods.
A flexible DAQ system for the Timepix4 ASIC in the 4DPHOTON project
arXiv:2607.14847v1 Announce Type: new Abstract: This paper presents the design and implementation of a flexible data acquisition system13 developed for the Timepix4 ASIC within the 4DPHOTON project. The system is based on a modular14 FPGA-centric architecture combining high-speed serial readout, Ethernet-based data transport,15 and deterministic multi-board synchronization. The hardware platform includes a scalable stack16 composed of commercial FPGA carrier boards, FMC-based interface electronics, detector-specific17 chipboards, and a dedicated Trigger Logic Unit for synchronized operation.18 The firmware architecture separates control and data paths, enabling independent configuration19 and high-throughput acquisition through UDP over 10 GbE links. The design supports zero-back-20 pressure operation toward the ASIC and allows adaptation of the readout bandwidth to different21 experimental conditions. Synchronization between multiple DAQ systems is achieved through22 a common clock and trigger distribution network, experimentally demonstrated with sub-100 ps23 precision.24 The system has been developed to support detector characterization, laboratory measurements,25 and beam-test campaigns for the 4DPHOTON detector concept. Hardware organization, firmware26 architecture, synchronization strategy, and performance measurements are presented.
LQCDMaster: Agentic Scientific Computing for Lattice Quantum Chromodynamics Research
arXiv:2607.15001v1 Announce Type: cross Abstract: Lattice quantum chromodynamics (LQCD) provides a first-principles framework for computing hadronic observables, but its practical use remains limited by the substantial expertise required to turn research motivation into reliable computing workflows. Here we present \textsc{LQCDMaster}, a tool-augmented, skill-guided and domain-specialized scientific computing agent that converts natural-language LQCD research tasks into executable PyQUDA computing workflows, including measurement scripts, job-submission artifacts, execution logs and numerical outputs. The system combines agentic planning, expert-annotated LQCD skills and a deterministic Wick-contraction tool to constrain the algebraically fragile components of code generation. We evaluate \textsc{LQCDMaster} on a benchmark at the forefront of scientific research, comprising 70 LQCD computing tasks, with observables covering local and nonlocal two-point functions, Wilson loops, meson and baryon three-point functions. The generated workflows exactly reproduce expert-written implementations in 63 of 70 tasks at machine precision, with three additional discrepancies attributable to convention mismatches. Across representative observables, the agent reduces implementation time from hours to minutes while preserving end-to-end numerical validation. Further, we present a typical case of \textsc{LQCDMaster}-driven exploration: a lattice computation of light-cone distribution amplitudes with diagonal Wilson-line, a quantity accessible with standard methods but never before computed, and computation of the spectrum of proton, deuteron, triton, hyperon, hyperdeuteron and hypertriton. This work pioneers the paradigm of agentic scientific computing by automating the end-to-end scientific computing workflows in lattice QCD research, lowering its barrier and facilitating the exploration and verification of non-standard scientific ideas.
SLT 2026 REAL-TSE Challenge: Real-world Target Speaker Extraction from Conversational Recordings
arXiv:2607.15198v1 Announce Type: cross Abstract: We introduce the REAL-TSE Challenge, an IEEE SLT 2026 satellite challenge on target speaker extraction~(TSE) from real conversational recordings. Given a multi-speaker mixture and one or more enrollment utterances from a target speaker, participating systems must recover only the target speech. Unlike simulated read-speech benchmarks, REAL-TSE evaluates Mandarin and English recordings that contain natural overlap, reverberation, noise, channel mismatch, and conversational dynamics. The challenge defines two complementary tracks: an Online track for low-latency streaming extraction and an Offline track for full-context processing. Systems are evaluated with Token Error Rate (TER), Speaker Similarity (SpkSim), DNSMOS, and target-speaker activity F1. This overview paper describes the task definition, datasets, baselines, evaluation protocol, submitted systems, condition-wise findings, and lessons for future real-world TSE benchmarks.
Sulfur photochemistry observationally traces mantle redox states of rocky planets
arXiv:2607.15204v1 Announce Type: cross Abstract: Volatile outgassing from planetary interiors controls the composition of rocky exoplanets' secondary atmospheres. However, observations indicate that disequilibrium processes, such as photochemistry and vertical transport, can strongly alter the chemical structure of Hot Jupiters. Which process dominates under different types of rocky planets, and how outgassing and photochemistry jointly determine the atmospheric composition, remain open questions. Sulfur species are promising tracers of interior-atmosphere coupling because their atmospheric abundances are sensitive to both mantle redox state and stellar irradiation. The PROTEUS planetary interior-atmosphere evolution modelling framework is coupled to two chemical models, FastChem and VULCAN, for post-processed chemistry calculations. We run a grid of planetary evolution simulations spanning diverse mantle redox states, instellation fluxes, and Solar versus M-star host-star spectra. For each case, we compare atmospheric compositions under thermochemical equilibrium, only vertical transport, and vertical transport plus photochemistry. The bulk atmospheric composition remains controlled by the redox state of the mantle and outgassing history, even when disequilibrium chemistry is included. Reduced mantles produce atmospheres rich in H2, and oxidised mantles are dominated by CO2. Photochemistry affects the upper atmosphere, strongly depleting neutral volatiles and enhancing radicals, especially for highly irradiated cases. SO2 is strongly enhanced at intermediate-to-oxidised redox states. Synthetic emission spectra show that photochemical SO2 can generate absorption features at 4 um and at 7.3 / 8.7 um, reaching ~60 ppm and ~100 ppm, before sequentially returning to the outgassed signatures of ~30 ppm and ~50 ppm for the oxidised mantle redox state. These signatures are detectable with JWST, motivating targeted observational campaigns.
Counterfactuals for Feature-Weighted Clustering
arXiv:2607.14719v1 Announce Type: new Abstract: Counterfactual explanations provide local, interpretable insight by identifying changes to an input that would alter its assigned outcome. Although well established in supervised learning, their extension to clustering is less direct, since cluster assignments are unlabeled and governed by the geometry of the partition. This paper introduces VoICE, a Voronoi-Induced Counterfactual Explainability framework for feature-weighted $k$-means clustering. Rather than treating cluster change as a crossing of a single pairwise centroid boundary, VoICE formulates counterfactual generation as projection onto the full weighted Voronoi region of a target cluster, incorporating feature weights directly into both the clustering geometry and the counterfactual objective to yield least-cost and parsimonious explanations under actionability constraints. Target regions are further intersected with data-derived bounds and homothetically contracted towards their centroids, limiting extrapolation and boundary sensitivity. VoICE consistently produces valid target-cluster membership, across several benchmark datasets, where the leading pairwise baseline does not.