Forskningsradar

Science Journals

Peer-reviewade publikationer — 53080 artiklar

Two-step growth of (In,Ga)N pseudo-substrates on GaN templates by plasma-assisted molecular beam epitaxy
arXiv:2607.15748v1 Announce Type: cross Abstract: (In,Ga)N layers are grown by plasma-assisted molecular beam epitaxy on GaN templates. We introduce a two-step protocol that involves switching the growth conditions from initially N-stable to metal-stable. Reflection high-energy electron diffraction as well as scanning electron and atomic force microscopy reveal that the first step results in a rough intermediate surface with open pits, whereas the final surface is smooth. The narrow linewidth of the photoluminescence band indicates an excellent compositional homogeneity of the upper layer. Its in-plane lattice constant is determined to be $\approx$3.26 \AA from X-ray diffraction measurements. This combination of favorable properties makes these layers attractive as pseudo-substrates for the growth of red-emitting (In,Ga)N light-emitting diodes. In particular, the approach presented here does not require any complex external processing and is, thus, scalable and economical.
VTAP Gripper: Synergizing Fingertip Sensing and a Visuo-Tactile Active Palm for Dexterous In-Hand Manipulation
arXiv:2607.15448v1 Announce Type: new Abstract: This paper presents a tactile-reactive gripper that integrates a Visuo-Tactile Active Palm (VTAP) and compliant, reconfigurable fingers equipped with tactile array sensors. The design exploits structured finger-palm synergy and multi-modal perception to achieve both robust grasping and fine manipulation. The actuated bi-modal palm seamlessly combines long-range visual localization with contact-rich tactile feedback, substantially extending the system's manipulation capability. To bridge the embodiment gap between human hand motion and the heterogeneous three-finger structure, we further propose a staged, gesture-conditioned retargeting framework for dexterous teleoperation. Extensive experiments validate the system across a range of challenging tasks: reactive grasping of YCB and fragile objects, in-hand syringe reorientation and plunger actuation, singulation of clustered objects down to 3 mm in diameter, and vision-tactile peg-in-hole insertion. Results demonstrate that high manipulation performance can be achieved through coordinated finger-palm interaction and multi-modal sensing, without resorting to high degrees of freedom anthropomorphic designs. The VTAP gripper and its retargeting framework offer a practical reference architecture for dexterous gripper design, manipulation, and contact-rich data collection in support of learning-based approaches. Project webpage: https://yuhochau.github.io/vtap/.
Near-field Dressing of Thermal Emission
arXiv:2607.15622v1 Announce Type: new Abstract: Radiative heat transfer at subwavelength distances is generally understood as enhanced energy exchange mediated by photon tunnelling between neighboring bodies. While near-field interactions can dramatically increase mutual heat transfer, whether they also modify the thermal radiation emitted by the bodies themselves remains an open question. Here we experimentally show that near-field electromagnetic coupling reshapes far-field thermal emission through a distance-dependent dressed emissivity. Using a dual-probe calorimetric platform, we independently monitor the radiative balance of two borosilicate microspheres over separations ranging from 120 micrometers to a few hundred nanometers, spanning the transition from the far field to the near field. Nanowatt-resolved differential radiometry reveals asymmetric heat fluxes and a non-monotonic response of the hotter sphere, demonstrating that thermal radiation is governed not only by emitter-bath interactions but also by coupling to the surrounding photonic environment. By analyzing the total power exchanged between the coupled system and the external thermal bath, we directly extract a dressed emissivity and show that near-field interactions renormalize the far-field thermal emission of the pair through a redistribution of the electromagnetic modes available to thermal fluctuations. These observations provide direct experimental evidence that thermal emitters are dressed by their electromagnetic environment, establishing a thermal analogue of the Purcell effect.
Polynomial-Based Solutions to Targeting Problems for Onboard Applications
arXiv:2607.15801v1 Announce Type: cross Abstract: This paper solves the targeting problem focusing on accuracy, computational efficiency, and reliability. The trajectory optimization problem is first recast as a polynomial optimization problem (POP) by leveraging differential algebra to compute high-order Taylor expansions of the nonlinear dynamics and constraints. Moment-sum-of-squares (SOS) optimization is then utilized to solve this POP. A convex formulation based on a second-order expansion of the dynamics is also proposed. For impulsive targeting, the moment-SOS and convex approaches are compared against traditional nonlinear programming (NLP) solvers and map inversion techniques. Results indicate that the moment-SOS approach provides solutions as accurate as traditional NLP, but with the critical advantage of guaranteeing convergence to the global optimum under mild assumptions. Furthermore, the method excels at handling large maneuvers and long propagation times, conditions in which standard linear approximations rapidly degrade. To demonstrate its versatility, the methodology is extended to a continuous low-thrust station keeping (SK) scenario in the Earth-Moon Circular Restricted Three-Body Problem. The algorithm's performance is then evaluated in the presence of significant state errors. The ability to directly handle non-convex constraints and recast complex, nonlinear dynamics into formulations with reliable convergence properties makes the moment-SOS approach suitable for autonomous onboard applications.
Relevant and Irrelevant: A Renormalization Group Analysis of Transformer Attention
arXiv:2607.15449v1 Announce Type: new Abstract: Using the language of Wilsonian renormalization group theory (RG), we treat the Transformer's attention mechanism as a perturbation of the trained MLP residual-stack fixed point and ask whether it constitutes a relevant, marginal, or irrelevant operator. We derive a fixed-point shift formula and obtain four testable predictions for the fixed-point geometry, effective rank profile, layer specificity, and perturbation decay spectrum. Testing these on synthetic Markov chain sequences with controlled correlation length, we find: (1) For large chains(long correlation), attention is strongly relevant: it closes a residual loss gap the MLP cannot bridge and drives a phase transition in representation space, with effective rank jumping above input dimensionality at layer 1 and stabilizing at a high-dimensional plateau. (2) For short chains(short correlation), attention is irrelevant: the Transformer converges to the same loss and fixed-point geometry as the MLP, though it contracts perturbations faster. (3) The transition is dominated by the first-layer head (L0H0), which accounts for more than 4 times the representational shift of any subsequent head, consistent with the prediction that the relevant operator acts before the MLP begins integrating out positional variation. (4) Perturbation decay experiments reveal a regime reversal: in the long correlation regime the Transformer selectively preserves slow Markov modes (5.4 times the dynamic range in decay length vs. 1.3 times for the MLP); in the short correlation regime it suppresses all modes faster than the MLP, with no spectral selectivity. Together, these results show that the relevance of attention is not a property of the architecture but of the spectral structure of the data-generating process, and that a first-order RG perturbation framework provides a predictive account of that difference.
Aggregation of Statistical Evidence under Exchangeability
arXiv:2607.15823v1 Announce Type: cross Abstract: We study aggregation of statistical evidence under unknown and potentially complex dependence using group-invariance. Building on permutation-based constructions that treat transformed datasets as exchangeable units, we aggregate evidence across statistics for each transformed dataset and calibrate the resulting aggregates across transformations. We develop a finite-sample power and adaptivity theory for this framework, together with extensions to sequential and data-dependent aggregation that preserve validity. For single-batch aggregation, which uses one collection of transformed datasets for both standardization and calibration, we show that the critical values uniformly improve on deterministic calibrations valid under arbitrary dependence, including Bonferroni correction, while adapting to the unknown dependence structure. We also introduce a sequential alpha-spending version that permits early rejection when evidence is strong, and a two-batch extension that separates standardization from calibration to accommodate learned aggregation rules and reduce computation. Applications to adaptive nonparametric testing and conformal prediction illustrate how these results sharpen existing aggregation methods.
Recovery of a Measure-valued Source in the Heat Equation from Sparse Boundary Measurements
arXiv:2607.15841v1 Announce Type: cross Abstract: This article is devoted to the inverse source problem of uniquely determining a measure-valued source from sparse boundary measurements. The measurements considered consist of flux observations over a time interval at two distinct points on the boundary of the domain. The main objective of this work is to extend the existing literature on inverse source problems from sparse boundary measurements, which has so far been limited to point sources or L2 sources, to the identification of a general class of Radon measures. Our approach combines several analytical tools, including regularity properties, boundary representations, and the time analyticity of solutions to the diffusion equation with singular sources. Our theoretical analysis is complemented by a numerical study of the problem. In particular, we investigate the reconstruction of point sources and of a source supported on a curve, and present numerical experiments illustrating the recovery of such sources from sparse boundary flux measurements.
A cubical formalisation of topos causal models: intervention, sheaf gluing, and the intuitionistic do-calculus
arXiv:2607.15629v1 Announce Type: new Abstract: Topos causal models recast causal inference inside a topos: a causal world is a presheaf, an intervention is a characteristic map into the subobject classifier, and reasoning is carried out in the intuitionistic internal language. We give the first machine-checked account of this 1-topos core, in Cubical Agda, over a previously verified probability monad and do-calculus. We build the classifier of sieves and realise the intervention $\mathrm{do}(X := x_0)$ as a characteristic map with its classification theorem; prove the sheaf gluing of independent mechanisms, which the source asserts but never proves; and machine-check the Kripke-Joyal forcing clauses of the internal language. In the modal layer we find and repair a gap: the three standard Lawvere-Tierney axioms do not force a closure operator. With the missing law restored, we exhibit the double-negation topology as a concrete instance and show that interventions and Pearl's rules are stable under every topology. Transportability of a counterfactual across a cover of regimes then coincides with this $j$-stability, understood as invariance across the cover. We further add a phenomenon the programme does not consider: a machine-checked contextuality obstruction, where pairwise-consistent local data admit no global model. The development assumes no axioms and typechecks under Agda's --safe flag, with the ordered field discharged concretely at $\mathbb{Q}$; the scope is the presheaf (1-topos) fragment, with type-level sheafification and the directed lift left to future work.
On the Structure of Address in Multi-Party Dialogue: From Discrete Labels to Continuous Levels
arXiv:2607.15648v1 Announce Type: new Abstract: In multi-party dialogues between a dialogue system and multiple users, identifying to whom an utterance is addressed is a key challenge. Prior work has typically treated addressee detection as a multi-class classification task, selecting a single label representing an individual participant or the group. This formulation assumes that address is inherently discrete and has primarily been used for predicting turn-taking. In this paper, we revisit this assumption by analyzing address as a continuous phenomenon. Using a multi-party human dialogue corpus annotated by multiple annotators, we construct both binary address labels derived from majority-vote addressee labels and continuous address levels inferred from annotator judgments using a latent-variable model. We then examine how these representations relate to turn-taking as well as listener behaviors, including gaze and backchannels. Our results show that, in addition to turn-taking, both gaze and backchannels are associated with address. Furthermore, models using continuous address levels achieve better predictive fit than those using discrete labels, suggesting that address may exhibit graded structure. Finally, we discuss the future directions of addressee detection research based on the findings of this study.
Distributional Matching for Vector Quantization: A Unified Theoretical and Empirical Framework
arXiv:2607.15933v1 Announce Type: new Abstract: The effectiveness of modern visual representation learning and autoregressive models critically depends on vector quantization (VQ), which discretizes continuous feature representations using a learnable codebook. Despite its widespread use, existing VQ methods often suffer from training instability and codebook collapse, arising from gradient mismatch induced by the straight-through estimator and the under-utilization of code vectors. In this work, we show that both issues can be traced to a fundamental mismatch between the distributions of feature vectors and code vectors, leading to inefficient representation and information loss. Building on this observation, we propose a distributional matching framework for vector quantization. We introduce principled criteria for desirable VQ behavior and demonstrate through theoretical analysis and empirical evaluation that aligning feature and code vector distributions provides a unifying mechanism for mitigating training instability and codebook collapse. We instantiate this framework using a Wasserstein-based objective with an efficient closed-form under a mild Gaussian approximation, and further show that a nonparametric alternative based on maximum mean discrepancy yields comparable performance. Extensive experiments on visual tokenization benchmarks support the effectiveness and robustness of the proposed approach.
Scalable Open-Source Visuotactile Sensor for 6-Axis Contact Wrench Estimation in Tensegrity Robots
arXiv:2607.15633v1 Announce Type: new Abstract: This paper presents a scalable, open-source visuotactile sensing system for tensegrity robots that enables six-axis wrench estimation and contact detection. The proposed endcap sensor integrates an elastomeric shell, a 3D-printed thermoplastic polyurethane (TPU) interface, and a rigid base housing an embedded camera and LED illumination ring. A novel gyroid-infill bonding technique is introduced to form a durable elastomer-TPU interface without adhesives, yielding a lightweight and modular design compatible with large-scale tensegrity structures. A tactile-to-wrench neural network maps shear vector fields to six-dimensional force and torque measurements. Experimental results demonstrate accurate and stable wrench estimation with a mean squared error (MSE) of 0.1531 on static validation data and out-of-domain generalization under dynamic motion. Furthermore, full-system integration on a 12 kg tensegrity robot confirms the sensor's ability to reliably identify ground contacts. The system substantially improves the practicality of tactile feedback for tensegrity robots, offering a low-cost, reproducible, and physically interpretable pathway toward contact-aware proprioception and state estimation. Open source files are available at \href{https://github.com/Jonathan-Twz/tensegrity-gelfoot}{github.com/Jonathan-Twz/tensegrity-gelfoot}
Before the Action: Benchmarking LLMs on Prospective Hypothesis Discovery
arXiv:2607.15766v1 Announce Type: new Abstract: Large language models (LLMs) excel at answering pre-specified questions, yet their ability to navigate the open-ended, pre-conclusion stage of discovery remains largely unmeasured. We introduce Prospective Hypothesis Discovery (PHD), which asks models to autonomously construct grounded, discriminative, and testable hypothesis spaces from inconclusive evidence, including anomalous observations and fragmented records, to guide subsequent investigation. To evaluate this capability, we introduce HypoArena, comprising HypoData, a benchmark of 988 cases across six scientific and analytical domains, and HypoEval, an evaluation framework for open-ended hypothesis sets. To construct HypoData at scale, we propose Retrospective Context Regression, a Forge--Audit pipeline that reconstructs pre-conclusion contexts from completed expert documents by removing explicit conclusions, target hypotheses, and retrospective causal attributions while preserving the factual substrate. Because PHD admits multiple valid outputs, HypoEval combines bidirectional pairwise judgments with Bradley--Terry--Davidson aggregation for ranking and six-dimensional rubric scoring for diagnosis. Experiments on 15 frontier LLMs reveal clear capability stratification and model-dependent effects of structured analytical skills, with gains for several lower-performing models on HypoArena but regressions for other systems, including a top-performing model. Compared with absolute rubric scoring, arena evaluation resolves finer-grained differences among models, with aggregated rankings showing strong agreement with human experts and an independent judge. Together, these results support treating PHD as a distinct target for evaluating how LLMs formulate investigative directions when final conclusions are withheld. Our code and data are publicly available at github.com/SKYLENAGE-AI/HypoArena and github.com/SKYLENAGE-AI/HypoArena.
A Projected Drift-Randomized Milstein Method for SDEs with Non-differentiable and Super-linear Drift Coefficients
arXiv:2607.15934v1 Announce Type: new Abstract: We propose a projected drift-randomized Milstein (PRM) method for stochastic differ ential equations with non-differentiable and super-linearly growing drift coefficients. The method extends the randomized Milstein approach beyond the globally Lipschitz setting by incorporating a drift projection into the randomized quadrature approximation. Moreover, unlike existing first-order Milstein-type methods for SDEs with super-linearly growing drift coefficients, the proposed method does not require spatial differentiability of the drift coefficient. Under suitable polynomial Lipschitz and one-sided Lipschitz conditions on the drift, together with standard regularity assumptions on the diffusion, we establish a one-step mean-square stability estimate and derive the required local residual bounds. These esti mates yield first-order strong convergence of the PRM method in the L2-sense. Numerical experiments confirm the theoretical convergence rate and demonstrate the applicability of the method to SDEs with non-differentiable and super-linearly growing drifts.
Vortex formation around islands in random waves
arXiv:2607.15938v1 Announce Type: new Abstract: Wave vortices are fundamental topological features of interference fields, occurring at nodal points where the wave amplitude vanishes. A distinct class of vortices can instead form around it islands or `holes' in two-dimensional wavefields, where the wave intensity remains finite and may even peak at the boundary. In particular, such vortices occur in M2 ocean tides around New Zealand, Madagascar, Iceland, and Svalbard, yet the conditions governing their appearance have remained elusive. Here we develop a statistical theory of vortices around islands in random two-dimensional wavefields, with and without the Coriolis effect, and test it experimentally. We determine the probabilities of vortices with different topological charges as functions of island size and Coriolis parameter. We find that island-bound vortices emerge with unexpectedly high probability, approaching 50% in non-rotating systems and nearly 100% in rotating systems. Moreover, for a broad range of parameters, the presence of a subwavelength island dramatically enhances vortex formation compared with homogeneous random wavefields. Our results explain the formation of tidal vortices around ocean islands of particular sizes (~0.1 of the characteristic wavelength) and establish a general mechanism for generating localized high-intensity vortices around defects in diverse wave systems, from water waves to nanophotonic structures.
Knowledge-Guided Cross-Modal Fusion for Adult-to-Pediatric ECG Transfer via Label-Conditioned Contrastive Alignment
arXiv:2607.15928v1 Announce Type: new Abstract: Adult and pediatric electrocardiogram (ECG) interpretation relies on age-sensitive criteria, and models pretrained mainly on adult ECGs often transfer poorly to pediatric populations when pediatric labels are scarce. Existing multimodal ECG--text methods typically align waveforms and text at the global sample level, entangling evidence from co-occurring diagnoses and limiting transfer under this gap. We propose Pediatric-Adult ECG Alignment via Cross-modal Enhancement (PEACE), a knowledge-guided framework pretrained on the largely adult MIMIC-IV ECG corpus. PEACE describes each diagnosis along rhythm, morphology, and ST--T axes and, per recording, composes only positive-label descriptors into three axis tokens and a fused embedding. A label query network (LQN) uses diagnostic labels as queries to cross-attend over ECG tokens and axis tokens, while label set aware bidirectional contrastive learning (LSBC) aligns pooled ECG features with the fused embedding when recordings share diagnoses. Curriculum adaptive fusion (CAF) gates alignment strength according to smoothed classification loss and training progress, limiting disruption during early optimization. The knowledge branch is used only for training supervision; inference uses ECG signals alone. On ZZU-pECG, PEACE reaches macro average AUCs of 59.39%, 81.74%, and 91.56% under zero-shot, 50-shot, and full fine-tuning, with the clearest gains over foundation and knowledge-pretraining baselines under limited supervision; versus domain adaptation initializations, zero-shot improves substantially while 50-shot AUC is comparable to DANN. After fine-tuning on PTB-XL, PEACE reaches 96.90% macro average AUC over nine harmonized labels. Ablations confirm that label-conditioned knowledge alignment, rather than global text fusion, is the key driver of pediatric transfer gains.
Physics-Based Deep Spatiotemporal Hyperlocal Radar Nowcasting with a Multi-Variable U-Net for High-Resolution Precipitation Forecasting
arXiv:2607.16080v1 Announce Type: new Abstract: Precipitation nowcasting over the immediate 10-90 min period is important for flood management and real-time decision-making in urban regions. Conventional short-range forecasting with high-resolution numerical weather prediction requires frequent data assimilation, model initialization, and spin-up, introducing computational latency. Machine learning provides an alternative by learning storm evolution directly from high-frequency observations and producing forecasts quickly after training. This is particularly relevant for Mumbai, India, where monsoon convection, land-sea interactions, and localized intense rainfall make short-term prediction difficult. Here, we develop a compact radar-only nowcasting framework that combines multi-elevation reflectivity, Doppler radial velocity, and radial-velocity-gradient proxy features within an encoder-decoder U-Net. Using the most recent radar volume scan, the model predicts 12 future composite reflectivity fields at 7.5-min intervals up to 90 min lead time. The derived velocity magnitude, divergence-like, directional-shear, and vorticity-like channels represent kinematic signatures associated with convergence and boundary interactions without requiring full wind-field retrieval. A high-reflectivity attention module improves sensitivity to convective cores, and physics-guided attribution examines whether the learned sensitivities are meteorologically meaningful. The model is trained using Mumbai Doppler radar observations from May to August 2023 and evaluated on temporally independent events. At 90 min lead time, Critical Success Index values are 0.437, 0.332, and 0.193 for $\geq$10, $\geq$20, and $\geq$30 dBZ thresholds, respectively. Compared with persistence, the model gives lower RMSE and higher spatial correlation at longer lead times. Once trained, it runs on a standard computer, generating nowcasts within seconds for real-time use.
Counting in logarithmic space
arXiv:2607.15881v1 Announce Type: cross Abstract: We study the class $\#\mathsf{L}$ of functions counting accepting paths of non-deterministic log-space Turing machines and construct methods to prove containment in $\#\mathsf{L}$. We prove that a large number of classical combinatorial and number theoretic functions belong to this class: classical functions from enumerative combinatorics (multinomial coefficients, Catalan numbers, linear extensions of trees, Stirling numbers, etc), algebraic combinatorics (number of standard Young tableaux, etc), discrete geometry, number theoretic functions, representation theoretic multiplicities in a large class of cases. We show that $\mathrm{GL}_2$-plethysm coefficients of bounded length outer partition can be counted by log$^2$-space polytime verifiers. We pose numerous questions and conjectures on $\#\mathsf{L}$ containment and its generalizations, that suggest venues for conditionally disproving $\#\mathsf{P}$-completeness. While studying which combinatorial functions are in $\#\mathsf{P}$ provides a formal way of (dis)proving the existence of combinatorial interpretations, the lower class $\#\mathsf{L}$ serves as an analogue for functions computable in polynomial time.
Which Hyperparameters Matter? A Game-Theoretic Framework for Interpretable Hyperparameter Sensitivity Analysis
arXiv:2607.15884v1 Announce Type: cross Abstract: This work presents a game-theoretic framework for interpretable hyperparameter-objective interaction analysis rather than proposing a new optimization algorithm. In the proposed framework, Shapley Effects are employed for global sensitivity analysis, while Pareto front sets are utilized to identify effective hyperparameter configurations and support early-stage model evaluation. The resulting analysis reveals which players (hyperparameters) are most influential with respect to different objectives in a given game (application). Consequently, the proposed framework provides interpretable insights into objective-aware hyperparameter interactions, enabling practitioners to guide subsequent optimization, reduce the search space, and perform early-stage model evaluation. The effectiveness of the proposed framework is demonstrated using three distinct neural network architectures across different problem domains under multi-objective settings.
Ultrametric organization of energy landscapes on random Erd\H{o}s--R\'enyi graphs: topological origin of barrier hierarchy
arXiv:2607.15902v1 Announce Type: cross Abstract: We investigate the ultrametric organization of energy landscapes defined on sparse random Erd\H{o}s--R\'enyi graphs. Each graph vertex is assigned a random free energy from a uniform distribution over an interval of width $\Delta F$, and the kinetics are modeled by a Markov process with Kramers transition rates. Using spectral decomposition of the rate matrix, we construct a kinetic Mahalanobis metric between basins of attraction. Computational experiments for graphs with $V=5000$ vertices and $E=5000$ edges show that the degree of nontrivial ultrametricity increases monotonically from $\approx42\%$ for $\Delta F=10$ kJ/mol to $\approx96\%$ for $\Delta F=1000$ kJ/mol. We prove a limit theorem: as $\Delta F\to\infty$, the logarithmic asymptotics of this metric converge pointwise to the classical single-linkage ultrametric. For finite $\Delta F$, corrections from suboptimal paths are exponentially suppressed with increasing $\Delta F$, so that the metric becomes asymptotically ultrametric. Our results suggest that ultrametricity is a universal property of sparse, locally tree-like networks with rugged energy landscapes in the limit of large energy spreads.
A phenomenological multiscale framework for orientational interactions and viscoelasticity in migrating epithelial monolayers
arXiv:2607.15975v1 Announce Type: cross Abstract: Collective migration of epithelial monolayers emerges from the interplay between mechanical interactions and biochemical signalling. Here, we present a phenomenological mechanobiological framework linking cell-scale orientational interactions to tissue-scale mechanics. We distinguish reversible and irreversible head-on and glancing collisions, showing that reversible interactions store orientational mechanical energy while preserving collision geometry, whereas irreversible interactions dissipate energy and alter cell orientation. The balance between energy storage and dissipation governs collective migration, mechanical feedback, and density-dependent processes including cell jamming and live cell extrusion. These interactions regulate cell elasticity, contractility, and adhesion, thereby modifying epithelial surface tension and the effective viscoelastic response of the monolayer. We quantify these effects using orientational interaction potentials, an effective second virial coefficient, and dimensionless measures of stored and dissipated orientational energy. The relative contribution of these mechanisms increases with cell packing density, becoming dominant near the jamming transition. This framework provides a constitutive interpretation connecting collision-induced orientation dynamics with emergent epithelial rheology and suggests how density-dependent interaction regimes shape collective migration and tissue viscoelasticity.
Is That Really My X-Ray? Measuring Internet-Exposed DICOM Services in the Presence of Deception
arXiv:2607.15839v1 Announce Type: new Abstract: DICOM is the dominant protocol for exchanging medical images, yet many Internet-facing deployments lack basic security controls, exposing sensitive patient data to unauthorized access. Accurately measuring this exposure is complicated by honeypots, network telescopes, and other measurement artifacts that inflate published estimates. This paper presents a noise-aware study of Internet-facing DICOM services, combining active IPv4-wide scanning with passive honeypot deployments. We scan common DICOM ports using an ethics-constrained probe limited to association negotiation and C-ECHO. Then, we introduce a reproducible false-positive filtering method that identifies honeypots, telescopes, malformed responders, and other deception artifacts, reducing apparent exposure by 39%. After filtering, we identify 3,979 vulnerable DICOM deployments, all of which lack encryption and accept unauthenticated connections; 1,782 run outdated or deprecated software, and 1,551 carry known remotely exploitable vulnerabilities. Follow-up scans reveal that approximately 50% of exposed services show no evidence of maintenance over 5 months of observation, and that responsible disclosure led to only modest, short-term remediation. On the passive measurement, we deploy Dicompot and find that its raw session logs overstate activity by up to 83% due to generic TCP noise being incorrectly logged as DICOM sessions. After filtering, we observe a reconnaissance gap: Most actors issuing C-ECHO probes never escalate to data exfiltration, consistent with sophisticated actors fingerprinting the honeypot and disengaging before proceeding. Our results show that DICOM exposure measurements can be significantly distorted by deception and logging artifacts, but once corrected, reveal widespread risks to patient privacy and healthcare security that existing deception systems are insufficient to capture.
Retraining Seeks Stable Signals
arXiv:2607.15623v1 Announce Type: cross Abstract: Predictive models deployed at scale influence future data, a phenomenon called performativity. And there is always one way to cope: Train the model on new data, deploy it again, and repeat. This process, called retraining or repeated risk minimization, creates a feedback loop between model and data that real-world learning systems can't avoid. Results on performative prediction shed light on this dynamic: If the model's influence on the data is small, retraining reaches a fixed point. What remains open is why fixed points should naturally exist, and what governs retraining when the model's influence is strong. In this work we develop a new perspective on retraining -- the stable signal principle -- that addresses these questions. We start from the assumption that the prediction target has at least some small model-independent component, a stable signal, such as the intrinsic quality of an item. We prove that when a nonzero stable signal exists, repeated risk minimization, suitably regularized, converges geometrically to the direction of this stable signal. This is true even if the model's influence on the target is arbitrarily large relative to the stable signal. Regularization emerges naturally as a force to control performativity, rather than to promote generalization, revealing a new facet of an old concept. We extend the analysis to a broad family of affine retraining operators under arbitrary model-induced feature changes, heterogeneous time-varying effects, and nonlinear responses. The stable signal perspective also applies to data feedback loops in language modeling, providing new explanations for the stability of language model training from model-generated data.
IMBench: A Benchmark for Intuitive Robotic Manipulation
arXiv:2607.15641v1 Announce Type: new Abstract: Humans combine reasoning and motor control to solve complex manipulation tasks under diverse constraints. They build an understanding of the physical world that helps them convert reasoning into actions and quickly adapt to new scenes, tasks, and rules. We refer to this capability as intuitive manipulation. Existing benchmarks fail to capture this integration: they evaluate physical reasoning in isolation from execution, or measure policy performance without requiring explicit reasoning. We introduce IMBENCH, a benchmark designed to evaluate intuitive manipulation as an integrated capability spanning perception, physical reasoning, action generation, and iterative execution. Our tasks require models to infer task-relevant physical structure and generate feasible action sequences under explicit constraints, including contact-rich manipulation, tool use, and multi-stage dependencies. We introduce a benchmark of 35 tasks, 14K filtered trajectories, and scalable tools for generating diverse scenarios. Experiments reveal a consistent gap: vision language models show partial physical reasoning ability but fail to produce executable plans, while state-of-the-art vision-language-action models struggle to satisfy task constraints and generalize across scenarios. These results identify intuitive manipulation as a missing axis in current foundation models and generalist robot policies, and position IMBENCH as a step toward evaluating and enabling more integrated, adaptive physical intelligence.
Difference-Based Relational Learning for Zero-Shot Object-Goal Visual Navigation With Direct Sim-to-Real Transfer
arXiv:2607.15642v1 Announce Type: new Abstract: End-to-end deep reinforcement learning (DRL) for zero-shot object-goal visual navigation remains challenged by the sim-to-real gap, particularly variations in object appearance and restricted camera field-of-view (FoV). This letter proposes a Temporal Difference-Relational Network (T-DRN) for robust zero-shot sim-to-real transfer. T-DRN combines a Siamese difference-based feature extractor, which computes relational difference between the target and observed objects to produce domain-independent representations, with a dual-frame temporal buffer that preserves short-term object continuity under narrow FoV. Extensive experiments in AI2-THOR demonstrate that T-DRN improves zero-shot generalization in terms of success rates over strong baselines. Furthermore, T-DRN is systematically validated on a physical wheeled robot, demonstrating robust performance under real sensing and actuation constraints and supporting the feasibility of direct sim-to-real transfer.
Negative quantum friction in nanoscale water flows: the Wigner picture
arXiv:2607.16011v1 Announce Type: cross Abstract: We explore the phenomenon of "quantum" friction based on a single-particle model patterned after the Wigner equation describing electrons flow in a solid wall confining nanoscale water flows. The numerical simulations show a clear signature of negative quantum friction, namely a net momentum transfer from the electrons in the solid wall to the flowing water molecules. Such net momentum transfer results into a sizeable reduction of the water friction, up to forty percent, depending on the strength of the coupling between classical and quantum fluctuations. Our results offer the prospect of a theoretical framework bridging classical and quantum description by using continuum kinetic theories and particle-based simulations.