Forskningsradar

Science Journals

Peer-reviewade publikationer — 54780 artiklar

Mixed Finite Elements for Geometrically Exact Beams using Discontinuous Rotations and Discrete Curvature
arXiv:2605.04573v3 Announce Type: replace Abstract: We propose a novel mixed finite-element formulation for geometrically exact (Simo--Reissner) beams that introduces the moment vector as additional independent field. The specific mixed form allows for an element-local, discontinuous approximation of rotations, which is key to a simple and efficient discretization framework. The concept of discrete curvature provides a mathematically consistent treatment of rotation discontinuities. For linear constitutive laws, the mixed form is derived via a Legendre transform of the curvature-related strain energy. Objectivity is retained at the discrete level by interpolating relative rotations through a multiplicative split of the rotation field; path-independence is inherent to the total Lagrangian setting and verified numerically. Several benchmarks demonstrate optimal rates of convergence and accuracy, irrespective of the beam's slenderness and order of approximation. Notably, the lowest-order element entirely avoids rotation interpolation by employing element-constant rotations only.
Sub-Gaussian Concentration and Entropic Normality of the Maximum Likelihood Estimator
arXiv:2605.07107v3 Announce Type: replace Abstract: It is well known that, under standard regularity conditions, the maximum likelihood estimator (MLE) satisfies a central limit theorem and converges in distribution to a Gaussian random variable as the sample size grows. This paper strengthens this classical result by developing several stronger forms of asymptotic normality for the normalized MLE. With additional assumptions on the score, we first establish sub-Gaussian tail bounds and convergence of all moments for the normalized estimation error. We then prove an entropic central limit theorem for a smoothed version of the estimator, showing convergence in relative entropy to the limiting Gaussian law. When the Fisher information of the normalized estimate is bounded, or its density has bounded first derivative, we further show that the smoothing can be removed, yielding entropic normality of the MLE itself. The proofs develop auxiliary tools that may be of independent interest, including exponential consistency bounds, high-moment estimates, and entropy-control arguments for the estimator.
Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning
arXiv:2607.08647v1 Announce Type: new Abstract: As autonomous agents are increasingly deployed across diverse operational contexts, aligning their behavior with human intent demands reward functions that remain robust to such changes rather than overfitting to any single environment. Inverse reinforcement learning (IRL) provides a principled way to infer such objectives from human feedback. However, existing analyses of optimal teaching approaches for IRL focus on single-environment, demonstration-only settings, leaving underexplored how heterogeneous feedback modalities and environment dynamics jointly constrain reward functions that generalize across multiple environments. Because demonstrations in one MDP entangle reward information with that environments specific structure, the resulting rewards frequently fail to generalize when the agent is deployed in a new setting. We first analyze how different feedback modalities constrain rewards, showing that, in the unlimited-data regime, comparisons impose strictly stronger global constraints than other modalities. Beyond this theoretical analysis, we introduce a hierarchical machine teaching algorithm for reward learning that operates across multiple MDPs. The algorithm first greedily selects informative environments that expose complementary reward constraints, then strategically queries low-cost feedback within those environments. Empirically, our method achieves substantially lower regret and stronger generalization to held-out environments than uniform teaching baselines under identical feedback budgets, demonstrating the importance of multi-environment, multi-modal teaching for learning dynamics-robust reward functions.
Secure Decentralized Federated Learning via Gossip and Virtual Voting
arXiv:2607.08651v1 Announce Type: new Abstract: Decentralized federated learning (DFL) removes the central server by letting nodes exchange model updates through peer-to-peer gossip, but existing gossip-based methods often lack provenance finality and resilience to Byzantine or lazy participants. Ledger-assisted federated learning (FL) improves auditability, yet blockchains, shards, or settlement committees can reintroduce global coordination costs that conflict with DFL locality. This paper proposes \emph{gspDAG-FL}, a secure DFL framework that derives consensus from the same gossip history used to disseminate models. Nodes exchange model payloads only with neighbors, while full nodes collect event certificates and receiver-endorsed accepted gossip proofs, reconstruct a compact Topology directed acyclic graph (DAG), and run Hashgraph-style virtual voting followed by compact full-node certificates. Finality is over unique model-origin tuples, not identical local parameter states. To improve resilience, gspDAG-FL combines payload validation, accepted-proof validation, and private semantic audit before aggregation. We formalize the adversarial setting, prove safety and conditional liveness of the control plane, and give a convergence guarantee for certified perturbed gossip under time-varying effective mixing. Experiments on MNIST classification and Penn Treebank language modeling, using fair held-out validation/audit data and networks up to \(N=100\), show that gspDAG-FL achieves learning quality close to validation-based ledger FL while reducing coordination bottlenecks, improving throughput, and maintaining high invalid-origin detection under mixed Byzantine and lazy participation.
BiNoMaP: Learning Category-Level Bimanual Non-Prehensile Manipulation Primitives
arXiv:2509.21256v3 Announce Type: replace Abstract: Non-prehensile manipulation, encompassing ungraspable actions such as pushing, poking, pivoting, and wrapping, remains underexplored due to its contact-rich and analytically intractable nature. We revisit this problem from two perspectives. First, instead of relying on single-arm setups or favorable environmental supports (e.g., walls or edges), we advocate a generalizable dual-arm configuration and establish a suite of Bimanual Non-prehensile Manipulation Primitives (BiNoMaP). Second, departing from prevailing RL-based approaches, we propose a three-stage, RL-free framework for learning structured non-prehensile skills. We begin by extracting bimanual hand motion trajectories from egocentric video demonstrations. Since these coarse trajectories suffer from perceptual noise and morphological discrepancies, we introduce a geometry-aware post-optimization algorithm to refine them into executable manipulation primitives consistent with predefined motion patterns. To enable category-level generalization, the learned primitives are further parameterized by object-relevant geometric attributes, primarily size, allowing adaptation to unseen instances with significant shape variations. Importantly, BiNoMaP supports cross-embodiment transfer: the same primitives can be deployed on two real-world dual-arm platforms with distinct kinematic configurations, without redesigning skill structures. Extensive real-robot experiments across diverse objects and spatial configurations demonstrate the effectiveness, efficiency, and strong generalization capability of our approach.
Mapping fat-water separated R1, R2*, and proton density fat fraction with the multi-echo MP2RAGE sequence
arXiv:2509.23136v2 Announce Type: replace Abstract: Purpose: To develop a technique for joint measurement of fat and water-specific longitudinal relaxation rates (R1f and R1w), effective transverse relaxation rate (R2*), and proton density fat fraction (PDFF) combining the Multi-Echo Magnetization Prepared Two Rapid Acquisition of Gradient Echoes (ME-MP2RAGE) sequence and fat-water separation. Theory and Methods: R1f and R1w were calculated with fat-specific and water-specific MP2RAGE signals. R2* and PDFF maps were obtained from fat-water separation applied to the second RAGE block. Sequence parameters optimization was performed via Cram\'er-Rao lower bounds theory, and we designed four protocols with different combinations of number of echoes and readout gradient schemes (I: 3 echoes unipolar, II: 6 echoes unipolar, III: 6 echoes bipolar, and IV: 10 echoes bipolar). We tested and validated these protocols with numerical simulations, phantom and in vivo experiments. In phantoms, we compared ME-MP2RAGE measurements with inversion recovery spin-echo (IR-SE) global R1 and 3D Fast Low Angle Shot (3D FLASH) R2* and PDFF. In vivo, we scanned the lower leg and neck of a healthy volunteer. Results: Numerical simulations showed accurate quantification of relaxation rates with mean relative bias < 3% and PDFF with mean bias < 0.003 using protocol ME-MP2RAGE IV (10 echoes bipolar). Phantom experiments showed excellent agreement with IR-SE and 3D FLASH measurements. In vivo, measurements in the lower leg and neck were consistent with literature values. Conclusion: We proposed an accurate method for simultaneous quantification of R1f, R1w, R2*, and PDFF from a single acquisition with the ME-MP2RAGE sequence.
Open-ended Multi-agent Autocurricula via Visual Inspection of Policies with Multi-modal LLMs
arXiv:2607.08193v1 Announce Type: new Abstract: Open-ended curricula in Reinforcement Learning (RL) aim to train generally-capable agents by identifying tasks that facilitate learning increasingly complex skills. A major challenge when designing such curricula is assessing task difficulty relative to the agent's current learning progress. While previous work has explored using scalar task scores or textual summaries of the agent's behavior, here we study a different approach: directly inspecting policy behavior via recorded episode videos. We introduce a simple yet effective instantiation of this approach which leverages a Video Language Model (VLM) to both process these videos and provide curriculum recommendations, which we call Visual Inspection of Policies (VIP). Since videos can naturally contain any number of controllable agents, we empirically study VIP on the StarCraft Multi-Agent Challenge (SMAC). We show that even with a lightweight and openly accessible VLM (VideoLLaMa2-7B), VIP can use policy videos to generate more effective curricula than both its text-only ablation and methods that rely on scalar task scores.
Formal Mechanisms for Market Stability in Self-Interested Agent Societies: A Marketplace Simulation Study
arXiv:2607.08652v1 Announce Type: new Abstract: Self-interested agents, left unconstrained, tend toward defection in repeated social dilemmas, causing cooperative gains from trade to collapse. This paper investigates what formal mechanisms, layered on top of unrestricted communication, are sufficient for a society of such agents to maintain market stability, and how resilient those mechanisms are to adversarial attack. We instantiate the research question as a multi-agent marketplace simulation where 18 LLM agents (DeepSeek-V3) with complementary production specialties must trade within a constrained social network to obtain utility. We conduct two experimental phases: (1) a mechanism comparison across eight conditions under progressive troll injection over 200 rounds, identifying Mediation as the top-performing mechanism; and (2) adversarial red-teaming of Mediation using iteratively prompt-optimised LLM-driven trolls, finding that the best attack (v6) reduces honest-agent utility by 13.3% but cannot collapse the market. Mediation enables recovery even under sustained adversarial pressure. We define adversarial robustness as a mechanism's ability to sustain positive honest-agent utility under optimised attack, and find that Mediation is robust: it can be bent but not broken.
LFPL: Revisited and Mechanized
arXiv:2605.12893v4 Announce Type: replace Abstract: Hofmann (1999) introduced the functional programming language LFPL to characterize the functions computable in polynomial time using an affine type system. LFPL enables a natural programming style, including nested recursion, and has inspired the development of type systems for automatic cost analysis, linear dependent type theories, and efficient memory management in functional programming languages. Despite its prominence, there does not exist a self-contained presentation, let alone a full mechanization, of LFPL and its core metatheory. This article presents a modern account and mechanization of LFPL and its metatheory with the goal of being self-contained and accessible while streamlining the strongest-known soundness and completeness results. The soundness proof works with the language LFPL+, which extends LFPL with additional language features. The proof is novel, adapting a technique by Aehlig and Schwichtenberg (2002) to construct explicit polynomials that bound the cost of an LFPL+ expression with respect to a big-step cost semantics. The completeness proof shows that LFPL programs can simulate polynomial-time Turing machines while only relying on restricted forms of linear functions and lists. It has the same structure as the original proof by Hofmann (2002) but greatly simplifies the core argument with a novel stack-like data structure that is implemented with first-class functions and lists. The mechanization includes the full soundness and completeness proofs, and serves as one of the first case studies of mechanized metatheory in the recently developed proof assistant Istari.
Asynchronous Federated Continual Segmentation with Evolving Clients and Label Spaces
arXiv:2503.15414v3 Announce Type: replace-cross Abstract: Federated learning seeks to foster collaboration among distributed clients while preserving the privacy of their local data. Traditional federated learning methods typically assume a fixed setting, where participating clients, client data, and learning objectives remain unchanged. However, in real-world scenarios, a federation may evolve over time, with changes in both its client composition and target label space. In this evolving federated setting, conventional round-wise model aggregation becomes inflexible, as each federation update requires repeated communication, repeated local computation, and synchronized participation from all accumulated clients. To address this limitation, we propose CA-MMDS, a continual multiple-model distillation framework for federated continual segmentation with asynchronous clients and evolving label spaces. Instead of repeatedly aggregating model parameters from all clients, CA-MMDS maintains a server-side archive of client models and updates the global model through proxy-based distillation from multiple archived local models. When new clients join or existing clients evolve, only the newly added or updated local models need to be uploaded, while unchanged clients can remain offline and continue to contribute through their archived models. This design substantially reduces communication and computation costs while enabling flexible asynchronous cooperation among evolving clients. Using multi-class 3D abdominal CT segmentation as an application task, we demonstrate that CA-MMDS efficiently incorporates evolving client knowledge while achieving competitive segmentation performance.
MasFACT: Continual Multi-Agent Topology Learning via Geometry-Aware Posterior Transfer
arXiv:2605.17361v2 Announce Type: replace Abstract: Multi-agent systems (MAS) powered by large language models (LLMs) have emerged as a powerful paradigm for complex problem solving, where performance critically depends on the underlying inter-agent communication topology. However, existing topology generation methods mainly optimize for isolated tasks, while real-world deployments involve streams of evolving tasks, requiring previously effective collaboration patterns to be retained and reused rather than rediscovered or overwritten. We identify a previously underexplored failure mode, \emph{topology forgetting}, in which adapting to new tasks shifts the topology generator away from communication structures required by earlier tasks. This issue stems from cross-task misalignment in both agent-level functional semantics and relational communication structures. To address this challenge, we propose \textbf{\textsc{MasFACT}}, a geometry-aware posterior transfer framework that preserves and reuses historical collaboration knowledge as transferable topology priors. We transfer these priors across task-specific agent spaces through Fused Gromov-Wasserstein optimal transport and perform PAC-Bayes-guided conservative posterior adaptation to balance task-specific plasticity with structural stability. Experiments across class-, domain-, and task-level continual settings demonstrate that \textsc{MasFACT} consistently improves average accuracy while reducing topology forgetting compared to strong topology generation and replay-based baselines, and can be seamlessly integrated with different MAS topology generators.
WebSwarm: Recursive Multi-Agent Orchestration for Deep-and-Wide Web Search
arXiv:2607.08662v1 Announce Type: new Abstract: Large language model (LLM)-based web search agents are transforming information seeking from simple factoid question answering into complex, deep-and-wide search and research-oriented tasks. A single ReAct-style agent is constrained by one long trajectory and limited context, making it difficult to handle depth and coverage simultaneously. Existing multi-agent systems improve search coverage through parallel execution and aggregation, but still exhibit clear limitations in recursive depth, collaboration adaptability, and evidence-grounded expansion. We propose WebSwarm, a progressive recursive delegation framework that jointly constructs task decomposition, recursive expansion, and agent collaboration during inference. WebSwarm dynamically instantiates agentic search nodes, each coupling a local objective with a search mode that specifies how the node should organize search and collaboration. Each node can either solve its objective itself or further delegate child nodes; after solving, it returns evidence and results upward, enabling parent nodes to further expand, revise, or aggregate the search process. To guide this process, WebSwarm first probes how task-relevant information is organized on the web to ground subsequent node expansion, and reuses process-level experience across homogeneous sibling nodes. Experiments on BrowseComp-Plus, WideSearch, DeepWideSearch, and GISA show that WebSwarm consistently outperforms single-agent and multi-agent baselines on deep, wide, and interleaved deep-and-wide tasks. Further analyses of ablation, task difficulty, web tool efficiency, and model generalization explain WebSwarm's effectiveness and provide insights for multi-agent search systems.
TRM-Raft: A Byzantine-Resistant Raft Consensus via Integrated Trust and Reputation Model
arXiv:2607.08666v1 Announce Type: new Abstract: Internetware envisions autonomous software entities collaborating over the open Internet. Raft consensus is widely adopted for its simplicity and performance in distributed coordination, e.g., service registries and blockchains. However, Raft assumes crash faults only, making it vulnerable to Byzantine behaviors like election forgery and log tampering. Existing BFT protocols incur high overhead, while ad-hoc hardening lacks unified defense. We propose \textbf{TRM-Raft}, a Byzantine-resistant enhancement that non-intrusively integrates a Blockchain-based Trust and Reputation Model (B-TRM) into the consensus core. It quantifies multi-dimensional node behaviors, applies adaptive penalties distinguishing accidental faults from malice, and embeds reputation into leader election and log replication. A reputation-aware election penalizes term/index forgery, excluding low-reputation nodes from leadership. A Schnorr-signature-based mechanism lets followers verify log integrity; tampering triggers reputation decay and leader replacement. Evaluated on Hyperledger Fabric in a realistic Internetware setting, TRM-Raft keeps malicious leader ratio below 5\% even with 40\% Byzantine nodes, with <10\% throughput loss and <5\% latency increase over vanilla Raft. TRM-Raft offers a lightweight, practical trustworthiness path for Internetware systems relying on Raft.
Optimal Debiased Inference on Privatized Data via Indirect Estimation and Parametric Bootstrap
arXiv:2507.10746v3 Announce Type: replace-cross Abstract: We design a debiased parametric bootstrap framework for statistical inference from differentially private data. Existing usage of the parametric bootstrap on privatized data ignored or avoided handling possible biases introduced by the privacy mechanism, such as by clamping, a technique employed by the majority of privacy mechanisms. Ignoring these biases leads to under-coverage of confidence intervals and miscalibrated type I errors of hypothesis tests, due to the inconsistency of parameter estimates based on the privatized data. We propose using the indirect inference method to estimate the parameter values consistently, and we use the improved estimator in parametric bootstrap for inference. To implement the indirect estimator, we present a novel simulation-based, adaptive approach along with the theory that establishes the consistency of the corresponding parametric bootstrap estimates, confidence intervals, and hypothesis tests. In particular, we prove that our adaptive indirect estimator achieves the minimum asymptotic variance among all ``well-behaved'' consistent estimators based on the released summary statistic. Our simulation studies show that our framework produces confidence intervals with well-calibrated coverage and performs hypothesis testing with the correct type I error, giving state-of-the-art performance for inference in several settings.
Perceptually Lossless Tactile Texture Synthesis with Compact Spectral Envelope Models
arXiv:2605.23804v2 Announce Type: replace Abstract: Modern audio-visual media rely on compact representations for efficient storage and transmission, whereas realistic digital touch still depends on high-resolution tactile recordings. Existing approaches for representing tactile signals constrain manipulation and limit the generation of new content. Here, we introduce two compact representations, spectral beta and spectral slope, that capture the temporal spectral structure of finger-surface friction signals while preserving perceptually relevant information. Spectral beta models spectral skewness using a two-parameter beta distribution, whereas spectral slope approximates the spectrum with an asymmetric bandpass filter defined by low- and high-pass orders. We evaluated these representations in a perceptual study with 14 participants using five virtual textures rendered on a friction-modulation display and compared them with physical textures and high-fidelity reproductions of recorded signals. Spectral beta achieved perceptual similarity ratings comparable to those of the original high-fidelity reproductions. Regression analysis further showed that matching spectral energy across nine critical frequency bands was the strongest predictor of perceived realism. Together, these findings suggest that tactile texture perception depends primarily on fundamental temporal spectral patterns and that modeling these patterns is sufficient for perceptually realistic rendering. These results establish an efficient and scalable framework for haptic compression, communication, and synthetic texture generation.
Mapping the strong-to-weak coupling crossover in polymer-film microcavity lasers
arXiv:2510.13384v3 Announce Type: replace Abstract: Organic semiconductors are particularly attractive for polaritonics due to their large exciton binding energies and oscillator strengths. Among them, the ladder-type conjugated polymer poly(paraphenylene) is distinguished by its rigid backbone, narrow exciton linewidth, high photoluminescence quantum yield, and enhanced photostability, making it an excellent candidate for organic polariton devices. While polariton lasing has been reported in various organic systems, systematic studies of the transition from polariton lasing to conventional photon lasing within a single, well-controlled material platform remain limited. Here, we present planar organic microcavities incorporating MeLPPP as the active medium, in which continuous tuning of the effective cavity length within a single device enables us to map the strong-to-weak coupling transition across five distinct cavity-mode orders. We demonstrate an approximately eighteen-fold increase in the lasing threshold when crossing from polariton to photon lasing. We further establish a quantitative framework in which the spectral dependence of the threshold governs a universal V-shaped blueshift of the emission energy across both coupling regimes. Finally, we show that vibron-mediated exciton relaxation, previously identified in the strong-coupling limit, persists across the crossover: lasing-threshold minima track the vibron resonances throughout the coupling transition.
Inverse Transfer and Coherence in Rotating Stratified Flow with Clouds and Phase Transitions
arXiv:2607.08673v1 Announce Type: new Abstract: Inverse energy transfer to large-scale coherent structures in idealized models of geophysical flows has been of interest for over four decades. Extensive knowledge exists regarding inverse transfer in rotating and stratified dry dynamics, characterized by the Rossby number and a single dry Froude number. The current study includes effects of water and phase changes, with dynamics characterized by the Rossby number and two Froude numbers for unsaturated and saturated environments. Using numerical computations with random forcing, inverse energy transfer is examined for a model with a Boussinesq dynamical core, incorporating water vapor and liquid water in the limit of asymptotically-fast cloud microphysics. Besides kinetic energy, total energy includes buoyant potential energies from each phase, and latent moist energy responsible for potential energy transfer at phase boundaries. The rotation and stratification terms are large and comparable, such that the dry version of the evolution equations is dominated by inverse transfer of pseudo potential vorticity(PV). For fixed Rossby and dry (unsaturated) Froude numbers, compared to dry dynamics, there is a reduction in energy transfer rate, associated with the larger Froude number of saturated regions. The upscale transfer to moist PV is influenced by nonlinear waves at lowest order resulting from nonlinear buoyancy near phase interfaces. These nonlinear waves lead to coherent updrafts and downdrafts roughly aligned with fuzzy, large-scale phase boundaries identified by the time average of a cloud indicator function. Statistical relationships between phase boundaries, updrafts/downdrafts and moist PV are explored in flow regions dominated by moist PV-vortices.
OREN: Octree Residual Network for Real-Time Euclidean Signed Distance Mapping
arXiv:2510.18999v3 Announce Type: replace Abstract: Reconstructing signed distance functions (SDFs) from point cloud data benefits many robot autonomy capabilities, including localization, mapping, motion planning, and control. Methods that support online and large-scale SDF reconstruction often rely on discrete volumetric data structures, which affects the continuity and differentiability of the SDF estimates. Neural network methods have demonstrated high-fidelity differentiable SDF reconstruction but they tend to be less efficient, experience catastrophic forgetting and memory limitations in large environments, and are often restricted to truncated SDF. This work proposes OREN, a hybrid method that combines an explicit prior from octree interpolation with an implicit residual from neural network regression. Our method achieves non-truncated (Euclidean) SDF reconstruction with computational and memory efficiency comparable to volumetric methods and differentiability and accuracy comparable to neural network methods. Extensive experiments demonstrate that OREN outperforms the state of the art in terms of accuracy and efficiency, providing a scalable solution for downstream tasks in robotics and computer vision.
Active control of the peak value of the Hanbury Brown-Twiss effect using coherent light by lensless holographic projection
arXiv:2510.20421v2 Announce Type: replace Abstract: Computer-generated holography enables projection of target patterns onto designated planes, providing deterministic control over the probability density function of the projected light intensity. Here, we introduce an active control scheme for the peak value of the Hanbury Brown--Twiss effect, $g^{(2)}(0)$, utilizing lensless holographic projection with coherent light. Notably, single-frame holographic projection yields a markedly different $g^{(2)}(0)$ from its multiframe-averaged counterpart due to the presence of coherent speckle noise. With the coherent speckle noise suppression, we derive an analytical expression $g^{(2)}(0)$ on holographic projection plane, revealing that it is determined by the target coherence length, its statistics, and the numerical aperture of projection system. Our experimental results show good agreement with the theoretical analysis, confirming the joint influence of these factors. By employing dynamic sparse target patterns, we achieve a maximum $g^{(2)}(0)$ of $39.77$. Numerical simulations, benchmarked against experimental measurements, reveal that coherent speckle noise enhances $g^{(2)}(0)$ through mutual superposition with the target pattern, leading to a joint modulation of intensity fluctuations. In summary, by manipulating multiple controllable parameters, we establish a robust strategy for tailoring $g^{(2)}(0)$, paving the way for advanced applications in speckle imaging and optical metrology.
TrajAudit: Automated Failure Diagnosis for Agentic Coding Systems
arXiv:2605.26563v2 Announce Type: replace Abstract: Agentic systems have been widely studied to automate coding tasks such as bug fixing and feature implementation. As these systems increasingly operate on complex codebases, understanding where and why they fail becomes essential for iterative refinement and operational reliability. Existing automated failure diagnosis approaches leverage \textit{task execution trajectories}, yet they struggle with trajectories produced by repository-level coding agents due to two key properties. First, these trajectories are often long, spanning many execution steps, making it difficult for LLMs to track the causal chain of failure over the execution history. Second, these trajectories are laden with noise, containing substantial low-signal observations such as redundant program structures and verbose code context, which can interfere with LLM reasoning. To address these challenges, we propose \textit{TrajAudit}, an automated failure diagnosis framework specifically for trajectories produced by repository-level coding agents. TrajAudit employs an investigator agent supported by two modules: one reduces failure-irrelevant noisy context through semantic saliency folding, and the other derives preliminary diagnostic guidance from test failure reports as prior knowledge to help LLMs focus on likely failure regions. The investigator agent can further invoke tools to inspect folded content on demand, enabling a focused investigation without losing access to the full trajectory context. We also introduce \textit{RootSE}, a benchmark of 102 real-world instances from repository-level coding tasks, each annotated with the earliest decisive error step and a justification. Experiments on RootSE show that TrajAudit outperforms the strongest baselines by 10.8\% and 21.6\% in exact failure localization accuracy in the with- and without-reference settings, respectively, demonstrating its effectiveness.
Residual energy in weakly compressible turbulence with a mean guide field
arXiv:2512.11973v2 Announce Type: replace-cross Abstract: The energy distribution is a fundamental property of magnetohydrodynamic (MHD) turbulence. In strongly magnetized turbulence energy imbalances arise and are quantified by the residual energy: $E_r~=~(E_{kin}~ - ~E_{mag})$; $E_{kin}$ and $E_{mag}$ stand for the volume-averaged kinetic and magnetic energy, respectively. We explore the properties of $E_r$ in weakly compressible MHD turbulence in the presence of an initially strong (guide) magnetic field, investigating how the driving mechanism and the magnetic field strength affect the cascade of $E_r$. We run a suite of direct numerical simulations with the PENCIL code. The sonic Mach number is approximately equal to 0.1 in all simulations, whereas the plasma beta varies. We drive turbulence by either injecting velocity or magnetic fluctuations at large scales and study the power spectra of kinetic, magnetic, density, and $E_r$. Magnetically driven simulations show locally imbalanced Alfv\'enic fluctuations and a $\propto k^{-3/2}$ cascade, consistent with the dynamic alignment theory. In the inertial range, $E_r \approx$ 0. Kinetically driven simulations give rise to a $\propto k^{-1}$ scaling, consistent with weakly interacting modes that preserve a high level of coherence throughout the inertial range. Residual energy is positive at all scales of the inertial range. The spectral slope of the $E_r$ cascade steepens systematically with increasing magnetization, varying from approximately -1 at $\beta = 0.3$ to between -2.0 and -5/3 at $\beta = 4.0$. The energy partition in weakly compressible turbulence is strongly influenced by the forcing mechanism, even when the global sonic and Alfv\'enic Mach numbers are comparable across simulations.
Multi-Sender Bayesian Persuasion with Imperfect Information
arXiv:2607.08675v1 Announce Type: new Abstract: We study a multi-sender Bayesian persuasion problem with one receiver and several strategic senders. The underlying ground state has multiple components, each privately observed by a different sender, while the receiver holds a common prior over the joint state space. Senders simultaneously choose signaling policies, and the receiver takes an action based on the posterior induced by the signals; each is sampled independently from the sender's signaling policy. We analyze the game induced by the receiver's straightforward policy, which selects a receiver-optimal action at every posterior. In particular, we characterize the senders' best responses under the straightforward policy and identify conditions on the prior that induce a fully informative equilibrium; i.e., truthfully reporting the ground truth is an equilibrium strategy for every sender. These conditions capture cases in which senders' incentives are sufficiently aligned to enable full revelation without additional commitment from the receiver. The important contribution of this paper is to analyze games induced by a more general (possibly randomized) class of action policies that the receiver commits to before senders choose their signaling strategies. We show that this commitment power fundamentally changes the problem. In particular, we show that for any prior over the joint state space, the receiver can construct action policies that maximize her payoff while ensuring a fully informative equilibrium.
Global Magnetohydrodynamic Simulations of Monster Shocks in Neutron Star Magnetospheres
arXiv:2602.21290v3 Announce Type: replace-cross Abstract: Waves launched from the neutron star surface or inner magnetosphere propagate through the magnetosphere as small perturbations, but can grow relative to the background magnetic field and steepen into ``monster shocks'' -- ultra-relativistic magnetized shocks which can power high-energy emission. Such shocks can develop around isolated magnetars, merging binaries, and collapsing neutron stars. They occur in magnetically dominated plasma and are described by relativistic magnetohydrodynamics (MHD). We present global relativistic MHD simulations of monster shocks in unperturbed and perturbed (``wrinkled'') backgrounds with a global dipolar geometry. Our simulations confirm analytical predictions for equatorial shocks and provide new insight into the behavior of oblique shocks off the equator. Simulations where the shock is formed through Alfv\'{e}n mode to fast mode conversion are also presented, demonstrating the generic nature of the monster shock mechanism. We explore how the presence of additional modes in the magnetosphere modifies the shock behavior. Modes of comparable amplitude can fragment the shock front, substantially reduce the magnetization, produce localized enhancements in the Lorentz factor relative to an unperturbed dipole background, and intermittently generate additional shocks along a line of sight.
Drift of interfaces in forced stably-stratified turbulence and the role of vertically-sheared helical structures
arXiv:2607.08678v1 Announce Type: new Abstract: Experimental investigations of forced stably stratified turbulence (SST) have shown that the step-like density profile, made of well-mixed density layers and sharp interfaces alternating along the gravity direction, undergo a slow coarsening dynamics with either decay or merging of interfaces. In this Letter, we focus on the coarsening dynamics phenomenon, by means of Direct Numerical Simulations of forced SST at moderate resolutions, and very long temporal integration. We show that the vertical drift and merging of interfaces is associated to the emergence of spatially-uniform, vertically-sheared helical structures that break the mirror-symmetry of the system. When these are absent, interfaces decay is observed instead. %how the kinetic energy excursions observed at $Fr=0.076$, occurring in parallel to vertical drift of interfaces, are due to the emergence of spatially-uniform, vertically-sheared helical structures that break the mirror-symmetry of the system. This is absent at larger $Fr=0.22$, where interface decay is observed instead. A dynamical correspondence between helicity dissipation rate by buoyancy effects and the vertical buoyancy flux allows to establish a (causal) connection between the chiral structures and the vertical movement of interfaces leading to merging.
Multi-Resolution Feature Stem for Diabetic Retinopathy lesion segmentation
arXiv:2607.08679v1 Announce Type: new Abstract: Diabetic Retinopathy (DR) is a leading cause of preventable blindness worldwide, requiring automated lesion segmentation using deep learning models for early detection and monitoring. However, DR lesions vary dramatically in size from tiny microaneurysms to large hemorrhages and exudates. This variability creates conflicting demands on the model architecture and input resolution, posing a challenge for effective design. This work investigates the impact of input resolution on different lesion types. Through systematic experimentation with multiple architectures (U-Net, UNet++, Vision Transformers, DeepLabV3+) at $512 \times 512$ and $1024 \times 1024$ resolutions, we identify a critical, counter-intuitive phenomenon where increasing input resolution has opposing effects on different lesion types. We demonstrate that while higher resolution is essential for resolving fine-grained microaneurysms, it can unexpectedly degrade performance on larger hemorrhages. This finding challenges the common assumption that higher resolution is uniformly beneficial. To address this, we propose a novel Multi-Resolution Feature Stem, an input-level pyramid integrated with a UNet++ backbone. This architecture processes multiple scales in parallel, capturing fine-grained details without sacrificing contextual information. This work contributes crucial empirical evidence of this complex, resolution-dependent behavior and a practical, parameter-efficient architecture that successfully resolves this trade-off.