Forskningsradar

Science Journals

Peer-reviewade publikationer — 54780 artiklar

ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models
arXiv:2502.09696v3 Announce Type: replace Abstract: Large Multimodal Models (LMMs) exhibit shortfalls when interpreting images and, by some measures, have poorer spatial cognition than young children or animals. Despite this, they attain high scores on many popular visual benchmarks, with headroom rapidly eroded by model progress. This creates a need for difficult benchmarks that remain relevant for longer. We introduce ZeroBench - a lightweight visual reasoning benchmark curated using adversarial filtering to be "impossible" for frontier LMMs at its original release, with initial SotA scores of 0% pass@1 and pass^5. We track progress on ZeroBench over the subsequent year, observing SotA reaching 6% pass^5 and 19% pass@5, indicating the potential longevity of the benchmark. We evaluate 46 LMMs on ZeroBench, compare performance to a human baseline, analyse strengths and weaknesses, chart a year of progress in visual capabilities, and publicly release ZeroBench at https://zerobench.github.io.
Ethics and EU AI Act in Cases of Work Disability Risk and Alzheimer's Disease Risk Prediction
arXiv:2607.05402v1 Announce Type: new Abstract: Improvements in AI technologies have made it feasible to develop new types of medical AI tools. However, these tools raise new kinds of questions, especially in relation to the ethics and AI Act compliance. We analyzed two cases of AI tools developed to predict medical risks, the risk of work disability (case A) and the risk of getting Alzheimer's disease (case B). We observed both cases using the ethical AI and the EU AI Act as frameworks, noted that they classify as high-risk systems, and that bringing them from the research environment to production would require a lot of work and compliance due to the related regulation.
Inverse-designed photonic interfaces beyond eigenmode expansion limits
arXiv:2607.06243v1 Announce Type: new Abstract: Photonic integrated circuits (PICs) enable optical systems with dramatically increased performance, cost-effectiveness, and scalability through enhanced light-matter interactions, high-density integration, and mass production. Due to the significant mode mismatch between various integrated photonic platforms and optical fibers, spot-size conversion interfaces with low-loss, compact footprint, and high manufacturability are essential. Conventional spot-size converters based on intuitive designs often require multi-layer tapering structures and tiny waveguide tips to adiabatically expand the eigenmodes. These rigid design constraints commonly lead to large device footprints and the requirements of multiple high-precision lithography steps. In this paper, we overcome these limitations using inverse design methods, which optimize the coupling efficiency over a large parameter space beyond traditional eigenmode evolution limits. Specifically, we demonstrate efficient and ultra-compact photonic interfaces on the thin-film lithium niobate (TFLN) platform, where the partially etched rib waveguides and non-vertical sidewalls have previously hindered the achievement of low-loss waveguide tapers in single-layer configurations. Our inverse-designed photonic structures achieve simulated and experimentally measured coupling efficiencies as low as 1 dB and 3 dB per facet between TFLN waveguides and lensed/ultra-high numerical aperture (UHNA) fiber, with broad 1-dB bandwidths exceeding 120 nm. The inverse-designed interfaces are highly compatible with standard TFLN PIC components and require only a single high-resolution lithography step. More importantly, the design concept transcends traditional eigenmode evolution theories and is broadly applicable to a variety of material platforms and application scenarios.
The Minimum Dominating Set Problem on Bipartite Circle Graphs: Complexity and Approximation
arXiv:2607.06251v1 Announce Type: new Abstract: A circle graph is the intersection graph of a set of chords in a circle. A dominating set of a graph $G=(V,E)$ is a subset $D\subseteq V$ such that every vertex in $V\setminus D$ is adjacent to at least one vertex of $D$. Computing a minimum dominating set is known to be NP-hard on circle graphs. In this paper, we study the minimum dominating set problem on bipartite circle graphs, namely, circle graphs admitting a chord representation in which the chords can be partitioned into two color classes such that no two chords of the same color intersect. We prove that the problem remains NP-hard for this restricted graph class by a reduction from Planar Monotone 3-SAT. On the positive side, we present a polynomial-time 2-approximation algorithm and develop a polynomial-time approximation scheme (PTAS) based on local search.
Dirac Spin Liquid Candidate in a Rydberg Quantum Simulator
arXiv:2602.14323v3 Announce Type: replace-cross Abstract: We experimentally investigate a frustrated spin-exchange antiferromagnet in a quantum simulator, composed of N = 114 dipolar Rydberg atoms arranged into a kagome array. Motivated by a recent theoretical proposal of a gapless U(1) Dirac spin liquid ground state, we use local addressing to adiabatically prepare low-energy states. We measure the local polarization and spin-spin correlations over this adiabatic protocol, and observe our system move from a staggered product state, through an intermediate magnetic crystal, and finally into a disordered, correlated liquid. We estimate the entropy density of this atomic liquid to be similar to that of frustrated magnetic insulators at liquid nitrogen temperatures. We compare the correlations in our liquid to those of a simple, parameter-free ansatz for the Dirac spin liquid, and find good agreement in the sign structure and spatial decay. Finally, we probe the static susceptibility of our system to a local field perturbation and to a geometrical distortion. Our results establish Rydberg atom arrays as a promising platform for the preparation and microscopic characterization of quantum spin liquid candidates.
A Unique Normal Form for Tensor Trains over Arbitrary Fields
arXiv:2607.06271v1 Announce Type: new Abstract: Tensor trains (or Matrix-Product States) are a data structure used in many fields of computer science and physics. They were recently shown to generalise binary decision diagrams when used over the 2-element Galois field, prompting the question of their reducibility in such a context, when the standard approach, over real or complex number, is not amenable to finite fields. We provide here a unique normal form and associated polynomial-time reduction strategy for tensor trains over arbitrary fields. We also show how to directly extract a normal form out of a full tensor, how to get the leading index and value of a normal form, and an upper bound on the size of a fully-reduced tensor train relative to a naive storage of the full tensor. On the one hand, this work strengthens the use of tensor trains as a relevant formal tool. On the other hand, from the perspective of tensor networks, it extends the formalism to more general settings than the well-studied real and complex fields, and crucially provides the first tensor train form with the uniqueness property.
Matter-wave Induced Transparency
arXiv:2607.03820v2 Announce Type: replace-cross Abstract: Electromagnetically induced transparency suppresses optical absorption through destructive interference, playing a central role in light-matter interaction and quantum information science. We report matter-wave induced transparency, where atomic collisional interactions induce transmission through a lossy molecular potential for the incident atomic scattering waves. Using cesium Bose-Einstein condensates and modulation-induced Feshbach resonances, we realize a three-level atom-molecule coupled system with unprecedented flexibility. Under the dark state condition, a narrow and tunable transparency window appears within a broad dissipative collisional resonance. The transparency window linewidth is controlled by modulation-induced coupling. And scattering pathways are selectable via multifrequency Floquet modulation. These results establish an interference-based route for exploring programmable nonequilibrium and non-Hermitian physics, steering quantum chemistry and precision measurements.
Robust Bayes-Assisted Conformal Prediction
arXiv:2607.04236v2 Announce Type: replace-cross Abstract: Bayes-assisted conformal prediction combines the strengths of Bayesian modelling with exact, distribution-free frequentist coverage guarantees. Although conformal validity is preserved even when the Bayesian working model (BWM) is misspecified, the size of the resulting prediction sets can degrade substantially when the prior is poorly aligned with the observed data. We address this limitation by introducing RoBAS (Robust Bayes-Assisted Shrinkage): a Bayes-assisted framework for constructing robust nonconformity scores, with two instantiations: one induced by a heavy-tailed BWM, and a closed-form empirical Bayes shrinkage score. The resulting scores adapt to the quality of the working information encoded in the prior: when this information is reliable, they exploit it to produce efficient prediction sets; when it is weak or inaccurate, they revert to the Distance-To-Average (DTA) score, a robust non-informative baseline. We evaluate the proposed scores on tabular and image regression tasks where the training distribution may differ from the calibration and test distributions, while the calibration and test data themselves remain exchangeable. We find that they are competitive with widely used scores in the absence of such shift, while substantially reducing interval widths in shifted settings.
Collective Cognition in Hybrid Groups: A Network Science Synthesis
arXiv:2607.05593v1 Announce Type: new Abstract: The growing integration of AI agents into human teams calls for a principled understanding of how collective intelligence emerges in hybrid systems. Recent frameworks clarify how attention, memory, and reasoning differences shape human-AI interaction at the individual and dyadic levels, but a formal account of how these differences scale to group-level dynamics is lacking. Most network science has examined either human-only or multi-agent AI-only systems, leaving open how its findings and parametrizations translate to hybrid groups. This chapter synthesizes network science, collective cognition, and multi-agent systems through the lens of attention, memory, and reasoning. We review how task environments, group topologies, agent-level processes, and incentive structures shape collective outcomes in human-only and AI-only networks, then examine how these results extend to hybrid settings, conceptualizing hybrid networks as heterogeneous human-AI nodes and links with distinct individual and transactive constraints. Our comparative analysis identifies which network effects are robust across agent types and which require revision, and highlights configurations that were peripheral in single-type traditions, such as human gatekeepers of AI sub-networks, but become structurally central in hybrid teams. Integrating a cognitive systems perspective with network science, we clarify how established exploration-exploitation and efficiency-redundancy trade-offs may operate differently in hybrid teams, and conclude with implications for organizational design, governance, and the responsible development of hybrid intelligence systems.
Bright solitons in hybrid-dispersion photonic crystal microresonators
arXiv:2607.06406v1 Announce Type: new Abstract: Bright dissipative Kerr solitons in optical microresonators provide chip-scale sources of ultrashort pulses and frequency combs. Their properties are defined by the cavity dispersion for which fundamentally conflicting requirements exist: short pulses and broadband spectra require weak dispersion, whereas strong dispersion is associated with predictable dynamics. Here, we resolve this conflict by introducing a localized strong-dispersion section spanning several modes around the pump resonance within an otherwise weakly dispersive system. We implement this hybrid-dispersion scheme in a photonic crystal microresonator and reveal a new soliton attractor of backward-propagating solitons, accessible at low pump power in a thermally stable manner within the blue-detuned regime. The conflicting requirements for broadband spectra and low-noise single-soliton formation are reconciled, even in microwave-repetition-rate resonators, which otherwise are prone to uncontrollable multi-soliton formation. These results highlight the potential to achieve previously incompatible characteristics in nonlinear photonic systems through hybrid-dispersion attractor shaping.
Terahertz-driven four-wave mixing at glass surfaces: Probing vibrational resonances and structural regimes
arXiv:2607.06417v1 Announce Type: new Abstract: Disordered materials such as glasses exhibit complex structural dynamics that are challenging to probe with conventional spectroscopies. We demonstrate that terahertz-driven four-wave mixing (FWM) at glass surfaces provides direct access to low-frequency vibrational modes and structural evolution in amorphous solids. Applied to a compositional series of PbO-silicate glasses (20-54 mol% PbO), this technique resolves distinct contributions from collective Boson-peak excitations and Pb-O / Si-O network stretching modes, and tracks their systematic evolution across structurally distinct compositional regimes. The dominant vibrational frequency blueshifts with PbO content, reflecting the progressive evolution of the Pb$^{2+}$ network role from silicate-modifier to ward network-former. A pronounced enhancement of the FWM signal near 44 mol% PbO coincides with the emergence of medium-range Pb-Pb correlations, while in-plane-to-out-of-plane FWM intensity ratio ($I_{\rm SS}/I_{\rm PS}$) tracks $\chi^{(3)}$ tensor anisotropy tied to Pb$^{2+}$ lone-pair spatial correlations. The non-monotonic peak in both observables at 44 mol% PbO - a composition where NMR finds no change in local Pb-O coordination and Pb-O-Pb free-oxide linkages are negligible - provides direct evidence that a collective lone-pair reorganization occurs in the medium-range structure independently of nearest-neighbor bonding. These results establish terahertz-driven FWM as a bulk-sensitive, near-surface depth-confined ($\sim$50 nm) nonlinear spectroscopy sensitive to vibrational and electronic structural fingerprints inaccessible to linear infrared, Raman, and terahertz time-domain probes.
A robust and versatile parallel FFT-based mechanical solver for general non-periodic and periodic boundary conditions
arXiv:2607.05929v1 Announce Type: new Abstract: General boundary conditions are implemented within a fast Fourier transform framework for linear and non-linear mechanical problems using small or finite transformation formulations. In the context of parallel computing (distributed memory), we present a framework that enables the combination of periodic and non-periodic (Dirichlet or Neumann) boundary conditions. Taking advantage of the link between non-periodic boundary conditions and the symmetries of the relevant components of the fluctuation displacement and stress fields, discrete trigonometric transforms are employed to adapt the classical Moulinec-Suquet fast Fourier transform approach. The present study employs an original displacement-based fixed-point algorithm in combination with a convergence acceleration method in order to solve boundary value problems. Finite difference approaches are used to build the discrete Green operators associated with a pre-conditioner (reference material), whose choice depends on the loading type and the small or finite transformation frameworks. The newly developed double tetrahedron scheme is employed to investigate non-periodic problems. Outcomes are compared to those of the classical hexahedral scheme. The robustness and computational efficiency of the presented parallel solver is demonstrated through numerical experiments of non-trivial loading scenarios (tension, bending, normal-mixed loading, torsion-bending), complex and densely discretized microstructures and diverse behavior laws (elasticity, isotropic plasticity, crystal plasticity), within small and finite transformation frameworks.
Plainbook: Data Science, in Plain Language
arXiv:2607.05717v1 Announce Type: new Abstract: Jupyter Notebooks have become widely adopted in data science, as they allow the sharing of reproducible computational analysis. They are, however, accessible only to people who understand computer code. To reach the broader audience of scientists interested in data analysis and computation, but unfamiliar with code, we introduce Plainbook, notebooks centered on natural language rather than code. Plainbook is based on two principles: promote the natural language descriptions, and verify the values. In plainbook, the natural language descriptions are preserved, rather than the resulting code; the code is generated automatically from the cell descriptions. As natural language is read top to bottom, Plainbook adopts a linear execution semantics, in which cells are guaranteed to be executed in the order in which they appear; there is no "hidden state" or out-of-order execution as in Jupyter. To allow users who may not understand code to verify the correctness of the computation, we have built into Plainbook verification mechanisms centered on values and value inspection. These include mechanisms that focus on individual cells, akin to unit tests, as well as global mechanisms. Both the linear execution semantics, and the verification mechanisms, are underpinned by a snapshot kernel that caches execution states and makes execution and verification efficient.
On the Convergence Analysis of DCA
arXiv:2211.10942v2 Announce Type: replace-cross Abstract: Difference-of-Convex (DC) programming, which seeks to minimize a function expressed as the difference of two convex functions, arises in a wide range of applications in machine learning, signal processing, and operations research. A classical and widely used algorithm for solving DC programs is the Difference-of-Convex Algorithm (DCA). In this paper, we revisit DCA from a distinctly DC-specific perspective. We first separate well-definedness from asymptotic convergence and introduce an additional assumption ensuring the solvability of the DCA subproblems, which clarifies why the choice of DC decomposition matters. We then develop a Lyapunov-descent-regularity framework in which the descent estimate is read directly from the convex subproblems and the regularity estimate is verified from DCA optimality conditions. This yields global convergence of the iterates $\{x^k\}$ for both standard and convex-constrained DC programs under either the classical Lojasiewicz subgradient inequality or the broader Kurdyka-Lojasiewicz (KL) property. We further explain how stronger regularity regimes, such as the Polyak-Lojasiewicz (PL) condition, fit into the same framework and sharpen the resulting convergence rates. Consequently, we obtain finite-time, linear, and sublinear rates for objective values and iterates in a way that cleanly separates well-definedness, DCA-specific structure, and KL/PL regularity, and that is readily transferable to DCA-type variants.
Unicode TAG-Block Concealment of Tool-Metadata Payloads in the Model Context Protocol: An Approval-View Fidelity Gap Across Three Independent Server Implementations
arXiv:2607.05744v1 Announce Type: new Abstract: The Model Context Protocol (MCP) is the dominant way coding agents discover and invoke external tools. A server advertises each tool through a tools/list handshake that returns a name, a natural-language description, and a JSON input schema. The client renders this metadata once, in a one-time approval dialog, and then injects it verbatim into the model's context on every subsequent turn. Nothing in the protocol requires the rendered approval view and the bytes delivered to the model to match. We isolate that gap as a single structural mechanism, concealment encoding, and show with a model-free, protocol-free analysis that Unicode's TAG block (U+E0000 to U+E007F) has no assigned glyph in any mainstream terminal, chat, or IDE renderer, so a payload written in it is absent from what a human reviewer sees while surviving byte-for-byte into the model's tokenizer. We then measure whether this mechanism actually defeats today's client-side defenses, building a proof-of-concept that speaks the real MCP JSON-RPC/stdio protocol against a genuine client and server. Across 5 distinct MCP metadata surfaces we implement 8 concrete techniques with a deterministic, protocol-level harness. All 8/8 techniques deliver an attacker-controlled payload into the model's context, 4/8 evade a representative string-matching sanitizer, and exactly as the mechanism analysis predicts, only the TAG-block encoding (1/8) is invisible in the human approval view while still reaching the model verbatim. MCP forces re-approval for 0/8 techniques even under a time-of-check to time-of-use rug-pull. To test whether these outcomes are a property of the protocol or an artifact of one server codebase, we re-implement the catalogue against 3 independently developed Python MCP server libraries and find total agreement across all 32 cross-library outcome cells. The baseline sanitizer flags 0 of 25 benign descriptions.
Single-photon polarization tomography with an integrated metal-superconductor nanowire array
arXiv:2607.06047v1 Announce Type: new Abstract: Light polarization is a primary degree of freedom for encoding quantum information. The scaling up of photonic quantum networks and computer architecture depends crucially on its precise characterization. This is typically achieved by placing external waveplates, polarizers, moving mounts, and recently metasurfaces, on top of the detectors. All these solutions complicate integration and scaling. Here we break convention with traditional architecture and present a monolithic, self-aligned metal-superconductor nanowire single photon detector (M-SNSPD) possessing intrinsic full polarization selectivity. Gold nanowires, co-fabricated atop NbTiN superconducting nanowires within the same lithographic footprint, act as polarization-selective plasmonic metamaterials inducing resonant absorption in the NbTiN. U-shaped wires provide linear polarization selectivity, while S-shaped meanders distinguish circular polarization, while retaining the high-count rates and low dark count rates of conventional SNSPDs. By arranging them into a four-pixel array we realize simultaneous projection onto four polarizations and demonstrate continuous polarization state tomography with an ensemble average fidelity exceeding 98%. Our approach opens new avenues towards scalable detector arrays with integrated plasmonic functionalities, for single photon polarimetry, imaging and spectroscopy.
Hypothesis-driven Model Expansion under Uncertainty for Open-World Robot Planning
arXiv:2607.06501v1 Announce Type: new Abstract: We consider an open-world planning setting in which service robots must operate in unknown environments with incomplete knowledge of objects and actions. Traditional closed-world approaches with pre-programmed knowledge bases fail when robots encounter unexpected situations and tasks, posing a fundamental challenge for autonomous knowledge expansion in human environments. In this work, we propose an open-world planning framework that enables robots to automatically generate, verify, and update hypotheses about their abstract world models. Our key insight is to explicitly maintain uncertainty-aware knowledge expansion and integrate hypothesis verification into goal-reaching planning. The framework leverages foundation models to generate initial hypotheses over states and transitions, and applies automated planning to produce action sequences that jointly address hypothesis verification and task execution. Through iterative execution and refinement, the robot expands its knowledge by incorporating verification feedback from the foundation models when hypotheses prove incorrect. Extensive experiments in simulated and real-world environments demonstrate that our framework enables autonomous knowledge expansion and effective operation in open-world settings. These results indicate that integrating uncertainty-aware model expansion from robot foundation models with planning advances the practical deployment of household service robots.
High-Accuracy Semi-Analytical Method for Solving the Problem of Electromagnetic Wave Scattering by Arbitrary Ensembles of Parallel Circular Cylinders
arXiv:2607.06517v1 Announce Type: new Abstract: A method is proposed for solving the two-dimensional problem of electromagnetic wave scattering by a cluster of an arbitrary number of parallel, infinitely long, homogeneous, non-overlapping right circular cylinders. The cylinders may have arbitrary radii and complex permittivities, and their axes, while remaining parallel, may occupy arbitrary positions in the transverse plane. The solution is constructed using an analytical expansion of the electromagnetic field in cylindrical harmonics. Multiple scattering is taken into account by Graf's addition theorem, which leads to a system of linear equations for the expansion coefficients. This system is solved numerically with condition number monitoring and, when necessary, extended-precision arithmetic, followed by a multistage verification of convergence. The method provides numerically verified solutions with controlled accuracy over a wide range of parameters, including densely packed subwavelength configurations. As an example, scattering of a normally incident, linearly polarized monochromatic plane wave by a subwavelength cluster of three identical aluminum nanocylinders (nanowires) is studied. The scattering, absorption, and extinction cross sections, as well as the scattering indicatrix, are computed and analyzed. Streamlines of the Poynting vector field are constructed, demonstrating redistribution of the energy flux between the cylinders of the cluster and the formation of localized regions of field enhancement near their surfaces.
Rethinking Indic AI from a Lens of Cultural Heritage Preservation
arXiv:2607.06544v1 Announce Type: new Abstract: As Artificial Intelligence (AI) makes inroads into different parts of the Indian subcontinent, there is significant interest in studying how AI impacts the linguistic and cultural foundations of this civilization. AI is seen as a ''double-edged sword'' where on the one hand, it can enable access and inclusion for a large population, on the other, it can homogenize worldviews and exclude underrepresented languages and worldviews. In this paper, we try to characterize this problem by addressing the extensive characteristic nature of Indian linguistics and the way they closely connect to cultural practices and worldview. We then perform a longitudinal survey of how Natural Language Processing (NLP) techniques have evolved in this space, tracing the historical development of Indic NLP, covering key milestones, methodological shifts, and resource creation efforts. In addition, the paper also examines the structural and sociolinguistic characteristics of Indian languages, such as rich morphology, complex scripts and grammar rules, diglossia, and large dialectal variation, and explains how these create unique challenges for building AI foundation models. We then discuss the growing role of Indic foundation models and analyze how these models address these long-standing resource and representation gaps. Finally, we propose a research direction called 'Culture Sensing', which re-imagines AI based on hermeneutic reasoning. Culture Sensing aims to address open problems such as ensuring equitable performance across low-resource languages and producing outputs that are culturally meaningful. By bringing together past work, current techniques, and emerging trends, this paper outlines research directions that can guide the next phase of Indic NLP and contribute to the development of more robust and inclusive Indic foundation models.
PerCaM-Health: Personalized Dynamic Causal Graphs for Healthcare Reasoning
arXiv:2605.07267v2 Announce Type: replace Abstract: Personalized healthcare decisions require reasoning about how physiological and behavioral variables influence an individual patient over time. Existing temporal causal discovery methods are poorly matched to this setting: cohort-level models provide stable but non-personalized structures, while per-patient discovery is unreliable because individual trajectories are short, noisy, irregular, and non-stationary. This creates a fundamental gap between population-level causal modeling and the patient-specific, time-varying mechanisms needed for intervention reasoning. We introduce PerCaM-Health, a framework for learning personalized dynamic causal graphs from longitudinal health data. The framework learns a knowledge-guided population temporal graph, then conservatively adapts and evolves it using patient-specific temporal evidence and rolling-window updates, producing interpretable and auditable graph sequences. By coupling these graphs with temporal structural equations, the framework enables patient-level counterfactual queries, such as estimating short-horizon outcome changes under hypothetical behavioral interventions. Experiments on a semi-synthetic dynamic health benchmark show that PerCaM-Health improves graph recovery, dynamic edge tracking, and intervention direction accuracy compared to cohort-level, per-patient, and non-personalized temporal baselines. These results demonstrate that jointly modeling personalization and temporal evolution yields more reliable causal structure and intervention reasoning.
Reproducible Validation of Voucher-Based L2 Interoperability: Diagnosing an ERC-4337 Compatibility Issue in an EIL SDK Implementation
arXiv:2607.05914v1 Announce Type: new Abstract: Ethereum Layer-2 (L2) ecosystems improve scalability but also fragment users, liquidity, gas funding, and execution across rollups. Consequently, cross-rollup interoperability is not only a bridging problem but also a wallet, execution, and validation problem. Ethereum Interop Layer (EIL) proposes a voucher-based architecture in which users create voucher requests on an origin chain and redeem XLP-signed vouchers on a destination chain. When reproducing the evaluated SDK version in a controlled local environment, we observed a compatibility issue in the \texttt{UserOperation} path: paymaster-related data can differ after signing, preventing a stable comparison between the user-authorized representation and the representation later inspected by the local validation flow. This paper presents a reproducible two-L2 validation framework and a controlled compatibility mitigation for that issue. We build a deterministic local testbed over Arbitrum- and Optimism-style development chains, deploy the core paymaster and bridge-related components, implement mock bundlers and event-driven XLP providers, and introduce a sanitized paymaster-data handling path together with a compatible multichain account wrapper. Using this framework, we execute the core voucher lifecycle from request creation to destination-chain voucher redemption and asset release. The contribution is an empirical diagnosis of an implementation-level compatibility barrier, a bounded mitigation that restores controlled end-to-end execution, and an inspectable validation artifact for studying voucher-based interoperability. The work does not claim a new interoperability protocol, universal wallet compatibility, or production readiness; it identifies the remaining gaps toward standard-account validation, one-signature multichain authorization, and full dispute-settlement support.
Lingering Authority: Revocable Resource-and-Effect Capabilities for Coding Agents
arXiv:2606.22504v1 Announce Type: cross Abstract: Coding agents often receive broad tool access for an entire task, even when a resource is needed only for one subgoal. We call this gap lingering authority: a temporary resource/effect capability remains exposed after the episode that justified it has closed. PORTICO is a reference monitor for revocable capabilities exposed to the planner. It compiles an explicit task contract into initial capabilities, grant rules, trusted closure predicates, and global deny rules. A request-grant-invoke lifecycle materializes expansions as opaque, epoch-bound handles. Closure removes those handles from the next planner interface and rejects stale replay before side effects. The monitor assumes mediated tools and a sound typed catalog. In controlled coding-agent tasks, PORTICO records no executed contract-forbidden effects in the evaluated runs, while controlled grants recover boundary work blocked by a fixed narrow envelope. A non-revoking comparator receives the same initial envelope and the same grants at the same turns. On the closure slice, both systems match task success, scope compliance, and all pre-closure decisions; PORTICO then rejects 10/10 post-closure reuses, while the comparator permits 10/10. A deterministic stale-write audit records 0/6 versus 6/6 executed forbidden effects. Scripted traces and six live model traces over file writes, git mutation, and network egress show the same split. In a four-episode same-policy diagnostic, broad request exposure preserves zero executed forbidden effects but raises blocked proposals from 67 to 84. Frozen real-repository runs, with commits and traces recorded, exercise the same lifecycle on real project layouts.
Factorizable joint shift revisited
arXiv:2601.15036v4 Announce Type: replace Abstract: Factorizable joint shift (FJS) represents a type of distribution shift (or dataset shift) that comprises both covariate and label shift. Recently, it has been observed that FJS actually arises from consecutive label and covariate (or vice versa) shifts. Research into FJS so far has been confined mostly to the case of categorical labels. We propose a framework for analysing distribution shift in the case of a general label space, thus covering both classification and regression models. Based on the framework, we generalise existing results on FJS to general label spaces and present and analyse a related extension to label distribution estimation of the expectation maximisation (EM) algorithm for class prior probabilities. We also take a fresh look at generalized label shift (GLS) in the case of a general label space.
Shape Over Intensity: Directional Topological Encoding for False Positive Reduction in Intracranial Aneurysm Detection
arXiv:2607.05317v2 Announce Type: replace Abstract: Automated detection of intracranial aneurysms (IAs) from CT angiography (CTA) is severely hindered by high false-positive rates. Convolutional neural networks (CNNs) rely on local pixel intensities, causing systematic confusion between saccular aneurysms and vascular bifurcations - a problem especially acute for small lesions (<3 mm), where detection sensitivity falls below 60%. We propose a plug-and-play, topology-aware false-positive reduction framework evaluating the Smooth Euler Characteristic Transform (SECT) - a directional representation encoding global 3D vascular geometry independently of intensity - against persistence-based summaries (Persistence Images and Landscapes), tested on a stratified subset of the RSNA 2025 dataset. SECT achieves an AUC of 0.943, substantially outperforming direction-agnostic methods (AUC ~0.68), and exhibits a clinical performance inversion: it excels on the sub-3 mm cohort, maintaining 0.943 AUC and 78.5% sensitivity at 95% specificity. The representation is also scanner-agnostic, achieving 0.927 mean AUC under leave-one-scanner-out (LOGO) validation across four manufacturers. By capturing asymmetric geometric invariants rather than intensity profiles, SECT reliably resolves the primary structural confounder in IA detection, positioning it as a robust downstream filter for hybrid deep-learning diagnostic pipelines.
Claimed or Attested? A Commit-Signature Dataset and Identity Trust Tiers across the World of Code
arXiv:2607.06194v1 Announce Type: new Abstract: An author string in a git commit is free text the committer typed, so identity resolution over a global commit corpus rests on a claim that nothing in the commit verifies. A cryptographically signed commit is different: it binds the commit to a key the committer controls, and when that key ties back to a real-world identity the git identity becomes attested rather than merely claimed. We release the first commit-signature axis for the World of Code (WoC), extracted for the V2604 collection. The signature travels in the commit object's gpgsig header and is already carried, unparsed, in the commit-message field of the WoC commit tables, so the axis is a scan over existing tables rather than a re-read of the object database. Over the V2604 corpus of 5,866,595,698 commits, 17.59% carry a signature (PGP dominant at 98.96%, with a growing minority of SSH and X.509/sigstore signatures), or 1,031,721,316 signed commits. We release the per-commit signature map c2sigFull, a key-to-author graph gated so that shared organization and continuous-integration keys are separated from person keys, and A2trust, a per-identity attestation tier (unsigned, signed, real-world-bound, cross-corpus attested) that extends the published A2cls identity-class dataset. The signature axis is a precision anchor, not a coverage layer: signed commits skew toward recent and security-conscious developers, a population that overlaps the scholarly authors a bibliography join targets. We use the person keys to build a cryptographically grounded alias gold that calibrates the heuristic WoC alias map independently of hand-labeled pairs, and to attach an attestation provenance to science-to-software identity links. All artifacts are released as a self-contained, in dependently hosted replication package keyed to the WoC V2604 collection.