arXiv:2607.13491v1 Announce Type: new
Abstract: Looped Transformers scale sequential computation by applying a compact stack of physical blocks for multiple rounds, increasing unrolled depth without increasing stored parameters. This reuse changes the residual-scaling problem: in an untied Transformer, each residual branch receives and applies its own parameter update, whereas in a looped Transformer one shared update aggregates gradients from repeated visits and is read back by those same visits in the next linearized forward pass. We formalize this tied-depth effect through a first-order perturbation bound controlled by a visit-alignment coefficient $\kappa_R$. The bound recovers the DeepNorm exponent when visits decorrelate, but in the conservative aligned regime it requires the exponent to increase from $1/4$ to $1/2$ as loop count grows at fixed physical depth. The resulting method, \textbf{DeepLoop}, keeps the Post-LN DeepNorm architecture and sets $\alpha=(2N)^{1/2}$ and $\beta=(8N)^{-1/2}$ for unrolled depth $N$. On GPT-style looped language models at GPT-2 small and GPT-2 medium scale, DeepLoop is neutral when no physical block is revisited and improves validation loss and downstream accuracy once recurrent depth is activated. These results show that stable recurrent depth requires residual scaling rules that account for parameter visits, not only nominal layer count.
Science Journals
arXiv:2607.13360v1 Announce Type: new
Abstract: A neural network is trained to learn the mass exchanges between vapour, liquid and ice phases in atmospheric convection. The network is trained on convection resolving output from a regional configuration of the LFRic model with a multi-moment microphysics parameterisation (CASIM). The loss function for learning these phase exchanges is formulated under the assumptions of thermal and mechanical equilibrium (same temperature and pressure for all phases), and mechanical dis-equilibrium (different Gibbs free energies for all phases). The network outputs determine the exchanges of vapour, liquid and ice so as to conserve mass, and the resulting change in entropy is determined from the network outputs so as to conserve energy. The neural network is implemented in a thermodynamically consistent manner within a 2D vertical slice discontinuous Galerkin model of a moist, non-hydrostatic atmosphere in order to simulate the formation of three-phase clouds for convection at sub-km resolution. The results are compared to those from a physics based representation of three-phase moist processes at thermodynamic equilibrium.
arXiv:2607.13546v1 Announce Type: new
Abstract: Mobile crowdsensing (MC) recruits mobile users to perform sensing tasks using their smartphones, enabling large-scale applications such as traffic monitoring and environmental sensing. A fundamental challenge is online worker recruitment under uncertainty, where the platform must learn workers' sensing performance while operating with a limited budget. Existing learning-based MC recruitment methods typically assume that each worker's sensing quality is stationary with a fixed mean over time. In practice, however, worker performance often improves with experience and eventually stabilizes, while the incurred sensing cost can be unknown in advance due to time-varying device and context states. In this paper, we study a budget-constrained online recruitment problem in which the platform selects one worker in each round, observes the sensing quality and incurred cost, where the expected sensing quality of each worker increases with experience and eventually converges to a plateau, and repeats until the budget is exhausted. We formulate this problem as a structured bandit model where each worker's expected reward evolves according to an unknown increasing-then-converging function of its participation count, and each worker has an unknown expected cost. We develop a cost-aware online learning framework that jointly learns evolving reward trajectories and heterogeneous costs, detects performance saturation, and allocates the limited budget to maximize long-term sensing utility. We provide theoretical performance guarantees and validate the proposed approach through extensive experiments, demonstrating consistent improvements over baselines that ignore experience-driven dynamics or assume known costs.
arXiv:2607.13520v1 Announce Type: new
Abstract: In this paper, we propose and study $r$-minimal codes with respect to $\mathbf{P}$-support, where $\mathbf{P}=(\Omega,\preccurlyeq_{\mathbf{P}})$ is a poset defined on the coordinate set of the ambient space $\mathbf{H}$. $r$-Minimal $\mathbf{P}$-codes are natural extensions of Hamming metric minimal codes that have been extensively studied in the literature. We characterize $r$-minimal $\mathbf{P}$-codes in terms of the notion so called cutting $r$-blocking maps, which generalizes the well-known equivalence between minimal Hamming metric codes and cutting blocking sets. We also give a necessary and sufficient condition for $r$-minimality in terms of $(\mathbf{P},\omega)$-weight defined on $\mathbf{H}$, where $\omega:\Omega\longrightarrow\mathbb{R}^{+}$ is an arbitrary weight function. This leads to a generalization of the well-known Ashikhmin-Barg criterion for Hamming metric minimal codes. We then prove two existence results for $r$-minimal $\mathbf{P}$-codes, both for general $\mathbf{P}$ and for the special case that $\mathbf{P}$ is a disjoint union of chains. When $\mathbf{P}$ is hierarchical, we characterize $r$-minimal $\mathbf{P}$-codes in terms of $r$-minimal Hamming metric codes. Finally, we characterize cutting $r$-blocking sets induced by hierarchical posets with two levels, which further enables us to answer a question raised in Hyun, Kim, Wu and Yue \cite{28}.
arXiv:2607.13527v1 Announce Type: new
Abstract: Recent video generation models (VGMs) have made substantial progress in visual fidelity, yet their ability to follow long, compositional instructions remains insufficiently evaluated. Existing evaluation protocols often rely on prompts that are short and semantically shallow, with limited atomic constraints and weak spatio-temporal dependencies. They also frequently depend on costly human evaluation or handcrafted vision pipelines, while providing little diagnostic insight into which instruction constraints succeed or fail. To address this gap, we propose VGIF-Score, a highly automated and interpretable framework for evaluating instruction following in video generation. VGIF-Score consists of two complementary components: an objective completion branch that parses prompts into a Spatio-Temporal Directed Acyclic Graph (ST-DAG) and performs dependency-aware QA with short-circuit diagnostics, and a subjective satisfaction branch that uses instruction-conditioned AutoRubric to assess cinematography, visual purity, motion smoothness, and physics adherence. Together, these components produce a unified score that captures both objective completion and perceptual satisfaction. We instantiate this framework on VGIF-Bench, a benchmark of 223 long, structurally entangled prompts paired with approximately 4.3K fine-grained evaluation items. Experiments on 14 proprietary and open-source VGMs across more than 3K generated videos show that VGIF-Score provides reliable, interpretable, and diagnostically useful evaluation of video generation instruction following. The code will be available at https://github.com/PRIS-CV/VGIF-SCORE.
arXiv:2607.13554v1 Announce Type: new
Abstract: As quantum computing technology continues to mature, the US National Institute of Standards and Technology (NIST) has outlined a migration timeline for Post-Quantum Cryptography (PQC), recommending the deprecation of certain elliptic curve cryptography (ECC) by 2030. Furthermore, privacy-sensitive application scenarios, such as vehicular communications, require the use of anonymous certificates. However, existing anonymous certificate schemes are still largely based on ECC. Therefore, this study proposes a multivariate cryptography-based anonymous certificate scheme, aiming to design quantum-safe anonymous certificates suitable for privacy-sensitive application services. The proposed multivariate cryptography-based anonymous certificate scheme is supported by rigorous mathematical proofs and illustrated with computational cases.
arXiv:2607.13446v1 Announce Type: new
Abstract: Combinatorial optimization problems are central to many challenges in logistics, finance, engineering, and the life sciences, yet they remain among the most computationally demanding. Many of these problems can be mapped onto the Ising model, in which binary spins interact through a network of couplings, and solutions correspond to low-energy, ideally ground-state, spin configurations. Photonic Ising machines have the potential to be fast and energy-efficient heuristic solvers of optimization problems by leveraging the low latency, high bandwidth, and inherent parallelism of optics. However, current photonic implementations remain limited in scalability, connectivity, reconfigurability, and time-to-solution, preventing their use in many practical applications. In this perspective, we examine the current landscape of photonic Ising machines, discuss the challenges and limitations of existing platforms, and identify the scientific and technological advances needed to realize large-scale systems. These developments could establish photonic Ising machines as useful hardware platforms for practical optimization.
arXiv:2607.13455v1 Announce Type: new
Abstract: High-quality teleoperation datasets are costly to collect, particularly for hard tasks. We observe that many tasks exhibit directional asymmetry: completing the forward hard task is difficult, whereas reversing it by relaxing or disrupting the environment is comparatively easy. This suggests that reversed easy-task trajectories can serve as a scalable supervision signal for the hard task, reducing the cost of manual demonstration collection. However, reversed data can be noisy, and directly training on it may yield suboptimal policies. To enable largely automated acquisition and effective use of reversed data, we propose a teleoperation-cost effective framework for hard policy learning via temporal reversal of easy tasks, consisting of three key components: a closed-loop data collection pipeline that alternates between hard-task and easy-task policies to autonomously reset the environment and generate diverse trajectories; a hierarchical data refinement pipeline that temporally inverts easy-task rollouts and filters low-quality motion using kinematic priors and a critic-guided advantage filter; and an iterative policy learning method that trains the hard-task policy using both initial reversed easy-task demonstrations and the filtered reversed data in a continuous online learning loop. By combining automated collection, hierarchical refinement, and iterative learning, our method enables scalable, reliable training of complex, high-precision manipulation tasks. Across two simulated benchmarks and real-robot experiments, we demonstrate that our method improves hard-task success rates with higher data efficiency and more stable training compared to reversal-based and reinforcement-learning baselines, without requiring extensive hard-task teleoperation.
arXiv:2607.13458v1 Announce Type: new
Abstract: Scene Text Recognition (STR) remains challenging due to the diversity of text appearances, including curvature, rotation, and perspective distortion. Recent Transformer-based approaches perform well but usually rely on one-dimensional positional encodings that ignore the 2D spatial structure of text images. Axial 2D extensions of Rotary Position Embedding (RoPE) exist for vision Transformers, but they assume roughly square, isotropic image content and apply the rotation only within encoder self-attention. Scene text violates both assumptions: crops are markedly anisotropic, and STR models are encoder-decoder, so the decoder must relate its queries to the encoder's 2D layout through cross-attention. We introduce 2D-RoPE-STR, which adapts axial 2D-RoPE to this setting through (1) an anisotropic row/column dimension allocation matched to the aspect ratio of text, and (2) an extension of the rotary coupling into encoder-decoder cross-attention, letting autoregressive decoding steps attend to encoder tokens by their 2D layout, a setting not addressed by prior encoder-only formulations. Both changes are essentially parameter-free and require no architectural redesign beyond the positional-encoding module. We further introduce a diagnostic protocol (a controlled ablation pair isolating only the positional encoding, an image-level net-win disagreement analysis, and encoder attention visualization) that identifies where and why relative 2D position helps: curved, rotated, and perspective-distorted layouts where reading order departs from a straight horizontal line. On six standard benchmarks (IIIT5K, SVT, ICDAR 2013, ICDAR 2015, CUTE80, SVTP), gains concentrate on exactly these irregular layouts, with ablations isolating each design choice against 1D RoPE and 2D sinusoidal and learnable alternatives.
arXiv:2607.13573v1 Announce Type: new
Abstract: Maneuvering target tracking in three-dimensional space remains a challenging problem due to complex motion dynamics and model mismatch. To address this, this paper proposes a hybrid model/data-driven algorithm named IMMNet, which integrates the interpretable structure of the interacting multiple model (IMM) algorithm with learnable neural components. Unlike end-to-end black-box methods, the proposed IMMNet algorithm not only can preserve the Bayesian inference mechanism that is essential for real-time radar applications, but also can adaptively learn motion patterns and noise characteristics from data. Extensive experiments demonstrate that the proposed IMMNet algorithm consistently outperforms the existing algorithms across various scenarios, validating it as a robust, interpretable, and practical solution for maneuvering target tracking.
arXiv:2607.13492v1 Announce Type: new
Abstract: Neural implicit representations have emerged as a powerful paradigm for 3D reconstruction. However, high-fidelity indoor surface reconstruction remains a significant challenge, primarily due to the pronounced \emph{geometric heterogeneity} of indoor scenes. Large texture-less planar regions typically require stronger regularization to suppress high-frequency artifacts, while thin structures demand sharper, more adaptive representations to mitigate the spectral bias of multi-layer perceptrons (MLPs) and prevent over-smoothing. Existing approaches often rely on spatially indiscriminate prior supervision and a scene-global SDF-to-density transformation, which constrains their ability to balance planar smoothness and detail preservation. In this paper, we propose CASA-SDF (Curriculum-Aware Spatial Adaptation for SDF), a unified framework that addresses this challenge via complementary adaptations of supervision and representation capacity. Specifically, Hybrid Spatially-Adaptive Uncertainty Annealing (SAUA) fuses semantic and photometric uncertainties to construct a pixel-wise curriculum for monocular prior supervision. This strategy maintains regularization in reliable regions while attenuating unreliable supervision early in training to enable data-driven photometric refinement. Meanwhile, Curvature-Aware Locally Adaptive Density Transformation (CALADT) progressively modulates the sharpness of the SDF-to-density mapping via a curvature proxy to enhance the representation of thin structures. Extensive experiments on benchmark indoor datasets demonstrate that CASA-SDF improves surface completeness and detail recovery on high-frequency structures, without compromising the stability of planar surfaces.
arXiv:2607.13584v1 Announce Type: new
Abstract: Spiking Neural Networks (SNNs) trained through unsupervised Spike-Timing-Dependent Plasticity (STDP) have been explored as solutions to visual loop closure problems, driven by the prospect of efficient on-device inference on neuromorphic devices. State-of-the-art STDP-based models deliver high classification accuracy but fail to reach the high Recall at 100% Precision (R@100P) needed for reliable autonomous navigation. We present a discrete, tensor-native implementation of the STDP-based SNN-VPR pipeline using PyTorch with snnTorch and evaluate it on a 100-place Nordland dataset using 15 independently-trained networks. The contribution of three decisions in the implementation is investigated. First, we show how to perform neuron assignment with a closed-form, deterministic tensor pipeline and show that it provides significantly higher R@100P than a standard argmax procedure. However, some of this gain comes from implementation differences compared to prior continuous-time models, which we measure independently. Second, ablation in isolation shows that state reset after each query helps improve R@100P regardless of the way neurons are assigned. Third, velocity-compensated sliding window aggregation over k consecutive frames reaches R@100P = 100.00% at k = 5 for constant-velocity traversal and an additional 0.20 ms latency. Taken together, these findings show the impact of inference stage design decisions in STDP-based SNN-VPR on recall precision, although the separate contribution of each mechanism and implementation differences is only partially disentangled and needs further examination.
arXiv:2607.13574v1 Announce Type: new
Abstract: We develop a convergent scheme to train neural networks involving analytic activation functions based on gradient flows. Convergence properties are guaranteed by Lojasiewicz theory. The main advantage of this approach is its simplicity of implementation. The coefficients of the network are approximated by solving a system of ordinary differential equations. We test the method by constructing residual neural network approximations of solutions of parametric problems. The dependence of the solutions of simple ordinary differential equations on a few parameters is correctly reproduced. The solutions of inverse problems involving wave constraints which depend on a few parameters can be reasonably approximated, even in regions in which the problem is severely ill posed.
arXiv:2607.13682v1 Announce Type: new
Abstract: Radiative Gaussian splatting has made sparse-view CT reconstruction fast, but existing methods output point estimates with no notion of where the reconstruction can be trusted. We exploit a property of transmissive X-ray imaging that RGB splatting cannot claim -- projection and voxelization are strictly linear in the per-Gaussian densities -- to equip radiative Gaussians with a variational density posterior whose predictive variance propagates in closed form, exactly, in a single forward pass, in both volume space ($\sigma^2(x)=\sum_i g_i(x)^2 s_i^2$) and projection space ($\mathrm{Var}[I_p]=\sum_i w_{i,p}^2 s_i^2$). We present the first systematic calibration study for Gaussian-splatting CT (Spearman / AUSE / ECE with temperature scaling), showing that the resulting per-voxel uncertainty ranks true reconstruction error on 14 of 15 scenes of the official benchmark across three view budgets -- 9 of 15 additionally meeting our magnitude-calibration target after a single temperature -- while the perturbation-ensemble heuristic of concurrent work, transplanted to voxel space under the same protocol on our development scenes, does not (rank correlation as low as $-0.08$). We then dissect why uncalibrated acquisition scores can nevertheless select acceptable views, identifying three regimes -- flat (isotropic, balanced), pathological (degenerate coverage), and anisotropic -- and showing, in controlled single-scene testbeds, that principled uncertainty earns a measurable premium only in the last, motivating a coverage-gated, maturity-scheduled acquisition policy; the same calibrated posterior further points toward a dose-adaptive stopping rule, whose experimental validation we leave to future work.
arXiv:2607.13542v1 Announce Type: new
Abstract: Topological structuring of light inevitably leverages on optical coherence to ensure that the imparted spatial phases are preserved, requiring highly coherent sources or coherence engineering embedded in the design. Now we show that thermal light can be spatially engineered to carry optical topologies in the form of Skyrmions. Such topologies are immune to time averaged decoherence, a fact we leverage on in reverse to create metasurface mediated incoherent topologies from a thermal source. The pristine nature of our measured Skyrmions validates the approach, while simulations reveal how coherence management in the metasurface design would further enhance the functionality. Remarkably, the generation stage inherits robustness from the topology, remaining immune to material and fabrication defects. Our work reports the first topologies from purely thermal light, opening a path to exploiting topology in ubiquitous everyday light sources.
arXiv:2607.13606v1 Announce Type: new
Abstract: Dynamic wetting plays a fundamental role in the dynamics of droplets and bubbles at solid surfaces by influencing contact line motion and interfacial evolution. In this work, three representative wetting-controlled benchmarks, namely droplet splashing, bubble coalescence at solid surfaces, and bubble dynamics under shear flow, are investigated using a three-dimensional volume-of-fluid framework coupled with a recently developed dynamic wetting model based on contact line velocity reconstruction method [19]. The model is first validated against experimental observations and literature data for droplet splashing and bubble coalescence. It accurately reproduces the transient contact line evolution, splashing morphology, and coalescence dynamics. In particular, dynamic wetting suppresses the premature bubble detachment predicted by static wetting models and yields substantially improved agreement with experimental observations. In shear flow, contact angle hysteresis and contact line dissipation strongly influence bubble deformation, sliding, and detachment. These results demonstrate that accurate treatment of dynamic wetting is essential for predicting wetting-controlled droplets and bubbles involving rapid contact line motion, strong interfacial deformation, and topology changes.
arXiv:2607.13613v1 Announce Type: new
Abstract: Centroid neural network (CentNN) is an unsupervised competitive learning algorithm in which centroid splitting is triggered only after strict local stabilization, often leading to prolonged low-movement training phases before model expansion. This report proposes FastCentNN, an accelerated variant that addresses this inefficiency by introducing an early splitting strategy based on the total centroid movement per epoch, which serves as a training entropy proxy. As a result, FastCentNN reduces unnecessary reassignment epochs while preserving the original winner-loser learning dynamics. FastCentNN supports both absolute and stage-relative movement thresholds, allowing the splitting criterion to remain either fixed or adaptive throughout training. Experiments on some benchmark datasets show that FastCentNN consistently achieves clustering quality comparable to CentNN while reducing runtime by up to 16% on synthetic 2D datasets and about 5% on high-dimensional datasets. FastCentNN therefore provides a practical and efficient drop-in replacement for CentNN, retaining its online adaptive learning behavior while offering a simple and interpretable speed-stability trade-off through configurable splitting thresholds.
arXiv:2607.13589v1 Announce Type: new
Abstract: Boron neutron capture therapy (BNCT) requires knowledge of patient-specific $^{10}$B concentration for accurate dose estimation, yet no established method provides real-time boron-sensitive information during irradiation. Backscattered thermal neutrons carry a $^{10}$B-dependent intensity modulation through the $^{10}$B(n,$\alpha$)$^{7}$Li reaction, documented in BNCT treatment rooms for three decades but not yet developed as a measurement signal. This paper uses Monte Carlo simulation to assess the feasibility of backscattered thermal neutrons as a measurement channel for $^{10}$B concentration. A thin $^{nat}$LiF-converter detector placed at the beam exit captures the composite forward-plus-backscatter field; differential imaging against a $^{10}$B-free baseline isolates the $^{10}$B-dependent component, quantified by the fractional reduction in the $^{6}$Li capture rate, termed Relative Detector Signal Reduction (RDSR). In homogeneous phantoms, RDSR shows linear concentration dependence ($R^2 = 0.997$) with a practical depth limit of approximately 6 cm. Edge-response analysis yields a diffusion-limited FWHM of 32-176 mm over 1-5 cm depth, with weak concentration dependence. In a voxelized patient phantom across 12 boron configurations, the $^{6}$Li capture cross-section provides intrinsic thermal neutron energy selectivity that preferentially weights the band where $^{10}$B absorption is concentrated. Region-of-interest integration achieves counting-statistics sensitivity below 10 ppm; the systematic detection floor (~22-28 ppm at $\pm$1% baseline uncertainty) identifies baseline-reference precision as the dominant constraint. The modeled detector produces limited dose perturbation (+11.6% treatment-time increase). These results establish the physical basis for a boron-sensitive backscattered neutron measurement concept in BNCT.
arXiv:2607.13306v1 Announce Type: new
Abstract: Standard Maxwellian plasmas exhibit a mathematical \textit{rigidity}, possessing insufficient degrees of freedom to support electrostatic double layers (DLs) and yielding only soliton solutions. This study investigates the hypothesis that the formation of DLs is a generic consequence of breaking this structural rigidity through parametric perturbation. By introducing two independent continuous control parameters, $\delta_1$ and $\delta_2$, into the electron distribution, we demonstrate that DLs are a structural property of any plasma model that relaxes the strict Maxwellian constraint. Through a Gardner small-amplitude expansion, we analytically prove that a perturbation must modify both the quadratic and cubic density coefficients to decouple the nonlinear structure and generate physical, supersonic double layers, deriving small-amplitude acoustic-limit threshold conditions of $\delta_1 > 1$ and $\delta_2 > 7/3$. We show that these theoretical boundaries broaden for large-amplitude, nonlinear structures. By mapping the exact existence regions of DLs in phase space, we demonstrate how higher-order terms relax the weak-amplitude limits, confirming that the Maxwellian state represents a singular point where the DL solution collapses.
arXiv:2607.13607v1 Announce Type: new
Abstract: Algorithmic collusion among pricing algorithms has raised concerns about sustained supra-competitive prices and their implications for social welfare. Existing work has largely focused on the probability that reinforcement-learning algorithms converge to cooperative strategies, typically under the assumption that exploration vanishes over time. Motivated by the observation that algorithms deployed in practice are likely to continue exploring in order to remain adaptive to changing environments, we study learning dynamics under constant exploration. In this setting, the relevant question is no longer whether an algorithm converges to a particular strategy profile, but rather what fraction of time the algorithms spend playing cooperative strategies. Even in the benchmark case of the repeated Prisoner's Dilemma with one-period memory, this yields high-dimensional stochastic learning dynamics, for which a complete analytic treatment is intractable. We show that cooperative strategies can be dominant in this time-averaged sense and derive a boundary predicting when such dominance arises, based on the expected dynamics of the Q-learning process. Extensive simulations show that this boundary is a strong predictor for non-defection-dominated behaviour under epsilon-greedy Q-learning.
arXiv:2607.13326v1 Announce Type: new
Abstract: Public communication of science has become a central component in the relationship between science and society. However, the media exposure of research staff has given rise to new forms of hostility, especially in digital environments. The study Experiences of researchers who interact with the media and social networks in Spain [Science Media Centre Spain, 2024] investigates - via a survey (N=237) - the incidence and typology of these attacks in the Spanish context. More than half of the research staff (51.05%) reported experiencing negative incidents, with a higher prevalence among women (56.9%) than men (46.2%). The attacks differ by gender: women face more challenges regarding their scientific capacity and sexist remarks, whereas men are more frequently targeted over their professional integrity. These dynamics reveal structural gender biases that affect the wellbeing and legitimacy of female scientists, emphasising the need for institutional policies with a gender perspective.
arXiv:2607.13369v1 Announce Type: new
Abstract: We present xChk, a reference identity provider for Bring Your Own Identity (BYOI): users enroll via heterogeneous proofs (government KYC, corporate SSO, WebAuthn/FIDO2, professional networks, live verification, longitudinal activity, behavioral signals) and disclose them as portfolio claims in standard OAuth 2.0 / OpenID Connect (OIDC) tokens, while each relying party applies its own sufficiency policy - the IdP transports claims and may evaluate an RP-supplied evidence policy for consent, but does not adjudicate access. Enrollment depth varies by modality (some paths are user-initiated; org KYB and officer binding are operator-assisted).
xChk also supports human-in-the-loop attestation for high-risk actions: humans can initiate attestations directly (browser UI / POST /api/attestations), and AI agents acting under those principals can trigger the same gateway via scope-gated authorize/attest - hash-chained human approvals on a shared verification graph (humans via OIDC; agents via API keys). A production deployment at https://in.xchk.io ships both initiation paths with bilateral RP evaluation at consent; one documented relying party (https://crabbyed.com, Appendix B) exercises Login with xChk.
arXiv:2607.13372v1 Announce Type: new
Abstract: Discourse data are the primary empirical basis of grammar writing in field linguistics, but producing interlinearized text is notoriously expensive - on the order of one hour of work per minute of recording. For endangered languages, where the time remaining to verify analyses with native speakers is itself limited, automating parts of the interlinearization workflow has direct documentary value. We implement a full neural annotation pipeline (morpheme segmentation, POS tagging, glossing) for Irabu Ryukyuan using deliberately small, transparent BiLSTM-CRF models, and evaluate it under a realistic hard constraint: approximately one hour of fully annotated discourse as the entire supervised resource. Two factors of the annotation itself are manipulated: its richness (with or without a POS tier) and its quantity (training budgets from 6 to 47 minutes). Gold POS improves grammatical glossing by +4.4 (SD 0.7) points (significant in all 5 seeds), and the gain grows as data shrink (+11.6 points at a quarter of the data); a POS tier more than halves the amount of glossed data needed to reach a given accuracy. In a fully automatic pipeline this gain is not yet realized: the tagger still errs on 12% of morphemes, and an incorrect POS misleads the glossing model more than no POS at all. The value is latent rather than lost: degrading gold POS with controlled noise shows the gain returning as tagger accuracy rises, with break-even near our tagger's current 88% and +1.6 to +3.2 points recovered at 92-96%. We conclude with a concrete recommendation for documentation practice: annotate quadrilinearly - text, POS, gloss, translation.
The impact of objective interactions on the performance of massive objective optimization algorithms
arXiv:2607.13377v1 Announce Type: new
Abstract: Many-objective optimization has been a field of interest over the past two decades and several evolutionary optimization algorithms have been introduced to tackle these problems; yet two fundamental questions remain underexplored: (i) What happens when the number of objectives grows beyond the typical many-objective regime of about fifteen and becomes massive? (ii) How do problem characteristics, such as the nature of interactions between objectives, influence algorithmic performance? To answer these questions we employ a diagnostic benchmark suite that allows control over problem characteristics and can be scaled to extremely high objective counts. Using this framework we evaluate several state-of-the-art evolutionary algorithms including NSGA-II, NSGA-III, MOEA/D and lexicase selection across a range of dimensionalities and diagnostic problem landscapes. Our experiments reveal that problem characteristics significantly affect algorithm performance. In particular, the nature of interactions between objectives appears important. These results highlight the importance of understanding these properties before selecting an algorithm for a specific problem. We also show that lexicase selection, an algorithm originally designed for genetic programming, compares favorably with state-of-the-art many-objective optimization algorithms while avoiding the dependence on predefined reference directions.
arXiv:2607.13420v1 Announce Type: new
Abstract: Clicks on homepage marketing blocks are driven by a dual-mechanism of content interest and access habits. However, habitual clicks often create Pseudo-Positives in marketing slots, where position advantage masks mediocre content quality, leading to biased recommendation ecosystems. We propose a framework called Orthogonal Disentanglement of Access habits (OrDA) to purify interest signals. OrDA utilizes a dual-tower structure with a gated allocation layer to adaptively route features and minimize interference. To ensure rigorous separation, we employ orthogonal regularization to constrain the latent interest and habit manifolds to be geometrically perpendicular. OrDA performs causal intervention (do-calculus) during inference to rank items solely by purified interest scores. Empirical online evaluations on large-scale datasets demonstrate that OrDA effectively eliminates access-habit bias, outperforming state-of-the-art methods in predictive accuracy. Online AB test 5.64% shows user click-through rates (UCTR) improvement on the Zhima homepage marketing block, Zhima rent-floor recommendation.