Forskningsradar

Science Journals

Peer-reviewade publikationer — 54780 artiklar

Induced-Minor-Closed Classes have Linear, Square-Root, or Sub-Polynomial Tree-Independence
arXiv:2607.12090v1 Announce Type: cross Abstract: An independent set in a graph $G$ is a set of pairwise non-adjacent vertices. A tree decomposition of $G$ is a pair $(T, \chi)$ where $T$ is a tree and $\chi : V(T) \rightarrow 2^{V(G)}$ is a function satisfying two axioms: for every edge $uv \in E(G)$ there is an $x \in V(T)$ such that $\{u,v\} \subseteq \chi(x)$, and for every vertex $u \in V(G)$ the set $\{x \in V(T) | u \in \chi(x)\}$ induces a non-empty and connected subtree of $T$. The sets $\chi(x)$ for $x \in V(T)$ are called the bags of the tree decomposition. The tree-independence number of $G$ is the minimum taken over all tree decompositions of $G$ of the maximum size of an independent set of the graph induced by a bag of the decomposition. A graph $H$ is an induced minor of a graph $G$ if a graph isomorphic to $H$ can be obtained from $G$ by vertex deletions and edge contractions. We prove that for every $t\in\mathbb{N}$ there exists an $\epsilon > 0$ such that every graph $G$ either contains the complete bipartite graph $K_{t,t}$ or the wall $W_{t\times t}$ as an induced minor, or has tree-independence at most $O(2^{O((\log n)^{1-\epsilon})})$. This leads to algorithms with running time $2^{n^{o(1)}}$, for a wide range of problems on $\{K_{t,t}, W_{t\times t}\}$-induced minor free graphs. Our result is a substantial generalization of existing bounds for the tree-independence and tree-width on various graph classes, and a partial resolution of the conjecture of Chudnovsky, E S, and Lokshtanov [Arxiv, 2025] that $\{K_{t,t}, W_{t\times t}\}$-induced minor free graphs have poly-logarithmic tree independence number. The generality comes at the cost of a sub-polynomial, rather than poly-logarithmic upper bound. Our result leads to a complete classification of induced-minor closed classes into ones that have sub-polynomial tree-independence, tree-independence equal to $\tilde{O}(\sqrt{n})$, and linear tree-independence.
Antiproof: Synthesizing Vulnerability Detectors and Proofs of Exploitability
arXiv:2607.12316v1 Announce Type: new Abstract: Discovering vulnerabilities before attackers exploit them requires high recall and reliable automatic validation, but existing approaches struggle to achieve both without prohibitive cost. We present Antiproof, an end-to-end vulnerability discovery system that combines neuro-symbolic detector synthesis for high-recall discovery with proof-of-exploitability oracles for automatic validation. Antiproof learns and iteratively refines static detectors from vulnerability datasets, then validates candidates by verifying whether executable proofs demonstrate concrete attacker capabilities. Evaluated on BountyBench and our curated KEVBench dataset, Antiproof detects 64 of 66 vulnerabilities, improving recall by more than 60 percentage points over static-analysis and neuro-symbolic baselines. In a scan of 50 widely deployed systems, Antiproof uncovered several hundred previously unknown vulnerabilities. We are responsibly disclosing all confirmed zero-days and have received 12 CVE assignments to date, including remote code execution vulnerabilities in Ray, SGLang, vLLM, and LiteLLM that could allow attackers to take over LLM training and inference systems.
The Sound of Absence: Audio-Language Embedding Models Struggle with Negation
arXiv:2607.12290v1 Announce Type: cross Abstract: Audio-language embedding models such as CLAP are widely evaluated on matching present sound events, but rarely on negation. We show this affirmation-only evaluation hides a key limitation: these models fail to encode negated sound concepts, mapping affirmative and negated captions to nearly identical representations. To expose this blind spot, we introduce NegEval-Audio, a framework that converts existing datasets into two negation-aware tasks, Retrieval-Neg and Multiple-Choice Negation (MCQ-Neg), to probe whether models distinguish present from absent events. On AudioCaps and Clotho, performance degrades sharply under negation, with negation-type MCQ accuracy falling far below chance, and the failure persists even for a recent multimodal LLM-based embedding model. While a training-free steering method improves MCQ-Neg, it yields marginal gains for Retrieval-Neg. This indicates that affirmation bias is a fundamental flaw in the representation geometry, necessitating explicit negation-aware training objectives.
Discrimination of Tectonic, Swarm and Impact-related Marsquakes using Spectral Characteristics
arXiv:2607.12493v1 Announce Type: cross Abstract: The NASA InSight mission observed over 2000 marsquakes in the course of its three year mission. These quakes varied in magnitude between 1.5 and 4.5, as well as in spectral content. We present a simple framework to describe the spectral characteristics of all observed marsquakes, based on source process; propagation through the mantle or crust; and local, receiver-side amplification. We assign to each quake an objective measure of its amplitude, as well as the spectral decay created by the duration of the rupture and the dampening of high frequencies due to visco-elastic attenuation. Together, this allows us to obtain characteristic patterns of the whole marsquake dataset, e.g. in terms of event magnitudes, source size, and - for quakes caused by meteoritic impacts - crater size. We show that a significant fraction of all marsquakes - the high-frequency quakes - form a swarm that is likely not caused by tectonic processes in rocks. Our analysis allows separation of the whole marsquake catalogue into three event classes, of tectonic quakes, meteoritic impacts, and swarm events. We finally conclude that the largest marsquake, S1222a, most likely belongs to the group of meteoritic impacts.
Constructing mode-resolved quantum optical models for emitters in photonic crystals
arXiv:2607.12603v1 Announce Type: cross Abstract: Recent advances are enabling quantum emitters to interact with photonic crystals, whose electromagnetic modes exhibit complex dispersion relations, spatial mode structure, and polarization textures. However, modeling light-matter behavior in these systems faces a persistent trade-off: electromagnetic approaches based on Maxwell-equation solvers provide realistic vectorial descriptions but are difficult to integrate with quantum many-body and non-perturbative methods, whereas simplified quantum-optical lattice models are tractable but typically rely on scalar and spatially independent light-matter couplings that miss essential features of these structured photonic environments. Here, we introduce a constructive framework to derive quantum-optical lattice descriptions that overcome this trade-off. Combining symmetry-constrained tight-binding constructions with numerically computed photonic band structures and field profiles, our method yields minimal, symmetry-enforced lattice Hamiltonians that reproduce the target photonic dispersion while retaining the mode-resolved (position- and polarization-dependent) structure of the light-matter coupling. We show that these models recover Green's-function-based emitter dynamics in the perturbative regime, while providing access to non-perturbative quantum dynamical simulations beyond emitter-only descriptions. As a proof of principle, we apply the framework to a two-dimensional photonic crystal and show that it captures polarization-dependent directional emission inaccessible to scalar models, while enabling the analysis of non-Markovian light-matter dynamics and entanglement. Our results provide a practical bridge between classical electromagnetic simulation tools and quantum-optical many-body and non-Markovian modeling in photonic crystal settings.
Strong order one-half convergence of a coupled tamed Euler--Peano scheme for reflected stochastic differential equations with super-linearly growing coefficients
arXiv:2607.12415v1 Announce Type: new Abstract: We study strong numerical approximations for reflected stochastic differential equations in possibly unbounded convex domains with super-linearly growing drift and diffusion coefficients. Under a coupled monotonicity condition and polynomial local Lipschitz assumptions, we first establish the well-posedness of the reflected SDE and derive uniform moment bounds for its solution. We then introduce a coupled tamed Euler--Peano scheme, in which the drift and the squared diffusion coefficient are tamed by a common factor and the resulting Euler--Peano path is corrected through the Skorokhod problem. This common taming factor preserves the drift--diffusion coercivity structure and yields uniform moment estimates for the numerical solution. We prove strong convergence of order $1/2$ for both the constrained state process and the boundary regulator, thereby recovering the standard Euler-type strong order in this reflected setting. Numerical experiments for a reflected stochastic Ginzburg--Landau type system illustrate the constraint preservation of the scheme and support the theoretical convergence rate.
Contrasting statistical patterns in melodic and molecular evolution reveal distinctive constraints in a culturally evolving system
arXiv:2607.12673v1 Announce Type: cross Abstract: Evolved sequences can be used to infer the rules of evolution. Orally transmitted folk melodies are evolved sequences whose similarity to protein sequences (one-dimensional, drawn from a limited alphabet) invites application of bioinformatics methods to study cultural evolution. A major obstacle is that melodies encode rhythm, which breaks some assumptions of standard sequence-alignment algorithms. We develop a rhythm-aware alignment method and apply it to \num{40000} Irish dance tune variants, enabling the first large-scale automated melodic alignment. Four canonical bioinformatics analyses -- mutability, substitution matrices, positional conservation, and covariance -- reveal patterns distinct from those of molecular evolution, revealing the forces that shape each domain: biochemical and biophysical constraints for proteins; memory, motor, and social biases for melodies. Together the results show that bioinformatics provides a powerful framework -- conceptual as much as algorithmic -- for studying cultural evolution. Although the cultural transmission of music has been discussed for centuries, here we show how to analyze it at large scale.
When Close Enough Is Not Enough: Autoregressive Drift in Quantum Circuit Synthesis
arXiv:2607.12780v1 Announce Type: cross Abstract: Quantum circuit optimization for fault-tolerant computing requires exact functional equivalence while minimizing expensive non-Clifford resources such as T gates. We study this problem using a compact 44.8M-parameter encoder-decoder transformer with structured circuit tokenization, evaluating on parameterized circuits (2-6 qubits) and Clifford+T circuits (3-6 qubits). On parameterized circuits, a hybrid approach -- structure from the transformer, angles from classical optimization -- achieves median fidelity 1.000 on 3-6 qubit circuits. On Clifford+T circuits, where all gates are discrete and no post-processing is possible, the model learns valid syntax and accurate T-Count statistics, yet exact equivalence degrades sharply with target length -- from 88% on circuits with <=9 gates to near zero beyond 26 gates. We trace this failure to autoregressive drift: early-token divergence cascading irrecoverably through left-to-right decoding. Two levers partially mitigate the drift: inference-time strategies that generate multiple candidates and select via equivalence verification raise exact-match rates from 7% to 22.5%, while scaling training data by 2.5x pushes them to 39.5%. Yet the degradation with target length persists -- even with more data, exact equivalence drops from 94% on short circuits to under 4% beyond 26 gates. The contrast between settings is our central finding: when approximate outputs can be rescued by post-processing, the transformer succeeds; when exact discrete correctness is required, autoregressive drift limits reliability, with both inference-time search and data scaling as effective levers while training-side fine-tuning and model-level diversification are not.
Supernova Neutrinos and the Origin of Biomolecular Homochirality
arXiv:2607.12813v1 Announce Type: cross Abstract: We investigate the role of parity-violating interactions between supernova neutrinos and chiral molecules in nearby interstellar molecular clouds as a potential source of biomolecular homochirality. We introduce neutrino interactions into the autocatalytic chemical reactions in a far-from-equilibrium noise-induced system. These interactions create a directional bias between L and D enantiomers in the racemization reactions, which is amplified by autocatalysis and stochastic fluctuations. We solve the stochastic equations within the Ito sense to obtain the dynamics of the probability distribution of the enantiomeric excesses, offering an astrophysical scenario for the delivery of homochirality seeds to Earth by meteorites. In spite of the weak interactions of supernova neutrinos, our framework introduces an amplification mechanism to yield a considerable enantiomeric excess of more than $10\%$, in agreement with the latest chemical analysis of the meteorites. Moreover, we scan over the parameter space of the model, inferring from the observational values in order to explore the window of opportunity to generate the initial seeds of homochiral states in interstellar molecular clouds.
ARDepth: Auto-regressive Monocular Depth Estimation with Progressive Visual Conditioning
arXiv:2607.12433v1 Announce Type: new Abstract: Diffusion models have recently become the dominant paradigm for monocular depth estimation (MDE). However, they implicitly assume that depth can be recovered as a globally smooth field through iterative denoising, which does not explicitly reflect the piecewise and scale-dependent organization of scene geometry. In practice, geometric structure emerges progressively across spatial scales, where coarse layout, surfaces, and boundaries are constructed in a hierarchical manner. Motivated by this observation, we introduce ARDepth, which formulates depth estimation as structured auto-regressive generation. Instead of recovering depth through global refinement, ARDepth progressively constructs depth representations as spatial resolution increases. To support this generative process, we introduce Scale-Progressive Conditioning (SPC) to inject multi-scale visual features at each generation stage, and Semantic-Aware Guidance (SAG) to provide scene-level semantic priors that enhance global structural consistency. Together, these designs enable the model to capture fine-grained local details while maintaining coherent global geometry. Empirical results demonstrate that our approach achieves strong performance and produces structurally consistent depth predictions across scales, validating auto-regressive generation as a promising alternative paradigm for geometric modeling.
A Quantum Computing Approach to Track Reconstruction in Strip-Type Detectors
arXiv:2607.12821v1 Announce Type: cross Abstract: This study investigates the use of quantum annealing for particle track reconstruction in strip-type gaseous detectors. In such detectors, ghost hits and multiple hit combinations can turn pattern recognition into a combinatorial optimization problem. We formulate two reconstruction subproblems as quadratic unconstrained binary optimization problems. The first subproblem selects detector hits associated with a single photon track inside a localized candidate region. The second subproblem selects cluster triplets from different detector layers so that multiple track candidates can be handled within a single quantum processing unit(QPU) submission. The proposed formulations are tested using simulated DAMSA detector events. For the single track hit selection task, the QPU based reconstruction gives position and angular resolutions close to those obtained with a Kalman based reconstruction. In the simultaneous association task, valid cluster triplets are first extracted from the QPU samples and then connected using an association rule based on graph connectivity to construct track candidates. The DAMSA event topology studied here has low pileup and is dominated by the two photon signal from axion-like particle(ALP) decay. In this setting, the results show that the QUBO formulations can reproduce local reconstruction decisions. This provides a practical basis for further studies of reconstruction methods that combine quantum and classical computing in more complex tracking environments.
MambaPSA: A Mamba-based Replacement for C2PSA in YOLO26
arXiv:2607.12681v1 Announce Type: new Abstract: State space models (SSMs), notably Mamba, have recently emerged as efficient alternatives to self-attention with linear computational complexity. We investigate the integration of Mamba into YOLO26, the latest non-maximum suppression (NMS)-free object detection framework, by proposing MambaPSA, a lightweight Mamba-based replacement for the C2PSA block at the end of the backbone. To complement this study, we additionally insert a bidirectional Vision Mamba (BiViM) module at the P3, P4, and P5 levels of the neck. Experiments on PASCAL VOC 2007+2012 show that MambaPSA reduces parameters by 2.9%, FLOPs by 12.1%, and improves CPU inference throughput by 17.6% (from 17 to 20 FPS) with negligible accuracy change (-0.1 mAP50:95), while the P4 BiViM placement yields the best accuracy gain (+0.9 mAP50:95). These results suggest that SSMs offer a favorable efficiency-accuracy trade-off when replacing attention-based blocks in NMS-free lightweight detectors.
Stability Analysis of Grid-Following and Grid-Forming Converters Connected to Generators
arXiv:2607.12697v1 Announce Type: new Abstract: This work presents an examination of the main interactions between grid-following (GFL) and grid-forming (GFM) voltage source converters (VSCs) and synchronous generators (SGs), capturing the dynamics of a real power grid and pointing out the limitations of considering an ideal one for stability studies. Eigenvalue trajectories and participation factors are studied to perform in-depth small-signal analyses. Specifically, the GFL and GFM converters are compared in different grid strength scenarios by varying their rating powers and the grid short circuit ratio. Then, time-domain simulations of the non-linear and the developed linear systems are run to validate the mathematical findings from the stability analysis. The results reveal that the stability of VSCs-dominated grids, either in GFL or GFM mode, is strongly affected by both the grid strength and the VSC power, due to the coupling between the VSC control and the SGs.
The log log jam in Gaussian state tomography
arXiv:2607.12983v1 Announce Type: cross Abstract: Unlike in finite dimensions, quantum information in continuous-variable systems has the peculiar feature that without imposing physical constraints, the sample complexity of state tomography can be unbounded. Remarkably, this is even the case for state-of-the-art protocols for learning Gaussian states, which have finite-dimensional descriptions: the best known rates scale with $\log \log E$, where $E$ is the energy of the system. We prove this is not an artifact of existing analyses, but a fundamental limitation of the measurements used. We show: (1) Any protocol that uses Gaussian measurements, even entangled or adaptively chosen ones, must incur a $\log \log E$ dependence. This answers an open question posed by a number of previous works. (2) There is a smooth tradeoff between the number of rounds of adaptivity and the energy dependence, and we give a matching protocol achieving this interpolated rate. (3) With highly entangled, non-Gaussian measurements, one can learn $n$-mode pure Gaussian states with $O(n^2 / \epsilon^2)$ samples, independent of $E$. This answers an open question posed by Chen et al. (4) A simple protocol based on the single-copy canonical phase POVM of Holevo and Helstrom learns single-mode pure Gaussian states with $O(1/\epsilon^2)$ samples, again independent of $E$. Our results clarify the role of energy in bosonic state tomography and shed new light on the intriguing interplay between adaptivity, entanglement, and magic in quantum learning.
Countability versus Computability
arXiv:2406.08493v3 Announce Type: replace Abstract: The concept of {\em countable sets} is attributed to Georg Cantor, who established the distinction between countable and uncountable sets in 1874. The concept of {\em computable sets} emerged in the 1930s through the foundational work on computing models by \Godel, Church, and Turing. In this paper, we investigate the connection between countability and computability. A {\em counting bijection} of a set $S$ is a bijection from the set of natural numbers to $S$. We say $S$ is {\em enumerable} if it is either finite or admitting a computable counting bijection. Our initial investigation shows that a set $S$ is enumerable if and only if it is computable. This equivalence offers new insights into set theory and computability theory. We further show that a set is countable if and only if it admits a {\em counting order}, which is a well order satisfying the {\em proximal} property. Based on this concept, we provide a procedure whose existence gives a necessary and sufficient condition for a set to be countable. This procedure is an algorithm if and only if the set is computable. A counting bijection $f$ is {\em increasing} if $f(x)>f(y)$ whenever $x>y$. We prove that an infinite set $S$ of natural numbers is definable in first-order arithmetic if and only if $S$ has an increasing counting bijection. This result has a significant implication: the standard proof that every set $S$ of natural numbers is countable is invalid. This is because the existing proof establishes that $S$ has an increasing counting bijection, which (by our result) would imply that $S$ is definable in first-order arithmetic. This leads to a contradiction with Tarski's undefinability theorem when $S$ is the set of \Godel\ numbers of the true arithmetic sentences.
Beyond Perceptual Distance: Discrepancy Assessment on Deep Representation for Out-of-Distribution Detection with Diffusion Model
arXiv:2409.10094v3 Announce Type: replace Abstract: Out-of-Distribution (OoD) detection aims to justify whether a given sample is from the training distribution of the classifier-under-protection, i.e., In-Distribution (InD), or from an unknown out distribution. Recent researches have leveraged Diffusion Models (DMs) for OoD detection due to their powerful distribution modeling capability. Given an input image, an InD-pretrained DM produces a corresponding InD-aligned counterpart, which serves as a generative reference for comparison. However, existing DM-based methods typically assess this underlying discrepancy through visual-level distances in the raw image space, which may be misaligned with the distributional discrepancy relevant to OoD detection. In this work, we investigate the fundamentals of discrepancy assessment in DM-based OoD detection, asking how the discrepancy between an input and its DM-generated counterpart should be formulated, and in which representation spaces and with which metrics it should be measured. To this end, we propose to assess the discrepancy in a classifier-relative manner by exploiting the representation spaces of the classifier-under-protection, whose training on InD data encodes rich task-relevant InD knowledge. In particular, we quantify two types of discrepancy: feature-level covariate discrepancy in deep feature representations and logit-level concept discrepancy in output logits, enabling effective differentiation between InD and OoD samples. Moreover, a subspace-based strategy is devised to refine representations of the DM generation to promote discrepancy assessment. Together, these designs form our novel detection framework, namely DDR. Extensive experiments on the challenging large-scale ImageNet-1K dataset demonstrate the superior detection performance of DDR over both DM-based and non-DM-based methods.
Cross-Core Inference Offload as an Operating-System Service on Dual-Core Microcontrollers
arXiv:2607.12620v1 Announce Type: new Abstract: Dual-core MCUs are asymmetric: on NXP's MCXN947, the second Cortex-M33 has no FPU, DSP extension, TrustZone, or MPU. We treat the asymmetry as a design input in the Phase 3 dual-core architecture of SynapticOS, an open-source Zephyr-based runtime: the AI runtime (models, NPU/DSP, scheduler) lives on the capable core, and the application core reaches inference only via a message-based OS service -- a remote system call. The transport is a pair of lock-free single-producer/single-consumer rings in shared SRAM: one writer per index, free-running 32-bit counters, ordering by data-memory barriers alone (the platform has no cross-core atomics). Because ring state is shared, a rebooting application core rejoins unaided. Requests carry priority classes, errors and timeouts propagate to the caller, and tensors stage zero-copy in a shared slot -- a 27 KB frame cannot exist twice in 64 KB of RAM. Measured on the FRDM-MCXN947 (both cores 150 MHz): the application core boots in 1,514 us and completes the handshake in 2,554 us, bit-identical over 11 boots; round trips are 15 us typical / 81 us worst-case (50 us budget); pushes cost 25 cycles; a 1,913-serve two-model soak had zero errors (stub-NPU latencies bracket transport, not silicon). An MPU region on the runtime core guards the application core's RAM (fault-injection verified); protection is one-directional -- the application core has no MPU, and ARMv8-M cannot block privileged reads. Two hardware-revealed defects are reported: releasing the second core into erased flash wedges the whole chip and its debug port (now prevented by a ROM-API blank check), and a Zephyr flash-driver Kconfig silently disarmed the devicetree MPU guard (now programmed at runtime). Firmware is 98.9 KB flash (runtime core) and 32.4 KB (application core, 42.6 of 64 KB RAM); 108 tests in 13 suites pass 100%. Apache 2.0: https://github.com/Dimitrios-Kafetzis/SynapticOS
Quantum Weakest Preconditions Revisited: Pre-expectations for Expected Runtime Analysis
arXiv:2607.12532v1 Announce Type: new Abstract: Quantum weakest preconditions are a fundamental tool for program verification of quantum programs. Many variations have been reported in the literature. We revisit quantum weakest preconditions from the perspective of expected runtime analysis of quantum programs and introduce a novel pre-expectation framework that enables to reason about the preconditions of quantum programs without the need of an upper bound. This is particularly interesting for quantum programs involving reward statements. The overall goal is to analyze runtime behavior even in the case of programs with potentially infinite expected runtime. This paper presents several ways to do so, e.g., a program transformation such that the expected runtime of a quantum program can be expressed using the weakest pre-expectation calculus with rewards.
Learning-based Probabilistic Load Forecasting with Post-hoc and In-model Uncertainty
arXiv:2607.12730v1 Announce Type: new Abstract: Smart-building load forecasters are often trained offline on dense, multivariate, high-frequency data, but deployment may provide only hourly, feature-limited inputs. Missing features must then be reconstructed, and their errors can propagate through the model. If this input uncertainty is not reflected, prediction intervals may become miscalibrated, affecting demand-response scheduling. Our work examines where uncertainty should be placed once inference inputs are reconstructed. We develop a unified one-day-ahead probabilistic forecasting framework that aligns temporal resolution, reconstructs the unavailable inputs, and derives causal features, and we compare a modular post-hoc residual-quantile scheme with an integrated in-model quantile-learning scheme. The comparison uses three mid-scale Deep Learning (DL) backbones: recurrent, hybrid recurrent, and attention-based Temporal Fusion Transformer (TFT) models, under identical inputs, forecasting horizon, preprocessing rules, and training budgets. Results show that uncertainty placement is backbone-dependent. Integrated quantile learning is most reliable with the TFT, yielding 2.2-3.6% MAPE and 28-83W RMSE on the labeled test window, while producing intervals about 5x narrower than the modular intervals at the closest-to-nominal coverage level. Diebold-Mariano tests support the TFT ranking and the mixed behavior of the recurrent backbones. A reconstruction-sensitivity test shows that reconstructed inputs increase the Quantile Score (QS) by 106% while interval width remains nearly unchanged, indicating that the model does not automatically absorb reconstruction-induced uncertainty. Robustness checks against non-DL baselines and seasonal hold-out weeks support this ranking. Our results expose the limits of post-hoc residual quantiles when inference depends on reconstructed inputs.
A Comprehensive Evaluation of Deep Learning Object Detection Models on Heterogeneous Edge Devices
arXiv:2409.16808v3 Announce Type: replace Abstract: Modern applications such as autonomous vehicles, intelligent surveillance, and smart city systems increasingly require object detection on resource-constrained edge devices. Yet, there is still limited understanding of how different object detection models behave across heterogeneous edge devices and under varying scene complexity. In this paper, we benchmark YOLOv8 (Nano, Small, Medium), EfficientDet Lite (Lite0, Lite1, Lite2), and SSD (SSD MobileNet V1, SSDLite MobileDet) on Raspberry Pi 3, 4, 5 with/without Coral TPU accelerators, Raspberry Pi 5 with AI HAT+, Jetson Nano, and Jetson Orin Nano. We evaluate energy consumption, inference time, and accuracy, and further examine how accuracy changes with the number of objects in the input image. The results reveal clear trade-offs among accuracy, latency, and energy efficiency across model-device combinations. SSD MobileNet V1 achieves the lowest latency and energy consumption but the lowest accuracy, whereas YOLOv8 Medium achieves the highest accuracy at higher computational cost. TPU-based Raspberry Pi devices improve the efficiency of SSD and EfficientDet Lite while reducing YOLOv8 accuracy. Orin Nano offers the most favorable overall balance across most model families. The object-count-based analysis further shows that models achieve more similar accuracy on simpler images, while the accuracy gap widens as scene complexity increases.
Growing a Tail: Increasing Output Diversity in Large Language Models
arXiv:2411.02989v2 Announce Type: replace Abstract: How diverse are the outputs of large language models when diversity is desired? We examine the diversity of responses of several language models to questions with multiple possible answers, comparing them with human responses. Our findings suggest that models' responses are highly concentrated, reflecting narrow, mainstream outputs, in comparison to humans, whose responses exhibit a much longer-tail. We examine three simple and practical ways to increase output diversity: 1) increasing generation randomness via temperature sampling; 2) prompting models to answer from diverse perspectives using a single prompt; 3) aggregating outputs from several models. We find that these interventions, especially when combined, can substantially increase output diversity, although single-model outputs generally remain less diverse than the human baseline. We discuss potential implications of these findings for future work in AI policy and governance that wishes to preserve cultural diversity, an essential building block of a democratic social fabric.
MolMiner: Toward Controllable, 3D-Aware, Fragment-Based Molecular Design
arXiv:2411.06608v3 Announce Type: replace Abstract: We introduce MolMiner, a fragment-based, geometry-aware, and order-agnostic autoregressive model for molecular design. MolMiner supports high-dimensional conditional control over twelve physicochemical and structural properties from partial specifications, constructs molecules via symmetry-aware fragment attachments, and conditions each generation step on force-field-relaxed three-dimensional geometry of the partial structure. Conditional control emerges without auxiliary property losses. On targeted property windows, conditioning lifts hit rates by up to 5.25x over unconditional generation and 3.5x over the training distribution itself -- overriding the model's intrinsic biases -- at the cost of a small reduction in unconditional distributional fidelity. MolMiner unifies dynamic geometry, symmetry handling, order-agnostic generation, and scalable multi-property conditioning within a single framework.
A new randomized CholeskyQR based on LU decomposition with partial pivoting
arXiv:2412.06551v5 Announce Type: replace Abstract: CholeskyQR has received considerable attention in recent years for its efficiency and simplicity in computing QR decomposition of the tall-skinny $X \in \mathbb{R}^{m\times n}$ with $m \ge n$ and $\mbox{rank}(X)=n$. Leveraging matrix sketching from randomized linear algebra, randomized CholeskyQR (RCholeskyQR) has been proposed to accelerate the computation by reducing the dimension of the problems. In this work, we propose RCLUPP, a new randomized CholeskyQR-type algorithm based on LU decomposition with partial pivoting (LUP decomposition). By taking LUP decomposition and the thin HouseholderQR on the sketched matrix, RCLUPP significantly improves the applicability and efficiency compared to LU-CholeskyQR2 (LC2). We present a rigorous rounding error analysis of RCLUPP, with a sharper bound of residual compared to those in the existing works. Comparative studies demonstrate that RCLUPP outperforms CholeskyQR2, Shifted CholeskyQR3 (SCholeskyQR3), and LC2 in terms of applicability while maintaining competitive accuracy and efficiency. A variant, RCLUPPr, performs LUP decomposition directly on $X \in \mathbb{R}^{m\times n}$, offering exceptional robustness and numerical stability for the ill-conditioned scenarios, which exceeds that of RCLUPP and RCholeskyQR. Numerical experiments on the synthetic and real-world matrices validate the theoretical results.
Nanoparticle Arrays for Efficient Organic Light-Emitting Diode Emission Management
arXiv:2607.12635v1 Announce Type: new Abstract: OLEDs are increasingly applied in illumination and displays because they offer excellent color quality, are mechanically flexible, and are self-emissive. However, their usage is limited by low external quantum efficiency (EQE) and efficiency roll-off at high driving voltages. These limitations, together with demands for smaller pixels and device sizes in emerging technologies, motivate innovations that increase efficiency and allow replacing external optical elements with embedded solutions. Here, we demonstrate enhanced outcoupling as well as directional and polarization control of OLED emission, based on collective surface lattice resonances of plasmonic nanoparticle arrays that are embedded in the active layers of four different state-of-the-art OLED structures. Both square arrays and more complex lattices producing flat bands are demonstrated to guide the light to directions and polarizations determined by their optical modes. We show that by the design of the array geometry and the OLED structure, spectral and angular enhancement of the electroluminescence (EL), up to 30 %, can be achieved. Our results verify that surface lattice resonances of nanoparticle arrays offer a robust and versatile embedded solution for tailoring the OLED emission, as well as exciting prospects for efficiency increase if combined with narrow-spectrum emitters.
GeoFovea-GS: Geometry-Aware Cross-Layer Gaussian Splatting for Wireless Aerial VR
arXiv:2607.12641v1 Announce Type: new Abstract: Wireless aerial virtual reality (VR) aims to provide immersive access to large-scale scenes, but high-resolution view generation and delivery are jointly constrained by limited bandwidth, latency, and power. 3D Gaussian Splatting (3DGS) can reduce the payload by rendering views from compact pose information, yet its geometry errors may cause severe VR quality degradation. Existing channel-aware or pixel-level resource allocation schemes fail to capture such geometry-sensitive distortion. To address this issue, this paper proposes GeoFovea-GS as a geometry-aware cross-layer framework for communication-efficient wireless aerial VR. A foveated geometry-aware distortion metric is developed to characterize photometric rendering error, geometric inconsistency, and view-dependent perceptual importance in a unified form. Based on this metric, the joint selection of pose-only 3DGS rendering and image/tile correction transmission is formulated as a cross-layer optimization problem under wireless constraints. A lightweight value-of-information scheduler is further developed to allocate communication resources to regions that are both geometry-critical and perceptually important. Experiments on real-world 3DGS scenes demonstrate that GeoFovea-GS achieves superior immersive rendering quality with substantially reduced transmission cost.