Forskningsradar

Science Journals

Peer-reviewade publikationer — 53899 artiklar

Debiasing Text-to-Image Evaluation via Implicit Cultural Alignment Reward Modeling
arXiv:2607.15740v1 Announce Type: new Abstract: As Text-to-Image (T2I) systems rapidly advance, evaluating the cultural authenticity of synthesized content has become increasingly important for fair and trustworthy generative AI. Existing T2I evaluation metrics and multimodal judges often rely on visual-semantic representations that underrepresent implicit cultural norms, leading to biased preference judgments and the omission of fine-grained cultural cues. In addition, visual question answering (VQA)-based evaluators typically depend on autoregressive text generation, which limits their scalability for real-time reward modeling. To address these limitations, we introduce an Implicit Cultural Alignment Reward Model built upon a lightweight 4.2-billion-parameter Multimodal Large Language Model (MLLM). Our framework integrates an Implicit Cultural Probe with a Skip-connection Cross-Attention (SkipCA) mechanism, enabling late-stage semantic features to directly attend to early-stage visual representations and better preserve culturally salient details. Evaluations on 3,323 challenging and carefully curated image pairs from the CulturalFrames benchmark show that our approach achieves 80.54% pairwise accuracy, with Pearson and Kendall correlation coefficients of 0.546 and 0.377, respectively, outperforming representative vision-language metrics and MLLM-based evaluators. Moreover, by bypassing autoregressive text generation, our model processes each evaluation in 0.21 seconds under our local inference setup, achieving a $10\times$ speedup over standard VQA-based evaluators. These results suggest that the proposed reward model can provide an efficient and culturally aware scalar signal for preference optimization pipelines such as Reinforcement Learning from Human Feedback and Direct Preference Optimization.
Orbis 2: A Hierarchical World Model for Driving
arXiv:2607.15898v1 Announce Type: new Abstract: Current world models operate at a single level of abstraction, with most prioritizing perceptual fidelity while lacking the spatial reasoning and semantic understanding required for real-world downstream tasks. We present a hierarchical driving world model that factorizes future prediction across two levels operating at distinct temporal and abstraction scales: a high-level predictor that forecasts coarse scene structure over extended temporal horizons, and a low-level generator that produces detailed predictions conditioned on the high-level output. This decomposition yields high perceptual fidelity while also capturing strong spatial and semantic representations. We further show that pretraining with a diffusion forcing objective yields substantially richer internal representations than the standard teacher forcing objective, while teacher forcing -- predicting only the next frame from clean context -- produces more stable autoregressive rollouts. We therefore introduce a generic two-stage training paradigm that pretrains the model with diffusion forcing and fine-tunes with teacher forcing, combining the representational benefits of the former with the rollout stability of the latter. Our approach achieves state-of-the-art results across the standard suite of driving world model evaluations on established benchmarks, including long-horizon generation fidelity, steering responsiveness evaluated on counterfactual scenarios, and internal representation quality. Project page with code, demo, checkpoints and qualitative results: https://lmb-freiburg.github.io/orbis2.github.io/
Growing Hypergraphs with Homophily
arXiv:2607.16046v1 Announce Type: new Abstract: There are many extant models of hypergraphs with interactions governed by attribute-based homophily between nodes, but most assume independence between edges conditional on node parameters. Relaxing this assumption, we study a mechanistic model of growing hypergraphs in which edge formation is influenced by both previous edges and binary node labels. Edges form in this model as noisy copies of previous edges, where the transmission of nodes from one edge to the next depends on multiple homophilic mechanisms between labels. These homophilic mechanisms give rise to tunable assortative structure in the hypergraph. We derive a power law for the degree distribution in this model and describe the long-term dynamics of the joint distribution of labels contained in edges. Our model defines a likelihood over a labeled hypergraph, allowing us to use standard maximum-likelihood techniques to structure algorithms. We estimate the model parameters on synthetic and real data via (stochastic) expectation maximization. These estimates give statistically-principled descriptions of the operation of homophily in empirical polyadic systems. We also demonstrate an approach to community detection via simulated annealing which, though computationally expensive, achieves competitive results on both synthetic data and certain empirical data sets known to be challenging to community detection techniques based on the edge-independence assumption. Our findings highlight the benefits of incorporating edge- and label-dependence in higher-order modeling and data analysis, and point to several directions for future work.
PASs-MoE: Mitigating Misaligned Co-drift among Router and Experts via Pathway Activation Subspaces for Continual Learning
arXiv:2601.13020v2 Announce Type: replace Abstract: Continual instruction tuning (CIT) requires multimodal large language models (MLLMs) to adapt to a stream of tasks without forgetting prior capabilities. A common strategy is to isolate updates by routing inputs to different LoRA experts. However, existing LoRA-based Mixture-of-Experts (MoE) methods often jointly update the router and experts in an indiscriminate way, causing the router's preferences to co-drift with experts' adaptation pathways and gradually deviate from early-stage input--expert specialization. We term this as Misaligned Co-drift, which blurs expert responsibilities and exacerbates forgetting. To address this, we introduce the pathway activation subspace (PASs), a LoRA-induced subspace that reflects which low-rank pathway directions an input activates in each expert, providing a capability-aligned coordinate system for routing and preservation. Based on PASs, we propose a fixed-capacity PASs-based MoE--LoRA method with two components: PAS-guided Reweighting, which calibrates routing using each expert's pathway activation signals, and PAS-aware Rank Stabilization, which selectively stabilizes rank directions important to previous tasks. Experiments on a CIT benchmark show that our approach consistently outperforms a range of conventional continual learning baselines and MoE--LoRA variants in both accuracy and resistance to forgetting, without increasing model parameters. Our code is publicly available at https://github.com/yueluoshuangtian/PASs-MoE.
Sym2Real: Symbolic Dynamics with Residual Learning for Data-Efficient Adaptive Control
arXiv:2509.15412v2 Announce Type: replace Abstract: We present Sym2Real, a fully data-driven framework for highly data-efficient adaptation of low-level controllers. Although symbolic regression is data-efficient, its role in real-world control has been limited due to its sensitivity to measurement noise, which corrupts the equations and leads to model degradation when fitted directly on real-world data. Sym2Real addresses this limitation by 1) learning first from low-fidelity simulation, where noise-free trajectories allow symbolic regression to identify the underlying dynamics, and 2) using a small amount of real-world data for targeted residual adaptation to bridge the sim-to-real gap. Using only about 10 trajectories, we achieve robust control of both a quadrotor and a racecar in the real world, without expert knowledge or simulation tuning. Through experimental validation on both platforms, we demonstrate consistent data-efficient adaptation across 6 out-of-distribution sim2sim scenarios and successful sim2real transfer across 5 real-world conditions. More information can be found at http://generalroboticslab.com/Sym2Real
Video = World + Event Stream
arXiv:2607.15038v2 Announce Type: replace Abstract: We present Wan-Streamer v0.3, which reframes our native-streaming interaction model under a single organizing view: a video is a world plus an event stream. The world is the persistent context in which a video unfolds, including the environment, scene, subjects, ambient acoustic conditions, voice characteristics, and other relatively stable conditions. The event stream is everything that changes over time within that world, including scene or environmental changes, subject behavior, speech, and other sounds. This yields a general-purpose pretraining task over large amounts of real video: given a world and incoming input, predict how the world moves, changes, and responds in real time. The resulting competence can be specialized to a broad family of real-time downstream tasks. We instantiate it on real-time full-duplex audio-visual interaction, where the event stream is the agent's speech together with free-form behavior. Functionally, the model's multimodal understanding process is vision-language-action-like: it maps multimodal user input to language-form speech and behavior actions. Wan-Streamer v0.3 preserves the v0.2 operating point: 640x368 video at 25 FPS, a 160 ms streaming unit, approximately 200 ms model-side response latency, and approximately 550 ms total interaction latency under a 350 ms bidirectional network budget.
Mitigation of Initial Transients in Total-f Gyrokinetic Turbulence Simulations Using Neoclassically Relaxed Distribution Function
arXiv:2607.15072v2 Announce Type: replace Abstract: Total-f five-dimensional gyrokinetic simulations are essential for self-consistent studies of multi-scale, multiphysics transport in the edge region of diverted tokamak plasmas. However, conventional initialization with a local Maxwellian distribution often generates large-amplitude transients, particularly geodesic acoustic modes (GAMs). These transients are especially severe in the plasma edge because of steep profile gradients, strong radial electric fields, and high safety factors, and they increase the computational time required to reach a saturated turbulent state. To address this problem, we present a new initialization scheme for the total-f XGC code that uses a relaxed particle distribution obtained from a computationally inexpensive axisymmetric simulation. Before the distribution is transferred to the full turbulence simulation, phase-space smoothing is applied to reduce particle noise while preserving its neoclassical structure. Applications to the Cyclone Base Case and an ASDEX Upgrade I-mode discharge demonstrate substantial suppression of transient GAMs, reduced particle noise, and a significant reduction in time to solution.
Long-Context Fine-Tuning with Limited VRAM
arXiv:2607.15105v2 Announce Type: replace Abstract: Parameter-efficient fine-tuning reduces model and optimizer memory, but dense attention still makes long training sequences expensive. We combine Hierarchical Global Attention (HGA) with segment-wise backpropagation and tiered KV storage. Only the active segment remains differentiable in VRAM; older KV is detached into RAM or NVMe, and HGA loads a bounded set of exact historical tokens for each query block. On Qwen3-8B with 4-bit QLoRA and PG19, dense training on a 16 GB Quadro RTX 5000 fits 2,048 tokens but fails at 4,096, whereas HGA reaches 16,384 tokens with 15.28 GB peak VRAM. Under evaluation the same adapter runs through 131,072 tokens on this card; VRAM is not constant but grows gently with the resident chunk summaries, so RAM and NVMe capacity set the practical limit beyond these lengths. At the shared 2K training length, HGA-trained and dense-trained adapters obtain 2.7405 and 2.7383 nat under the same dense-attention readout, while the stock model obtains 2.9541. At this boundary HGA training is already marginally faster (217.75 vs. 207.02 tokens/s), and the HGA-to-dense throughput ratio improves from 1K to 2K; because HGA keeps the attended historical set per token approximately constant while dense work per token grows, we expect this lead to widen as context grows. Dense attention is used for the main quality and retrieval comparisons so that they measure the learned weights and remain compatible with standard generation frameworks. HGA can also be used for retrieval and generation; an optimized production-grade serving implementation is under development.
T^2MLR: Transformer with Temporal Middle-Layer Recurrence
arXiv:2607.15178v2 Announce Type: replace Abstract: Transformer reasoning is limited by autoregressive decoding, which repeat edly compresses rich hidden computation through token space and makes it difficult for intermediate reasoning states to persist across time. We in troduce Transformers with Temporal Middle-Layer Recurrence (T2MLR), a transformers-based latent reasoning architecture that fuses a cached middle layer representation from the previous token directly into an earlier layer of the current token position, enabling abstract intermediate computation to persist across decoding steps with little inference overhead. Across natural-language pretraining and multi-hop reasoning finetuning, T2MLR consistently outperforms data- and parameter-matched Transformer base lines. Moreover, applying recurrence to only a localized middle-layer block (as little as 20% of the network) often outperforms full-layer recurrence. Im portantly, T2MLR does not require pretraining from scratch: retrofitting the recurrent pathway into an existing pretrained 1.7B Transformer and briefly finetuning substantially improves math reasoning, lowering the barrier to practical adoption. These results suggest that effective latent reasoning in Transformers does not require looping over all layers as in previous works, but can instead emerge more strongly from targeted middle-layer recurrence.
Can We Trust Item Response Theory for AI Evaluation?
arXiv:2607.15190v2 Announce Type: replace Abstract: AI benchmarks increasingly leverage item-level statistical models, particularly item response theory (IRT), to estimate model capabilities, rank systems, select informative examples, and diagnose benchmark quality. However, AI benchmark data often departs from the data regime of human testing, for which standard IRT estimation tools were originally developed: benchmarks typically involve fewer evaluated models, far more items, and capability distributions that may be skewed, clustered, or multimodal. We examine how these regime mismatches challenge the reliability of IRT modeling for AI evaluation. Using item parameters and capability distributions derived from six widely used LLM benchmarks, we simulate response matrices under three common IRT models and compare four estimation tools used in recent benchmark studies: marginal maximum likelihood, Markov chain Monte Carlo, variational inference, and a neural pseudo-Siamese estimator. Across 18,000 simulation conditions, we systematically evaluate computational feasibility, scalability, and the reliability of IRT inferences about model rankings, predicted performance, and item characteristics. Results show that classical estimators can become infeasible in large benchmark settings, whereas scalable estimators can produce unreliable item-level and ranking inferences with small or non-normally distributed model sets. This study identifies when latent trait models reliably support or risk distorting AI benchmarking claims, and what sample sizes and diagnostics are needed for trustworthy use.
Beyond Success Rate: Cost-Aware Evaluation of Offensive and Defensive Security Agents
arXiv:2607.15263v2 Announce Type: replace Abstract: Security-agent evaluations commonly measure peak offensive capability under generous inference budgets, emphasizing vulnerability discovery, exploit development, penetration testing, and CTF completion. Such measurements are useful but incomplete: in operational security, every reasoning step, tool call, telemetry query, and enrichment request consumes budget. We evaluate language-model security agents through this cost-success lens on offensive Cybench challenges and defensive Splunk BOTS v1 investigation challenges. Instead of reporting only best-case success, we compare models at fixed cost levels and decompose performance by inference spend and tool spend. Our results show distinct scalingregimes for red- and blue-team tasks. Offensive CTF performance improves with additional test-time compute, and scaled open-weight models can approach frontier proprietary systems while remaining cost-competitive. Defensive SOC investigation does not scale in the same way: success depends more heavily on disciplined tool use, telemetry navigation, and selective enrichment than on raw reasoning budget alone. We argue that security-agent benchmarks should measure economic efficiency and operational fit alongside task success. Cost-aware, SOC-native evaluations provide a clearer picture of which models are practically useful today and where defensive agents still need to improve. We present an interactive website with our results https://evals.frontier.security.
Minimax and Bayes Optimal Best-Arm Identification
arXiv:2506.24007v5 Announce Type: replace-cross Abstract: This study investigates minimax and Bayes optimal strategies for fixed-budget best-arm identification. We consider an adaptive procedure consisting of a sampling phase followed by a recommendation phase, and we design an adaptive experiment within this framework to efficiently identify the best arm, defined as the one with the highest expected outcome. In our proposed strategy, the sampling phase consists of two stages. The first stage is a pilot phase, in which we allocate samples uniformly across arms to eliminate clearly suboptimal arms and to estimate outcome variances. Before entering the second stage, we solve a Gaussian minimax game, which yields a sampling policy and a decision rule. In the second stage, samples are allocated according to this policy. After the sampling phase, the procedure enters the recommendation phase, where we select an arm using the decision rule. We prove that this single strategy is simultaneously asymptotically minimax and Bayes optimal for the simple regret, and we establish upper bounds that coincide exactly with our lower bounds, including the constant terms. The lower bounds hold against every adaptive experiment and for every fixed number of arms, and the strategy attains them without knowing the outcome distributions or the prior.
Nutation Damping from Core-Mantle Boundary Topography
arXiv:2507.01671v3 Announce Type: replace-cross Abstract: Periodic gravitational forcing by the Moon and Sun produces small oscillations in Earth's rotation known as nutations. Nutations are amplified by a resonance with a natural motion of the liquid core called the Free Core Nutation, whose amplitude is limited by friction-like processes at the core--mantle boundary. Previous studies have attributed this damping to the dissipation of electric currents induced in the lower conducting mantle, but, given current knowledge of the lower mantle, electromagnetic coupling appears insufficient to fully account for the observed lag. We show that additional dissipation arises from the interaction of the tidal flow inside the core with the topography of the core--mantle boundary, which excites internal waves that extract energy and momentum from the flow. Adapting a theory originally developed for tides over seafloor topography, we find that the observed damping can be fully accounted for by a topography of typical amplitude $\sim$5~km dominated by features of wavelength $\sim$1500~km. The dissipation is highest when the upper core is neutrally buoyant. Such amplitudes are larger than typical inferences from global seismic studies but are not ruled out, given regional seismic evidence for kilometer-scale features and the sparse constraints at these scales.
3-Colouring Planar Graphs
arXiv:2507.03163v2 Announce Type: replace-cross Abstract: We show that every $n$-vertex planar graph is 3-colourable with monochromatic components of size $O(n^{4/9})$. The best previous bound was $O(n^{1/2})$ due to Linial, Matou\v{s}ek, Sheffet and Tardos [Combin. Probab. Comput., 2008].
Demonstration of deuterium's enhanced sensitivity to symmetry violations governed by the Standard-Model Extension
arXiv:2507.07473v2 Announce Type: replace-cross Abstract: We have performed hyperfine spectroscopy of two transitions in ground-state deuterium and searched for violations of CPT and Lorentz symmetry that would manifest as sidereal variations of the observed transition frequencies. Several nonrelativistic proton coefficients of the Standard-Model Extension framework have been addressed. The spin-independent coefficients with momentum power k=2,4 are constrained for the first time. Bounds on spin-dependent coefficients are improved by exploiting a sensitivity enhancement originating from the relative momenta of the nucleons in the deuteron. The best previous constraints by hydrogen maser measurements are surpassed by 4 and 14 orders of magnitude for coefficients with k=2 and 4, respectively.
Unified ab initio quantum-electrodynamical density-functional theory for cavity-modified electron-phonon-photon coupling in solids
arXiv:2603.24095v2 Announce Type: replace-cross Abstract: Quantum-electrodynamical density-functional theory (QEDFT) provides a first-principles framework for describing materials coupled to quantized electromagnetic fields. While QEDFT has successfully captured cavity-induced modifications of electronic structures in atoms and molecules, a fully self-consistent and accurate framework to simulate and predict the structural, phonon-related, polarization and optical response of periodic solids in optical cavities has remained elusive. Here, we introduce a unified QEDFT approach that combines collective light-matter coupling parameter in the electronic ground state, density functional perturbation theory for phonons, and real-time time-dependent QEDFT for optical excitations. This framework enables ab initio calculations of cavity-modified electronic and phononic dispersions, Born effective charges, dielectric tensors, and both resonant and non-resonant optical absorption spectra. Using wurtzite gallium nitride (GaN) in an optical cavity as a case study, we demonstrate that the quantized vacuum field reshapes electronic, phononic and polarization properties, producing experimentally accessible signatures in the dielectric function and absorption spectra. These results establish QEDFT as a general first-principles platform for predicting and exploring cavity-modified quantum materials.
Random Access Codes: Explicit Constructions, Optimality, and Classical-Quantum Gaps
arXiv:2604.21274v3 Announce Type: replace-cross Abstract: A random access code (RAC) encodes an $L$-bit string into a $k$-bit message, $L>k$, so that any requested bit can be recovered with high probability; a quantum RAC (QRAC) uses $k$ qubits instead. We give a geometric characterization of optimal classical $(L,k)$-RACs under average and worst-case decoding criteria. The average criterion is reduced to choosing $2^k$ representatives in $\{0,1\}^L$, while the worst-case criterion is reduced to a minimax problem over $2^k$ points in $[0,1]^L$ with a distance-like objective. This framework proves optimality for several parameter families, with many optimal constructions arising from standard infinite families of binary linear codes. It also yields two explicit classical--quantum separations. First, for every $L>1$, we construct a $(L,1)$-QRAC whose average decoding success probability strictly exceeds the optimal classical value. Second, for the family $(2^k-1,k)$, we prove worst-case optimality of a classical RAC and construct a QRAC with strictly larger worst-case success probability. For the family $(L,L-1)$, the framework identifies a classical RAC that is average-case optimal and, under a stated conjecture, also worst-case optimal. The same viewpoint further recovers explicit $(L,L-1)$-QRACs attaining a previously conjectured upper-bound value.
Collective amplification and anisotropic narrowing of alignment signals in cesium vapor under strong spin exchange near zero magnetic field
arXiv:2605.13466v3 Announce Type: replace-cross Abstract: We present the results of an experimental study of the anomalous anisotropy of alignment signals in cesium vapors under strong spin-exchange conditions near zero magnetic field with linearly polarized optical pumping. We show that the anisotropy of the Hanle resonances in the plane perpendicular to the pump beam increases with concentration: in one direction the widths remain broadened by spin-exchange, whereas in the other they approach the spin-exchange relaxation free limit. With a further increase in concentration, additional nonlinear effects arise, such as signal amplification, bistability, hysteresis, and memory. To explain these effects we construct a illustrative theoretical model incorporating spontaneous polarization effects under strong spin exchange conditions. The model qualitatively shows that the ultra-narrow alignment resonances may originate from quadrupole anisotropy arising from the projection of spontaneous transverse orientation onto the detection axis. The unique properties of these resonances, such as their extremely small width and magnetic field-controlled bistability with a long-term memory effect, make them promising for use in quantum sensing and information.
Seasonal Statistics of Shannon Rate in a Dynamical Poisson-Voronoi Cellular Network
arXiv:2605.16560v3 Announce Type: replace-cross Abstract: In this work we consider a dynamical cellular communication network in which mobile base stations (BSs) are modeled as a homogeneous Poisson point process on $\mathbb{R}^2$. Each base station moves at a constant speed in a random direction. A typical user connects to the nearest base station and it experiences variable signal and interference powers depending on the distance of all the stations. Along the motion of the stations, the user swaps its serving station, and such an event is called a {\em handover}. We are interested in the performance evaluation of the system under some classical and tropical metrics of interest at different time of events, inducing handovers, maximal proximity of serving station, nearest interferer at closest or farthest distance with respect to the user or at any typical time epoch. The main results of the paper are closed or integral form expressions for the basic metrics of interest, in particular coverage probability and Shannon rate at these epochs. We can make an analogy with ``seasons'' based on the fluctuations of signal and interference power. Strong or mild signal or interference power correspond to different seasons of Shannon rate along the evolution of the system. We also provide a complete comparison study of the metrics at interest at these epochs.
Instability in Complex Oscillator Networks: Limitations and Potentials of Network Measures and Machine Learning
arXiv:2402.17500v2 Announce Type: replace-cross Abstract: A central question of network science is how functional properties of systems emerge from their structure. For networked dynamical systems, structure is typically captured through network measures. We investigate the relationship between these measures and stability metrics across non-linear and linear oscillators, as well as real-world power grid topologies and dynamics. We find that this relationship is highly sensitive to the underlying ensemble: minor changes in the networks considered, such as going from mean degree 6 to mean degree 8, can invert the correlation between a network measure and stability. We also investigate network measures as inputs for machine learning, as well as Graph Neural Networks (GNNs) as predictors of stability. Both GNNs and the non-linear combination of many network measures can accurately predict stability within a given ensemble, yet both can fail when the ensemble changes. We conclude that neither approach reliably identifies the underlying structural causes of instability.
A Clinically Validated Foundation Model for Comprehensive Lung Pathology Interpretation
arXiv:2605.25878v2 Announce Type: replace-cross Abstract: Pathological assessment guides lung cancer diagnosis, treatment selection, and prognostic evaluation, yet current CPath approaches rely on task-specific models for isolated objectives. Although pan-cancer foundation models offer versatility, they lack subspecialty-level depth and have not been evaluated across clinical workflows or prospectively validated in real-world settings. We introduce PulmoFoundation, a multi-center, prospectively validated, randomized controlled trial (RCT)-evaluated foundation model for comprehensive lung pathology assessment across pre-operative, intra-operative, and post-operative care. Built upon Virchow2 via subspecialty-specific pretraining using ~40,000 diagnostic H&E-stained whole-slide images (WSIs), PulmoFoundation was systematically evaluated on ~26,000 WSIs across 32 clinically relevant tasks. In addition to accurately predicting molecular markers and patient survival, our model achieves clinical-grade performance in core diagnostic tasks across biopsy, frozen section, and surgical resection slides. In a registered prospective study of 1,357 patients across 11 diagnostic tasks, our model achieved an average AUC of 92.3%. Using pre-specified triage thresholds, PulmoFoundation could reduce additional second-review burden for 68.8% of biopsies and 83.0% of frozen sections, and defer 44.5% of IHC stain orders, with PPVs of 1.000, 0.991, and 0.966. Beyond prospective validation, we conducted a crossover RCT with eight pathologists, in which AI assistance improved diagnostic accuracy across 5,264 case-reader pairs (91.7% w/ AI vs. 83.2% w/o AI). AI assistance also reduced median diagnostic time by 18.3%, increased diagnostic confidence by 9.0%, and improved inter-rater agreement from moderate (kappa = 0.55) to substantial (kappa = 0.76). Together, these evaluations support PulmoFoundation as a clinically validated decision-support system for lung pathology.
A Proof in Coq that Core Logic is not Paraconsistent
arXiv:2606.05953v3 Announce Type: replace-cross Abstract: Tennant claims that his Core logic $\mathbb{C}$ is paraconsistent. It means that the sequent of the First Lewis Paradox, i.e. $\lnot A, A \vdash B$ is declared false, and its corresponding antisequent, called `Claim~1', i.e. $\lnot A, A \nvdash B$ true, as in minimal logic $\mathbf{M}$. This paper proves that Claim~1 entails a contradiction in $\mathbb{C}$, so that, to preserve consistency, the Core logician must reject the claim that his system is paraconsistent. The proof is purely logical, in four steps within a five-rule fragment $\mathcal{F}$ of $\mathbb{C}$ and its refutation system in the sense of Lukasiewicz and Goranko; the Appendix certifies every step in Coq -- with no axiom assumed and every commitment displayed as a named hypothesis -- and the same certification is replayed independently in Lean~4.
VTLoc: Learning-based Tactile Contact Localization in Visual Point Clouds
arXiv:2607.16146v1 Announce Type: new Abstract: Vision and touch are complementary modalities essential for robotic perception and manipulation. While vision provides global object context, touch offers precise local information at contact points. Integrating these modalities for contact localization, i.e., predicting the location of touch on an object's surface, poses significant challenges due to the need for accurate spatial alignment between tactile data and visual geometry. To address this challenge, we propose VTLoc, a novel visual-tactile framework that localizes contact points from tactile readings using a 3D point cloud as visual input. VTLoc introduces two key components: a geometric multi-modal alignment module, which reconstructs a pseudo-point cloud from fused visual-tactile features and aligns it with the visual point cloud to enforce spatial consistencies across modalities; and an iterative localizing updater, which iteratively refines the predicted contact location using fused visual-tactile features. Evaluated on a new benchmark of 100 real-world objects, VTLoc improves single-touch contact localization by reducing local-to-global correspondence ambiguity.
A Kalman Filter-Assisted Data-Predictive SAR ADC With Reduced Switching Energy for Low-Power Applications
arXiv:2607.16139v1 Announce Type: cross Abstract: The proliferation of Internet of Things (IoT) devices and wearable health monitors has created an urgent demand for ultra-low-power analog-to-digital converters (ADCs). Successive approximation register (SAR) ADCs are widely used in such applications, yet their energy efficiency remains constrained by the sequential bit-by-bit switching of the capacitive DAC (CDAC). The high-weight most significant bit (MSB) transitions dominate the total switching energy, and the rigid N -cycle conversion flow imposes a hard lower bound on latency per sample.This paper presents a Kalman filter-assisted data-predictive SAR ADC that replaces the first four comparator-driven decisions with a recursive state estimator. The Kalman filter predicts the 4 MSBs from the complete conversion history before each cycle begins, enabling simultaneous parallel switching of the MSB capacitors. This eliminates redundant CDAC transitions, shortens the quantization cycle by four clock periods, and reduces switching energy by approximately 50%. An optimized 4-bit MSB switching scheme further suppresses residual switching at the hardware level. The ADC, designed in a 180-nm CMOS process, supports configurable dual-mode operation, toggling between a conventional mode and the Kalman-driven predictive mode for robustness under erratic inputs. At 20 MS/s and a 1.8-V supply, the predictive mode reduces total power consumption by 50.3% (from 1.96 mW to 0.975 mW), with a measured SNR/SFDR of 57.88/74.51 dB at 504 kHz, confirming its suitability for energy-constrained wireless sensor networks.
Breakdowns for Human-Machine Creative Reflexivity
arXiv:2607.15866v1 Announce Type: new Abstract: Generative AI (GenAI) works via goal-directed computation, which differs fundamentally from human creative processes. This poses challenges for the intelligent support of creative experiences. We propose ``breakdowns'' as opportunities for the exchange of perspectives between human and machine. Breakdowns disrupt a flow and force us to consciously evaluate our ``being-in-the-world''. Between human and machine, breakdowns can function as openings for collaborative creative reflection. We are currently studying human-human creative interactions, to identify the markers of these inter-subjective openings, and to understand how they are used in a co-creative process. We present preliminary findings on breakdowns as a design principle for creativity support, prioritising human creative agency and meaningful reflection over automated content generation.