Forskningsradar

Science Journals

Peer-reviewade publikationer — 54780 artiklar

Online TCP Acknowledgment under General Delays
arXiv:2604.13428v2 Announce Type: replace Abstract: In a seminal work, Dooly, Goldman, and Scott (STOC 1998; JACM 2001) introduced the classic Online TCP Acknowledgment} problem: a sequence of $n$ packets arrives over time, and the objective is to minimize both the number of acknowledgments sent and the total delay experienced by the packets. They showed that a natural greedy algorithm, which acknowledges when the delay of pending packets equals the acknowledgment cost, is $2$-competitive. Online TCP Acknowledgment is the canonical online problem with delay, capturing the fundamental tradeoff between reducing service cost through batching and the delay incurred by pending requests. Prior work has largely focused on richer service-cost models, e.g., Multi-Level Aggregation. However, besides the work of Albers and Bals (SODA 2003), which studies maximum delay and similar objectives, not much is known beyond the sum of delay costs of requests. In this work, we study Online TCP Acknowledgment under two generalized delay-cost models. In the batch-aware model, each batch incurs a delay cost that depends on the packet delays within that batch. For the max-over-batches objective, we show that greedy remains $2$-competitive. For the sum-over-batches objective, the picture changes sharply: greedy is $\Omega(n)$-competitive, and the optimal deterministic competitive ratio is $\Theta(\log n)$. Our upper bounds only require the batch delay function to be monotone. In the batch-oblivious model, the delay cost is a function of the global packet-delay vector. We show that greedy is $2$-competitive for continuous submodular delay costs, and more generally under a weaker zero-coordinate diminishing-marginals condition. This yields $2$-competitive algorithms for ordered norms. Using the submodular-norm approximation of Patton, Russo, and Singla, we also obtain an $O(\log n)$-competitive algorithm for arbitrary symmetric norms.
TrustVLA: Mechanism-Guided Inference-Time Defense Against Vision-Language-Action Backdoors
arXiv:2607.12571v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are deployed through pipelines that end users cannot audit, and a poisoned VLA can behave normally on clean observations while a small visual trigger redirects a long-horizon robot policy before any failure becomes observable. Existing vision or language defenses rarely explain what a triggered VLA representation looks like or how to recover behavior without retraining. We study this gap through two independently proposed VLA attacks from groups with distinct injection strategies, BadVLA and INFUSE; the latter persists after downstream clean adaptation. Across the evaluated poisoned models, we identify a recurring internal mechanism: a \emph{compact causal footprint}, namely a small visual support that is attention-seeded, spatially compact, and \emph{causal} in a precise sense -- masking it returns a clean-calibrated evidence-evolution score to the normal operating region. This footprint motivates TrustVLA, a mechanism-guided inference-time defense that adapts the Dirichlet evidence framework from trusted classification to monitor per-token, per-layer epistemic uncertainty in VLA policies. With only a small clean calibration set, TrustVLA (i)~detects abnormal evidence evolution, (ii)~localizes the compact support by counterfactual mechanism-score drop, and (iii)~recovers the observation by localized inpainting. Across OpenVLA/LIBERO and $\pi_{0.5}$ transfer evaluations, TrustVLA reduces attack success while preserving clean-task performance, providing a retraining-free, mechanism-guided defense for visual-triggered VLA backdoors.
Low-Latency Neural Models for Real-Time Music Enhancement
arXiv:2607.12872v1 Announce Type: new Abstract: Music recordings and live streams are often affected by noise, reverberation, spectral imbalances, or artifacts that degrade listening quality. While speech enhancement has matured into a well-defined research area, music enhancement is less established because musical signals combine overlapping sources, wide bandwidths, strong dynamics, and intentional production effects. We study real-time music enhancement under strict causal and low-latency constraints. We formulate the task around recovery of the intended produced mix from acoustic and production-oriented degradations, adapt compact causal networks to music, and compare speech-derived real-time baselines, an external music-denoising model, an offline restoration reference, and a music-specific MusicFilterNet-MS variant. On the tested hardware, all causal models run faster than real time, but improvements depend strongly on the dataset, degradation type, and metric family; under several objective criteria, indiscriminate enhancement can worsen the degraded input. The main contribution is therefore a benchmark and an analysis rather than a universal best model: real-time music enhancement is feasible, but robust improvement requires degradation-aware modeling, stereo-aware processing, identity-preserving correction, and evaluation beyond a single objective score.
Evaluation Bias and Epistemic Inequality in Global Software Development
arXiv:2607.12563v1 Announce Type: new Abstract: This paper examines fairness and accountability in global software development by focusing on how competence is assessed and valued across unequal regional contexts. We compare software engineers from East Africa (Rwanda and Uganda) and Northwestern Europe (Sweden and the Netherlands), regions that are increasingly connected but embedded in asymmetric technological, economic, and institutional structures. Despite the rapid growth of African technology ecosystems, empirical evidence on everyday engineering practice and evaluation in these contexts remains limited. We present findings from an on-site mixed-methods study with 48 software engineers across four countries. The study combines programming, system design, and code review tasks with semi-structured interviews. Our results reveal consistent gaps between measured performance and perceived competence. Senior engineers in Rwanda often performed at a level comparable to that of their European peers, yet European participants systematically underestimated the competence of East African engineers. We also observed differences in communication styles and organizational practices across regions, reflecting distinct but complementary ways of working.
How Agentic Is Agentic Commerce? A Population-Scale Measurement of x402 Adoption and Authenticity
arXiv:2607.12575v1 Announce Type: new Abstract: AI agents are said to be forming an economy in which they pay, on their own, for the data, APIs, and compute they consume. x402, which settles a stablecoin payment on-chain for each purchase, is the most widely deployed protocol for this, and its hundreds of millions of settlements are read as proof that the economy has arrived. We show the count cannot be read as adoption: it is the one metric an interested party can manufacture almost for free, since the facilitator sponsors the gas and nothing on-chain marks who controls a payment. We give the first population-scale measurement of x402 on Base, supplemented with a coarser Solana census. Identifying settlements from their on-chain event and resolving the true payer through the meta-transaction layer, we sort each by what its trace can prove via a payment graph. Over a 280-day window Base carries 136{,}708{,}672 settlements worth \$44{,}121{,}383.81, concentrated on every axis we measure (payer, recipient, and value Gini all above 0.98), yet 21.20\% are fictitious and 63.78\% internal settlement within a linked cluster. What is genuinely independent is bounded: it lies between the \$187{,}861.35 that demonstrably reaches a nameable service and the \$20{,}258{,}746.09 (45.92\% of value) not provably manufactured. Finally, we resolve the count's manufacturable component, a coherent operator-driven economy, star-shaped, machine-timed, and gas-subsidized. Settlement count measures manufacturability, not adoption.
An Omnilingual-ASR-Based Speech-LLM System for the 2nd MLC-SLM Challenge
arXiv:2607.12468v1 Announce Type: new Abstract: We describe our submission to Task 1 of the 2nd MLCSLM Challenge: a cascaded diarization-then-recognition system that combines DiariZen-Large-s80 (WavLM-Large) segmentation, CAM++ embedding-based two-speaker clustering, and a LoRA-adapted omniASR LLM 7B v2 recognizer, with no oracle segmentation or speaker labels at test time. On the official Development set (150 conversations, 21 language/accent categories) the system attains a macro tcpMER of 29.27%, versus 79.15% for the official baseline; on the Evaluation set it scores 50.23%. We also analyze two engineering choices that substantially affect tcpMER. First, embedding-based speaker clustering outperforms an end-to-end-style alternative that assigns speakers from ASR <sc> turn markers alone. Second, overlap-aware segmentation, although intended to raise diarization recall, increases tcpMER because overlapped speech is transcribed twice.
High-Frequency Gravitational Wave Constraints from Precision Spectroscopy
arXiv:2607.12617v1 Announce Type: cross Abstract: Gravitational waves affect the propagation of electromagnetic waves in laser cavities, modulating the frequency of emitted photons. We use this effect to search for high-frequency gravitational waves between 100 kHz and 100 MHz using optical precision spectroscopy. Our limits constrain much of this frequency range for the first time. We discuss future improvements of the technique, which we expect to enhance the sensitivity by eight orders of magnitude, and to extend the frequency coverage up to at least 1 GHz.
Optimal Assembly of Repurposed Lithium-Ion Battery Packs under Cell Heterogeneity and Screening Uncertainty
arXiv:2607.12951v1 Announce Type: new Abstract: The growing supply of retired electric vehicle batteries presents an opportunity for second-life stationary energy storage, but assembling heterogeneous retired cells into reliable packs is challenging due to substantial variation in capacity, DC internal resistance (DCIR), and self-discharge. This paper proposes a robust optimization framework for cell-to-pack assembly of second-life batteries. A topology-screening stage first identifies minimum-cell series-parallel configurations satisfying inverter and energy requirements, reducing the dimensionality of the subsequent assignment problem. For each candidate topology, a mixed-integer linear program selects cells and assigns them along the series string, enforcing power, voltage, and energy requirements as hard constraints while minimizing a normalized, weighted sum of DCIR spread, capacity spread, and self-discharge imbalance. Additionally, measurement uncertainty in capacity and DCIR is modeled as bounded intervals to guarantee feasibility under worst-case parameter deviations. The framework is evaluated on four heterogeneous inventories for a 10 kW/10 kWh stationary backup application. The proposed method satisfies all feasibility requirements in every case, while single-metric sorting heuristics each fail on at least one inventory. Relative to the best single-metric baseline by objective value, it reduces the normalized mismatch objective by 76-87%, demonstrating that jointly optimizing cell matching with application-level feasibility requirements improves heterogeneous second-life pack assembly under screening uncertainty.
Traceable In Situ Microwave Power Measurement at the Cryogenic Device Plane in a Dilution Refrigerator
arXiv:2607.12751v1 Announce Type: new Abstract: Accurate knowledge of the microwave power delivered to a cryogenic device under test (DUT) is essential for the characterization and operation of superconducting quantum circuits. However, this information is difficult to obtain inside dilution refrigerators because of distributed attenuation, impedance mismatch, switch-path repeatability, and temperature-dependent microwave components. This paper presents an in situ measurement method for RF power at the cryogenic device plane. The method uses a custom variable temperature stage (VTS) as a cryogenic thermal-transfer element. The TVS is alternately heated by a four-wire DC heater and by microwave power dissipated in a 20 dB pass-through attenuator. By fitting the thermal transients and comparing the corresponding steady-state temperatures, the absorbed microwave power is inferred from a directly measured DC electrical power through an AC/DC substitution procedure. The finite reflection and transmission of the attenuator are then accounted for by cryogenic two-port scattering-parameter measurements based on a switch-assisted Short--Open--Load--Reciprocal calibration, so that the result is referred to the DUT reference plane. The system is demonstrated in a dilution refrigerator with powers between -43 and -58 dBm at the DUT input plane. The demonstrated relative standard uncertainty ranges from about 2% at -43.9 dBm to about 40% at -57.6 dBm. The proposed approach combines thermal RF power transfer, cryogenic S-parameter correction, and uncertainty evaluation in a measurement architecture compatible with quantum-device experiments, providing a practical route toward traceable microwave-power calibration at millikelvin stages.
RFMSR: Residual Flow Matching for Image Super-Resolution
arXiv:2607.12753v1 Announce Type: new Abstract: Image super-resolution (ISR) has witnessed remarkable progress with diffusion models and flow matching. The dominant text-to-image (T2I) based approaches leverage large-scale foundation models as generative priors, achieving impressive perceptual quality but at the cost of massive model sizes and prohibitive training expenses. Recent flow-matching-based vision-only approaches have made significant strides; however, they adopt standard flow formulations that transport from a pure Gaussian prior to the data distribution, discarding the rich structural information already present in the low-quality (LQ) input. Furthermore, existing single-step acceleration techniques often forfeit the model's multi-step inference capability. In this paper, we propose Residual Flow Matching for Image Super-Resolution (RFMSR), a vision-only framework that centers the source distribution at the LQ latent, reducing transport distance and preserving structural priors throughout the flow trajectory. We further introduce a two-phase training strategy: Phase I pretrains the velocity field via conditional flow matching, while Phase II applies end-to-end supervision to the single-step prediction while retaining the velocity loss across all timesteps, achieving high-quality single-step generation without sacrificing multi-step refinement. Extensive experiments demonstrate that RFMSR achieves comparable or even superior perceptual quality compared to state-of-the-art (SOTA) methods. The source code is available at https://github.com/Faze-Hsw/RFMSR.
Interpretable and Verifiable Hardware Generation with LLM-Driven Stepwise Refinement
arXiv:2606.19387v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved remarkable success in software development. However, they are susceptible to hallucinations, meaning that they can introduce subtle semantic and logical errors. Due to the high stakes in chip design and manufacturing, hardware engineers are still reluctant to rely on LLMs for register-transfer level (RTL) generation. In this paper, we propose a hardware generation framework that combines the creativity and broad knowledge of LLMs with the explainability and mathematical rigor of formal methods. Specifically, we devise a set of transformation rules that cover various design decisions and hardware features. By iteratively applying these rules, an LLM agent can convert a design specification into an RTL program with guaranteed correctness. Experimental results demonstrate the effectiveness and efficiency of the framework.
Single-Shot High-Energy Muon and Particle Radiography with a Multi-GeV Laser-Wakefield-Accelerator-Driven Source
arXiv:2607.12984v1 Announce Type: new Abstract: We report the first demonstration of single-shot particle radiography using a 1-10 GeV laser-wakefield-generated beam of muons, pions, and neutrons. The test objects were imaged ~15 m from the beam source, through dense lead shielding followed by the walls of a building and a truck. The muon content of the beam was directly confirmed using large volume scintillator-based detectors, which recorded particle decay events with timing delays consistent with the muon lifetime. Simulations confirm that the high energy component of the beam transmitted through the test object is nearly entirely composed of muons, directly showing their highly penetrative nature, with a single-shot fluence equivalent to >8 hours of integration of cosmic ray muons near the horizon. Our work establishes single-shot high-energy particle radiography with a laser-wakefield-accelerator-driven source.
ChunkFlow: Towards Continuity-Consistent Chunked Policy Learning
arXiv:2607.12992v1 Announce Type: new Abstract: Vision-language action (VLA) models increasingly adopt chunked action heads to satisfy real-time constraints; however, this introduces boundary jitter: overlapping regions between consecutive chunks often yield inconsistent predictions, degrading temporal coherence and the task success rate. Existing methods, such as inference-time blending, merely reweight mismatched proposals without correcting underlying errors, leading to residual accumulation under biased or noisy histories. We propose ChunkFlow, a seam-aware training-and-execution framework for chunked policies that aligns chunk structure with boundary execution. It partitions each chunk into frozen, editable, and future zones, applies deterministic overlap blending at execution, and trains raw predictions with seam and first- and second-order continuity losses. History corruption and scheduled sampling improve robustness to executed-history errors, while an AWAC fine-tuning stage adapts the policy without removing these structural regularizers. Under mild smoothness assumptions, pre-blending seam discrepancies provably decay with increasing overlap. Experiments on CALVIN, LIBERO, and real robots show an improved success-stability trade-off with low-latency inference. Project page: https://cytoderm-ai.github.io/chunkflow.
Real-time fall detection based on vision for low-power edge platforms
arXiv:2607.12909v1 Announce Type: cross Abstract: Falling detection is vital for elderly care and intelligent surveillance; however, prevailing vision-based approaches predominantly frame it as static pose classification or discrete temporal pattern matching, fundamentally overlooking the instability dynamics of the human support system. This paper proposes a physics-informed falling detection framework that recasts falling as a stability-loss event in a coupled dynamical system. We introduce a novel dual-LTC architecture comprising a Center-of-Mass (CoM) subsystem and a Base-of-Support (BoS) subsystem, both instantiated as Liquid Time-Constant (LTC) neural networks to continuously model inertial trajectory evolution and ground-contact adjustment through adaptive time constants, Physical interpretability of falling motion. A learnable coupling module emulates physical interaction between the two subsystems, while a Stability Manifold classifier operates in the joint latent space to detect boundary crossing via Lyapunov-inspired stability metrics. Complementary counterfactual trajectory projection and Time-to-Collision (TTC) estimation further enable irreversibility assessment and early warning. The architecture is designed to support a three-state prediction paradigm (Normal, Falling, Fallen); in this preliminary study, we validate the core stability discrimination capability on a two-class dataset (Normal vs. Falling), leaving the full three-state temporal transition to future work. Unlike conventional CNN--RNN pipelines, the proposed formulation encodes continuous-time mechanical inertia, yielding a sub-50K-parameter network capable of real-time inference on resource-constrained edge devices. Extensive experiments demonstrate competitive accuracy with superior physical interpretability, validating its efficacy for low-compute visual fall detection.
Conditional Enhancement of Radical Pair Dynamics via Chiral State Preparation
arXiv:2605.22130v2 Announce Type: replace Abstract: Chiral-induced spin selectivity (CISS) has been shown to enhance magnetic sensitivity in radical pair mechanism (RPM) models under specific Hamiltonian conditions, yet whether these enhancements persist across a broader parameter space remains untested. We incorporate the CISS effect as a spin-dependent initial state and recombination operator and systematically evaluate the spin dynamics of a model radical pair across a comprehensive parameter sweep of the RPM Hamiltonian. We characterise the orientational response through symmetric and antisymmetric decomposition of the yield distribution under field reversal, providing a direct quantitative signature of CISS-induced symmetry breaking. Our analysis demonstrates that CISS does not function as a generic amplifier of magnetic sensitivity. Claimed enhancements are conditional on the relative alignment of the internal hyperfine and dipolar interaction axes, arising specifically under conditions of non-collinear internal interactions. Extension to a two-nucleus model confirms that these enhancements are sensitive to nuclear spin. CISS-induced effects observed in the single-nucleus model are substantially suppressed when a second collinear nucleus is introduced, with the exception of the hyperfine axis rotation sweep where non-collinear tensor misalignment drives a robust antisymmetric response. These findings indicate that the conditions for CISS-enhanced magnetoreception are more stringent than previously demonstrated, requiring highly ordered and rigid molecular geometries to sustain the effect.
Learning Latent Energy-Based Models via Interacting Particle Langevin Dynamics
arXiv:2510.12311v2 Announce Type: replace-cross Abstract: We develop interacting particle algorithms for learning latent variable models with energy-based priors. To do so, we leverage recent developments in particle-based methods for solving maximum marginal likelihood estimation (MMLE) problems. Specifically, we provide a continuous-time framework for learning latent energy-based models, by defining stochastic differential equations (SDEs) that provably solve the MMLE problem. We obtain a practical algorithm as a discretisation of these SDEs and provide theoretical guarantees for the convergence of the proposed algorithm. Finally, we empirically validate the effectiveness of our method on synthetic and image datasets and demonstrate that using a particle based approach offers significant improvement in computational efficiency.
PRISM Edit: One Vector for All Temporal Answers
arXiv:2607.11327v2 Announce Type: replace Abstract: Model editing keeps large language models (LLMs) up to date without retraining, but temporal facts expose a limitation of the prevailing locate-and-edit paradigm: an update is not always a replacement. When a fact changes, the new answer should become current while the old answer may remain correct in historical time contexts. Building on this insight, we use causal tracing to show that LLMs already support this distinction via a two-stage internal computation: early MLP layers retrieve a time-agnostic subject representation, and later layers modulate it with temporal context to yield the time-correct answer. Motivated by this finding, we introduce PRISM Edit, which optimizes a single polysemous representation across temporal contexts and leverages the model's inherent modulation pathway to route it to temporally correct predictions, without any architectural modification. We evaluate on TimeConflict, a new temporal editing benchmark we introduce, and on temporally augmented CounterFact. PRISM Edit improves over the best baseline by +23.3 Temporal Consistency (TC) and +33.7 Current Relative-time Score (CRS) on average while being more than 2x faster. Code and data are publicly available at https://github.com/AnonymousStudy972/PRISM-Edit.
Attention-Discounted Adaptive Sampler for Masked Diffusion Language Models
arXiv:2606.10829v2 Announce Type: replace Abstract: Masked diffusion language models can reduce inference steps by revealing multiple tokens per denoising iteration, but this parallelism is fragile: positions that are individually confident may be unsafe to commit together when their predictions are coupled. Existing training-free samplers such as Top-\(k\), Fast-dLLM, and EB-Sampler mainly control how many tokens to reveal, while often ranking candidates by token-wise scores that ignore interactions within the selected set. We propose ADAS, a training-free reranking rule for parallel masked diffusion decoding. ADAS leaves the base sampler's stopping rule unchanged and modifies only subset construction: it greedily discounts a candidate when it attends strongly to already selected positions whose predictions remain uncertain. Unlike graph-constrained methods that turn attention into hard compatibility constraints, ADAS keeps attention continuous and uses it as a soft marginal penalty. Across LLaDA-8B-Base and Dream-7B-Base on GSM8K, MATH500, HumanEval, and MBPP, plugging ADAS into Top-\(k\), Fast-dLLM, and EB-Sampler improves low-NFE performance at matched denoiser evaluations by \(9.11\) and \(10.46\) percentage points on average, respectively, with \(3.1\%\) per-forward runtime overhead. These results show that soft attention-discounted reranking is a simple and modular way to improve quality in highly parallel decoding for masked diffusion language models.
Double negation stable h-propositions in cubical sets
arXiv:2209.15035v2 Announce Type: replace-cross Abstract: We give a construction of classifiers for double negation stable h-propositions in a variety of cubical set models of homotopy type theory and cubical type theory. This is used to give some relative consistency results: classifiers for double negation stable propositions exist in cubical sets whenever they exist in the metatheory; the Dedekind real numbers can be added to homotopy type theory without changing the consistency strength; we construct a model of homotopy type theory with extended Church's thesis, which states that all partial functions with double negation stable domain are computable.
Imputation-free transformer learning enables robust Alzheimer's disease prediction and calibrated uncertainty quantification across heterogeneous clinical cohorts
arXiv:2607.11656v2 Announce Type: replace-cross Abstract: Accurate diagnostic classification and disease-severity prediction for Alzheimer's disease are hampered by the incompleteness and heterogeneity of real-world clinical data. Left unaddressed, these barriers prevent reliable disease modelling and hinder effective clinical evaluation. Conventional imputation strategies introduce systematic bias, distort inter-feature relationships, and yield overconfident predictions, limitations especially consequential in diagnostic settings. Here, we propose NITROGEN, an imputation-free transformer that jointly models within-patient feature dependencies and between-patient relational structure through masked and intersample attention, enabling robust multimodal learning directly from partially observed records. We trained NITROGEN on ADNI (N=7858 scans), and evaluated it on two independent cohorts: OASIS-3 (N=2675 scans) and AIBL (N=1286 scans). Across cohorts and diagnostic and cognitive score prediction tasks, NITROGEN showed robust calibration and uncertainty quantification advantages over tree-based ensemble methods, while maintaining competitive discriminative performance. Cross-cohort and cross-method analyses identified cortical thickness in the temporal pole, age, and APOE genotype as important, though not individually sufficient, features for AD classification. We further introduced a modality-aware uncertainty adjustment that augments predictive uncertainty proportionally to the importance of absent modalities, enabling calibrated confidence when diagnostic information is unavailable. Together, our results show that imputation-free attention learning preserved meaningful discrimination under cohort shift, revealing expected degradation on more distributionally different cohorts, and demonstrate that evaluating models along calibration, interpretability, and cross-cohort reliability, not accuracy alone, is essential for clinical deployment.
Intermodal entanglement in a quantum optical model of HHG due to the back-action on the driving field
arXiv:2603.01315v3 Announce Type: replace-cross Abstract: Preparation of nonclassical light with special quantum properties is essential for quantum technologies. High-harmonic generation (HHG) is a process which not only enables the creation of attosecond pulses but also has the potential to generate light with intricate quantum properties. In a recent experiment [PRX Quantum 5. 040319], nonclassical inter-harmonic correlations have been measured from a HHG source between low-order harmonics. In this work, we theoretically investigate entanglement between different harmonics within an effective, phenomenological quantum optical model. This model implements a significant degree of simplification regarding the processes within the target material, treating the material through susceptibilities, as it is usual in quantum optics. Such an approach yields a general description of HHG in the few-harmonic generation regime, permitting the implications that can be derived within it to hold broadly within the domain of validity. We find that entanglement is produced as a result of the often neglected back-action. We can qualitatively reproduce experimentally measured nonclassicalities, which suggests that intermodal entanglement can, to an extent, be considered a universal phenomenon associated with HHG, rather than a result of using specific material targets.
Demonstrating TOFFEE: A Learned System for Synthesizing Data Agent Trajectories at Scale
arXiv:2607.06233v2 Announce Type: replace Abstract: LLM-powered data agents are playing an increasingly important role in data-driven decision making. However, existing data agents struggle to generalize to unseen data environments and analytical workflows, especially in heterogeneous enterprise settings. This creates a growing need for synthesizing high-quality data agent trajectories that capture complex analytical workflows for given data environments. Such trajectories support two key downstream uses: they can serve as supervised finetuning (SFT) data that adapts data agent models to the target domain, and as in-context learning (ICL) demonstrations to guide general-purpose LLMs in unfamiliar data environments. Thus, we introduce TOFFEE, a system for synthesizing high-quality data agent trajectories from given data environments via Monte Carlo Tree Search (MCTS) with adaptive model selection and cross-task prefix reuse. We show that TOFFEE can effectively generate scalable trajectory data for complex analytical tasks across heterogeneous environments. In this demonstration, we present the system framework of TOFFEE, including its task pool construction, trajectory explorer, and learned cost model. We also introduce the web interface of TOFFEE and its workflow, and demonstrate two end-to-end scenarios: trajectory synthesis for data agent finetuning, and demonstration-augmented data agent reasoning.
Epistemic Stance Flexibility Probing: Measuring Prompt-Conditioned Register Shift in Large Language Models
arXiv:2607.12739v1 Announce Type: new Abstract: A language model may be asked either what experts believe about a contested claim or what it believes about the claim itself. A trustworthy conversational agent should distinguish these two requests and respond in different epistemic registers: neutral attribution in the first case and stance expression in the second. Whether such a shift occurs-and whether it occurs coherently-is not directly assessed by existing benchmarks for accuracy, instruction following, or safety. We introduce ESFP, a behavioral benchmark that treats the contrast between externally attributed and self-attributed prompts as the fundamental unit of measurement. ESFP consists of 104 carefully controlled items spanning six epistemic categories and five phrasing templates, and evaluates model responses along four complementary dimensions: lexical self-attribution, representation-level responsiveness to role framing, sentence-level stance content density assessed by an LLM judge panel, and cross-condition stance consistency. Evaluating eight frontier models from five vendors, we find that epistemic flexibility is largely orthogonal to general model capability: a 27B open-weight model matches the strongest proprietary systems, the flagship model of one family underperforms its lightweight counterpart, and reasoning-optimized models do not consistently exhibit higher flexibility. Stance content density provides the strongest signal, while surface-level lexical markers such as 'I think' can change substantially without corresponding changes in expressed stance. We provide item-level bootstrap confidence intervals, weight-sensitivity analyses, and an explicit discussion of the interpretation limits of the composite score. ESFP measures a model's propensity to adapt its epistemic stance under changing attribution conditions, rather than a general competence measure.
Superimposed Transmission for Cooperative Cellular and Cell-Free Massive MIMO Systems
arXiv:2607.12709v1 Announce Type: new Abstract: This paper proposes a superimposed transmission strategy for cooperative cellular and cell-free massive MIMO systems. By classifying users into near and far, the base station transmits an additional data symbol for each near user, superimposed on the signals from distributed access points. Successive interference cancellation is employed at near-user receivers to decode both symbols. The proposed strategy achieves the highest peak spectral efficiency while maintaining fairness at the cell edge, thereby outperforming all the existing network configurations in system capacity.
High-frequency magnetotransport in LaMnO3 samples synthesized by microwave irradiation versus conventional heating
arXiv:2607.12689v1 Announce Type: new Abstract: Magnetoresistance of hole-doped LaMnO3 at frequencies above a few MHz has been seldom reported compared to numerous studies on magnetoresistance measured at direct current. Here, we contrast the high-frequency magnetoresistance in the frequency range f = 0.9-3 GHz in polycrystalline LaMnO3 (LMO) synthesized by microwave irradiation of oxide precursors in a microwave furnace (MW-LMO) and by conventional heating (CH-LMO) in an electrical furnace. The structure at room temperature changed from orthorhombic in CH-LMO to rhombohedral in MW-LMO. A combination of magnetization and resistivity studies suggest that the CH-LMO sample is a canted antiferromagnetic insulator below 140 K but the MW-LMO is a ferromagnetic metal with T_C = 240 K. While the high-frequency resistance of the CH-LMO sample at 300 K decreases gradually with increasing strength of the applied dc magnetic field for all f, a peak appears at H = H_r larger than 0 gauss when f equal or greater than 1.4 GHz in the MW-LMO sample and H_r increases linearly with f. We attribute the observed features in the high-frequency magnetoresistance of MW-LMO to current-driven resonant excitation of spins in the paramagnetic state. Its absence in the CH-LMO sample highlights the need for an optimum hole density to observe this effect.