Forskningsradar

Science Journals

Peer-reviewade publikationer — 53080 artiklar

Ciphertext- and Polynomial-Level Optimization for Fully Homomorphic Encryption
arXiv:2607.15750v2 Announce Type: replace Abstract: Fully homomorphic encryption (FHE) schemes such as RNS-CKKS enable privacy-preserving services through direct computation on encrypted data. While recent FHE compilers optimize FHE programs, they operate at the coarse-grained ciphertext level, where each ciphertext operation comprises a sequence of polynomial operations. At this granularity, the compilers miss polynomial-level optimization opportunities across ciphertext operations. This work presents Recifhe, a new multi-level compiler that supports both ciphertext-level and polynomial-level optimization. At the ciphertext level, Recifhe transforms a non-FHE input program into an FHE program by inserting ciphertext management operations and applies global optimizations. At the polynomial level, Recifhe eliminates redundant polynomial computations across ciphertext operations. Recifhe achieves a 1.25x speedup over ciphertext-level-only optimization.
DECODEM: Data Extraction from Corporate Organizational Documents via Enhanced Methods
arXiv:2607.15879v2 Announce Type: replace Abstract: Much empirical legal research depends on translating unstructured text into structured variables. In corporate governance research as elsewhere, this translation has traditionally relied on human coding of documents such as charters and bylaws, a process that is costly, difficult to scale, and often opaque. This paper introduces DECODEM, a set of benchmark datasets for evaluating the automated extraction of corporate governance variables from organizational documents. The benchmarks pair randomly sampled corporate charters and bylaws with high-quality human annotations covering a range of governance provisions commonly studied in empirical work. Using these datasets, the paper evaluates several large-language-model extraction pipelines that vary in prompt design, task decomposition, and document handling. The underlying task consists of a set of document-level binary classification problems, one for each governance variable. The results show that automated extraction is feasible at a high level of accuracy for many provisions, with median performance near the upper bound across approaches. At the same time, performance varies systematically across variables, with a small number of provisions accounting for most of the remaining errors. More elaborate prompting strategies and cascading pipelines do not consistently improve performance for frontier models, but substantially narrow the gap between frontier and efficiency-oriented models in some settings, suggesting that pipeline design can partly substitute for model capability. By providing a standardized benchmark and a systematic evaluation of extraction methods, the paper demonstrates that current frontier models can extract legally meaningful information from complex corporate documents with high accuracy and suggests an important future role for automated feature extraction in constructing corporate governance datasets.
Induction in Both Directions: A Mechanistic Analysis of In-Context Learning in Masked Diffusion Language Models
arXiv:2607.15893v2 Announce Type: replace Abstract: While the internal mechanisms of autoregressive (AR) transformers have been studied extensively, much less is known about diffusion language models (DLMs), an emerging alternative that generates text by iterative denoising. In this work, we study how DLMs implement induction, a mechanism behind in-context learning in which the model finds a repeated context and copies the token that followed it. Our analysis compares attention-only AR models and absorbing-mask DLMs with matched architectures. We find that DLMs learn a bidirectional induction circuit, where previous-token and next-token heads write local context into the residual stream and later induction heads use it to find and copy the answer from the matching source position. The circuit is direction-symmetric, working whether the source appears in the past or in the future. When only left context is visible, matching what an AR model sees, the DLM does not outperform its AR counterpart in induction capabilities. However, we observe it has stronger induction when both sides of the masked token are visible, pointing to bidirectional context access rather than a stronger one-sided mechanism. Beyond induction, we provide causal evidence that DLMs compute the global fraction of masked tokens and use it as an implicit timestep, even though they are given no explicit timestep embedding.
PIXIE: A Zero-Shot texture-invariant 6D pose estimation framework for unseen objects with assembly defects
arXiv:2607.16015v2 Announce Type: replace Abstract: 6D pose estimation remains a key challenge in robotics and computer vision, particularly in industrial environments. The deployment of currently available data-driven methods is often limited by resource-intensive data pipelines, reliance on textured 3D models, and sensitivity to geometric deviations caused by damages or assembly defects. We present PIXIE, a zero-shot framework that estimates the 6D pose of an object from an RGB image using only an untextured 3D model. Synthetic depth and normal maps are rendered from sampled reference viewpoints and matched to the query image via a pretrained cross-modality feature matcher. Matched keypoints are back-projected to obtain 2D--3D correspondences for PnP-based pose estimation. Relying exclusively on geometry makes the method inherently robust to lighting and texture variation, while correspondence filtering handles geometric deviations between the model and physical object. We evaluate on widely-used public benchmarks, reporting state-of-the-art results on texture-less objects without object-specific training, and introduce a novel dataset with assembly defects, texture variations, and occlusion to demonstrate real-world applicability.
When Does Muon Help Agentic Reinforcement Learning?
arXiv:2607.16169v2 Announce Type: replace Abstract: Muon is competitive with AdamW in large-scale pre-training, but its value for reinforcement-learning (RL) post-training remains unclear. We study vanilla Muon in sparse-reward agentic RL through matched single-seed comparisons with AdamW on ALFWorld using Qwen2.5-0.5B-Instruct. Under Group-in-Group Policy Optimization (GiGPO), applying Muon only to hidden weight matrices raises final-window validation success from 0.290 to 0.546 (+88%); high-rate AdamW controls retain no post-update success. The effect depends on the advantage estimator and learning rate. At 3e-5, Muon improves GRPO from 0.161 to 0.268, whereas GraphGPO's late-window gap narrows near saturation. At 1e-5, GraphGPO Muon reaches 0.901, raises normalized validation AUC from 0.399 to 0.556, and reaches 0.5 and 0.75 success 30 and 60 updates earlier, respectively. These exploratory results show that Muon can benefit agentic RL and motivate studying the policy optimizer, advantage estimator, and learning rate jointly.
Bidders' Responses to Auction Format Change in Internet Display Advertising Auctions
arXiv:2110.13814v4 Announce Type: replace-cross Abstract: We study actual bidding behavior when a new auction format gets introduced into the marketplace. More specifically, we investigate this question using a novel dataset on internet display advertising auctions that exploits a staggered adoption by different publishers (sellers) of first-price auctions (FPAs), instead of the traditional second-price auctions (SPAs). We analyze the auction format change using difference-in-differences regressions and a synthetic difference-in-differences estimator, which better handles pre-trends. The results show that revenue per sold impression (price) jumps considerably for treated publishers relative to control publishers, with increases ranging from 25% to 70% of the pre-treatment price level of the treated group. Moreover, for later auction format changes, the increase in price levels under FPAs relative to those under SPAs tends to dissipate over time, reminiscent of the revenue equivalence theorem, although the extent of this reversion depends on the specification. We view these results as suggestive of initially insufficient bid shading following the format change, as opposed to an immediate transition to a new Bayesian Nash equilibrium, with prices tending to decline in several specifications in a manner consistent with gradual adjustment in bidding behavior as bidders learn to shade their bids. Our work constitutes one of the first field studies on bidders'responses to auction format changes, providing an important complement to theoretical model predictions. As such, it provides valuable information to auction designers when considering the implementation of different formats.
Decoherence control of a single-photon optomechanical system in phase-sensitive reservoirs
arXiv:2111.05554v5 Announce Type: replace-cross Abstract: Recent advancements in strong single-photon optomechanical coupling also demand a deeper understanding of environmental interactions in this regime.In this regime the standard Lindblad master equation, which is derived in the eigenbasis of the bare optical and mechanical modes, misassigns the dephasing rate. We therefore use the Dressed-State Master Equation (DSME), which is formulated in the eigenbasis of the strongly coupled photon-phonon Hamiltonian, the dressed states of the system. This work investigates the impact of squeezed vacuum and thermal reservoirs on the decoherence of cavity photon Fock states in the strong coupling regime. We demonstrate that decoherence can be effectively controlled by tuning reservoir parameters, with the control mediated through a cavity dephasing term that becomes significant at high temperatures. The findings presented provide critical insights into reservoir engineering for precise control of quantum decoherence, advancing the understanding of strongly coupled optomechanical systems in engineered environments.
Nonconvex Matrix Factorization is Geodesically Convex: Global Landscape Analysis for Fixed-rank Matrix Optimization From a Riemannian Perspective
arXiv:2209.15130v3 Announce Type: replace-cross Abstract: We study a general matrix optimization problem with a fixed-rank positive semidefinite (PSD) constraint. We perform the Burer-Monteiro factorization and consider a particular Riemannian quotient geometry in a search space that has a total space equipped with the Euclidean metric. When the original objective f satisfies standard restricted strong convexity and smoothness properties, we characterize the global landscape of the factorized objective under the Riemannian quotient geometry. We show the entire search space can be divided into three regions: (R1) the region near the target parameter of interest, where the factorized objective is geodesically strongly convex and smooth; (R2) the region containing neighborhoods of all strict saddle points; (R3) the remaining regions, where the factorized objective has a large gradient. To our best knowledge, this is the first global landscape analysis of the Burer-Monteiro factorized objective under the Riemannian quotient geometry. Our results provide a fully geometric explanation for the superior performance of vanilla gradient descent under the Burer-Monteiro factorization. When f satisfies a weaker restricted strict convexity property, we show there exists a neighborhood near local minimizers such that the factorized objective is geodesically convex. To prove our results, we provide a comprehensive landscape analysis of a matrix factorization problem with a least squares objective, which serves as a critical bridge. Our conclusions are also based on a result of independent interest stating that the geodesic ball centered at Y with a radius 1/3 of the least singular value of Y is a geodesically convex set under the Riemannian quotient geometry, which as a corollary, also implies a quantitative bound of the convexity radius in the Bures-Wasserstein space. The convexity radius obtained is sharp up to constants.
Vector-Valued Gossip over $w$-Holonomic Networks
arXiv:2311.04455v2 Announce Type: replace-cross Abstract: We study the weighted average consensus problem for a gossip network of agents with vector-valued states. For a given matrix-weighted graph, the gossip process is described by a sequence of pairs of adjacent agents communicating and updating their states based on the edge matrix weight. Our key contribution is providing conditions for the convergence of this non-homogeneous Markov process as well as the characterization of its limit set. To this end, we introduce the notion of "$w$-holonomy" of a set of stochastic matrices, which enables the characterization of sequences of gossiping pairs resulting in reaching a desired consensus in a decentralized manner. Stated otherwise, our result characterizes the limiting behavior of infinite products of (non-commuting, possibly with absorbing states) stochastic matrices.
A figure-of-merit-based framework to evaluate photovoltaic materials
arXiv:2404.14732v3 Announce Type: replace-cross Abstract: I propose a general quantitative framework to evaluate the quality, track the historical development, and guide future optimization of photovoltaic (PV) absorbers at any development level, both lab-made and computer-simulated. The framework is centered around a PV figure of merit designed to include efficiency limitations that are not captured by classic detailed balance methods derived from the Shockley-Queisser limit. A more stringent set of figure-of-merit-driven efficiency limits are calculated for 28 experimentally synthesized PV absorbers and 10 PV computationally modeled absorbers. Among early-stage absorbers, this analysis reveals very large differences in their likelihood of achieving high PV efficiencies in the future. Since the proposed figure of merit is instantly evaluated from a single equation, it can be a suitable objective function for closed-loop research on PV materials in autonomous labs, while also providing a quantitative bridge between computationally determined material properties and PV efficiency.
Optimizing alphabet reduction pairs of arrays
arXiv:2406.10930v2 Announce Type: replace-cross Abstract: In our earlier paper, "2 CSPs all are approximable within a constant differential factor" (ISCO 2018, LNCS 10856), we introduced a family of combinatorial designs called 'alphabet reduction pairs of arrays' (ARPAs). These designs are parameterized by three integers $q,p,k$, with $p\leq q$ and $k\leq p$: $q$ is the size of the alphabet from which the arrays draw their entries; $p$ is the maximum number of distinct symbols allowed in a row of the second array; $k$ is the largest integer for which the two arrays coincide -- up to row permutations -- on any $k$-element subset of their columns. The first array must contain at least one occurrence of the word $0\ 1 \cdots\ q-1$ as a row. The idea is to cover as many occurrences of this word as possible using as few words as possible, each containing at most $p$ distinct symbols. ARPAs are related to the approximability of constraint satisfaction problems with bounded constraint arity ($k$-CSPs). In this context, we are particularly interested in ARPAs that maximize the frequency of the word $0\ 1 \cdots\ q-1$. We call such ARPAs 'optimal' and study them in this paper. To this end, we introduce a simpler family of combinatorial designs called 'Cover pairs of arrays' (CPAs), which can be viewed as partially defined ARPAs with Boolean entries. We prove that ARPAs and CPAs are equivalent with respect to maximizing the frequency of their target word. As a corollary of our proof, computing the frequency of the target word in optimal ARPAs reduces to solving a linear program in $q + p + 1$ continuous variables and $k + 1$ constraints. We also prove the optimality of previously known ARPAs for $p=k$ and provide optimal ARPAs for $k=1$ and $k=2$.
Gradient Span Algorithms Make Predictable Progress in High Dimension
arXiv:2410.09973v2 Announce Type: replace-cross Abstract: We prove that all 'gradient span algorithms' have asymptotically deterministic behavior on scaled Gaussian random functions as the dimension tends to infinity. This is a functional generalization of similar results for random quadratic functions and spin glasses. They explain the counterintuitive phenomenon that different training runs of many large machine learning models result in approximately equal cost curves despite random initialization on a complicated non-convex landscape. This 'predictable progress' phenomenon is exploited by the AutoML community: Since the optimization progress of a single run is already representative, multiple retries with the same hyperparameters are not necessary.
Adaptive High-Level Tight Control of Prostate Cancer: A Path from From Terminal Disease to Chronic Condition
arXiv:2410.16005v4 Announce Type: replace-cross Abstract: Metastatic prostate cancer is one of the leading causes of cancer-related morbidity and mortality worldwide. It is characterized by a high mortality rate and a poor prognosis. In this work, we explore how a clinical oncologist can apply a Stackelberg game-theoretic framework to prolong metastatic prostate cancer survival, or even make it chronic in duration. We utilize a Bayesian optimization approach to identify the optimal adaptive chemotherapeutic treatment policy for a single drug (Abiraterone) to maximize the time before the patient begins to show symptoms. We show that, with precise adaptive optimization of drug delivery, it is possible to significantly prolong the cancer suppression period, potentially converting metastatic prostate cancer from a terminal disease to a chronic disease for most patients, as supported by clinical and analytical evidence. We suggest that clinicians might explore the possibility of implementing a high-level tight control (HLTC) treatment, in which the trigger signals (i.e. biomarker levels) for drug administration and cessation are both high and close together, typically yield the best outcomes, as demonstrated through both computation and theoretical analysis. This simple insight could serve as a valuable guide for improving current adaptive chemotherapy treatments in other hormone-sensitive cancers.
The Illusion-Illusion: Vision Language Models See Illusions Where There Are None
arXiv:2412.18613v2 Announce Type: replace-cross Abstract: Illusions are entertaining, but they are also a useful diagnostic tool in cognitive science, philosophy, and neuroscience. A typical illusion shows a gap between how something `really is' and how something `appears to be', and this gap helps us understand the mental processing that led to how something appears to be. Illusions are also useful for investigating artificial systems, and much research has examined whether computational models of perception fall prey to the same illusions as people. Here, I invert the standard use of perceptual illusions to examine basic processing errors in current vision language models. I present these models with illusory-illusions, neighbors of common illusions that should not elicit processing errors. These include such things as perfectly reasonable ducks, crooked lines that truly are crooked, circles that seem to have different sizes because they are, in fact, of different sizes, and so on. I show that many current vision language systems mistakenly see these illusion-illusions as illusions. I suggest that such failures are part of broader failures already discussed in the literature.
Impossibility of Quantum Private Queries
arXiv:2501.12842v5 Announce Type: replace-cross Abstract: Symmetric private information retrieval is a cryptographic task allowing a user to query a database and obtain exactly one entry without revealing to the owner of the database which element was accessed. The task is a variant of general two-party protocols called one-sided secure function evaluation and is closely related to oblivious transfer. Under the name quantum private queries, quantum protocols have been proposed to solve this problem in a cheat-sensitive way: In such protocols, it is not impossible for dishonest participants to cheat, but they risk detection [V. Giovannetti, S. Lloyd, and L. Maccone, Phys. Rev. Lett. 100, 230502 (2008)]. We give an explicit attack against any cheat-sensitive symmetric private information retrieval protocol, showing that any protocol that is secure for the user cannot have non-trivial security guarantees for the owner of the database.
Electron dynamics induced by quantum cat-state light
arXiv:2501.16801v2 Announce Type: replace-cross Abstract: We present an effective theory for describing electron dynamics driven by an optical external field in a Schr\"{o}dinger's cat state. We show that the reduced electron density matrix evolves as an average over trajectories $\{\rho_\alpha\}$ weighted by the Sudarshan--Glauber $P$ distribution $P(\alpha)$ in the weak light--matter coupling regime. Each trajectory obeys an equation of motion, $\mathrm{i} \partial_t\rho_\alpha=\mathcal{H}_{\alpha} \rho_\alpha-\rho_\alpha\mathcal{H}_{\alpha}$, where an effective Hamiltonian $\mathcal{H}_{\alpha}$ becomes non-Hermitian due to quantum interference of light. The optical quantum interference is transferred to electrons through the asymmetric action between the ket and bra state vectors in $\rho_{\alpha}$. This non-Hermitian dynamics differs from the conventional one observed in open quantum systems, described by $\mathrm{i} \partial_t\rho=\mathcal{H}\rho-\rho \mathcal{H}^\dagger$, which has complex conjugation in the second term. We confirm that the reduced, trajectory-resolved effective theory agrees with full electron-photon simulations for the few-electron Dicke model, thereby validating the interferential non-Hermitian description in the weak-coupling regime.
Aspects of Spatially-Correlated Random Fields: Extreme-Value Statistics and Clustering Properties
arXiv:2501.17936v2 Announce Type: replace-cross Abstract: Rare events of large-scale spatially-correlated exponential random fields are studied. The influence of spatial correlations on clustering and non-sphericity is investigated. The size of the performed simulations permits to study beyond-$7.5$-sigma events (one in $10^{13}$). As an application, this allows to resolve individual Hubble patches which fulfil the condition for primordial black hole formation. It is argued that their mass spectrum is drastically altered due to co-collapse of clustered overdensities as well as the mutual threshold-lowering through the latter. Furthermore, the corresponding non-sphericities may imply possibly large changes in the initial black hole spin distribution.
On Erlang ODE approximations of differential equations with distributed time delays
arXiv:2502.12984v5 Announce Type: replace-cross Abstract: In this paper, we propose a general approach for approximate simulation and analysis of delay differential equations (DDEs) with distributed time delays based on methods for ordinary differential equations (ODEs). The key innovation is that we 1) propose an Erlang mixture approximation of the kernel in the DDEs and 2) use the linear chain trick to transform the resulting approximate DDEs to ODEs. We refer to this as the Erlang ODE approximation of the DDEs, and we prove that the Erlang mixture approximation converges for continuous and bounded kernels if the number of terms increases sufficiently fast. Furthermore, we show that if the kernel is also exponentially bounded, the Erlang ODE approximation can be used to assess the stability of the steady states of the original DDEs and that the solution to the ODE approximation converges. Additionally, we propose an approach based on bisection and least-squares estimation for determining optimal parameter values in the approximation. Finally, we present numerical examples that demonstrate the accuracy and convergence rates of the approximations and the efficacy of the proposed approach for bifurcation analysis and Monte Carlo simulation. The numerical examples involve a modified logistic equation, chemotherapy-induced myelosuppression, and a point reactor kinetics model of a molten salt nuclear fission reactor.
CorStitch: Accessible Video Stitching for Coral Monitoring
arXiv:2505.00462v2 Announce Type: replace-cross Abstract: We develop CorStitch, a video-stitching software that automates the process of stitching underwater videos for coral reef monitoring and assessment. It offers a free, flexible, and user-friendly alternative to existing stitching software that may be difficult to access. CorStitch utilizes a Fourier-based image registration algorithm to stitch the central horizontal strips of successive frames of down-looking belt and dive transect videos, generating georeferenced and marked mosaics. Tests show that CorStitch can produce high-quality mosaics that are comparable to those generated by existing stitching software, with the added advantage of being open-source and accessible to users with varying levels of technical expertise. CorStitch has the potential to supplement the coral reef monitoring efforts of local communities, government agencies, and non-governmental organizations.
Optimal $\mathbb{H}_2$ Control with Passivity-Constrained Feedback: Convex Approach
arXiv:2505.10811v2 Announce Type: replace-cross Abstract: We consider the $\Set{H}_2$-optimal feedback control problem, for the case in which the plant is passive with bounded $\Set{L}_2$ gain, and the feedback law is constrained to be output-strictly passive. We show that this problem distills to a convex, infinite-dimensional optimal control problem, in which the optimization domain is the Youla parameter for the closed-loop system. We devise truncated, finite-dimensional optimizations to find sub-optimal controllers, and lower bounds on the optimal objective. Furthermore we show that both these optimizations converge to the optimal objective of the original infinite-dimensional problem as their respective domains are increased. The idea is demonstrated on a simple vibration suppression example.
Spin-current correlations in photoionization of chiral molecules
arXiv:2505.23460v5 Announce Type: replace-cross Abstract: Chirality-induced spin selectivity (CISS) refers to phenomena where molecular chirality governs spin polarization. While symmetry simply requires chiral molecules to support spin-vector correlations, we show that CISS is fundamentally a conditioned measurement of these correlations. We illustrate this principle for spin-resolved one-photon ionization of a randomly oriented ensemble of chiral molecules. We introduce and quantify the phenomenon of enantio-sensitive locking of the photoelectron current to its spin, thereby providing a complete description of spin-conditioned photoelectron currents in one-photon ionization.
Learning MMSE Filters for OFDM Channel Estimation: Attention Transformer Gains at Linear Inference
arXiv:2506.00452v5 Announce Type: replace-cross Abstract: In orthogonal frequency division multiplexing (OFDM), accurate channel estimation is crucial. Classical signal processing-based approaches, such as linear minimum mean-squared error (LMMSE) estimation, often require second-order statistics that are difficult to obtain in practice. Recent deep neural network (DNN)-based methods have been introduced to address this, but they often suffer from high inference complexity. This paper proposes an Attention-aided MMSE (A-MMSE), a model-based DNN framework that learns the linear MMSE filter via the Attention Transformer. Once trained, the A-MMSE performs channel estimation through a single linear operation, eliminating nonlinear activations during inference and thus reducing computational complexity. To improve the learning efficiency of the A-MMSE, we develop a two-stage Attention encoder that captures the frequency and temporal correlation structure of OFDM channels. We also introduce a rank-adaptive extension that adjusts the filter rank at deployment time, enabling efficient operation under resource-constrained receivers. Numerical simulations show that A-MMSE consistently outperforms baseline methods across a wide range of signal-to-noise ratio (SNR) conditions. In particular, the A-MMSE and its rank-adaptive extension provide an improved performance-complexity trade-off.
Kernels of trace operators via fine continuity
arXiv:2507.04536v2 Announce Type: replace-cross Abstract: Given a closed subset $\Gamma$ of $\mathbb{R}^n$ that is the support of a measure $\mu$, we study the kernels of trace operators from fractional Sobolev spaces $H_p^\alpha(\mathbb{R}^n)$ into the space of $\mu$-equivalence classes of functions on $\Gamma$. We characterise these kernels as the closure of $C_c^\infty(\mathbb{R}^n\setminus \Gamma)$ in $H_p^\alpha(\mathbb{R}^n)$, provided quasi continuous representatives of elements of $H_p^\alpha(\mathbb{R}^n)$ have the following key property: they vanish quasi everywhere on $\Gamma$ if and only if they vanish $\mu$-almost everywhere on $\Gamma$. We establish that this key property holds if the measures satisfy localized upper density conditions. Such measures need not be doubling, in particular the set $\Gamma$ may be a finite union of closed sets having different Hausdorff dimensions. We provide corresponding results for spaces $H_p^\alpha(\Omega)$ on domains $\Omega\subset \mathbb{R}^n$ satisfying a weakened version of the measure density condition. We observe that the above key property is essential for the convergence of Galerkin integral equation methods, based on integration with respect to the measure $\mu$, for certain BVPs in the complement of $\Gamma$.
Sequential Attention-based Sampling for Histopathological Analysis
arXiv:2507.05077v5 Announce Type: replace-cross Abstract: Deep neural networks are increasingly applied in automated histopathology. Yet, whole-slide images (WSIs) are often acquired at gigapixel sizes, rendering them computationally infeasible to analyze entirely at high resolution. Diagnostic labels are largely available only at the slide-level, because expert annotation of images at a finer (patch) level is both laborious and expensive. Moreover, regions with diagnostic information typically occupy only a small fraction of the WSI, making it inefficient to examine the entire slide at full resolution. Here, we propose SASHA -- Sequential Attention-based Sampling for Histopathological Analysis -- a deep reinforcement learning approach for efficient analysis of histopathological images. First, SASHA learns informative features with a lightweight hierarchical, attention-based multiple instance learning (MIL) model. Second, SASHA samples intelligently and zooms selectively into a small fraction (10-20\%) of high-resolution patches to achieve reliable diagnoses. We show that SASHA matches state-of-the-art methods that analyze the WSI fully at high resolution, albeit at a fraction of their computational and memory costs. In addition, it significantly outperforms competing, sparse sampling methods. We propose SASHA as an intelligent sampling model for medical imaging challenges that involve automated diagnosis with exceptionally large images containing sparsely informative features. Model implementation is available at: https://github.com/coglabiisc/SASHA.
Dynamic Output-Feedback Controller Synthesis for Dissipativity and $H_2$ Performance from Noisy Input-State Data
arXiv:2507.06788v3 Announce Type: replace-cross Abstract: In this paper we propose dynamic output-feedback controller synthesis methods for discrete-time linear time-invariant systems. The synthesis goal is to achieve dissipativity with respect to a given quadratic supply rate or a given $H_2$ performance level. It is assumed that the model of system dynamics is unknown, expect for the disturbance term. Instead, we have a recorded trajectory of the control input and the state, which can be corrupted by an unknown but bounded disturbance. The state data is used only for the purpose of controller synthesis, while the designed controller is output feedback controller, i.e., the full state is not used for control in real time. The presented synthesis method is formulated in terms of linear matrix inequalities parametrized by a scalar variable, while in noiseless case it reduces to linear matrix inequalities. Within the considered setting, the synthesis procedure is non-conservative.