论文广场 - AcademicHub

01.

arXiv (CS.CL) 2026-06-19 DOI: arXiv:2305.14985

IdealGPT: Iteratively Decomposing Vision and Language Reasoning via Large Language Models

作者:

Haoxuan You ↗Rui Sun ↗Zhecan Wang ↗Long Chen ↗Gengyu Wang ↗Hammad A. Ayyubi ↗Kai-Wei Chang ↗Shih-Fu Chang ↗

The field of vision-and-language (VL) understanding has made unprecedented progress with end-to-end large pre-trained VL models (VLMs). However, they still fall short in zero-shot reasoning tasks that require multi-step inferencing. To achieve this goal, previous works resort to a divide-and-conquer pipeline. In this paper, we argue that previous efforts have several inherent shortcomings: 1) They rely on domain-specific sub-question decomposing models. 2) They force models to predict the final answer even if the sub-questions or sub-answers provide insufficient information. We address these limitations via IdealGPT, a framework that iteratively decomposes VL reasoning using large language models (LLMs). Specifically, IdealGPT utilizes an LLM to generate sub-questions, a VLM to provide corresponding sub-answers, and another LLM to reason to achieve the final answer. These three modules perform the divide-and-conquer procedure iteratively until the model is confident about the final answer to the main question. We evaluate IdealGPT on multiple challenging VL reasoning tasks under a zero-shot setting. In particular, our IdealGPT outperforms the best existing GPT-4-like models by an absolute 10% on VCR and 15% on SNLI-VE. Code is available at https://github.com/Hxyou/IdealGPT

阅读与讨论 → 访问原文 →

02.

arXiv (CS.CV) 2026-06-24 DOI: arXiv:2512.24731

EchoFoley: Event-Centric Hierarchical Control for Video Grounded Creative Sound Generation

作者:

Bingxuan Li ↗Yiming Cui ↗Yicheng He ↗Yiwei Wang ↗Shu Zhang ↗Longyin Wen ↗Yulei Niu ↗

Sound effects build an essential layer of multimodal storytelling, shaping the emotional atmosphere and the narrative semantics of videos. Despite recent advancement in video-text-to-audio (VT2A), the current formulation faces three key limitations: First, an imbalance between visual and textual conditioning that leads to visual dominance; Second, the absence of a concrete definition for fine-grained controllable generation; Third, weak instruction understanding and following, as existing datasets rely on brief categorical tags. To address these limitations, we introduce EchoFoley, a new task designed for video-grounded sound generation with both event level local control and hierarchical semantic control. Our symbolic representation for sounding events specifies when, what, and how each sound is produced within a video or instruction, enabling fine-grained controls like sound generation, insertion, and editing. To support this task, we construct EchoFoley-6k, a large-scale, expert-curated benchmark containing over 6,000 video-instruction-annotation triplets. Building upon this foundation, we propose EchoVidia a sounding-event-centric agentic generation framework with slow-fast thinking strategy. Experiments show that EchoVidia surpasses recent VT2A models by 40.7% in controllability and 12.5% in perceptual quality.

阅读与讨论 → 访问原文 →

03.

arXiv (CS.AI) 2026-06-24 DOI: arXiv:2606.24679

FlowPipe: LLM-Enhanced Conditional Generative Flow Networks for Data Preparation Pipeline Construction

作者:

Kunyu Ni ↗Lei Cao ↗Jie He ↗Xiaotong Zhang ↗Jianfeng Jin ↗Junyu Dong ↗Yanwei Yu ↗

arXiv:2606.24679v1 Announce Type: cross Abstract: Data preparation pipelines improve data quality in machine learning by transforming raw tables into learning-ready data through sequential cleaning and feature transformation operators. However, automatically constructing such pipelines is computationally difficult because operator sequences are combinatorial and end-to-end evaluation is expensive. Existing state-of-the-art (SOTA) Multi-DQN methods still face three key limitations: decoupled value estimators weaken long-horizon credit assignment, dataset context is only weakly injected into the policy, and exploration is inefficient in a sparse search space with many invalid states. To address these issues, we propose FlowPipe, a unified framework that formulates pipeline synthesis as conditional probabilistic flow generation over a directed acyclic graph. FlowPipe uses Conditional Generative Flow Networks (C-GFlowNets) with a Trajectory Balance objective to connect terminal validation rewards with early pipeline decisions. It further introduces Deep Semantic Modulation through Feature-wise Linear Modulation (FiLM), allowing LLM-derived logical priors to condition the policy's internal activations according to dataset semantics. In addition, FlowPipe incorporates failure awareness into the flow objective to avoid invalid states and concentrate search on high-potential regions. Experiments on two benchmark suites with 74 real-world datasets show that FlowPipe outperforms SOTA baselines, improving accuracy by 11.96% on average and achieving 12.5x faster training convergence. Source code is available at https://github.com/KunyuNi/FlowPipe.

阅读与讨论 → 访问原文 →

04.

arXiv (CS.AI) 2026-06-17 DOI: arXiv:2606.17074

Surveying GenAI-based Automation in Printed Circuit Board Design and Test

作者:

Sahana Srinivasan ↗Benjamin Turnbull ↗Hammond Pearce ↗

arXiv:2606.17074v1 Announce Type: cross Abstract: Generative artificial intelligence (GenAI) is increasingly used for applications in the hardware and software domains. It purports to reduce the manual effort involved in the development and testing of complex systems before release. Within the hardware space, most tasks have focused on design automation of integrated circuits, particularly with hardware description languages. However, other types of hardware also exist! In this survey, we instead examine how GenAI has been and is being across the printed circuit board (PCB) design life cycle. This includes everything from supply chains, system specification, circuit design, layout and optimisation, validation and test, and PCB assembly and distribution. Through this lens we present a taxonomy of discovered works, categorising them according to their intent and contributions. This survey also identifies key technical challenges that GenAI faces in this space, such as domain-specific data scarcity and limited support for integration with existing PCB tools. Finally, future research directions are discussed: our survey shows that there are many opportunities remaining when considering how GenAI may be integrated into various tasks in PCB design and test.

阅读与讨论 → 访问原文 →

05.

medRxiv (Medicine) 2026-06-25 DOI: HASH:e1a91365b7e19a589a63a1a8b8ad36b9

Clinician contributions to disparities in severity of illness trajectories among mechanically ventilated patients

作者:

Chesley ↗Yakusheva ↗Lu ↗Kohn ↗Belk ↗Scott ↗Halpern ↗Kerlin ↗

Rationale. Racial disparities in outcomes among patients with acute respiratory failure are well-described, but the contributions of clinicians to these disparities have not been evaluated. Objectives. Among mechanically ventilated patients, we evaluated racial disparities in severity of illness trajectories and adapted value-added modeling to quantify nurse and physician relationships with these disparities. Methods. In a retrospective cohort of mechanically ventilated patients across five hospitals between 2018 and 2022, we used generalized estimating equations to model the change in Laboratory-based Acute Physiology Score version 2 (LAPS) from the start to end of intensive care unit admission ({Delta}LAPS). Consistent with value-added modeling, we randomly allocated the cohort into development and testing partitions, and fit separate multiple linear regression models of {Delta}LAPS using concurrent nurse and physician assignments (determined at 4-hour intervals), patient race, and clinician-race interaction terms as fixed effects. Clinician-specific and clinician-race interaction coefficients were extracted to determine race-specific value-add for each clinician. We defined the race-contextual value-add difference (RCVAD) as a clinician-level measurement of the difference in that clinician's value-add between Black and White patients in their care; a positive RCVAD indicates a more favorable severity of illness trajectory for Black relative to White patients and vice versa. Measurement and Main Results. Among 6,555 distinct patients, 7,247 clinical encounters, 405 nurses, and 70 physicians, Black patients accounted for 2,926 (40%) encounters. Overall, Black patients had significantly less improvement in {Delta}LAPS than White patients (difference in LAPS decline = 2.26 [0.23, 4.29], p=0.029). In the development partition, median nurse RCVAD was -0.10 (interquartile range [IQR]: -1.17, 1.14) with 191 (47%) nurses having a positive RCVAD; median physician RCVAD was -0.18 (IQR: -1.34, 0.56) with 29 (41%) having a positive RCVAD. Conclusions. Black mechanically ventilated patients experience less improvement in severity of illness during intensive care unit admission than White patients. While the majority of physicians and nurses were associated with disparities-exacerbating illness trajectories, many other clinicians were associated with disparities-mitigating trajectories. Future work to understand practices associated with disparities-exacerbating and disparities-mitigating care profiles could inform interventions to reduce disparities overall.

阅读与讨论 → 访问原文 →

06.

arXiv (CS.CV) 2026-06-25 DOI: arXiv:2606.25277

An Integrated Hardware-Software Design for Low-Data Spatial Defect Detection in Robotic Visual Inspection with Hybrid Optoelectronic Neural Networks

作者:

Chaoqing Tang ↗Jiaxuan Li ↗Huanze Zhuang ↗Guiyun Tian ↗Chao Wang ↗Yihao Ouyang ↗Wenzhong Liu ↗

To address data overload and inefficient shape-level annotation in robotic visual inspection, this paper proposes a hardware-software integrated optoelectronic architecture. A non-imaging, low-data paradigm is established to minimize annotation dependency. First, a sensor-in-the-loop strategy reconfigures a Digital Micromirror Device (DMD) as a physical optical convolutional layer, enabling photonic-domain feature extraction that unifies sensing hardware and processing software. To suppress data volume at the source, a block-based compressed sensing strategy encodes spatial information into low-dimensional temporal signals, drastically reducing redundancy. Subsequently, to bypass laborious manual defect shape annotation, natural language descriptions guide the network to align with highly generalizable features from Contrastive Language-Image Pre-training (CLIP), steering the attention maps of the optoelectronic neural network toward defect shapes. Furthermore, a Localization Accuracy for Attention (LAA) metric is proposed to quantify shape-level defect localization performance. Experiments on transparent material defect detection validate the system's effectiveness. Parametric analysis reveals how measurement matrices, compression ratios, and block sizes affect accuracy. Results show that, compared to traditional imaging, the proposed architecture maintains equivalent accuracy while reducing data volume by 90% for Vision Transformers and computational workload by 60% for Convolutional Neural Networks. This low-data paradigm offers an efficient solution for industrial automation scenarios involving massive data streams, high acquisition costs, or constrained edge resources.

阅读与讨论 → 访问原文 →

07.

PLOS Computational Biology 2026-06-01 DOI: HASH:f8c540b3d3539b4f79eff973de90a67f

Histology-informed spatial domain identification through multi-view graph convolutional networks

作者:

Huihui Zhang ↗

by Huihui Zhang, Jiaxing Chang, Zirong Li, Yue Sun, Pinli Hu, Haoxiu Wang, Hang Yang, Yonglin Ren, Xingtan Zhang, Zehua Chen, Kok Wai Wong, Haojing Shao Identifying spatial domains is crucial in spatial transcriptomics, yet effectively integrating gene expression, spatial location, and histology remains challenging. We present STESH, a Spatial Transcriptomics clustering method that combines Expression, Spatial information and Histology. STESH extracts histological features using a convolutional neural network and generates expression, histology, spatial, and collaborative convolution modules for a multi-view graph convolutional network with a decoder and attention mechanism. We evaluated STESH on multiple tissue types and technology platforms. STESH consistently outperformed ten state-of-the-art methods, achieving superior clustering accuracy with the highest scores in adjusted Rand index, normalized mutual information, and Fowlkes-Mallows index.

阅读与讨论 → 访问原文 →

08.

bioRxiv (Bioinfo) 2026-06-24 DOI: HASH:ac68f4199cae0f640a7288adc3d4255b

InVitroGap: an open-source tool for automated quantification of wound closure in the in vitro scratch assay

作者:

ARYA ↗R. K ↗Sindhani ↗Dewala ↗S. R ↗Weight ↗C. J ↗Bukavina ↗

Abstract Background and Objective: Scratch assays are widely used to study wound closure in vitro, but quantitative image analysis remains constrained by manual variability, proprietary workflows, and tools requiring programming expertise. We developed InVitroGap, a Python-based application with a browser-accessible interface for automated quantification of scratch assay closure from sequential microscopy images. Methods: RCC-ER and Renca cells were seeded in 96-well ImageLock plates and scratched using a WoundMaker device for uniform linear wounds or a 200 uL pipette tip for crisscross wounds. Phase-contrast time-lapse images acquired at 0, 24, and 48 h with an IncuCyte SX5 system were independently analyzed using IncuCyte 2023A Rev2 and InVitroGap. The InVitroGap pipeline combines Gaussian smoothing, gradient-based texture mapping, adaptive percentile thresholding, and morphological post-processing to quantify wound confluence and relative wound density (RWD). Agreement was evaluated using paired comparisons, Pearson and Spearman correlations, Bland-Altman analysis, and mean absolute error (MAE). Results: InVitroGap measurements closely tracked IncuCyte outputs across both cell lines, with no significant between-method differences (p > 0.05), strong pooled correlations (R square = 0.964 for RWD; R square = 0.983 for wound confluence), and small mean biases (absolute bias [≤] 1.64%). The tool successfully processed crisscross wounds from brightfield image series, and a complete four-timepoint series was analyzed in approximately 10 seconds, with robust performance across distinct cell morphologies and wound geometries. Conclusions: InVitroGap provides a transparent, computationally efficient, and platform-independent alternative for scratch assay analysis, delivering performance comparable to commercial systems while remaining freely accessible at https://invitrogap.vercel.app/.

阅读与讨论 → 访问原文 →

09.

arXiv (quant-ph) 2026-06-16 DOI: arXiv:2606.16623

Fuzzy-processing quantum computation

作者:

Yan-Xiong Du ↗

arXiv:2606.16623v1 Announce Type: new Abstract: Quantum computation has attracted numerous attentions and develops rapidly in the recent decades. To against the decoherence and the control errors upon the qubits, quantum error corrections are adopted. Such approaches require lots of redundant qubits, accurate measurement and timely feedback. Here we investigate a new framework of quantum computation that is associated with fuzzy processing. It will benefit significantly from three aspects: the fuzzy recognition of qubit states reduce the required gate fidelity; the fuzzy encoding encodes the information of the qubits into a distribution of probability, suppressing the fluctuations in the output of long quantum circuits; the fuzzy feedback offers a more efficient way to control the qubits when precision information of quantum states are absent. Furthermore, the fuzzy processing can be integrated into quantum error correction, eliminating the need for immediate correction operations. The proposed scheme will be fairly suitable for the solution of decision problems, which has significant applications in the optimization problems and control problems.

阅读与讨论 → 访问原文 →

10.

arXiv (quant-ph) 2026-06-16 DOI: arXiv:2606.16424

Real-space spectral functions of three-dimensional billion-size topological non-Hermitian matter with tensor networks

作者:

Yitao Sun ↗Jose L. Lado ↗Guangze Chen ↗

arXiv:2606.16424v1 Announce Type: cross Abstract: Non-Hermitian systems host a wide range of unconventional topological phenomena while large-scale simulations in finite three dimensional systems remain challenging because of the rapidly growing number of sites. In particular, higher-order topological corner modes are often studied only in small lattices, where strong finite-size effects can mask their intrinsic behavior. Here, we develop a tensor-network framework that combines quantics tensor cross interpolation with the kernel polynomial method, enabling compact representations of large non-Hermitian tight-binding Hamiltonians and direct calculations of real-space spectral functions for systems exceeding one billion lattice sites. Using this approach, we investigate three-dimensional non-Hermitian higher-order topological insulators with with structured real-space geometries. The unprecedented system size enables direct access to the macroscopic regime and allows corner-mode spectral responses to be resolved in genuinely three-dimensional systems.By tuning the loss strength, we identify distinct in-gap corner modes across weak- and strong-loss regimes.Our results establish tensor-network algorithms as a powerful strategy to perform real-space spectral calculations in exceptionally large non-Hermitian systems.

阅读与讨论 → 访问原文 →

11.

arXiv (CS.CV) 2026-06-16 DOI: arXiv:2606.15837

Learning a Sampling-Free Variational DNN Plugin from Tiny Training Sets to Refine OOD Segmentation With Uncertainty Estimation

作者:

Jimut B. Pal ↗Suyash P. Awate ↗

Deep neural networks (DNNs) frequently fail to generalize to out-of-distribution (OOD) medical images because of variations in scanners and acquisition protocols. Retraining DNN models to address these distribution shifts is often impractical due to the high cost of acquiring and annotating new medical datasets. To address this, we introduce VarDeepPCA, a novel lightweight variational DNN framework designed to restore/refine degraded segmentation maps by leveraging intrinsic geometric priors. Unlike existing approaches that require target-domain data or extensive pre-training, our VarDeepPCA explicitly learns a distribution of valid anatomical geometries using only small in-distribution (ID) datasets. Theoretically, our novel variational learning framework leverages a reinterpretation of the softmax mapping to implicitly perform exact distribution modeling, thereby enabling computationally efficient, sampling-free learning and inference. This also enables VarDeepPCA to provide uncertainty estimates associated with its restored segmentation maps. We empirically validate our framework across 4 distinct clinical applications, using 14 publicly available datasets, involving segmentation of the myocardium, neuroretinal rim, prostate, and fetal head. Comparisons against 15 existing methods demonstrate that VarDeepPCA consistently restores segmentation maps produced by the existing methods on OOD data to (i) significantly improve anatomical plausibility of geometries and clinical utility of the segmentations, and (ii) significantly reduce errors, without needing any more training data than that used by existing methods.

阅读与讨论 → 访问原文 →

12.

arXiv (CS.LG) 2026-06-25 DOI: arXiv:2603.07221

Margin in Abstract Spaces

作者:

Yair Ashlagi ↗Roi Livni ↗Shay Moran ↗Tom Waknine ↗

arXiv:2603.07221v2 Announce Type: replace Abstract: Margin-based learning, exemplified by linear and kernel methods, is one of the few classical settings where generalization guarantees are independent of the number of parameters. This makes it a central case study in modern highly over-parameterized learning. We ask what minimal mathematical structure underlies this phenomenon. We begin with a simple margin-based problem in arbitrary metric spaces: concepts are defined by a center point and classify points according to whether their distance lies below $r$ or above $R$. We show that whenever $R>3r$, this class is learnable in any metric space. Thus, sufficiently large margins make learnability rely only on the triangle inequality, without any linear or analytic structure being necessary. Our first main result extends this phenomenon to concepts defined by bounded linear combinations of distance functions, and reveals a sharp threshold: there exists a universal constant such that whenever the margin is larger than this constant, the class is learnable in every metric space, while below it there exist metric spaces where it is not learnable at all. We then ask whether margin-based learnability can always be explained via an embedding into a linear space – that is, reduced to linear classification in some Banach space through a kernel-type construction. We answer this negatively by demonstrating a margin learnable class that cannot be embedded into any Banach space in which linear classification with margins is learnable.

阅读与讨论 → 访问原文 →

13.

arXiv (CS.AI) 2026-06-25 DOI: arXiv:2606.25705

GUI agent: Guided Exploration of User-Sensitive Screens

作者:

Aradhana Nayak ↗Mussadiq Nazeer ↗Wang Peng ↗Feng Liu ↗

arXiv:2606.25705v1 Announce Type: new Abstract: LLM agents are increasingly being used to automate tasks for users within an open GUI environment. They inevitably encounter screens containing user-sensitive information, for which takeover of task execution by the user is highly desirable or even necessary. State-of-the-art LLM-driven agents are usually fine-tuned to complete tasks regardless of the safety implications of their actions. This makes their real-world deployment difficult and adversely affects the reliability. Therefore, it is crucial to identify and categorize user-sensitive states and define user-sensitive queries. This dataset would be to engineers to recognize and request handover to the user in critical scenarios. This short paper develops an explorer agent that systematically explores the query space starting from one demonstrated task to identify queries that, if executed, would lead to user-sensitive states in a GUI environment.

阅读与讨论 → 访问原文 →

14.

medRxiv (Medicine) 2026-06-24 DOI: HASH:02d90d6d97a095f84592bd9424637023

Pembrolizumab, Temozolomide and HSPPC-96 Vaccine in Newly Diagnosed Glioblastoma Post-Chemoradiation: Results from a Multi-institutional, Phase 2, Randomized, Placebo-Controlled Trial

作者:

Ozer ↗B. H ↗Lindhorst ↗S. M ↗Merrell ↗R. T ↗Trevino ↗C. R ↗Rudnick ↗J. D ↗Avgeropoulos ↗N. G ↗…

Background: GBM is one of the most common and most aggressive brain tumors in adults, and upfront standard of care treatment has limited efficacy. Immune checkpoint inhibitor strategies have significantly improved outcomes in various solid tumors but have not proven effective in GBM, suggesting other strategies may be needed to realize their full potential. Methods: GBM patients were treated with upfront standard of care chemoradiation with temozolomide and pembrolizumab, followed by adjuvant temozolomide and pembrolizumab for six nine-week cycles. Depending on production of sufficient vaccine, patients were randomized into HSPPC-96 vaccine or placebo group (q4 weeks) while those with failed vaccine production continued on study unblinded as an ancillary group. The primary objective was overall survival at one year, and secondary endpoints were progression-free survival at six months, overall and progression-free survival, radiographic response, and tolerability by patient-reported outcomes and adverse event documentation. Results: 90 patients were screened, 32 were treated (8 vaccine, 9 placebo, 15 ancillary), and 26 were evaluable for radiographic responses prior to accrual termination. The study did not meet its primary endpoint of overall survival at one year (65.5% in vaccine group, 75% in placebo). Progression-free endpoints were mildly improved in the vaccine group but were not significant, and response rates were not significantly different. The regimen was well-tolerated and safe. Conclusions: Though limited by early discontinuation, these findings do not support the combination of pembrolizumab and HSPPC-96 vaccine with standard of care therapy. Trials Registration: ClinicalTrials.gov identifier: NCT03018288

阅读与讨论 → 访问原文 →

15.

arXiv (CS.AI) 2026-06-24 DOI: arXiv:2606.00618

Efficient Test-time Inference for Generative Planning Models with OCL Search

作者:

Robert Gieselmann ↗Mihai Samson ↗Federico Pecora ↗Jeremy L. Wyatt ↗

arXiv:2606.00618v2 Announce Type: replace Abstract: Generative models have emerged as a powerful paradigm for AI planning, yet their performance remains constrained by the training data distribution. One approach is to improve generated solutions during inference by scaling test-time compute. A more efficient alternative is to optimize the inference process itself. In this paper, we show that a modified version of a classical Open-Closed List (OCL) search provides just such an efficient inference procedure. Our algorithm synergizes two learned components: a generative model that performs fast rollouts from intermediate states and a heuristic model that prioritizes among candidate reasoning paths. Key contributions include novel exploration control mechanisms and integration of learned models within the OCL framework. Across multiple combinatorial planning domains, our approach outperforms both neurosymbolic search baselines and classical solvers in computational efficiency and solution quality.

阅读与讨论 → 访问原文 →

16.

arXiv (CS.CV) 2026-06-16 DOI: arXiv:2606.16690

PATCH: Action-Chunk-Conditioned Latent Patch Innovation Monitoring for Robot Manipulation

作者:

Yanan Zhou ↗Ranpeng Qiu ↗Yincong Chen ↗Jiajie Cui ↗Weiming Zhi ↗

Learning-based manipulation policies have made substantial progress in real-world robot manipulation, particularly for short-horizon action generation. However, deployment in open workspaces remains fragile under unexpected local scene dynamics, such as moving objects, transient occlusions, or disturbances near the intended motion. Existing runtime monitors often rely on global observation anomalies, policy uncertainty, or frame-level visual changes, and struggle to distinguish task-relevant execution risk from benign visual variation. We introduce PATCH, an action-chunk-conditioned latent patch innovation monitor for deployment-time intervention. Given the active action chunk, PATCH defines a projected execution corridor, predicts latent patch evolution inside it, and accumulates persistent residuals unexplained by the robot's own motion. These residuals form a localized intervention signal that allows PATCH-Router to pause execution, select an available recovery source, and resume the original policy once localized innovation subsides. Experiments on real robot rollout data show that PATCH produces more stable and context-relevant triggers than competing runtime monitors. Real-robot deployment further demonstrates monitor-driven intervention and policy resumption for disturbance-aware manipulation. Project Page: https://yananzhou5555.github.io/PATCH/.

阅读与讨论 → 访问原文 →

17.

arXiv (quant-ph) 2026-06-24 DOI: arXiv:2606.08916

Chemical tuning of magnetic ordering and cryogenic magnetocaloric response in zircon-type Gd1-xErxVO4

作者:

Ming Zeng ↗Muqing Su ↗Liang Ming ↗Xiaolong Yang ↗Wang Chen ↗Lingwei Li ↗Hai-Feng Li ↗

arXiv:2606.08916v2 Announce Type: replace-cross Abstract: Chemical substitution offers an effective route to tune magnetic ordering and magnetocaloric performance in rare-earth oxides for cryogenic refrigeration. Here we investigate the structural evo lution, magnetic properties, and magnetocaloric effect of polycrystalline zircon-type Gd1-xErxVO4 (x=0, 0.1, 0.25, 0.5, and 0.75). Powder X-ray diffraction confirms that all samples crystallize in the tetragonal zircon structure without detectable impurity phases. Substitution of Gd3+ by the smaller Er3+ ion produces a systematic lattice contraction and modifies the magnetic behavior of the rare-earth sublattice. In particular, the magnetic ordering temperature is suppressed from 3.65(2) K in GdVO4 to 2.76(2) K in Gd0.9Er0.1VO4 , accompanied by a weakening of the spin-flop-like field-induced anomaly observed in the parent compound. A low Er concentration correspondingly improves the low-temperature magnetocaloric performance, with Gd0.9Er0.1VO4 exhibiting a max imum magnetic entropy change of 45.1 J kg-1 K-1 for mu_0 Delta H=7T. These results demonstrate that weak Er substitution effectively tunes the competition among exchange interactions, dipolar coupling, and magnetic anisotropy, optimizing the balance between magnetic ordering and available spin entropy in zircon-type rare-earth vanadates, which is crucial for developing efficient cryogenic refrigeration materials.

阅读与讨论 → 访问原文 →

18.

arXiv (math.PR) 2026-06-24 DOI: arXiv:2509.24950

On domains of elliptic operators with distributional coefficients

作者:

Immanuel Zachhuber ↗

arXiv:2509.24950v2 Announce Type: replace-cross Abstract: We show how one can use recently gained insights from the study of singular SPDEs, more particularly the study of singular operators via the theory of Paracontrolled Distributions, to construct domains for (singular) elliptic operators. Formally we consider \[ A (u) = (1 - \Delta) u + \nabla V \cdot \nabla u + \xi u + {{div} (\rho u)}, \] where $V \in \mathcal{C}^{\delta}$, $\xi \in \mathcal{C}^{- 2 + \delta}$, $\rho \in \mathcal{C}^{- 1 + \delta}, {div} \rho = 0$} and which satisfy a structural assumption that is notably satisfied when $\xi$ is a sub-critical noise, see {[MvZ22]}. We also show that under this assumption, one can construct a continuous change of variables $\Theta$ which satisfies \[ A \Theta - (1 - \Delta) \in \mathcal{L} (H^{2 - \delta''} ; H^{\delta'}) \] which allows us to define $A$ rigorously and parametrise a domain. Moreover, for suitably regularised operators \[ A_{\varepsilon} (u) := (1 - \Delta) u + \nabla V_{\varepsilon} \cdot \nabla u + (\xi_{\varepsilon} + c_{\varepsilon}) \cdot u + {{div} (\rho_{\varepsilon} \cdot u)}, \] we show that for a strongly converging regularised change of variables $\Theta_{\varepsilon} \rightarrow \Theta$ we have \[ A_{\varepsilon} \Theta_{\varepsilon} \rightarrow A \Theta in \mathcal{L} (H^2 ; L^2) \] which in particular implies norm resolvent convergence to a limiting closed operator. Finally, we give a class of examples and show how to apply these results to prove strong analytical local well-posedness for a singular Schrödinger equation formally given by \[ i \partial_t u + (1 - \Delta) u + \nabla V \cdot \nabla u + \xi \cdot u = - | u |^2 u \] for singular $V, \xi$ and that its solution is the limit of the solution of the classical solutions of a regularised equation

阅读与讨论 → 访问原文 →

19.

PLOS Medicine 2026-06-25 DOI: HASH:8dd3b6fe33374758bd1176b2d7bf8252

Association of armed conflict and global measles cases: A structural equation modeling analysis of 193 countries from 2000 to 2023

作者:

Tyler Y. Headley ↗

by Tyler Y. Headley, Yesim Tozan Background Global armed conflict and population displacement are increasing, yet their association with population health remains poorly understood. We developed and tested four theoretical models linking armed conflict, population displacement, and socioeconomic development to measles burden across 193 countries from 2000 to 2023. Methods and findings We analyzed longitudinal country-level data comprising 4,632 country-year observations, combining fixed-effects panel regression and structural equation modeling (SEM). Observed variables included battle-related deaths (BRDs) and forcibly displaced population sizes, while socioeconomic development was modeled as a latent variable incorporating gross domestic product (GDP) per capita, life expectancy, and mean years of schooling. Outcomes were total measles cases and incidence per million population. All four constructed models demonstrated excellent fit (Comparative Fit Index [CFI] 0.991–0.996; Tucker–Lewis Index [TLI] 0.976–0.989; Root Mean Square Error of Approximation [RMSEA] 0.046–0.062). Higher contemporaneous BRDs were associated with higher measles cases (β = 0.17; 95% Confidence Interval [CI] [0.14, 0.20]; p

阅读与讨论 → 访问原文 →

20.

arXiv (CS.CL) 2026-06-25 DOI: arXiv:2606.25821

SARA: Unlocking Multilingual Knowledge in Mixture-of-Experts via Semantically Anchored Routing Alignment

作者:

Tianyu Dong ↗Yangyang Liu ↗Jiang Zhou ↗Xinwei Wu ↗Xiaohu Zhao ↗Hao Wang ↗Heng Liu ↗Linlong Xu ↗Longyue Wang ↗Weihua Luo ↗Shaolin Zhu ↗Deyi Xiong ↗…

Sparse Mixture-of-Experts (MoE) architectures have emerged as an increasingly influential paradigm as they offer a strategic balance between parameter scalability and computational efficiency. However, low-resource languages, which suffer from a scarcity of high-quality training data, often have their tokens routed to different experts than those predominantly activated by high-resource inputs, which limits cross-lingual expert sharing. This cross-lingual routing divergence consequently hinders their efficacy in multilingual contexts. To address this issue, we propose SARA (Semantically Anchored Routing Alignment), a framework designed to transfer specialized capabilities from high-resource languages as anchors to low-resource languages. SARA explicitly aligns the routing distribution of multilingual inputs with high-resource semantic anchors using a symmetric Jensen-Shannon (JS) divergence constraint. Unlike traditional distillation methods that operate on output logits, SARA directly aligns the internal routing distributions of MoE layers, encouraging mechanistic consistency in expert selection across languages. We conduct experiments on 2 LLMs across 5 low-resource languages and 3 benchmarks. Experiment results demonstrate that SARA outperforms standard instruction tuning, e.g., +0.8% on Qwen3-30B-A3B and +1.2% on Phi-3.5-MoE-instruct on Global-MMLU. Further analyses show that SARA effectively addresses performance bottlenecks in low-resource languages, providing a scalable pathway to enhance multilingual capabilities in sparse architectures.

阅读与讨论 → 访问原文 →

21.

arXiv (CS.AI) 2026-06-16 DOI: arXiv:2606.15523

AQ4SViT: An Automated Quantization Framework with Search Gating Policy for Compressing Spiking Vision Transformers

作者:

Rachmad Vidya Wicaksana Putra ↗Saad Iftikhar ↗Muhammad Shafique ↗

arXiv:2606.15523v1 Announce Type: cross Abstract: Spiking Vision Transformers (SViTs) have emerged as alternative low-power ViT models, but their large sizes hinder their deployments on resource-constrained embedded AI systems. To address this, state-of-the-art works proposed quantization techniques to compress SViT models, but their manual, human-guided approach needs a huge design time and power/energy consumption to find the appropriate quantization setting for each given network, making this approach not scalable for quantizing multiple networks. Toward this, we propose AQ4SViT, a novel automated quantization framework for SViTs that can provide quick quantization settings with good trade-offs between accuracy and memory. To achieve this, AQ4SViT employs the following key ideas: quantization search strategy that evaluates the quantization setting candidates while considering the accuracy constraint; and search gating policy that quickly evaluates and selects promising quantization candidates by leveraging membrane potential drift as a performance proxy. In the search gating policy, AQSViT employs two search algorithm variants to provide trade-off options: Greedy search, which performs fast but may lead to local optima; and Beam search, which performs slower but has better performance in finding global optima selection due to a wider search space. Experimental results show that AQ4SViT-Greedy quickly finds the appropriate quantization settings, achieving up to 6.6x faster search time and up to 82.5% memory saving compared to the state-of-the-art; while AQ4SViT-Beam further reduces the memory footprint by up to 90% compared to the state-of-the-art, but with 4.5x longer search time; all these results are obtained while maintaining high accuracy within 1.5% from the original/non-quantized models on the ImageNet dataset. These results highlight that AQ4SViT framework offers advancements toward SViT deployments on embedded AI systems.

阅读与讨论 → 访问原文 →

22.

arXiv (CS.CL) 2026-06-18 DOI: arXiv:2606.18893

Learning Robust Pair Confidence for Multimodal Emotion-Cause Pair Extraction

作者:

Zhuangzhuang Pan ↗Ning Dong ↗Yingna Su ↗Yan Xia ↗

Multimodal emotion-cause pair extraction (MECPE) requires reliable pair confidence over candidate pairs. Existing pair scorers commonly use pair-level cross entropy over valid candidates, which treats links mostly independently. This leaves the relative confidence geometry among competing causes under-constrained, allowing gold pairs to stay close to hard negatives or rely on incidental non-gold context. We study this vulnerability as pair-confidence brittleness and propose RPCL (Robust Pair Confidence Learning), a training-only framework for pair-confidence learning. RPCL encourages pair confidence to be both discriminative and stable: gold pairs are separated from row-wise hard negatives through a confidence-difference margin constraint, and clean pair predictions are aligned with predictions from a corrupted view where non-gold contextual utterance representations are partially corrupted. The original clean pair scorer and decoding pipeline are used unchanged at inference time. On ECF, MECAD, and MEC4, RPCL improves the three-seed mean Pair F1 over a matched base model by 2.58 to 2.83 percentage points in the full text-audio-video setting, and improves mean Pair AUPRC on all three datasets. Diagnostic analysis further shows larger gold-negative confidence gaps and lower margin-violation severity. These results suggest that explicitly shaping pair confidence is an effective training strategy for MECPE.

阅读与讨论 → 访问原文 →

23.

arXiv (quant-ph) 2026-06-11 DOI: arXiv:2305.04908

Tight Bounds for Quantum Phase Estimation and Related Problems

作者:

Nikhil S. Mande ↗Ronald de Wolf ↗

arXiv:2305.04908v3 Announce Type: replace Abstract: Phase estimation, due to Kitaev [arXiv'95], is one of the most fundamental subroutines in quantum computing. In the basic scenario, one is given black-box access to a unitary $U$, and an eigenstate $\lvert \psi \rangle$ of $U$ with unknown eigenvalue $e^{i\theta}$, and the task is to estimate the eigenphase $\theta$ within $\pm\delta$, with high probability. The cost of an algorithm for us is the number of applications of $U$ and $U^{-1}$. We tightly characterize the cost of several variants of phase estimation where we are no longer given an eigenstate, but are required to estimate the maximum eigenphase of $U$, aided by advice in the form of states (or a unitary preparing those states) which are promised to have at least a certain overlap $\gamma$ with the top eigenspace. We give algorithms and nearly matching lower bounds for all ranges of parameters. We show that a small number of copies of the advice state (or of an advice-preparing unitary) are not significantly better than having no advice at all. We also show that having lots of advice (applications of the advice-preparing unitary) does not significantly reduce cost, and neither does knowledge of the eigenbasis of $U$. We immediately obtain a lower bound on the complexity of the Unitary recurrence time problem, resolving an open question of She and Yuen~[ITCS'23]. Lastly, we study how efficiently one can reduce the error probability in the basic phase-estimation scenario. We show that a phase-estimation algorithm with precision $\delta$ and error probability $\epsilon$ has cost $\Omega\left(\frac{1}{\delta}\log\frac{1}{\epsilon}\right)$, matching an easy upper bound. This contrasts with some other scenarios in quantum computing (e.g., search) where error-probability reduction costs only a factor $O(\sqrt{\log(1/\epsilon)})$. Our lower bound uses a variant of the polynomial method with trigonometric polynomials.

阅读与讨论 → 访问原文 →

24.

medRxiv (Medicine) 2026-06-24 DOI: HASH:17bff8520611d99fd62607036a6e09be

Barriers and facilitators to diabetes management among adults and healthcare providers in a peri-urban Ugandan health facility: A qualitative study

作者:

Larissa ↗K. N. Y ↗Kooko ↗Musoke ↗Kisame ↗Komangoya-Nzonzo ↗A. D ↗Nakisita ↗Dandy ↗M. W. W ↗

Diabetes mellitus is an increasing public health challenge in Uganda and other low- and middle-income countries, where health systems face growing demands for chronic disease care. Although quantitative studies have documented poor glycemic control and health system constraints, less is known about how patients and healthcare providers experience diabetes management in peri-urban public health settings. This study explored barriers and facilitators to diabetes management among adults with diabetes mellitus and healthcare providers at a peri-urban health facility in Uganda. We conducted a qualitative descriptive study at Kasangati Health Centre IV, Wakiso District, Uganda, between February and March 2025. Data were collected through 15 in-depth interviews with adults living with diabetes mellitus and 8 key informant interviews with healthcare providers involved in diabetes care. Participants were purposively selected based on their experience with diabetes management and service delivery. Interviews were audio-recorded, transcribed verbatim, translated where necessary, and analyzed using a hybrid inductive-deductive thematic approach informed by the Theoretical Domains Framework. Five interrelated themes were identified: (1) institutional and environmental factors influencing access to diabetes care; (2) cognitive and informational factors influencing medication adherence; (3) social influences on diabetes management; (4) emotional experiences of patients and healthcare providers; and (5) self-management strategies and continuity of care. Across these themes, participants identified barriers including resource limitations, communication challenges, medication management difficulties, stigma, emotional distress, and weak follow-up systems. Facilitators included peer support, religious and community networks, health education, provider flexibility, and patient-developed adherence strategies. Diabetes management was influenced by interacting health-system, social, informational, and behavioural factors. Resource constraints, limited health literacy, stigma, and weak follow-up systems hindered effective management, while social support, health education, and patient self-management strategies facilitated continued engagement in care. Interventions that strengthen chronic care services, patient education, and community support may improve diabetes outcomes in similar resource-constrained settings.

阅读与讨论 → 访问原文 →

25.

arXiv (quant-ph) 2026-06-16 DOI: arXiv:2605.22424

Long-range nonstabilizerness of topologically encoded states from mutual information

作者:

David Aram Korbany ↗Tyler D. Ellison ↗David T. Stephen ↗Lorenzo Piroli ↗

arXiv:2605.22424v2 Announce Type: replace Abstract: We study long-range nonstabilizerness (LRN), namely the obstruction to remove nonstabilizerness with shallow-depth local quantum circuits. In one-dimensional settings, the mutual information between disconnected spatial regions has proven to be a powerful tool to diagnose LRN. In this work, we focus on encoded states of two-dimensional topologically-ordered systems, and explore the ability of the mutual information to serve as a diagnostic of LRN. Focusing on the concrete setting of lattice models defined on a torus, we show that information about LRN can be gained from the analysis of the mutual information between non-overlapping regions containing non-contractible loops, and of the change of such mutual information under modular real-space transformations. We exemplify this idea in the toric code and the non-abelian string-net model with doubled Fibonacci topological order. In the former case, we show that the mutual information provides a full classification, certifying LRN for all encoded non-stabilizer states. In the latter case, instead, our approach does not lead to a full classification, as it detects LRN for all states except from a finite subset with special transformation properties under the modular group. Finally, we discuss how our results on LRN constrain the logical gates that can be implemented fault-tolerantly on the torus.

阅读与讨论 → 访问原文 →

探索全球前沿学术脉络