• raseliarison
  • nirinA
  • adrien
  • blog
  • code
  • FAQ
  •  home  
  •  news  
    • arXiv
      • astro-ph
      • cond-mat
      • cs
      • eess
      • gr-qc
      • hep-ex
      • hep-lat
      • hep-ph
      • hep-th
      • math
      • math-ph
      • nlin
      • nucl-ex
      • nucl-th
      • physics
      • q-bio
      • quant-ph
      • stat
    • physics
      • phys.org
      • physics world
    • linux
      • kernel
      • slackware
    • nature
      • natcomputsci
      • natastron
      • natbiomedeng
      • nenergy
      • nnano
      • natmachintell
      • nbt
      • nmeth
      • natecolevol
      • nmicrobiol
      • ng
      • nchembio
      • natelectron
      • micronano
      • nphoton
    • bioRxiv
    • plos one
    • world
      • BBC
      • Al Jazeera
    • earth
      • earth observatory
      • weather
      • weather forecast
    • universe
      • apod
      • hubble
      • atel
      • nasa
  •  wiki  
  •  gemini  
  •  python  
  • q-bio updates on arXiv.org

    q-bio updates on the arXiv.org e-print archive.

    BreCol: Benchmarking Classical and Deep-Learning Methods for Microbiome-Based Cancer Detection

    oai:arXiv.org:2609.27207v1

    arXiv:2609.27207v1 Announce Type: new Abstract: DNA sequencing of the gut microbial community shows promise for cancer detection, but questions remain about the generalizability of results across studies. We propose BreCol, a benchmark of 2,040 16S rRNA gene sequencing runs across 26 studies spanning breast cancer, colorectal cancer, and healthy cohorts. Train-test splits are made within pre-2023 studies, while holdout evaluation uses studies from 2023 onward, reflecting temporal separation from training data. Classical models reach test/holdout AUCs of 0.77/0.60 for cancer diagnosis and 1.00/0.83 for cancer type prediction. We train the models on both cancer types simultaneously and find that colorectal cancer is often easier to detect than breast cancer. We also evaluate two deep learning models: HyenaDNA, a long-range sequence model that pools hidden states for classification, and SetBERT, a transformer that produces contextualized embeddings over sets of reads. Both deep learning models underperform the best classical methods on holdout data, though tuning training set size and the classification head yields modest gains. Our classical pipeline uses unsupervised clustering to derive features from tetramer frequencies, preserving within-run compositional signal and achieving near state-of-the-art performance without relying on taxonomic assignments. BreCol data and associated code are publicly available.

    https://arxiv.org/abs/2609.27207


    Topological Inference for Organoids

    oai:arXiv.org:2609.27668v1

    arXiv:2609.27668v1 Announce Type: new Abstract: The reproducibility of organ morphology and the extent to which computational models can predict morphogenesis remain difficult to quantify, particularly for organs with complex networks of fluid-filled lumina. Here, we combine Topological Data Analysis (TDA), biophysical simulation, and Bayesian inference to study lumen morphogenesis in pancreatic organoids. Lumen formation is governed by physical processes that are challenging to measure directly, including cell proliferation and luminal osmotic pressure. We simulate organoid development using a phase-field model and address the inverse problem of inferring these parameters from either time-lapse images or single morphological snapshots. Since lumen architectures vary substantially in size, structure, and connectivity, conventional geometric descriptors provide only a partial representation of their morphology. We therefore represent each organoid using SampEuler, a topological descriptor derived from the Euler Characteristic Transform (ECT). We first show that SampEuler captures morphological information encoded by established morphometrics. We then perform parameter inference using an approximate Bayesian computation (ABC) rejection framework with the SampEuler Wasserstein distance. Using synthetic organoids with known ground-truth parameters, our approach accurately recovers the osmotic pressure and the proliferation rate while revealing a compensatory trade-off between the two processes. Applied to experimental data from ten pancreatic organoids, the inferred posterior distributions are consistent with biological expectations. Together, these results establish a non-destructive, image-based pipeline for estimating otherwise inaccessible physical parameters governing lumen formation and highlight the potential of topological representations for linking complex biological morphology to mechanistic models.

    https://arxiv.org/abs/2609.27668


    AI-Driven Neural Surrogates for In Silico Design of Cognitive-Affective Neuromodulation Targets

    oai:arXiv.org:2609.27729v1

    arXiv:2609.27729v1 Announce Type: new Abstract: In neuropsychiatry, the primary goal is often not only to decode brain activity but to change it, for example to lessen a negative affective bias or an overly salient memory. Motivated by control theory, we develop an AI-driven neural-surrogate framework that proposes candidate representational changes and tests their predicted perceptual effects from snapshots of stimulus-evoked fMRI activity, without physical stimulation. The framework combines fMRI decoding, deep generative modeling, and constrained latent-space steering. Valence and memorability are used only as worked examples. Using more than 36,000 image-fMRI observations from four deeply sampled Natural Scenes Dataset participants, subject-specific models recovered coarse generative structure from visually responsive cortex (two-way identification, 0.79-0.88; chance, 0.5). Graded perturbations were reconstructed as images and evaluated with automated scorers and human ratings from 7,200 trials by 18 participants. In the primary VDVAE model, valence shifted from -0.61 to +1.03 SD and memorability from -1.34 to +1.45 SD; a later Versatile Diffusion refinement reduced or altered these effects. Across five perturbation levels, human valence ratings moved in the predicted direction under the linear time-correction model (mean slope, 0.038 SD per unit of alpha; 95 percent CI, 0.003-0.074; positive in 16 of 18 participants). Perceived memorability did not change reliably. Baseline agreement with the automated assessor was suggestive for valence (r = 0.30) and weak for memorability (r = 0.10). Extreme perturbations drifted from the original stimulus, so intended change must be weighed against loss of fidelity. These findings provide a falsifiable upstream method for designing and behaviorally testing candidate representational targets for future neuromodulation in psychiatry, while marking the limits of the present static approximation.

    https://arxiv.org/abs/2609.27729


    Spatial-Molecular Information Analysis Based on Histopathological Geometry: KL-Divergence Decomposition of Cell Distribution and Location-Dependent Expression States in PD-1 Immunohistochemistry

    oai:arXiv.org:2609.27755v1

    arXiv:2609.27755v1 Announce Type: new Abstract: Spatial biology enables molecular expression to be analyzed together with the location of cells in tissue. However, it is often difficult to distinguish whether cells preferentially accumulate near a histopathological boundary from whether their molecular expression states vary with that location. To establish informational scientific framework, we define boundary distance as R and a binary molecular expression state as G, and compare the observed joint distribution P\left(R,G\right) with Q\left(R\right)P\left(G\right), where Q\left(R\right) is a null distribution determined by tissue geometry. The resulting Kullback-Leibler divergence decomposes into D_{R}, which quantifies deviation of the spatial distribution of cells from the geometric null, and I\left(R;G\right), which measures how strongly molecular expression state depends on boundary distance. Idealized simulations and semi-synthetic validation using geometry reconstructed from a PD-1 IHC image confirmed that these two effects can be generated separately and recovered by the corresponding terms. Plug-in estimators showed positive finite-sample bias, supporting permutation-based inference. We then applied the framework to five publicly available PD-1 lung cancer IHC images from the Human Protein Atlas. In pooled analysis, I\left(R;G\right)=0.0302 nats and was significant in a stratified within-image label-permutation test (p=0.00040). This framework provides a quantitative way to distinguish where cells are located from how their molecular expression states vary with tissue location. Larger cohorts, independent datasets, additional cancer types, and other molecular markers will be required to establish generalizability and clinical utility.

    https://arxiv.org/abs/2609.27755


    Circadian Derived Features for Early Discrimination Across Insomnia Severity Levels: At Least 8 Weeks of Monitoring Are Needed for Clinically Meaningful Assessment

    oai:arXiv.org:2609.28252v1

    arXiv:2609.28252v1 Announce Type: new Abstract: Background: Wearable devices provide continuous, objective measures of daily activity and offer promise for assessing sleep disorders. However, the minimum monitoring duration needed to differentiate insomnia severity remains unclear. We investigated when wearable-derived behavioral features become informative for distinguishing Insomnia Severity Index (ISI) categories and examined the contribution of Activity Count (AC) and circadian-derived features. Methods: We analyzed wearable data from 2,305 participants in the Advancing Understanding of Recovery after Trauma (AURORA) study. Separate binary classifiers were developed for four ISI categories across six follow-up periods using AC and circadian-derived feature sets. Performance was evaluated using accuracy, F1-score, precision, recall, and AUROC. Results: Classification improved with longer monitoring. Across ISI categories, approximately eight weeks was the earliest time point at which wearable-derived features consistently achieved informative discrimination, with modest improvements thereafter. Participants without clinically significant insomnia were easiest to identify, reaching an AUROC of 0.693. Circadian-derived features performed comparably to, and in several cases better than, AC features, suggesting that the temporal organization of daily activity provides information beyond overall activity volume. Conclusions: Approximately eight weeks of longitudinal wearable monitoring may represent a practical minimum for differentiating ISI-defined insomnia categories. Longer monitoring provided only incremental improvements. Circadian behavioral features show promise as digital biomarkers for objective insomnia assessment.

    https://arxiv.org/abs/2609.28252


    Motif-Vocab: StatisticallyCalibrated Transcription-Factor-Identity Tokenization forGenomic Language Models

    oai:arXiv.org:2609.28386v1

    arXiv:2609.28386v1 Announce Type: new Abstract: Tokenization is a central design choice in genomic language models, yet most deoxyribonucleic acid (DNA) tokenizers use characters, fixed-length k-mers, or frequency-derived subwords without explicitly using prior information about the specificity of DNA-binding regulatory factors. We introduce Motif-Vocab, a biologically informed tokenizer that scans both DNA strands for statistically calibrated motif matches, emits transcription-factor (TF) identity tokens, and applies nucleotide, $k$-mer, or byte-pair encoding (BPE) to unmatched sequence. Motif-specific null distributions put position-weight matrices (PWMs) of different lengths and degeneracy on a common significance scale; deterministic overlap rules make the representation reproducible. In controlled Bidirectional Encoder Representations from Transformers (BERT) pretraining on two billion base pairs, real motif libraries outperform randomized-motif controls on 54 of 55 in-scope downstream tasks. On a motif-disjoint recognition task derived from DART-Eval Task 2, TF-specific tokens improve macro-F1 by 0.040 over a position-matched generic motif token and by 0.033 over a matched no-motif tokenizer (95\% bootstrap confidence interval: 0.027--0.038). Motif tokens also receive stronger attribution and produce larger occlusion effects than shuffled controls. Dense no-motif tokenizers remain strong general-purpose baselines, including a near-tie on the five-task BERT-base panel. Thus, Motif-Vocab is not a universal accuracy replacement; it is a targeted, interpretable inductive bias for motif-sensitive genomic modeling.

    https://arxiv.org/abs/2609.28386


    The Computational Value of Sensory-Aligned Receptive Fields Depends on Neuronal Expressivity

    oai:arXiv.org:2609.26940v1

    arXiv:2609.26940v1 Announce Type: cross Abstract: Biological sensory neurons have selective receptive fields organized along meaningful stimulus coordinates, such as frequency, motion direction, or retinotopic position. Such structure may arise from efficient coding and biological constraints on activity, connectivity, and wiring, as computational studies of simple neurons have shown across modalities. This raises a question: do structured receptive fields confer a computational advantage beyond resource efficiency itself, and does this advantage persist when individual neurons are highly expressive? We address this question in recurrent networks of Expressive Leaky Memory neurons, where we can independently vary neuronal complexity and the organization of feed-forward receptive fields. Across auditory and event-based visual classification tasks, receptive fields aligned with a task-relevant sensory coordinate improve test accuracy relative to budget-matched random receptive fields. This advantage disappears when sensory coordinates are scrambled, or when receptive fields follow task-irrelevant coordinates, showing that the benefit comes from alignment with task geometry rather than restricted connectivity alone. Increasing neuronal complexity reduces the performance advantage of structured receptive fields. Finally, generic synaptic sparsity regularization induces input selectivity and partially recovers performance, but remains substantially below explicitly structured receptive fields, suggesting that sparsity alone is insufficient to recover the full computational benefit of task-aligned receptive fields. Together, our results show that appropriate receptive fields can serve as a computational prior beyond sparsity itself, and that their value depends on the computational expressivity of individual neurons.

    https://arxiv.org/abs/2609.26940


    Spatially Resolved Nucleated Polymerization: A Free-Boundary Model of Protein Aggregation in Concentrated Solutions

    oai:arXiv.org:2609.27185v1

    arXiv:2609.27185v1 Announce Type: cross Abstract: Kinetic models of protein aggregation describe populations by size, not by spatial organization or morphology. We extend Lumry-Eyring nucleated polymerization to a model in which the monomer is a density field and each aggregate is a region bounded by a level set. Growth is a flux condition on the available sites of a surface. Condensation is a reaction between the bonding sites of two surfaces in contact, at a rate set by the bond rate and the contact geometry. The availability of those sites is a field on the interface, and its equilibrium value follows from Wertheim's perturbation theory. The collision efficiency and the Fuchs stability ratio are therefore computed, not fitted. In a well-mixed limit the model's spatial averages satisfy the rate equations term by term; the monomer fraction agrees to eight parts in ten thousand, a difference that arises from equating aggregate size with volume. The condensation kernel's exponent is $0.5806\pm0.0013$ against the $0.600\pm0.010$ fitted to a monoclonal antibody. The computed stability ratio reproduces thirteen of fourteen published conditions at twelve $k_BT$, but only with the bond rate at the top of its range. In a many-body box, aggregates merge at $2.2$ to $3.8$ times the two-body rate.

    https://arxiv.org/abs/2609.27185


    Benchmarking Active Spot Selection for Cost-Efficient Spatial Transcriptomics

    oai:arXiv.org:2609.27208v1

    arXiv:2609.27208v1 Announce Type: cross Abstract: Spatial transcriptomics (ST) measures gene expression in tissue context, but dense capture grids can be costly and may repeatedly sample morphologically similar regions. Most active learning strategies were developed for categorical labels and independent samples. We conduct a retrospective pool-based benchmark of active learning versus uniform Random sampling for ST, where expression vectors are high-dimensional and continuous and candidates are spatially correlated. Using two fully profiled public ST cohorts, we mask candidate expression vectors and simulate multi-round selection with uncertainty-based Monte Carlo dropout (MC-dropout) and temporal output discrepancy (TOD), and diversity-based CoreSet and TypiClust-inspired selection. We compare 160 completed configurations at 5%, 10%, 30%, and 50% of the fold-wide training spot pool under patient-level cross-validation, with a separate full-label reference. Within each budget, strategies share the selection schedule, morphology-to-expression predictor, and optimization protocol. We assess mean per-gene within-slide Pearson correlation coefficient (PCC), expression-cluster agreement, and Moran's I fidelity. On HER2-positive breast cancer, pooled mean PCC differences from Random across the four active strategies were -0.0176, -0.0117, +0.0056, and +0.0057 at 5%, 10%, 30%, and 50%, respectively. On cutaneous squamous cell carcinoma (cSCC), three strategies were below Random at 5%, and all four were below Random at 10%. On HER2-positive breast cancer, CoreSet and MC-dropout had lower PCC but higher expression-cluster agreement than Random at the two smallest budgets; this pattern did not reproduce on cSCC. Under the reported fixed training horizons, the evaluated active strategies do not consistently improve on Random at small budgets, and rankings depend on the evaluation measure.

    https://arxiv.org/abs/2609.27208


    Physiologically Informed Digital Auscultation for Pneumonia Detection in Long-term Care Residents

    oai:arXiv.org:2609.27222v1

    arXiv:2609.27222v1 Announce Type: cross Abstract: Pneumonia is difficult to diagnose in older long-term care residents; multimorbidity and atypical presentations obscure signs, motivating operationally efficient objective testing. We analyzed multi-channel digital stethoscope recordings from 185 Japanese residents (73 pneumonia, 112 symptomatic without), using radiologist-confirmed chest X-rays and clinician diagnoses as supervisory signals that train convolutional neural networks, multimodal fusion, and channel-based variants with time-domain Grad-CAM interpretability. Models were evaluated with repeated patient-level cross-validation showing models with X-ray supervision outperformed clinician supervision (F1 0.729, accuracy 0.783 vs. F1 0.637, accuracy 0.711). Additionally, a three-channel selection protocol maintained performance (F1 0.736; accuracy 0.803), with two mid-thoracic sites ranking highest and Grad-CAM attention overlapping adventitious sounds. These findings indicate automated multi-channel lung-sound analysis can aid long-term care pneumonia diagnosis, with X-ray supervision being more reliable than clinical, and fewer channels preserving performance while lowering acquisition times.

    https://arxiv.org/abs/2609.27222


    Discover, Falsify, Revise: Auditing Input-Use Claims from Source Code to Predictive Contribution in Agent-Discovered Cell Models

    oai:arXiv.org:2609.27234v1

    arXiv:2609.27234v1 Announce Type: cross Abstract: AI virtual cells aim to predict cellular responses to specified interventions, yet held-out predictive performance alone does not establish use of the supplied perturbation information. This prediction-claim gap matters in agentic model discovery, where language-model agents generate and revise predictors using score-based feedback. We introduce CELLAUDIT, which audits input-use claims by asking whether an input can enter the cited computation, whether fitted predictions depend on it, and whether that dependence improves prediction of observed response. On a paired morphology-transcriptomics perturbation benchmark (BBBC047), an agent-selected predictor attains a mean held-out Global Pearson correlation coefficient (PCC) of 0.3153 but remains invariant to compound replacement; a control-profile-only predictor reaches 0.3142. Source inspection identifies a compound-query pathway blocked by singleton key-value attention, and the invariance persists after refitting with disjoint control wells. In a stratified audit of 48 candidates across two linked tasks, 47 change predictions under compound replacement on both held-out folds, but only 20 show target-loss gains with intervals above zero on both folds. On BBBC047, falsification-guided revisions recover positive mean compound contributions while retaining gains over the control-profile-only baseline. In matched sci-Plex searches, audit-enriched feedback yields higher held-out performance and larger mean compound and dose contributions across five trajectories, although paired intervals span zero. Refitting fixed designs on an independently acquired cohort shows predictive generalization need not imply generalization of input-use claims: dose contribution persists, whereas support for compound identity does not. CELLAUDIT adds a falsification layer to agentic model discovery, moving from generate-score-revise toward discover-falsify-revise.

    https://arxiv.org/abs/2609.27234


    Pheno-GS: Phenoscape-scale Geodesic Sinkhorn

    oai:arXiv.org:2609.27633v1

    arXiv:2609.27633v1 Announce Type: cross Abstract: High-throughput single-cell data is now collected across large patient cohorts. Understanding patient-level heterogeneity from cellular-level data motivates phenoscaping: embedding each single-cell distribution as a "datapoint," with distances given by optimal transport (OT). Computing geometry-aware OT at this scale, between all pairs of patient datasets, remains an open challenge, since existing methods either rely on Euclidean ground metrics that distort manifold structure or fail under sparse, unevenly sampled, or large-scale data. We present \textbf{Pheno-GS} (Phenoscape-scale Geodesic Sinkhorn), which computes accurate, scalable geodesic transport distances under noisy, unbalanced, large-scale settings via three components: ($1$) graph connectivity regularization for well-defined geodesics on sparse/disconnected manifolds; ($2$) an unbalanced OT formulation via KL marginal penalties; and ($3$) a batched matrix algorithm computing all pairwise distances in one heat diffusion (over $200 \times$ faster than Geodesic Sinkhorn for $500$ distributions). We validate Pheno-GS on synthetic benchmarks and a CyTOF perturbation dataset.

    https://arxiv.org/abs/2609.27633


    Penguin data reanalyzed via Computational Taxonomy

    oai:arXiv.org:2609.28201v1

    arXiv:2609.28201v1 Announce Type: cross Abstract: We employ Computational Taxonomy (CT) to reanalyze the penguin data set penguins_lter by validating and addressing two biological issues: Sexual Size Dimorphism (SSD) and mate-selection criteria. Via Scientific Data Analysis (SDA) computing, CT constructs a Taxonomic Hierarchy by splitting Species first and then Sex, without involving Island, to achieve less complexity. This Taxonomic Hierarchy validates SSD as a branch comparison: (Species, Sex = Male)-vs-(Species, Sex = Female), upon which SDA explores all potential pieces of associative information from all covariate feature-sets, including interacting effects from order-2 to order-4, and then confirms them via their idiosyncratic reliability checks. The collective of confirmed information pieces are displayed on a heatmap platform to manifest underlying dynamics of SSD with explicit block-structured heterogeneity found within males and females. SSD dynamics is explained through mechanistic dependence pertaining to one chief factor consisting of up to 8 feature-sets: Body-Mass coupled by combinations of {Culmen-length,Culmen-depth, Flipper-length}, and two minor factors consisting of low-order combinations of {Culmen-length,Culmen-depth, Flipper-length}. Such Intra-Sex heterogeneity invalidates all Logistic regression modeling on SSD in the original paper. Further, we explore potential mate-selection criteria through the data-frame of Nest-ID within-species homogeneity.

    https://arxiv.org/abs/2609.28201


    Radiomics and artificial Intelligence for thyroid cancer diagnosis: Concepts, challenges, and solutions

    oai:arXiv.org:2404.07239v2

    arXiv:2404.07239v2 Announce Type: replace Abstract: Thyroid cancer is an increasing global health concern that requires advanced diagnostic methods. The application of AI and radiomics to thyroid cancer diagnosis is examined in this review. A review of multiple databases was conducted in compliance with PRISMA guidelines until October 2024. A combination of keywords led to the discovery of an English academic publication on thyroid cancer and related subjects. 368 papers were returned from the original search after 112 duplicates were removed. Relevant studies were selected according to predetermined criteria after 176 articles were eliminated based on an examination of their abstract and title. After the comprehensive analysis, an additional six studies were excluded. Among the 42 included studies, radiomics analysis, which incorporates ultrasound (US) images, demonstrated its effectiveness in diagnosing thyroid cancer. Various results were noted, some of the studies presenting new strategies that outperformed the status quo. The literature has emphasized various challenges faced by AI models, including interpretability issues, dataset constraints, and operator dependence. The synthesized findings of the 42 included studies mentioned the need for standardization efforts and prospective multicenter studies to address these concerns. Furthermore, approaches to overcome these obstacles were identified, such as advances in explainable AI technology and personalized medicine techniques. The review focuses on how AI and radiomics could transform the diagnosis and treatment of thyroid cancer. Despite challenges, future research on multidisciplinary cooperation, clinical applicability validation, and algorithm improvement holds the potential to improve patient outcomes and diagnostic precision in the treatment of thyroid cancer.

    https://arxiv.org/abs/2404.07239


    Differences in Neurovascular Coupling in Patients with Major Depressive Disorder: Evidence from Simultaneous Resting-State EEG-fNIRS

    oai:arXiv.org:2506.11634v2

    arXiv:2506.11634v2 Announce Type: replace Abstract: Neurovascular coupling (NVC), the relationship between neural activity and cerebral hemodynamic responses, remains poorly understood in major depressive disorder (MDD). To investigate alterations in NVC associated with depressive symptom severity, we simultaneously recorded resting-state electroencephalography (rsEEG) and functional near-infrared spectroscopy (fNIRS) in 206 participants, including 134 patients with MDD and 72 healthy controls stratified by age to disentangle disease-related alterations from age-associated effects. NVC in the prefrontal cortex (PFC) was characterized by the consistency and temporal lag between spontaneous electrophysiological peaks and corresponding hemodynamic responses. We found that age significantly enhances the NVC consistency (p < 0.05), whereas this relationship was altered in patients with MDD during both the oxygen-consumption phase and the subsequent hemodynamic compensation phase. Furthermore, among patients with MDD, the strength of this consistency decreased more significantly with greater illness severity (partial r = -0.336, p = 0.0601). These findings suggest that changes in neurovascular coupling effects are strongly associated with the severity of depression. By leveraging wearable neuroimaging techniques, this study provides multimodal evidence for altered neurovascular coupling in depression and highlights its potential as a biomarker for disease monitoring and recovery trajectories.

    https://arxiv.org/abs/2506.11634


    Functional Connectivity-Guided Band Selection for Motor Imagery Brain-Computer Interfaces

    oai:arXiv.org:2605.00746v2

    arXiv:2605.00746v2 Announce Type: replace Abstract: Reliable control in motor imagery brain-computer interfaces (MI-BCIs) requires the precise decoding of user-specific neural rhythms, which vary significantly across individuals. The Common Spatial Pattern (CSP) algorithm is a cornerstone of MI-BCI decoding, yet its performance depends strongly on the spectral range of the input EEG data. Although Filter Bank CSP (FBCSP) extends this as a data-driven decoding framework, its frequency sub-bands are predefined rather than selected using subject-specific physiological criteria. This paper presents a proof-of-concept study of static functional connectivity (FC)-guided band selection for MI-BCI, demonstrated using a conventional FBCSP-based pipeline. The proposed method identifies the most discriminative spectral bands by calculating phase-based connectivity across four sensorimotor channels using wPLI, PLV, and PLI. Nine bands in a 4-40 Hz filter bank are ranked by the effect size of their hemispheric coupling differences and pruned to the top K bands for feature extraction and classification via FBCSP and a Support Vector Regressor. This framework was tested for K values ranging from 1 to 8 across the BCI Competition IV-2a (n = 9) and OpenBMI (n = 54) datasets. Performance was benchmarked against standard nine-band FBCSP and random ablation to determine the minimum number of bands (K*) required to maintain accuracy within a 2% baseline equivalence zone. Results show FC-guided selection can outperform random ablation and achieve near-baseline performance while reducing required CSP fits by 22.2% to 77.8%. PLV enables the most aggressive dimensionality reduction by prioritizing the {\mu} and low-\b{eta} ranges, while wPLI demonstrates superior inter-session robustness by mitigating volume conduction. These findings establish FC-guided selection as a principled and interpretable alternative to heuristic filter bank designs.

    https://arxiv.org/abs/2605.00746


    Critical Flicker Fusion Frequency As An Experience-Restricted Constraint On Visual Temporal Resolution: What Does And Does Not Change It

    oai:arXiv.org:2607.29068v2

    arXiv:2607.29068v2 Announce Type: replace Abstract: Experience-dependent plasticity is fundamental to adaptive behaviour, yet the conditions under which basic sensory timing can be modified in adulthood remain poorly specified. Critical flicker fusion frequency (CFFF), the threshold at which flicker is perceived as continuous, is unusually informative here: what fails to change it is as well documented as what does. Reviewing the stability and training literature, we argue CFFF is best characterised neither as non-plastic nor as held near a physiological ceiling, but as modifiable only by a specific class of experience, not by amount. Repeated testing, cognitive training without temporal content, and incoherent-flicker exposure leave the threshold unchanged regardless of duration. By contrast, a narrow class of perceptual-learning paradigms -- pairing coherent directional motion with a task-relevant target -- reportedly raises the threshold substantially, with gains retained at one year in a small subsample. Gains are largest below typical values, as in amblyopia, and absent in normally sighted observers under the same protocol. Three qualifications apply: the evidence rests on small samples; the threshold-raising paradigms have been assessed almost exclusively with heterochromatic flicker photometry rather than luminance-defined CFFF, leaving construct equivalence unestablished; and a minority of untrained controls show comparable changes. We examine whether the restriction originates at peripheral, thalamocortical, or cortical levels; available data do not adjudicate between them. Functional arguments for why such a restriction might be adaptive are offered as rationales, not evidence. Proposed links to working-memory precision and metacognition are stated as predictions; the first test of them was largely negative.

    https://arxiv.org/abs/2607.29068


    Graph construction in QUBO-based recursive phylogenetic tree reconstruction

    oai:arXiv.org:2609.16640v2

    arXiv:2609.16640v2 Announce Type: replace Abstract: Molecular sequence data are used to reconstruct evolutionary relationships among taxa, but reconstruction accuracy depends not only on the tree-building method but also on how pairwise sequence relationships are represented. We evaluated sequence-to-affinity representations in a recursive normalized-cut (Ncut) framework whose graph-partitioning subproblems were formulated as quadratic unconstrained binary optimization (QUBO) models and solved using Simulated Bifurcation. Using simulated amino-acid and nucleotide datasets spanning multiple tree-generation settings and evolutionary divergence, we compared normalized bit-score affinities with representations derived from transformed sequence similarities and evolutionary distances, examined post-swap refinement, and used neighbor joining (NJ) as a distance-based comparator. Affinity representation substantially affected internal split recovery, particularly for nucleotide data. JC69-based local affinities maintained comparatively high accuracy as divergence increased, whereas normalized bit-score and BLAST-derived kernel representations declined more markedly. Post-swap refinement generally improved recovery, but not consistently across individual reconstructions. NJ achieved higher mean split recovery than corresponding recursive Ncut reconstructions for WAG and JC69 distances across all evaluated conditions, whereas recursive Ncut outperformed NJ for BLAST-derived logarithmic distances under some conditions. These results show that graph construction is an important determinant of recursive Ncut-based phylogenetic reconstruction. A representation that performs well within Ncut does not necessarily provide the most accurate use of the underlying pairwise distances. Pairwise representation, affinity transformation, optimization, and recursive tree construction should therefore be evaluated jointly.

    https://arxiv.org/abs/2609.16640


    Calibration and transfer in indicator-based assessments of artificial consciousness

    oai:arXiv.org:2603.27597v2

    arXiv:2603.27597v2 Announce Type: replace-cross Abstract: Research on artificial consciousness increasingly shifts evaluation from behaviour to internal architecture. Theory-based indicators are used to update probability assignments. This improves on behavioural tests but raises two distinct problems. First, these assignments cannot currently be calibrated against independently established artificial consciousness outcomes. Second, their evidential relevance is transferred from biological cases without independent support that indicator-consciousness relations remain stable across substrates. This commentary distinguishes calibration from transfer and adapts the iterative natural-kind strategy by proposing a preliminary, theory-relative comparative space for cross-substrate assessment.

    https://arxiv.org/abs/2603.27597


    ProteinJEPA: Latent prediction improves protein language model pretraining

    oai:arXiv.org:2605.07554v2

    arXiv:2605.07554v2 Announce Type: replace-cross Abstract: Protein language models are trained primarily with masked language modeling (MLM), which predicts masked amino-acid identities. Joint-embedding predictive architectures (JEPA) instead predict latent representations, but have not been applied to proteins. ProteinJEPA supplements MLM with a cosine loss for predicting the half-depth hidden states of a teacher given the unmasked sequence. On 19 tasks, with ESM2 at 35M and 150M parameters and three pretraining seeds, MLM+JEPA outperforms compute-matched and step-matched MLM-only continued training in 78 and 76 of 114 comparisons (14 losses, 22 ties). The median compute-matched gain is $+0.0106$ on structure- and homology-sensitive tasks versus $+0.0041$ elsewhere, led by SCOPe-40 retrieval and remote homology with improvements of 6.1 percentage points in Recall@1 and 2.7 points in accuracy, respectively. Gains on these tasks increase with model size from 8M to 150M. Against the off-the-shelf checkpoint, MLM+JEPA wins 81 of 114 comparisons (median $+0.0068$) without improving MLM loss. In random initialization the gain is smaller and replicates inconsistently across seeds ($p{=}0.059$). The same recipe improves the causal ProGen3 model, beating a compute-matched next-token-prediction control on 12 of 16 tasks. Ablations show that cosine loss beats mean squared error, while adding shallower targets removes most of the task gain. JEPA-only training collapses downstream performance: latent prediction complements MLM rather than replacing it. Code: https://anonymous.4open.science/r/protJepa-FF24

    https://arxiv.org/abs/2605.07554


    A Parameter-Free Few-Shot Evaluation for Elephant Vocalisation Classification

    oai:arXiv.org:2608.14824v2

    arXiv:2608.14824v2 Announce Type: replace-cross Abstract: We present a parameter-free episodic evaluation of nearest-centroid classification of elephant vocalisations on fixed pretrained embeddings, for the Elephant Voices (EV) and Linguistic Data Consortium (LDC) datasets. We ask not which embedding yields the best classifier trained on all labelled data, but how the simplest classifier performs as the number of exemplars per class varies. There are no learnable parameters, because each class is modelled as the mean of its support embeddings and each query is assigned to the nearest centroid under squared Euclidean distance. Evaluation covers the fixed Perch (ver. 1), Perch (ver. 2) and HuBERT (base, layer 2) embeddings, alongside mel frequency cepstral coefficient (MFCC) features, $N$-way $k$-shot, under the same stratified $K$-fold cross-validation protocol as the trained classifiers. None of these embedding models was trained to distinguish elephant call types. On the smaller EV dataset the centroid classifier is markedly data-efficient. Using Perch (ver. 1) or Perch (ver. 2) embeddings it overtakes in mean average precision (mAP) the fully-trained logistic regression (LR) baseline from one or two exemplars and the recurrent baseline from two. Over the reduced set of call types on which the strongly-supervised end-to-end baseline was trained, the centroid classifier using Perch (ver. 2) embeddings overtakes that baseline in mAP as well, from two exemplars. On the larger LDC dataset the recurrent baselines retain their advantage for all considered values of $k$. Only LR is overtaken, and only in mAP. Nearest-centroid classification is therefore preferable precisely when exemplars are few and the fixed embedding already separates the call types.

    https://arxiv.org/abs/2608.14824