The Mechanistic Invariance Test: Genomic Language Models Fail to Learn Positional Regulatory Logic

Bryan Cheng; Jasper Zhang

arXiv:2604.06549·q-bio.GN·April 9, 2026

The Mechanistic Invariance Test: Genomic Language Models Fail to Learn Positional Regulatory Logic

Bryan Cheng, Jasper Zhang

PDF

TL;DR

Genomic language models excel at tasks but fundamentally fail to learn the positional regulatory logic of gene regulation, instead relying on surface statistical correlations.

Contribution

Introduction of the Mechanistic Invariance Test (MIT), a benchmark to distinguish true positional understanding from statistical shortcuts in genomic language models.

Findings

01

Models fail to learn genuine positional regulatory logic.

02

Surface statistics dominate model predictions, not biological mechanisms.

03

A simple position-aware PWM outperforms billion-parameter models.

Abstract

Genomic language models (gLMs) have transformed computational biology, achieving state-of-the-art performance across genomic tasks. Yet a fundamental question threatens the foundation of this success: do these models learn the mechanistic principles governing gene regulation, or do they merely exploit statistical shortcuts? We introduce the Mechanistic Invariance Test (MIT), a rigorous 650-sequence benchmark across 8 classes with scrambled controls that enables clean discrimination between compositional sensitivity and genuine positional understanding. We evaluate five gLMs spanning all major architectural paradigms (autoregressive, masked, and bidirectional state-space models) and uncover a universal failure mode. Through systematic mechanistic probing via AT titration, positional ablation, spacing perturbation, and strand orientation tests, we demonstrate that apparent compensation…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.