Researchers in Taiwan, Korea and at Brookhaven Test a Quantum Attention Model on Merck's Alzheimer's Drug Verubecestat: arXiv 2610.04588, 3 October 2026
Quentir Medicine Monitor
Evidence-based insights for quantum medicine. Published by Quentir Systems LLC · October 8, 2026.

Merck stopped its large Alzheimer's trial of verubecestat early, a pill designed to block the enzyme BACE1, which helps produce the amyloid found in the brains of patients. The drug had been carefully refined molecule by molecule, yet in a trial involving 1,958 participants with mild-to-moderate disease it did not reduce cognitive or functional decline compared with placebo.
That chemistry now has a second life as a test set. On 3 October 2026 five researchers from National Yang Ming Chiao Tung University, KAIST, Taiwan's National Center for High-Performance Computing, Brookhaven National Laboratory and Taipei Medical University posted a study of quantum attention for drug discovery. They used the compound series that led to verubecestat as a case study, asking whether a model with a small quantum circuit inside it orders those molecules by potency the way the laboratory did. The same models were also scored on standard benchmarks, including whether a molecule can cross the blood-brain barrier.
The paper, "Variational Quantum Attention for Molecular Graph Learning," by Yu-Cheng Lin, Yu-Chao Hsu, Tai-Yue Li, Nan-Yow Chen and Samuel Yen-Chi Chen, is a preprint on arXiv (2610.04588) and has not yet been peer reviewed. For a pharmacologist, a hospital formulary lead or a research director weighing quantum computing claims in drug discovery, its value lies in how plainly it reports what changed and what did not.
Software that predicts molecular properties is already a routine filter in early drug discovery. Chemists draw large numbers of candidate structures, and models rank them for solubility, brain penetration or binding strength before anyone spends money on synthesis and testing. A poor ranking sends chemists after the wrong molecules. The question this paper asks is narrow and useful: what happens to such a model when one small part of it is replaced by a quantum circuit?
What Lin, Hsu, Li, Chen and Chen changed inside a graph attention network, using six or seven qubits
The researchers worked with a graph attention network, a common type of model in which atoms are points and chemical bonds are the lines between them. As the model passes information along the bonds, an attention step decides how much weight each neighboring atom deserves. The team replaced only that scoring step with a variational quantum circuit, a short sequence of adjustable quantum operations. The receiving atom, the neighboring atom and the bond between them are combined, encoded into the circuit and turned into an attention weight. Everything else in the model stays classical.
The design keeps the quantum part small. The combined atom-and-bond description has 64 or 128 numbers, which the authors encode into the amplitudes of six or seven qubits. A direct angle encoding would have needed 64 or 128 qubits. They call the result QGAT and compare it with GATv2, an otherwise identical classical model, and with six further classical baselines.
All quantum circuits in the paper were run in classical simulation on conventional computers; none ran on a quantum processor. The authors say this directly in their list of limitations. They also tested how the trained model behaves when simulated hardware noise is added afterward. Performance stayed close to the noise-free result at error rates up to 1 percent per operation. At stronger noise the loss depended on the task: sharp for solubility and BACE1 activity at 10 percent and above, gradual for others, and barely visible for blood-brain barrier prediction.
How QGAT scored on five drug discovery benchmarks: blood-brain barrier prediction rose from 0.6808 to 0.7157
The team tested five tasks: two classification sets, BACE (BACE1 binding, 1,513 molecules) and BBBP (blood-brain barrier penetration, 2,039 molecules); two regression sets for solubility and hydration energy, ESOL and FreeSolv; and a BACE1 activity set of 4,580 compounds built from the ChEMBL database. Molecules were split by core scaffold, so the test compounds had structures the model had not seen during training, a stricter test than a random split.
These figures come from the strongest circuit configuration for each task, averaged over five scaffold folds; no single configuration produced all of them. The clearest gain came on BBBP, where the mean ROC-AUC, a score in which 0.5 is chance and 1.0 is perfect, rose from 0.6808 (plus or minus 0.0111) for the classical twin to 0.7157 (plus or minus 0.0095) for QGAT, and it improved under all six circuit designs the team tried. QGAT also scored slightly higher on BACE, ESOL and the ChEMBL activity set, while the classical model kept a slightly lower error on FreeSolv. The quantum version used about 0.5 to 1.3 percent fewer trainable parameters. The authors call the effect task dependent, and their own ablations support that reading: no single circuit design won everywhere, and adding circuit layers did not steadily help.
This echoes an earlier experiment the Monitor read, in which researchers swapped quantum circuits into one component of a molecule generator at a time and found that the quantum molecule generator hit a chemistry limit on validity even as other scores improved. Both studies swap a single component, which makes it possible to see where a circuit helps and where it does not.
Quantum pillar: computing. Technology readiness: TRL 2 of 9. The quantum circuit exists only as a classical computer simulation trained on public chemistry datasets; it has not run on quantum hardware, and no compound it ranked has been made or tested because of it.
What the model saw in Merck's verubecestat series: a Spearman correlation of 0.8264 against 0.7954
The case study draws on the series Merck chemists published in 2016 on their way to verubecestat (MK-8931). None of the compounds appears exactly in the training data, although one analogue is a close match, with a similarity score of 0.92, so the authors call the series an external case study and do not claim it as a strict test on unfamiliar chemistry. Spearman correlation measures whether a model puts compounds in the right order of potency, which is what a chemist choosing the next molecule to make actually needs. QGAT reached 0.8264, against 0.7954 for the classical model. The classical model had a slightly lower numerical error, but the models were trained mostly on IC50 measurements while the series reports Ki values, so the authors rely on the ranking and give the error only as a reference.
The authors then looked at which parts of the molecule each model relied on. QGAT placed weight on the fluoropyridine group, the central amide, the ring system whose hydrogen bonds reach the enzyme's two catalytic aspartate residues, and the fluorophenyl group. Those are the regions seen engaging BACE1 in the crystal structure of the enzyme bound to verubecestat, Protein Data Bank entry 5HU1. In the attribution analysis, QGAT gave a positive contribution to the fluorine on the pyridine ring, matching Merck's finding that this substitution improved potency, where the classical model gave it a negative one.
The paper is candid about the limits. Both models missed the roughly fivefold potency gain from the fluorine on the phenyl ring. The model sees only the drug molecule, never the protein, so the authors treat the overlap with the binding site as a qualitative alignment only. On two further Alzheimer's compound series, QGAT led modestly on umibecestat, 0.692 against 0.655, and the two models were level on LY3202626, 0.788 against 0.784.
Why a better ranking of BACE1 inhibitors says nothing yet about Alzheimer's patients: the 2018 NEJM trial
Verubecestat shows the distance between a well-ranked molecule and a useful medicine. In the 78-week trial reported by Egan and colleagues in the New England Journal of Medicine in May 2018, patients took 12 mg or 40 mg a day or placebo. The paper notes that verubecestat lowers amyloid levels in the spinal fluid of patients, yet cognitive scores declined about equally in every group, and rash, falls, sleep disturbance and suicidal ideation were more common on the drug. The trial was stopped early for futility.
Ranking compounds by potency against BACE1 is one early step; whether lowering BACE1 activity helps patients is a separate question that only clinical trials answer. A model that orders those analogues more faithfully could at most help chemists choose molecules sooner. It cannot tell them whether the target itself will pay off in the clinic. Property prediction sits early in a long chain, and its value depends on whether the biology downstream is right. For clinicians and hospital buyers, the useful reading is that quantum attention in this paper is a modeling choice for medicinal chemistry, measured on public benchmarks, with no bearing yet on which treatments reach the ward.
What would move quantum attention closer to a medicinal chemistry team: hardware runs, protein-aware models and more compound series
The authors name the next steps themselves: models that include the protein and its binding site, quantitative tests of whether attention and attribution match known chemistry, and broader validation across chemical series and targets. A run of the same six- or seven-qubit circuits on real quantum processors, with the noise results compared against the simulation, would show whether the BBBP gain survives hardware. An independent group reproducing the scaffold-split results would carry more weight than any single preprint.
For now, the 3 October preprint shows that a small quantum circuit can stand in for a classical attention scorer in a drug discovery model at comparable accuracy, with one consistent gain on blood-brain barrier prediction and a better potency ranking on one historic Alzheimer's series.
Sources
Primary source: Yu-Cheng Lin, Yu-Chao Hsu, Tai-Yue Li, Nan-Yow Chen and Samuel Yen-Chi Chen, Variational Quantum Attention for Molecular Graph Learning, arXiv preprint 2610.04588, 3 October 2026. Also drawn on: Egan et al., New England Journal of Medicine, 2018; Scott et al., Journal of Medicinal Chemistry, 2016; and RCSB Protein Data Bank entry 5HU1. Readiness and implications are the Monitor's editorial assessments.