• HOME
  • NEWS
  • EXPLORE
    • CAREER
      • Companies
      • Jobs
    • EVENTS
    • iGEM
      • News
      • Team
    • PHOTOS
    • VIDEO
    • WIKI
  • BLOG
  • COMMUNITY
    • FACEBOOK
    • INSTAGRAM
    • TWITTER
Thursday, October 8, 2026
BIOENGINEER.ORG
No Result
View All Result
  • Login
  • HOME
  • NEWS
  • EXPLORE
    • CAREER
      • Companies
      • Jobs
        • Lecturer
        • PhD Studentship
        • Postdoc
        • Research Assistant
    • EVENTS
    • iGEM
      • News
      • Team
    • PHOTOS
    • VIDEO
    • WIKI
  • BLOG
  • COMMUNITY
    • FACEBOOK
    • INSTAGRAM
    • TWITTER
  • HOME
  • NEWS
  • EXPLORE
    • CAREER
      • Companies
      • Jobs
        • Lecturer
        • PhD Studentship
        • Postdoc
        • Research Assistant
    • EVENTS
    • iGEM
      • News
      • Team
    • PHOTOS
    • VIDEO
    • WIKI
  • BLOG
  • COMMUNITY
    • FACEBOOK
    • INSTAGRAM
    • TWITTER
No Result
View All Result
Bioengineer.org
No Result
View All Result
Home NEWS Science News Technology

CNN or Transformer? AI Duel Reveals Best Way to Read X-ray Patterns

by
October 8, 2026
in Technology
Reading Time: 5 mins read
0
CNN or Transformer? AI Duel Reveals Best Way to Read X-ray Patterns

CNN or Transformer? AI Duel Reveals Best Way to Read X-ray Patterns

Share on FacebookShare on TwitterShare on LinkedinShare on RedditShare on Telegram

Every year, the world’s factories and incinerators bury mountains of value in plain sight. Ashes from municipal waste plants, slags from metallurgical furnaces, and shredded residues from dead batteries and old electronics all contain minerals that could be recovered, reused, or safely locked away—but only if engineers know exactly what crystalline phases lurk inside them. A new study published in Results in Engineering by Yi Zhang and colleagues now shows that the choice of artificial intelligence architecture can make or break that knowledge, in a head-to-head contest between two of deep learning’s most celebrated designs.

The team, based at the AI Production Network Augsburg in Germany, set out to answer a deceptively simple question: when a machine reads an X-ray diffraction pattern, is a convolutional neural network or a Vision Transformer better at working out both which minerals are present and how much of each one the sample contains? Until now, no one had compared the two architectures under truly identical conditions on this task, because earlier studies had split the problem into separate stages—first identifying phases, then estimating their fractions—which made fair comparison impossible.

X-ray diffraction is the gold standard for quantitative phase analysis. When X-rays strike a finely ground powder, the crystal planes inside each mineral scatter the beam into a fingerprint of peaks at characteristic angles. The trouble is that real industrial residues produce nightmarish spectra: quartz and cristobalite, two forms of silicon dioxide, scatter at overlapping angles; cuprite and tenorite, both copper oxides, share diffraction features; and zirconia contaminating samples from grinding media produces a dense thicket of peaks that obscures everything beneath it. The traditional answer, Rietveld refinement, works well but demands expert users, careful initial parameters, and hours of computation for every sample—a bottleneck that simply cannot scale to the torrent of waste streams requiring characterization.

Deep learning promises to collapse that bottleneck, but the field has been divided over how to feed diffraction data to a neural network. Some researchers convert the one-dimensional intensity curve into a two-dimensional image and apply image-recognition networks, borrowing mature architectures from computer vision. That approach, however, discards spectral resolution through downsampling and can smear away the fine peak structure that carries the quantitative information. Zhang’s team instead treated the diffraction pattern as what it truly is: a one-dimensional sequence of intensities over diffraction angle, preserving its native structure.

Both contenders in the study shared the same multi-task skeleton. A feature extractor digests the raw spectrum into a compact embedding, which then feeds two parallel heads: a classification head that predicts whether each of seven target phases is present, and a regression head that estimates the weight fraction of every phase, constrained to sum to one. Crucially, the regression output is masked by the classification predictions, so phases judged absent automatically receive zero fraction. The models trained end-to-end on a composite loss combining binary cross-entropy for classification with mean squared error for regression. The convolutional network stacked three one-dimensional convolutional layers with shrinking kernel sizes to capture local peak shapes, while the Vision Transformer chopped the spectrum into patches and applied self-attention, letting every part of the spectrum communicate with every other part regardless of distance.

Because real labeled measurements are scarce and expensive, the researchers built a hybrid data strategy. They simulated tens of thousands of X-ray patterns from crystallographic data for seven phases—cristobalite, cuprite, graphite, copper, quartz, tenorite, and zirconia—using pseudo-Voigt peak profiles, exponentially decaying backgrounds, amorphous humps, and phase-specific noise calibrated so that each phase became measurably harder to detect, mimicking real laboratory degradation. The noise parameters were not guessed; they were optimized so that a template-based classifier’s detectability threshold degraded by roughly five percentage points per phase, a rigorous way of ensuring the simulated data posed genuine difficulty. On top of this synthetic mountain of 50,000 patterns sat a deliberately small set of just 105 real laboratory measurements of four reference phases, reflecting the data-poor reality of industrial practice.

The results, obtained under identical training data, loss functions, hyperparameter tuning, and evaluation metrics, tell a nuanced story. On purely simulated data, both architectures were nearly flawless at classification, with accuracy above 0.999, but the CNN reached its best regression performance early and plateaued, while the Transformer kept improving as the dataset grew, only converging with the CNN at the full 50,000 samples. The real drama unfolded on laboratory measurements. There, the Vision Transformer delivered perfect phase classification across all four reference phases and cut the regression error dramatically—reducing root mean square error by 31 percent and mean absolute error by 35 percent compared with the CNN. The lone exception was quartz, where the convolutional network held its ground, likely because quartz’s sharp, well-resolved peaks reward the local, receptive-field approach of convolutions.

The transfer experiments added another twist. When models trained only on simulated data were applied directly to real measurements without any adaptation—a zero-shot test—the CNN actually transferred more robustly than the Transformer. But a modest amount of fine-tuning on real samples largely closed the simulation-to-real gap for both architectures, and the Transformer benefited most from every additional real sample, achieving its best results when 72 real measurements were used for adaptation. Visualizations of the models’ internal feature spaces using UMAP projections revealed why: the Transformer’s embeddings formed compact, well-separated clusters in which simulated and real spectra of the same phase already sat close together before any fine-tuning, whereas the CNN’s feature space remained diffuse and only partially realigned after adaptation. The convolutional network’s predictions, however, proved more stable across repeated random splits of the small real dataset, suggesting its strong built-in assumptions about local peak patterns act as a safeguard when data is scarce.

For practitioners, the message is refreshingly practical rather than doctrinaire. If labeled real measurements are available for fine-tuning, the Vision Transformer is the stronger choice, especially for materials dominated by broad, overlapping, or poorly crystalline features such as graphite’s smeared reflections. If the goal is zero-shot deployment on new instruments or new sample types with no adaptation data, the humble CNN remains a formidable competitor. Architecture selection, in other words, should follow the data budget and the spectral character of the target phases—not fashion.

The broader implications reach well beyond waste recycling. Quantitative phase analysis underpins everything from cement quality control to pharmaceutical polymorph screening to planetary geology, and any technique that removes the expert bottleneck from Rietveld refinement multiplies the throughput of entire laboratories. The Augsburg team is candid about the limits of the current work: only four of the seven simulated phases received experimental validation, the real dataset contains just 105 patterns, and the framework assumes a closed set of phases that may not hold in open-ended industrial screening. Future work will expand the real measurements, incorporate explicit domain adaptation techniques such as adversarial training, and add uncertainty quantification to flag low-confidence predictions. But the core finding stands as a milestone in applied machine learning: in the contest between convolution and attention, the winner depends on where you stand—and now, for the first time, scientists know exactly where that line is drawn.

Subject of Research: Deep learning architectures for quantitative phase analysis of X-ray diffraction patterns from industrial mineral residues

Article Title: A comparative study of convolutional neural networks and vision transformers for quantitative phase analysis from X-ray diffraction

Article References: Zhang, Y., Eulitz, S., Michalak, A., Vollprecht, D., & Mikelsons, L. (2026). A comparative study of convolutional neural networks and vision transformers for quantitative phase analysis from X-ray diffraction. Results in Engineering, 32, Article 113313. https://doi.org/10.1016/j.rineng.2026.113313

Image Credits: AI Generated

DOI: 10.1016/j.rineng.2026.113313

Keywords: X-ray diffraction, quantitative phase analysis, deep learning, convolutional neural networks, vision transformers, mineral residues, recycling, Rietveld refinement, transfer learning, simulation-to-real, multi-task learning, self-attention

News Source: Blake Davidson. (October 8, 2026). CNN or Transformer? AI Duel Reveals Best Way to Read X-ray Patterns. Scienmag.

Tags: convolutional neural networksdeep learningmineral residuesMulti-task learningquantitative phase analysisrecyclingRietveld refinementself-attentionsimulation-to-realTransfer Learningvision transformersX-ray diffraction
Share12Tweet7Share2ShareShareShare1

Related Posts

Nanoparticles That Act as Frequency Mixers Could Transform Background-Free Bioimaging

Nanoparticles That Act as Frequency Mixers Could Transform Background-Free Bioimaging

October 8, 2026
A Simple Blood-Cell Index May Help Flag Heart Risk in Kawasaki Disease

A Simple Blood-Cell Index May Help Flag Heart Risk in Kawasaki Disease

October 8, 2026

Sunlight Heats, the Sky Cools: Flexible Cell Turns Temperature Gap into Tripled Power

October 8, 2026

Physicists Unveil Universal Speed Limits on How Fast Quantum Systems Decay

October 8, 2026

POPULAR NEWS

  • Alloys That Shrink Their Own Grains: New PIX Mechanism Refines Metals With Heat Alone

    Alloys That Shrink Their Own Grains: New PIX Mechanism Refines Metals With Heat Alone

    29 shares
    Share 12 Tweet 7
  • Endurance Exercise Reshapes the Liver in Males and Females Through Distinct Molecular Routes

    29 shares
    Share 12 Tweet 7
  • Single Transcription Factor PU.1 Rapidly Converts Fibroblasts into Macrophage-Lineage Cells

    29 shares
    Share 12 Tweet 7
  • New Scale Measures How Ready Nurse Educators Really Are for the AI Era

    29 shares
    Share 12 Tweet 7

About

We bring you the latest biotechnology news from best research centers and universities around the world. Check our website.

Follow us

Recent News

Alloys That Shrink Their Own Grains: New PIX Mechanism Refines Metals With Heat Alone

Endurance Exercise Reshapes the Liver in Males and Females Through Distinct Molecular Routes

Single Transcription Factor PU.1 Rapidly Converts Fibroblasts into Macrophage-Lineage Cells

Subscribe to Blog via Email

Success! An email was just sent to confirm your subscription. Please find the email now and click 'Confirm' to start subscribing.

Join 85 other subscribers
  • Contact Us

Bioengineer.org © Copyright 2023 All Rights Reserved.

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • Homepages
    • Home Page 1
    • Home Page 2
  • News
  • National
  • Business
  • Health
  • Lifestyle
  • Science

Bioengineer.org © Copyright 2023 All Rights Reserved.