Articles

Implementing AI to Streamline Chromatographic Peak Picking

Peak picking can see an efficiency boost with AI, but what is required to get up and running?
Updated
Written byHolden Galusha
Register for free to listen to this article
Listen with Speechify
0:00
5:00

Traditionally, scientists have manually reviewed chromatograms to identify and quantify compounds in a sample. The process involves establishing a baseline to represent the signal in the absence of analytes, setting a threshold to differentiate true peaks from noise, and inspecting the chromatogram for peaks that meet these criteria. Identified peaks are validated by comparing their retention times, shapes, and spectral characteristics against known standards or reference materials.

As chromatographs shifted to record chromatograms digitally instead of printing them onto long rolls of paper, algorithms were introduced in the chromatography software to help identify peaks and minimize noise. The software performs a preliminary peak identification, and the user confirms its validity.

While these algorithms are helpful, their utility is limited—after all, algorithms are static and cannot be tailored to an organization’s unique data. This is where artificial intelligence (AI) can take the baton.

For a broader look at where AI is entering the chromatographic workflow — from automated method scouting and retention-time prediction through to gradient optimisation and method transfer — the guide to AI-assisted chromatographic method development provides the wider context this article sits within. For the full landscape of AI across analytical science, the practical guide to AI and machine learning for separation scientists is a useful starting point

Working in analytical science?

Register for a FREE Separation Science account to subscribe to the Separation Science Newsletter.

Subscribe for free

Comparing Manual vs AI Peak Picking

It’s important to acknowledge that every chromatography lab is different, so some may benefit greatly from AI while others may be hindered by it. Here are two breakdowns illustrating the pros and cons of manual and AI-powered peak picking, respectively:

Manual peak picking

ProsCons
Quality assurance: Scientists personally ensure their data is accurate and precise.Time-consuming: Peak picking is a meticulous task that demands time, focus, and experience.
Full context: Humans can examine data in full context, weighing factors that may not necessarily be quantified, in a way that AI cannot.Cognitive limits: While humans are excellent at recognizing patterns, they may not be as good as AI. Some chromatography data may contain patterns that only AI can perceive.

AI peak picking

ProsCons
Automated: AI automates the peak picking process, slicing turnaround time and saving scientists time and energy to focus on work that requires a human touch.No contextual understanding: AI models cannot know the inevitable nuances of a situation, which may result in output that seems appropriate on paper but does not necessarily reflect reality.
Accuracy: As mentioned, some data may contain patterns imperceptible to humans. In these cases, introducing AI can heighten accuracy as the AI will identify those patterns and account for them in the final output.Lack of transparency: You cannot step through an AI model’s “reasoning.” you can only infer how it arrived at a particular output. While this is sufficient for many applications, it requires scientists to exercise caution with AI.

Let’s say you’ve reviewed the pros and cons of both approaches, built a business case, and received approval to introduce AI in your peak picking process. What are the next steps?

What Do You Need to Implement AI?

To roll out an AI solution, you’ll need more than just chromatograms. There are two key ingredients to implementing AI: annotated, high-quality data and a platform to access the solution.

High-quality data

Data is the lifeblood of any AI solution. Commercial AI solutions generally come pre-trained on the vendor’s data set. While this is useful and sufficient for many use cases, labs should consider investing the time into collecting and annotating their own data, which will enable them to launch an AI solution tailored for them.

Annotation is the process of defining and differentiating between elements of interest in a data set, which will enable the AI model to learn those patterns and draw the same distinctions in future analyses.

After annotation, you can then further train the AI model on your lab’s data in a process known as fine-tuning. “Fine-tuning a model on a lab’s own data is highly beneficial as it increases accuracy by tailoring the model to the specific characteristics and nuances of the lab’s data,” says Edison Cerda, product manager, informatics at Agilent Technologies. “It also improves relevance,” Cerda continues, “ensuring the model’s outputs are more actionable for your specific use cases.” Annotating your lab’s data and fine-tuning an AI model on it is essential to your AI strategy.

The data consistency question connects directly to the upstream development workflow: methods developed with clean, well-annotated chromatographic data, as covered in the guides to automated method scouting and gradient optimisation, tend to produce training data that AI peak-picking models can learn from more reliably.

Continue reading below…
eBooksAbstract blue circle background
Evaluating Prep LC Phases for Selectivity and Cleaning Tolerance
A performance evaluation of hybrid silica stationary phases across 300 alkaline washes and scale-up fraction analysis.
Read More

Data platforms

Of course, all the data in the world means nothing if you can’t use it. You need a centralized platform that can (1) receive input to process data and (2) make the output readily available to users. This is not a straightforward process. “Integrating AI solutions with existing lab infrastructure and workflows can be complex and time-consuming,” Cerda says. Architecting the data pipelines that will enable the AI solution to receive data and then generate output requires technical expertise, deep collaboration with your organization’s IT staff, and a clear vision of how you will integrate the solution with your lab’s workflow.

Questions to consider for establishing data pipelines and access

How are you hosting your AI model, and how will your chromatography data make it to the model? Some chromatograph vendors offer AI peak picking solutions that are cloud-based and operate as a software-as-a-service, easily scaling with your lab’s usage. The downside is that subscribing to these services can be expensive, and they may only work with chromatographs made by that particular vendor, effectively locking you in to that vendor’s walled garden.

In contrast, there are also on-premises solutions available, such as the open-source project PeakBot. On-premises solutions do away with recurring expenses and have enhanced flexibility. They can analyze data from any chromatograph as long as that data is scrubbed and standardized. The trade-offs, however, lie in cost and convenience.

According to Cerda, high-performance servers are required to handle the data processing and model training of on-premises solutions. These servers represent a significant capital investment, though their recurring operating costs may be lower than the cost of subscribing to a cloud AI platform.

The other factor to consider is convenience. A cloud solution makes access easy —it’s either integrated with the chromatograph’s software or available via a web browser. Meanwhile, an on-premises server demands network bandwidth, security, and a workstation to access the platform. All these factors must be weighed against each other when deciding if you should opt for a cloud or on-premises solution.

Continue reading below…

After deciding on a solution, you must then set up the data pipeline to move data from the chromatograph to the AI model. Cloud-based solutions will handle this out-of-the-box, piping all data into one model via the internet. If hosting on-premises, the ideal setup will hinge on the number of chromatographs you have, your internal network infrastructure, physical proximity, and other factors. One simple, but manual and time-consuming, solution is to export data from each chromatograph to a USB drive, walk it to the workstation hosting the AI model, and carry on from there. Ideally, you will be able to automatically send data to the model via your lab’s internal network without exposing it to the internet. Consider hiring a lab informatics consultant to identify the best solution according to your needs and budget.

For scientists also evaluating where peak picking fits within a broader AI-assisted analytical workflow, the companion guides on machine learning for retention time prediction and AI-assisted method transfer cover adjacent stages of the method development process.


Share Your Expertise! Would you like to share your knowledge and news with the broader analytical community? Sign up to our contributor list today!

Frequently Asked Questions (FAQs)

  • What is the traditional method of peak picking in chromatography?

    Traditionally, scientists manually review chromatograms to identify and quantify compounds, which involves establishing a baseline, setting thresholds to differentiate true peaks from noise, and validating the identified peaks against known standards.

  • How do AI algorithms improve the peak picking process?

    AI algorithms automate the peak picking process, reducing turnaround time and allowing scientists to focus on more complex tasks that require human insight, while potentially increasing accuracy by identifying patterns that humans may not perceive.

  • What are the necessary components to implement AI in chromatography?

    To implement AI in chromatography, you need annotated, high-quality data and a centralized platform capable of processing that data and generating outputs, which may include either cloud-based or on-premises solutions.

  • What are the pros and cons of manual peak picking versus AI peak picking?

    Manual peak picking offers quality assurance and full context understanding, but is time-consuming and limited by human cognitive capacity. AI peak picking is automated and can enhance accuracy but lacks contextual understanding and transparency in its decision-making.

  • Why is data annotation important for AI in chromatography?

    Data annotation is crucial as it defines and differentiates elements within a data set, enabling the AI model to learn patterns effectively, which is essential for the model to provide more accurate and relevant outputs during analysis.

Add Separation Science as a preferred source on Google

Add Separation Science as a preferred Google source to see more of our trusted coverage

Published In

Redefining Analytical Boundaries magazine issue cover
March 2025

Redefining Analytical Boundaries

Welcome to Separation Science’s March 2025 issue, ‘Redefining Analytical Boundaries,’ released to coincide with Pittcon 2025. This issue serves as your guide to some of the latest breakthroughs and important discussions in analytical science.

Meet the Author(s):

  • Separation Science Placeholder Image
    Holden Galusha is the associate editor for Lab Manager. He was a freelance contributing writer for Lab Manager before being invited to join the team full-time. Previously, he was the content manager for lab equipment vendor New Life Scientific, Inc., where he wrote articles covering lab instrumentation and processes. Additionally, Holden has an associate of science degree in web/computer programming from Rhodes State College, which informs his content regarding laboratory software, cybersecurity, and other related topics. You can reach him at hgalusha@labmanager.com.View Full Profile

Here are some related topics that may interest you:

Loading Next Article...
Loading Next Article...