A recap of our live seminar, with the full recording.
On 15 May we hosted a free webinar that brought together deep technical insights and real world experience on how to build robust, sustainable and efficient bioinformatics workflows. The aim was clear: help teams scale their analyses, keep results reproducible and stay ahead of the accelerating pace of scientific discovery.
Also available directly on YouTube.
Dries De Maeyer
Single cell analysis simplified: insights from building a better bioinformatics workflow
Dries De Maeyer shared how the single cell analysis group at Johnson & Johnson matured its workflow strategy. The team began with highly flexible notebooks that supported rapid experimentation, but as data volumes increased and multimodal studies became standard, notebooks alone could not support scalability, consistency or long term reproducibility.
To address this, the group evolved step by step:
- notebooks
- scripts
- containerised components
- automated Nextflow workflows
Each step added more structure and stability. Containers ensured tool reproducibility over time, while Nextflow enabled parallel execution, consistent environments and large scale processing across thousands of samples.
Dries highlighted that success did not come from selecting a single best method, but from combining solid science with well designed, repeatable workflows. Today the J&J team delivers results within a day, maintains complete traceability and can reliably rerun analyses even years later. The outcome is a workflow ecosystem that is fast, predictable and built for regulated environments.
Toni Verbeiren
Building sustainable bioinformatics workflows and infrastructure
Toni Verbeiren followed with a deep dive into the tools and principles behind sustainable workflow engineering. Early collaborations with J&J revealed a key insight: workflow engines alone are not enough, teams also need a way to package, test, version and share components for workflows to remain maintainable.
This is exactly what Viash and the OpenPipeline ecosystem are designed for. Viash turns scripts and prototype code into production-ready building blocks that include:
- Clear inputs and outputs
- A containerised runtime
- Metadata and versioning
- Automated unit and integration tests
These components integrate into Nextflow workflows, forming modular analyses that run identically on a laptop, an HPC cluster, Kubernetes or a cloud environment. Viash Hub provides a structured catalogue where components can be published, discovered and reused, with support for both public and private instances.
This approach removes common pain points such as dependency problems, unstructured scripts, environment drift, unclear parameters and difficulty extending existing work. With versioned releases and containerised components, workflows evolve from prototypes into production-ready, tested tools.
Toni concluded with a clear message: sustainable bioinformatics requires platform thinking. With Viash and OpenPipeline, teams can build workflows that remain reproducible, maintainable and ready for real world scientific and clinical demands.