2. Denis Bauer,
PhD
Oscar Luo,
PhD
Rob Dunne,
PhD
Piotr SzulAidan O’BrienLaurence Wilson,
PhD
Adrian White
Andy Hindmarch
David Levy
Dan Andrews
Kaitao Lai,
PhD
Arash Bayat
PhD
John Hildebrandt
Mia Chapman
Ian Blair
Kelly Williams
Jules Damji
Gaetan Burgio
Lynn Langit
Jim Counts
Matthew Jones
Natalie Twine,
PhD
Prabha Pillay
Transformational Bioinformatics Team
www.csiro.au
Denis C. Bauer | @allPowerde
25. CSIRO Team Trains Other Researchers
Team creates reproducible
cloud environments
• AWS CloudFormation Templates for
EMR
• Setup screencasts for Databricks
and AWS
• Scripts and recommended
parameters
26. Next Steps
• VariantSpark on GCP
• Use GCP DataProc – compare to AWS EMR
• Use GCP GKE – compare to AWS EKS (K8)
• VariantSpark on Terra.bio
• First optimize container for GCP raw compute
• Write WDL for VariantSpark tool/workflow
• Publish on Dockstore as Tool/Workflow
• Publish example Jupyter notebooks
• Publish example Terra.bio VariantSpark
workflow