-
Bioinformatics Support
Expert consultation and pipeline development for genomics, transcriptomics, proteomics, and other omics workflows. Pre-installed bioinformatics software stack.
View detailsBioinformatics Support
Expert consultation and pipeline development for genomics, transcriptomics, proteomics, and other omics workflows. Pre-installed bioinformatics software stack.
Pipelines & Workflows: Out-of-the-box execution for Nextflow and Snakemake (including full nf-core suite compatibility via Singularity).
Compute Specialization: High-memory nodes (up to 2 TB RAM) for de novo assembly and single-cell RNA-Seq, plus GPU-accelerated nodes for AlphaFold and structural modeling.
Data Management: High-speed parallel storage coupled with automated archival tiers for raw FASTQ, BAM/CRAM, and VCF files.
-
Scientific Software
Curated library of scientific software modules: Python, R, MATLAB, Julia, TensorFlow, PyTorch, and hundreds of domain-specific tools via the Lmod module system.
View detailsScientific Software
Curated library of scientific software modules: Python, R, MATLAB, Julia, TensorFlow, PyTorch, and hundreds of domain-specific tools via the Lmod module system.
This comprehensive ecosystem of software tools supports advanced research across disciplines including bioinformatics, imaging, AI/ML, and data science. These tools are optimized for high-performance computing (HPC) environments and enable scalable, reproducible, and efficient scientific workflows.
In addition to HPC resources, researchers also have access to a broader collection of software applications through the Weill Cornell Medicine Library Scientific Software Hub. See Learn More link -
Training & Workshops
Regular training sessions on HPC fundamentals, job scheduling, parallel programming, and scientific computing best practices. Open to all WCM researchers.
View detailsTraining & Workshops
Regular training sessions on HPC fundamentals, job scheduling, parallel programming, and scientific computing best practices. Open to all WCM researchers.
Registration: Scientific Computing Training Series -
Cloud Bursting
Seamless extension of on-premise HPC to AWS and Google Cloud for peak demand workloads. Hybrid infrastructure management and cost optimization support.
-
Data & Workflow Services
Support for researchers to efficiently organize, move, and process their own data at scale.
View detailsData & Workflow Services
Support for researchers to efficiently organize, move, and process their own data at scale.
Workflow Automation & Pipeline Development – Guidance and templates for building automated, scalable, and reproducible workflows using Snakemake, Nextflow, CWL, or custom scripts.
Data Transfer & Synchronization – Best practices and tools for secure, high-speed data movement between HPC clusters, cloud resources, and collaborator sites, including parallel transfers and checkpointed syncing for very large datasets.
Containerized Pipelines – Support for Singularity/Apptainer containers to make workflows portable and reproducible across systems.
Job Orchestration & Monitoring – Tips for efficient scheduling, monitoring, and optimizing pipeline execution on HPC and hybrid cloud resources.
Data Management Guidance – Recommendations on organizing, documenting, and versioning datasets to support reproducibility, collaboration, and long-term accessibility. -
Open On-Demand (OOD) Access
Web-based portal that allows researchers to interactively access HPC resources, manage jobs, and visualize results directly from a browser.
View detailsOpen On-Demand (OOD) Access
Web-based portal that allows researchers to interactively access HPC resources, manage jobs, and visualize results directly from a browser.
Interactive Sessions – Launch RStudio, Jupyter notebooks, MATLAB, or terminal sessions on the cluster without using SSH.
File Management & Transfers – Upload, download, and manage large datasets directly via the portal.
Job Submission & Monitoring – Submit SLURM jobs, monitor progress, and check outputs with a graphical interface.
Visualization Tools – Use remote desktop sessions or GPU-backed visualization for large-scale imaging and simulations.
User-Friendly Interface – Designed for researchers who prefer point-and-click access over command-line operations.
Secure & Integrated – Single sign-on access that works with existing HPC authentication and preserves data security.
-
AWS
At Weill Cornell Medicine, Information Technologies & Services (ITS) provides tailored pathways to utilize Amazon Web Services (AWS) for research, clinical data, and administrative applications while maintaining institutional security and compliance.
View detailsAWS
At Weill Cornell Medicine, Information Technologies & Services (ITS) provides tailored pathways to utilize Amazon Web Services (AWS) for research, clinical data, and administrative applications while maintaining institutional security and compliance.
At Weill Cornell Medicine (WCM), cloud services via AWS are structured into three distinct support tiers to align with technical expertise and project complexity:
SSOD (Standard Solutions on Demand): Fully managed by a dedicated support team. Designed for low-complexity, standalone instances (EC2, S3, RDS) on shared accounts with minimal user effort required.
Managed Solutions: Managed by an extended Cloud Team (covering Cloud, Security, Ops, and white-glove hosting). Ideal for medium-to-high complexity production apps and sensitive data handling.
Co-Managed Solutions: You manage your own dedicated account and resources directly, backed by basic infrastructure and monitoring support from the Cloud Team for a minimal service fee -
HPC Cluster Access
High-performance computing clusters for CPU and GPU workloads. Submit jobs via SLURM scheduler with access to thousands of cores and terabytes of RAM.
-
Large-Scale Storage
High-capacity, high-speed storage solutions for large datasets including genomics, imaging, and simulation outputs. NFS and Lustre parallel file systems available.
View detailsLarge-Scale Storage
High-capacity, high-speed storage solutions for large datasets including genomics, imaging, and simulation outputs. NFS and Lustre parallel file systems available.
Large-Scale Storage Access – Leverage over 26 PB of high-performance storage across Lustre, GPFS, and NFS systems for genomics, imaging, simulation, and other large datasets.
-
Cloud Bursting
Seamless extension of on-premise HPC to AWS and Google Cloud for peak demand workloads. Hybrid infrastructure management and cost optimization support.
-
HPC Cluster Access
High-performance computing clusters for CPU and GPU workloads. Submit jobs via SLURM scheduler with access to thousands of cores and terabytes of RAM.