
The Indian Cancer Genome Atlas (ICGA) Portal is the official data visualization and exploration platform of the Indian Cancer Genome Atlas.
Built on the open-source cBioPortal framework, the Portal hosts curated, harmonized, clinically annotated, multi-omics cancer datasets generated through ICGA studies. It enables researchers to explore genomic, transcriptomic, proteomic, and associated clinical datasets through an intuitive, browser-based interface while operating within ICGA’s governed data access framework.
The Portal currently hosts datasets from the ICGA Breast Cancer Programme and will expand to additional cancer types as new studies are completed.
The ICGA Portal provides researchers with secure, browser-based access to curated and processed cancer genomics datasets.
The Portal enables exploratory and hypothesis-driven research through interactive visualisation and analytical tools while ensuring compliance with applicable ethical approvals, institutional governance policies, and regulatory requirements.
The Portal visualises highly processed, curated, harmonised, secondary-level cancer genomics datasets. It is intended for research use and is not designed to host raw sequencing data.
The Portal has been developed with philanthropic support from Strand Life Sciences Ltd.
The portal provides secondary-level processed datasets, including:
Genomic Data
Transcriptomic Data
Proteomic Data
Clinical Data
All datasets are de-identified and curated before release.
ICGA adheres to strict ethical and regulatory standards in data sharing.
The current release contains clinically annotated breast cancer datasets generated across participating ICGA centres.
The cohort includes:
Additional cancer types will be incorporated in future releases.
Metadata of the ICGA’s cohort on breast cancer patients
The ICGA follows a controlled access model governed by the Data Access Committee (DAC) in alignment with ICGA Data Policy and DBT PRIDE guidelines.
Digital Personal Data Protection (DPDP)
ICGA is committed to implementing a privacy-by-design governance framework consistent with the principles of the Digital Personal Data Protection Act, 2023 and the Digital Personal Data Protection Rules, 2025.
Accordingly:
ICGA data is accessible through two routes. Both routes require a completed application and approval by the DAC before access is granted.
Route A — ICGA Portal (cBioPortal)
Interactive, browser-based access to processed and visualised datasets within the secure ICGA portal environment. Please remember no raw data would be available here. Researchers can explore somatic mutation profiles, gene expression patterns, proteomics summaries, and associated clinical metadata using the portal’s built-in analysis tools. Data remains within ICGA’s secure infrastructure at all times. [The portal visualizes highly processed, curated, and harmonized multidimensional cancer genomics data. ]
Route B — AWS Controlled Access / Secure Computational Workspace
Programmatic access to data files — VCF files, MAF files, expression matrices, and processed proteomics outputs — for researchers requiring computational analysis beyond what the portal interface supports. Access is provided within ICGA-managed, India-based infrastructure. Applicants are responsible for all associated AWS infrastructure costs. Contact ICGA for details.
Both routes are subject to a single consolidated application reviewed by the DAC.
Note: Commercial and industry applications are subject to additional review including execution of a Commercial Data Licensing Agreement. Contact suveera@icga.co.in with details of your intent before submitting any data request.
Researchers seeking access must submit a single consolidated application that includes:
All applications are reviewed by the ICGA Data Access Committee (DAC). Incomplete applications will not be considered. Full and final approval is followed by the signing of a Data User Agreement (DUA) with ICGA.
In cases where there is clear scientific justification and upon DAC approval, researchers may be granted access to processed data files within an ICGA-approved secure compute environment. Data does not leave ICGA’s governed infrastructure unless the DAC exceptionally approves it. The modality of access will be determined by ICGA on a case-by-case basis following DAC’s review.
Such requests must demonstrate:
Eligible file types include processed somatic variant data (VCF/MAF), normalised expression matrices, and proteomics outputs.
The following are not available for external access at this stage:
Requests for the above will not be considered under the current policy framework.
Note:
Given the strict regulatory framework of the Digital Personal Data Protection (DPDP) Act, your application must comprehensively address the following:
All approved users must comply with the following conditions throughout the approved access period:
Suggested citation:“The results [published or shown] here are based, in whole or in part, on data generated by the Indian Cancer Genome Atlas (ICGA) Network: https://icga.in, https://icga.net.in“
ICGA is dedicated to advancing cancer research through a rigorous, end-to-end process that involves:
Due to legal and ethical considerations, ICGA is unable to accommodate requests for biological samples, analytes, or tissue materials. All cases within the ICGA programme have been consented exclusively for ICGA use, and the redistribution of materials to outside parties is prohibited. Additionally, the majority of tissue samples have been depleted through the multiple assays performed for ICGA research.
The portal provides access to processed, secondary-level data, including:
Yes, provided your specific access request was approved with download. Data can be downloaded from:
Users can also define custom cohorts (“virtual studies”) using clinical or genomic filters and download the corresponding datasets.
No.
The Portal contains processed, secondary-level datasets only.
FASTQ, BAM, CRAM and other raw sequencing files are not available.
Access to controlled datasets requires application through the ICGA Data Access Committee (DAC).
This includes:
For queries, contact:
No. Since the portal provides processed, gene-level summarized data rather than raw count matrices, workflows requiring raw count reprocessing are not supported directly from portal downloads.
The portal supports:
Users can query specific genes or define cohorts before exporting results.
No. The ICGA Portal is an independent instance of the cBioPortal platform and hosts ICGA-specific datasets only.
No. The interface is designed to be intuitive. However, familiarity with cBioPortal workflows may be helpful for advanced queries and cohort analyses.
Yes. Users can apply clinical and genomic filters to create custom cohorts (“virtual studies”) for downstream exploration and data export.
Users are required to acknowledge the Indian Cancer Genome Atlas (ICGA) in any publications, presentations, or outputs derived from ICGA data.
Suggested citation:
“The results published or shown here are based, in whole or in part, on data generated by the Indian Cancer Genome Atlas (ICGA) Network: https://icga.in and https://icga.net.in.”
Where applicable, users should also cite associated ICGA publications relevant to the dataset used.
For citation-related queries, contact:
Yes. ICGA datasets are periodically updated as new data is generated, processed, and curated.
Breast cancer study data, for example, is routinely updated.
Users should record:
for reproducibility and future reference.
For dataset version queries, contact: