PANDA Services
Data access and application support
PANDA supports investigators in identifying and accessing appropriate population-based cancer data sources, including SEER-Medicare, the Pennsylvania Cancer Registry (PACR), Pennsylvania Health Care Cost Containment Council (PHC4), National Cancer Database (NCDB), the UPMC Network Cancer Registry, and area-level or geospatial datasets. Support includes guidance on data source selection, feasibility, regulatory requirements, and data application processes. More information about the databases can be found here.
Study design and consultation
PANDA provides consultation on study design and analytic planning for studies using population-based cancer data. This may include defining study cohorts, identifying appropriate covariates and outcomes, developing analysis plans, and advising on the strengths and limitations of available datasets.
Data management and linkage
PANDA supports data management, cleaning, linkage, and analytic dataset creation. This can include preparing cancer registry data to answer a particular research question, integrating insurance claims data, assist with linkage of electronic health record derived, geospatial, or area-level variables that can be incorporated in downstream analyses.
Statistical analysis and reporting
PANDA provides analytic support for population-level cancer research, including descriptive analyses, regression modeling, outcomes analyses, tables and figures, and interpretation of findings for manuscripts, abstracts, presentations, and grant applications.
Geospatial and area-level analysis
PANDA supports integration and analysis of geospatial and area-level measures relevant to cancer surveillance, treatment access, exposures, and outcomes. This may include rurality, deprivation, provider availability, distance or travel-time measures, air pollution, and catchment-area relevant indicators.
Grant, manuscript, and presentation support
PANDA supports investigators who want to leverage population-based and area-level data for grants, manuscripts, abstracts, and presentations. This may include drafting data and methods sections, preparing preliminary data, developing budgets related to data access or analytic support, and creating tables or figures.
Data Sources
- Surveillance, Epidemiology, and End Results (SEER) Program
- SEER-Medicare Linked Data
- Pennsylvania Cancer Registry
- Pennsylvania Health Care Cost Containment Council Hospital Discharge Data
- National Cancer Database
- UPMC Network Cancer Registry
- Area-Level and Geospatial Datasets
- Additional data sources may be supported depending on project needs and feasibility. Contact PANDA to discuss potential data options.