Skip to content

Train model in Fast-Higashi and impute in Higashi with a much higher resolution #67

Description

@ypauling

Hi Ruochi,
Thanks for this great toolset!

The task I am trying to do with Higashi/Fast-higashi is to generate imputed contact matrices in high resolution (ideally 5k or 10k) in the pseudo-bulk level. I am wondering if this pipeline can be used for this goal:

  1. Use Fast-Higashi at low resolution (100k or even 500k) to generate embeddings and cell type annotations efficiently.
  2. Feed the pre-trained embeddings from Fast-Higashi into Higash model but with high resolution (5k or 10k).
  3. Extract imputation results from Higashi output.
  4. Use Merge2Cool tool to merge single cell high resolution imputed contact maps for a given cell type.

If this pipeline can work, I am also wondering if it saves resources at all or this is not different from using Higashi only for all steps.

We use snm3c for all our projects so that we usually create cell type annotations using the methylation data. Do you recommend running imputation/TAD/compartment calling at single cell level or directly using tools for bulk data on the pseudobulk merged contact files? I'd like to hear your thoughts and would appreciate any help you can provide.

Best,
Bing

Metadata

Metadata

Assignees

No one assigned

    Labels

    No labels
    No labels

    Projects

    No projects

    Milestone

    No milestone

    Relationships

    None yet

    Development

    No branches or pull requests

    Issue actions