
Great work!
I have a few questions to clarify some details, and would really appreciate your help:
1. Is the 3M of CRSD referring to the features extracted using different encoders on the RS5M dataset?
2. Is the original dataset obtained from https://huggingface.co/datasets/Zilun/RS5M/tree/main?
3. Where can the original dataset corresponding to the 1M of LRSD be obtained?
4. Which method is used to obtain the vector database for the final CRSD and LRSD?
Thank you for your time and help!