Dear @wingrune,
Thank you for sharing such brilliant work! I have a couple of questions regarding the Semantic Segmentation Loss:
- Absence of the target object in the input view: I am curious about how often the target object actually appears in the input image. When using PPO, the trajectory is random, meaning the target object might not exist within the view for the entire trajectory. If this happens, does the Semantic Segmentation Loss still function effectively?
- More advanced segmentation models: Table 4 shows that the influence of the segmentation mask is crucial. Have you considered trying stronger, foundational models like SAM or FastSAM to generate the binary masks?
Thank you for your time, and I look forward to your reply!
Dear @wingrune,
Thank you for sharing such brilliant work! I have a couple of questions regarding the Semantic Segmentation Loss:
Thank you for your time, and I look forward to your reply!