Papers/2608.12502
🧪 Test?View on arXiv

HIMEC: Directional Change Representation and Fixed-Interface Decoding for Remote Sensing Image Change Captioning

Aysha Ashra, Author 2, Author 3, Author 4, Author 5

multimodalimage captioningremote sensingdeep learning
2608.12502
Builder Relevance
80%
Aug 14

Abstract

HIMEC introduces a novel approach to remote sensing image change captioning by utilizing Directional Change Representation and fixed-interface decoding to improve semantic change description.

Reality Card

Core Claim

HIMEC achieves a CIDEr score of 142.81 on LEVIR-CC, outperforming traditional methods that use direct fused-feature memory.

Method / Result

Achieved a mean cosine distance of 0.69 on changed LEVIR-CC validation pairs.

Limitations

Findings are limited to the evaluated cascade, which may not generalize to other architectures.

Paper to code

Verified implementation resources so builders can test the paper’s claims instead of stopping at the abstract.

No verified implementation link has been attached yet. AIBuzzHub will keep this panel separate from unverified search results.
← Back to all papers