Papers › Reflective Decoding Network for Image Captioning
Reflective Decoding Network for Image Captioning
Lei Ke, Wenjie Pei, Ruiyu Li, Xiaoyong Shen, Yu-Wing Tai
State-of-the-art image captioning methods mostly focus on improving visual features, less attention has been paid to utilizing the inherent properties of language to boost captioning performance. In this paper, we show that vocabulary coherence between words and syntactic paradigm of sentences are also important to generate high-quality image caption. Following the conventional encoder-decoder framework, we propose the Reflective Decoding Network (RDN) for image captioning, which enhances both the long-sequence dependency and position perception of words in a caption decoder. Our model learns to collaboratively attend on both visual and textual features and meanwhile perceive each word's relative position in the sentence to maximize the information delivered in the generated caption. We evaluate the effectiveness of our RDN on the COCO image captioning datasets and achieve superior performance over the previous methods. Further experiments reveal that our approach is particularly advantageous for hard cases with complex scenes to describe by captions.
In Syntology Open this paper in Syntology's Atlas, the map of the papers in Syntology's graph and their citations.
Code
No code repository is listed for this paper in the archive or in Syntology's graph.
Code Syntology ran Syntology
Not run by Syntology. Nothing on this page verifies that the listed code works.
Tasks
1 archive task tag without a task page not shown.
Results from the paper archive 2025-07-28
| Task | Dataset | Model | Metric | Value | Rank at snapshot | Leaderboard | Report |
|---|---|---|---|---|---|---|---|
| Image Captioning | COCO (Common Objects in Context) | RDN | CIDEr | 125.2 | #5 of 17 | Archive leaderboard | report |
| Image Captioning | COCO Captions | RDN | BLEU-1 | 80.2 | #28 of 41 | Archive leaderboard | report |
| Image Captioning | COCO Captions | RDN | BLEU-4 | 37.3 | #28 of 41 | Archive leaderboard | report |
| Image Captioning | COCO Captions | RDN | CIDER | 125.2 | #28 of 41 | Archive leaderboard | report |
| Image Captioning | COCO Captions | RDN | METEOR | 28.1 | #28 of 41 | Archive leaderboard | report |
| Image Captioning | COCO Captions | RDN | ROUGE-L | 57.4 | #28 of 41 | Archive leaderboard | report |
Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections