AITopics | prophet attention

ProphetAttention: PredictingAttentionwithFutureAttention

Neural Information Processing SystemsFeb-7-2026, 13:45:42 GMT

However,foreachtimestepinthedecoding process, theattention based models usually use the hidden state of the current input to attend to the image regions.

artificial intelligence, machine learning, prophet attention, (19 more...)

Neural Information Processing Systems

Country:

Asia > China > Beijing > Beijing (0.04)
North America > Canada > British Columbia > Metro Vancouver Regional District > Vancouver (0.04)
Asia > China > Guangdong Province > Shenzhen (0.04)

Technology: Information Technology > Artificial Intelligence > Machine Learning (1.00)

Add feedback

Appendix of Prophet Attention

Neural Information Processing SystemsOct-2-2025, 04:30:34 GMT

CIDEr-c40, which is the default ranking score in the leaderboard, and rank the 1st. Compared with image captioning, the target of video captioning is the video clip, i.e., an ordered The dataset contain 10,000 video clips, and each video is paired with 20 annotated sentences. We use the official splits to report our results. CIDEr, which is built upon on n-gram matching, is used in our tests for performance evaluation. All re-implementations and our experiments were ran on V100 GPUs.

caption, machine learning, natural language, (18 more...)

Neural Information Processing Systems

Country: Asia > China (0.30)

Genre: Research Report > New Finding (0.35)

Technology:

Information Technology > Artificial Intelligence > Machine Learning > Neural Networks > Deep Learning (0.49)
Information Technology > Artificial Intelligence > Natural Language > Machine Translation (0.30)

Add feedback

13fe9d84310e77f13a6d184dbf1232f3-Paper.pdf

Neural Information Processing SystemsOct-2-2025, 04:30:27 GMT

artificial intelligence, machine learning, natural language, (21 more...)

Neural Information Processing Systems

Country: Asia > China (0.28)

Genre: Research Report (1.00)

Technology:

Information Technology > Artificial Intelligence > Natural Language (1.00)
Information Technology > Artificial Intelligence > Machine Learning > Neural Networks > Deep Learning (1.00)
Information Technology > Artificial Intelligence > Vision (0.99)

Add feedback

In the beginning, based on the Up-Down model, we have attempted to implement the Constant Prophet Attention

Neural Information Processing SystemsOct-2-2025, 04:23:20 GMT

We thank all the reviewers for the helpful comments. We will revise the paper to address your concerns. R1-Q1: The implementation seems straight-forward and the ablation analysis on the loss function. Thus we kept using L1 norm in the rest of experiments. We will conduct a systematic comparison between various loss functions in the next revision.

artificial intelligence, natural language, up-down model, (17 more...)

Neural Information Processing Systems

Technology: Information Technology > Artificial Intelligence > Natural Language (0.34)

Add feedback

Prophet Attention: Predicting Attention with Future Attention

Neural Information Processing SystemsOct-9-2024, 14:25:13 GMT

Recently, attention based models have been used extensively in many sequence-to-sequence learning systems. Especially for image captioning, the attention based models are expected to ground correct image regions with proper generated words. However, for each time step in the decoding process, the attention based models usually use the hidden state of the current input to attend to the image regions. Under this setting, these attention models have a deviated focus'' problem that they calculate the attention weights based on previous words instead of the one to be generated, impairing the performance of both grounding and captioning. In this paper, we propose the Prophet Attention, similar to the form of self-supervision.

future attention, image region, prophet attention, (1 more...)

Neural Information Processing Systems

Technology: Information Technology > Artificial Intelligence > Machine Learning (0.62)

Add feedback

Prophet Attention: Predicting Attention with Future Attention for Image Captioning

Liu, Fenglin, Ren, Xuancheng, Wu, Xian, Fan, Wei, Zou, Yuexian, Sun, Xu

arXiv.org Artificial IntelligenceApr-11-2023

Recently, attention based models have been used extensively in many sequence-to-sequence learning systems. Especially for image captioning, the attention based models are expected to ground correct image regions with proper generated words. However, for each time step in the decoding process, the attention based models usually use the hidden state of the current input to attend to the image regions. Under this setting, these attention models have a "deviated focus" problem that they calculate the attention weights based on previous words instead of the one to be generated, impairing the performance of both grounding and captioning. In this paper, we propose the Prophet Attention, similar to the form of self-supervision. In the training stage, this module utilizes the future information to calculate the "ideal" attention weights towards image regions. These calculated "ideal" weights are further used to regularize the "deviated" attention. In this manner, image regions are grounded with the correct words. The proposed Prophet Attention can be easily incorporated into existing image captioning models to improve their performance of both grounding and captioning. The experiments on the Flickr30k Entities and the MSCOCO datasets show that the proposed Prophet Attention consistently outperforms baselines in both automatic metrics and human evaluations. It is worth noticing that we set new state-of-the-arts on the two benchmark datasets and achieve the 1st place on the leaderboard of the online MSCOCO benchmark in terms of the default ranking score, i.e., CIDEr-c40.

artificial intelligence, machine learning, prophet attention, (20 more...)

arXiv.org Artificial Intelligence

2210.10914

Country: