AITopics | query feature

Dual-Path Temporal Decoder for End-to-End Multi-Object Tracking

Neural Information Processing SystemsJun-14-2026, 11:41:27 GMT

We present a novel end-to-end transformer-based framework for Multiple Object Tracking (MOT) that advances temporal modeling and identity preservation. Despite recent progress in transformer-based MOT, existing methods still struggle to maintain consistent object identities across frames, especially under occlusions, appearance changes, or detection failures. We propose a dual-path temporal decoder that explicitly separates appearance adaptation and identity preservation. The appearance-adaptive decoder dynamically updates query features using current frame information, while the identity-preserving decoder freezes query features and reuses historical sampling offsets to maintain long-term temporal consistency. To further enhance stability, we introduce a confidence-guided update suppression strategy that retains previously reliable features when predictions are unreliable. Extensive experiments on MOT benchmarks demonstrate that our approach achieves state-of-the-art performance across major tracking metrics, with significant gains in association accuracy and identity consistency. Our results demonstrate the importance of decoupling dynamic appearance modeling from static identity cues, and provide a scalable foundation for robust tracking in complex scenarios.

artificial intelligence, machine learning, natural language, (16 more...)

Neural Information Processing Systems

Genre:

Research Report > Experimental Study (0.93)
Research Report > New Finding (0.68)

Technology:

Information Technology > Artificial Intelligence > Representation & Reasoning (1.00)
Information Technology > Artificial Intelligence > Machine Learning > Neural Networks (1.00)
Information Technology > Artificial Intelligence > Natural Language > Information Retrieval > Query Processing (0.57)

Add feedback

Dual-Path Temporal Decoder for End-to-End Multi-Object Tracking

Neural Information Processing SystemsJun-9-2026, 21:42:58 GMT

We present a novel end-to-end transformer-based framework for Multiple Object Tracking (MOT) that advances temporal modeling and identity preservation. Despite recent progress in transformer-based MOT, existing methods still struggle to maintain consistent object identities across frames, especially under occlusions, appearance changes, or detection failures. We propose a dual-path temporal decoder that explicitly separates appearance adaptation and identity preservation. The appearance-adaptive decoder dynamically updates query features using current frame information, while the identity-preserving decoder freezes query features and reuses historical sampling offsets to maintain long-term temporal consistency. To further enhance stability, we introduce a confidence-guided update suppression strategy that retains previously reliable features when predictions are unreliable. Extensive experiments on MOT benchmarks demonstrate that our approach achieves state-of-the-art performance across major tracking metrics, with significant gains in association accuracy and identity consistency. Our results demonstrate the importance of decoupling dynamic appearance modeling from static identity cues, and provide a scalable foundation for robust tracking in complex scenarios.

artificial intelligence, natural language, proceedings, (5 more...)

Neural Information Processing Systems

Genre: Research Report > New Finding (0.60)

Technology: Information Technology > Artificial Intelligence > Natural Language (1.00)

Add feedback

Cross Attention Network for Few-shot Classification

Ruibing Hou, Hong Chang, Bingpeng MA, Shiguang Shan, Xilin Chen

Neural Information Processing SystemsApr-30-2026, 19:47:29 GMT

Few-shot classification aims to recognize unlabeled samples from unseen classes given only few labeled samples.

artificial intelligence, classification, machine learning, (17 more...)

Neural Information Processing Systems

Country: Asia > China (0.14)

Technology:

Information Technology > Artificial Intelligence > Machine Learning > Statistical Learning (0.95)
Information Technology > Artificial Intelligence > Vision (0.94)
Information Technology > Artificial Intelligence > Machine Learning > Neural Networks > Deep Learning (0.46)

Add feedback

Hybrid Mamba for Few-Shot Segmentation

Neural Information Processing SystemsMar-21-2026, 10:55:05 GMT

Many few-shot segmentation (FSS) methods use cross attention to fuse support foreground (FG) into query features, regardless of the quadratic complexity. A recent advance Mamba can also well capture intra-sequence dependencies, yet the complexity is only linear. Hence, we aim to devise a cross (attention-like) Mamba to capture inter-sequence dependencies for FSS. A simple idea is to scan on support features to selectively compress them into the hidden state, which is then used as the initial hidden state to sequentially scan query features. Nevertheless, it suffers from (1) support forgetting issue: query features will also gradually be compressed when scanning on them, so the support features in hidden state keep reducing, and many query pixels cannot fuse sufficient support features; (2) intra-class gap issue: query FG is essentially more similar to itself rather than to support FG, i.e., query may prefer not to fuse support features but their own ones from the hidden state, yet the success of FSS relies on the effective use of support information. To tackle them, we design a hybrid Mamba network (HMNet), including (1) a support recapped Mamba to periodically recap the support features when scanning query, so the hidden state can always contain rich support information; (2) a query intercepted Mamba to forbid the mutual interactions among query pixels, and encourage them to fuse more support features from the hidden state. Consequently, the support information is better utilized, leading to better performance. Extensive experiments have been conducted on two public benchmarks, showing the superiority of HMNet. The code is available at https://github.com/Sam1224/HMNet.

artificial intelligence, proceedings, support feature, (8 more...)

Neural Information Processing Systems

Technology: Information Technology > Artificial Intelligence (0.84)

Add feedback

HybridMambaforFew-ShotSegmentation

Neural Information Processing SystemsFeb-16-2026, 09:04:59 GMT

Manyfew-shot segmentation (FSS) methods use cross attention to fuse support foreground (FG) into query features, regardless of the quadratic complexity.

artificial intelligence, machine learning, segmentation, (18 more...)

Neural Information Processing Systems

Country:

Europe > Germany > Bavaria > Upper Bavaria > Munich (0.04)
Asia > Middle East > Israel > Tel Aviv District > Tel Aviv (0.04)

Technology:

Information Technology > Artificial Intelligence > Machine Learning (0.93)
Information Technology > Artificial Intelligence > Vision (0.68)

Add feedback

f7fef21d1fb3e950b12b50ad7f395e31-Paper-Conference.pdf

Neural Information Processing SystemsFeb-12-2026, 22:31:16 GMT

intermediate prototype, prototype, segmentation, (15 more...)

Neural Information Processing Systems

Country: Asia > China > Shaanxi Province (0.04)

Genre: Research Report (0.46)

Technology:

Information Technology > Artificial Intelligence > Vision (1.00)
Information Technology > Artificial Intelligence > Natural Language (1.00)
Information Technology > Artificial Intelligence > Machine Learning > Neural Networks > Deep Learning (0.68)
Information Technology > Sensing and Signal Processing > Image Processing (0.68)

Add feedback

6447714b83edcbed61dbe10371dd7ae5-Supplemental-Conference.pdf

Neural Information Processing SystemsFeb-12-2026, 19:41:18 GMT

artificial intelligence, machine learning, similarity, (19 more...)

Neural Information Processing Systems

Technology: Information Technology > Artificial Intelligence > Machine Learning (0.47)

Add feedback

Focus on Query: Adversarial Mining Transformer for Few-Shot Segmentation Yuan Wang

Neural Information Processing SystemsFeb-12-2026, 19:41:14 GMT

Previous works focus their efforts on exploring the support information while paying less attention to the mining of the critical query branch.

machine learning, natural language, segmentation, (17 more...)

Neural Information Processing Systems

Country:

Asia > Middle East > Israel (0.04)
Asia > China > Anhui Province > Hefei (0.04)

Genre: Research Report (0.68)

Industry: Health & Medicine (0.68)

Technology:

Information Technology > Sensing and Signal Processing > Image Processing (1.00)
Information Technology > Artificial Intelligence > Vision (1.00)
Information Technology > Artificial Intelligence > Machine Learning > Neural Networks (0.93)
Information Technology > Artificial Intelligence > Natural Language (0.70)

Add feedback

Few-ShotSegmentationviaCycle-Consistent Transformer

Neural Information Processing SystemsFeb-10-2026, 21:12:44 GMT

In this paper, we focus on utilizing pixel-wise relationships between support and query images to facilitate the few-shot segmentation task.

machine learning, natural language, segmentation, (19 more...)

Neural Information Processing Systems

Technology:

Information Technology > Artificial Intelligence > Natural Language (1.00)
Information Technology > Artificial Intelligence > Machine Learning > Statistical Learning (0.47)
Information Technology > Artificial Intelligence > Machine Learning > Neural Networks (0.47)

Add feedback

196f5641aa9dc87067da4ff90fd81e7b-AuthorFeedback.pdf

Neural Information Processing SystemsFeb-7-2026, 15:54:51 GMT

ACandallReviewers: We thank all reviewers. We discuss the motivation of usingS-based prior later. Asamatter offact,R4alsobrought upthis10 point, yet gave an accept score (7). Imposing15 a prior when available (could come from any source) is application-dependent and can indeed lead to enhanced16 performances, butis,again,notnecessary. Hence, empirical marginals should be closeˆpS(y) ˆpQ(y).

artificial intelligence, prediction, query feature, (1 more...)

Neural Information Processing Systems

Technology: Information Technology > Artificial Intelligence (0.34)

Add feedback

Filters

Collaborating Authors

query feature

Information about AI from the News, Publications, and Conferences

Automatic Classification – Tagging and Summarization – Customizable Filtering and Analysis

Dual-Path Temporal Decoder for End-to-End Multi-Object Tracking

Dual-Path Temporal Decoder for End-to-End Multi-Object Tracking

Cross Attention Network for Few-shot Classification

Hybrid Mamba for Few-Shot Segmentation

HybridMambaforFew-ShotSegmentation

f7fef21d1fb3e950b12b50ad7f395e31-Paper-Conference.pdf

6447714b83edcbed61dbe10371dd7ae5-Supplemental-Conference.pdf

Focus on Query: Adversarial Mining Transformer for Few-Shot Segmentation Yuan Wang

Few-ShotSegmentationviaCycle-Consistent Transformer

196f5641aa9dc87067da4ff90fd81e7b-AuthorFeedback.pdf