Goto

Collaborating Authors

 Asia


Prediction Market Philosophers Got What They Wanted. They're Not Happy About It

WIRED

Prediction Market Philosophers Got What They Wanted. Getting the future right is now big business. But at a festival in the Bay Area, forecasters worry that sports markets could take the whole industry down. On June 11, Kalshi released a buzzy ad featuring noted New York Knicks fan Timothée Chalamet. It was a zeitgeist-capturing moment for prediction markets, akin to the 2022 Super Bowl, when seemingly every commercial featured a celebrity shilling crypto.


Flow-Based Policy for Online Reinforcement Learning

Neural Information Processing Systems

We argue that in addition to training signals, enhancing the expressiveness of the policy class is crucial for the performance gains in RL. Flow-based generative models offer such potential, excelling at capturing complex, multimodal action distributions. However, their direct application in online RL is challenging due to a fundamental objective mismatch: standard flow training optimizes for static data imitation, while RL requires value-based policy optimization through a dynamic buffer, leading to difficult optimization landscapes.


US export ban on Anthropic's AI models further strains alliances

Al Jazeera

Artificial intelligence has become the latest issue to drive a wedge between the United States and its allies after US President Donald Trump ordered tech giant Anthropic to cut off foreign access to its powerful Mythos 5 and Claude Fable 5 AI models, citing national security concerns. The US issued the unprecedented order for all foreign nationals in and outside the US last week, promoting Anthropic to take the two AI models completely offline to ensure compliance. The two public versions of the model, Mythos 5 and Fable 5, were due to be released in early June. Anthropic said the US government did not provide a reason for the order, but that it was its "understanding" that the Trump administration believed it had become aware of a method of "jailbreaking" Fable 5. The Trump administration's ban immediately sent shockwaves across Europe, which is heavily dependent on US-developed AI.


GSAlign: Geometric and Semantic Alignment Network for Aerial-Ground Person Re-Identification

Neural Information Processing Systems

Aerial-Ground person re-identification (AG-ReID) is an emerging yet challenging task that aims to match pedestrian images captured from drastically different viewpoints, typically from unmanned aerial vehicles (UAVs) and ground-based surveillance cameras. The task poses significant challenges due to extreme viewpoint discrepancies, occlusions, and domain gaps between aerial and ground imagery. While prior works have made progress by learning cross-view representations, they remain limited in handling severe pose variations and spatial misalignment. To address these issues, we propose a Geometric and Semantic Alignment Network (GSAlign) tailored for AG-ReID. GSAlign introduces two key components to jointly tackle geometric distortion and semantic misalignment in aerial-ground matching: a Learnable Thin Plate Spline (LTPS) Module and a Dynamic Alignment Module (DAM). The LTPS module adaptively warps pedestrian features based on a set of learned keypoints, effectively compensating for geometric variations caused by extreme viewpoint changes.


Officer Chair Laptop Handbag

Neural Information Processing Systems

Existing 3D generation primarily emphasizes geometries and textures while neglecting physical-grounded modeling. Consequently, despite the rapid development of 3D generative models, the synthesized 3D assets often overlook rich and important physical properties, hampering their real-world application in physical domains like simulation and embodied AI. As an initial attempt to address this challenge, we propose PhysX, an end-to-end paradigm for physical-grounded 3D asset generation. 1) To bridge the critical gap in physics-annotated 3D datasets, we present PhysXNet - the first physics-grounded 3D dataset systematically annotated across five foundational dimensions: absolute scale, material, affordance, kinematics, and function description. In particular, we devise a scalable human-in-the-loop annotation pipeline based on vision-language models, which enables efficient creation of physics-first assets from raw 3D assets.


Sparse Meets Dense: Unified Generative Recommendations with Cascaded Sparse-Dense Representations

Neural Information Processing Systems

Generative models have recently gained attention in recommendation systems by directly predicting item identifiers from user interaction sequences. However, existing methods suffer from significant information loss due to the separation of stages such as quantization and sequence modeling, hindering their ability to achieve the modeling precision and accuracy of sequential dense retrieval techniques. Integrating generative and dense retrieval methods remains a critical challenge. To address this, we introduce the Cascaded Organized Bi-Represented generAtive retrieval (COBRA) framework, which innovatively integrates sparse semantic IDs and dense vectors through a cascading process. Our method alternates between generating these representations by first generating sparse IDs, which serve as conditions to aid in the generation of dense vectors. End-to-end training enables dynamic refinement of dense representations, capturing both semantic insights and collaborative signals from user-item interactions. During inference, COBRA employs a coarse-to-fine strategy, starting with sparse ID generation and refining them into dense vectors via the generative model. We further propose BeamFusion, an innovative approach combining beam search with nearest neighbor scores to enhance inference flexibility and recommendation diversity.


Ukraine is advancing. Now is the perfect time for Trump to strengthen NATO

FOX News

Trump's shift toward supporting Ukraine is promising, but Pentagon moves to reduce NATO forces risk undermining deterrence against Russia.


Elite colleges are losing America's trust. Community colleges can win it back

FOX News

Inflation, a tough economy, and AI threats to white-collar jobs have crushed trust in elite schools, creating opportunity for community colleges and certification programs.


Rethinking Protein Protein Interaction Prediction from Pairs to Graphs

Neural Information Processing Systems

Deep learning-based computational methods have achieved promising results in predicting protein-protein interactions (PPIs). However, existing benchmarks predominantly focus on isolated pairwise evaluations, overlooking a model's capability to reconstruct biologically meaningful PPI networks, which is crucial for biology research. To address this gap, we introduce PRING, the first comprehensive benchmark that evaluates PRotein-protein INteraction prediction from a Graph-level perspective. PRINGcurates a high-quality, multi-species PPI network dataset comprising 21,484 proteins and 186,818 interactions, with well-designed strategies to address both data redundancy and leakage. Building on this golden-standard dataset, we establish two complementary evaluation paradigms: (1) topologyoriented tasks, which assess intra and cross-species PPI network construction, and (2) function-oriented tasks, including protein complex pathway prediction, GO module analysis, and essential protein justification. These evaluations not only reflect the model's capability to understand the network topology but also facilitate protein function annotation, biological module detection, and even disease mechanism analysis. Extensive experiments on four representative model categories, consisting of sequence similarity-based, naive sequence-based, protein language model-based, and structure-based approaches, demonstrate that current PPI models have potential limitations in recovering both structural and functional properties of PPI networks, highlighting the gap in supporting real-world biological applications. We believe PRINGprovides a reliable platform to guide the development of more effective PPI prediction models for the community.


ForceVLA: Enhancing VLAModels with a Force-aware MoE for Contact-rich Manipulation

Neural Information Processing Systems

Vision-Language-Action (VLA) models have advanced general-purpose robotic manipulation by leveraging pretrained visual and linguistic representations. However, they struggle with contact-rich tasks that require fine-grained control involving force, especially under visual occlusion or dynamic uncertainty. To address these limitations, we propose ForceVLA, a novel end-to-end manipulation framework that treats external force sensing as a first-class modality within VLA systems. ForceVLA introduces FVLMoE, a force-aware Mixture-of-Experts fusion module that dynamically integrates pretrained visual-language embeddings with real-time 6-axis force feedback during action decoding. This enables context-aware routing across modality-specific experts, enhancing the robot's ability to adapt to subtle contact dynamics. We also introduce ForceVLA-Data, a new dataset comprising synchronized vision, proprioception, and force-torque signals across five contactrich manipulation tasks. ForceVLA improves average task success by 23.2% over strong π0-based baselines, achieving up to 80% success in tasks such as plug insertion. Our approach highlights the importance of multimodal integration for dexterous manipulation and sets a new benchmark for physically intelligent robotic control. Code and data will be released at website.