Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Jack Yang

Spatio-Temporal Joint Density Driven Learning for Skeleton-Based Action Recognition

May 29, 2025

Shanaka Ramesh Gunasekara, Wanqing Li, Philip Ogunbona, Jack Yang

Abstract:Traditional approaches in unsupervised or self supervised learning for skeleton-based action classification have concentrated predominantly on the dynamic aspects of skeletal sequences. Yet, the intricate interaction between the moving and static elements of the skeleton presents a rarely tapped discriminative potential for action classification. This paper introduces a novel measurement, referred to as spatial-temporal joint density (STJD), to quantify such interaction. Tracking the evolution of this density throughout an action can effectively identify a subset of discriminative moving and/or static joints termed "prime joints" to steer self-supervised learning. A new contrastive learning strategy named STJD-CL is proposed to align the representation of a skeleton sequence with that of its prime joints while simultaneously contrasting the representations of prime and nonprime joints. In addition, a method called STJD-MP is developed by integrating it with a reconstruction-based framework for more effective learning. Experimental evaluations on the NTU RGB+D 60, NTU RGB+D 120, and PKUMMD datasets in various downstream tasks demonstrate that the proposed STJD-CL and STJD-MP improved performance, particularly by 3.5 and 3.6 percentage points over the state-of-the-art contrastive methods on the NTU RGB+D 120 dataset using X-sub and X-set evaluations, respectively.

* IEEE Transactions on Biometrics, Behavior, and Identity Science (2025)

Via

Access Paper or Ask Questions

FullStack Bench: Evaluating LLMs as Full Stack Coders

Dec 03, 2024

Siyao Liu, He Zhu, Jerry Liu, Shulin Xin, Aoyan Li, Rui Long, Li Chen, Jack Yang, Jinxiang Xia, Z. Y. Peng(+7 more)

Figure 1 for FullStack Bench: Evaluating LLMs as Full Stack Coders

Figure 2 for FullStack Bench: Evaluating LLMs as Full Stack Coders

Figure 3 for FullStack Bench: Evaluating LLMs as Full Stack Coders

Figure 4 for FullStack Bench: Evaluating LLMs as Full Stack Coders

Abstract:As the capabilities of code large language models (LLMs) continue to expand, their applications across diverse code intelligence domains are rapidly increasing. However, most existing datasets only evaluate limited application domains. To address this gap, we have developed a comprehensive code evaluation dataset FullStack Bench focusing on full-stack programming, which encompasses a wide range of application domains (e.g., basic programming, data analysis, software engineering, mathematics, and machine learning). Besides, to assess multilingual programming capabilities, in FullStack Bench, we design real-world instructions and corresponding unit test cases from 16 widely-used programming languages to reflect real-world usage scenarios rather than simple translations. Moreover, we also release an effective code sandbox execution tool (i.e., SandboxFusion) supporting various programming languages and packages to evaluate the performance of our FullStack Bench efficiently. Comprehensive experimental results on our FullStack Bench demonstrate the necessity and effectiveness of our FullStack Bench and SandboxFusion.

* 26 pages

Via

Access Paper or Ask Questions

Joint Temporal Pooling for Improving Skeleton-based Action Recognition

Aug 18, 2024

Shanaka Ramesh Gunasekara, Wanqing Li, Jack Yang, Philip Ogunbona

Figure 1 for Joint Temporal Pooling for Improving Skeleton-based Action Recognition

Figure 2 for Joint Temporal Pooling for Improving Skeleton-based Action Recognition

Figure 3 for Joint Temporal Pooling for Improving Skeleton-based Action Recognition

Figure 4 for Joint Temporal Pooling for Improving Skeleton-based Action Recognition

Abstract:In skeleton-based human action recognition, temporal pooling is a critical step for capturing spatiotemporal relationship of joint dynamics. Conventional pooling methods overlook the preservation of motion information and treat each frame equally. However, in an action sequence, only a few segments of frames carry discriminative information related to the action. This paper presents a novel Joint Motion Adaptive Temporal Pooling (JMAP) method for improving skeleton-based action recognition. Two variants of JMAP, frame-wise pooling and joint-wise pooling, are introduced. The efficacy of JMAP has been validated through experiments on the popular NTU RGB+D 120 and PKU-MMD datasets.

* 2023 International Conference on Digital Image Computing: Techniques and Applications, DICTA 2023

Via

Access Paper or Ask Questions