03077nas a2200229 a 450000500170000000800410001703500150005810000210007324500870009426000670018130000420024849000260029050001220031650200580043852020360049665000350253270000370256770000480260471000640265281000580271685600730277420260818145444.0260219s20259999th u ms t 000 eng d a.b124780391 aNayeem, Jannatun10aLeveraging phonological clustering for word-level Bangla Sign language recognition aPathum Thani, Thailand :bAsian Institute of Technology,c2025 a63 leaves :bill.+e1 online resource1 aThesis ;vno.CS-25-01 aA thesis submitted in partial fulfillment of the requirements for the degree of Master of Science in Computer Science aThesis (M. Sc.) - Asian Institute of Technology, 2025 aBangla Sign Language (BdSL) recognition presents multifaceted challenges due to signer diversity and spatiotemporal variability. While flat classification pipelines are widely used, they often overlook the underlying phonological relationships among signs that can inform more structured and accurate recognition. To address this gap, we propose a hierarchical recognition framework that integrates phonological clustering into a Bidirectional Long Short-Term Memory (Bi-LSTM)-based sequence modeling pipeline. First, baseline classification{u2014}referring to a flat, non-clustered recognition approach{u2014}is performed us ing five Bi-LSTM configurations of increasing complexity to assess the trade-off between accuracy and model size. The resulting confusion matrices are analyzed to identify sign pairs with high misclassification rates, revealing underlying phonological similarities. Based on this analysis, a confusion-matrix driven clustering strategy is employed to group visually and phonologically similar signs. Cluster-specific feature engineering is then applied, and the same Bi-LSTM architecture is retrained separately within each cluster. Experiments are conducted on a curated 50-class subset of the SignBD-Word dataset. In the baseline setting, the most complex model (Bi-LSTM-1) achieves 91.10% accuracy with 2.6 million parameters. With the proposed confusion-matrix driven clustering architecture, four out of six clusters employing the lightweight Bi-LSTM classifiers outperform the baseline model, reaching up to 94.58% accuracy. Remarkably, the lightweight Bi-LSTM-4 model{u2014}with only 373K parameters (14.4% of Bi LSTM-1){u2014}achieves a weighted average accuracy of 92.60% on cluster-level classification, surpassing the baseline by1.5percentagepoints. Thisreflectsan85.6%reduction in a number of parameters, demonstrating that efficient cluster-specific feature engineering, which allows each model to capture nuanced patterns within each cluster can improve overall predictive accuracy. 0aSign languagexData processing0 aChantri Polprasert,eChairperson0 aMongkol Ekpanyapong,eExamination Committee2 aADB-Japan Scholarship Program (ADB-JSP),eScholarship Donor2 aAsian Institute of Technology.tThesis :vno.CS-25-01403Full-Textuhttp://203.159.5.9/ait-thesis/Viewer/viewer.php?id=B23610