An Intelligent Machine Learning-Based Framework for Scam Call Detection Using Textual and Behavioural Features
编号:85 访问权限:仅限参会人 更新:2026-07-25 21:37:21 浏览:22次 Online

报告开始:2026年07月30日 15:40(Asia/Kolkata)

报告时间:15min

所在会场:[S4] Computer Vision and Pattern Recognition [S4-2] Computer Vision and Pattern Recognition

视频 无权播放 演示文件

提示:该报告下的文件权限为仅限参会人,您尚未登录,暂时无法查看。

摘要
Telecom fraud has widespread financial costs, with total damages surpassing USD 38.95 billion in 2023. Traditional blacklist systems are struggling to keep pace with these adaptive fraud techniques. For instance, techniques such as Caller-ID spoofing and dynamic number rotation pose a major challenge. We propose an intelligent multimodal scam call detection framework, which unifies TF-IDF textual features with synthetically constructed behavioural call metadata such as call duration, call frequency, time-of-day, and attempts to call repeatedly. Our framework brings robustness to binary and multi-class classification of call fraud. The performance of three supervised classifiers, namely Logistic Regression (LR), Linear Support Vector Machine (SVM), and Random Forest (RF), is evaluated on a labeled dataset of 5,926 calls, which is split randomly and stratified in a ratio of 80% for the training set and 20% for the test set. The support of the SMOTE technique and the use of a weighting scheme to counteract class imbalance were applied. Linear SVM performed best and provided the following results for binary classification: accuracy = 99%, precision = 0.96, recall = 0.91, and F1-score = 0.94. In the four-class classification scheme, which includes Normal, Bank Fraud, Lottery Scam, and Emergency Scam, the linear SVM achieved a 98% weighted average performance. Among the behavioral features, call duration (0.081) and call frequency (0.060) ranked highest and found to be the most important and discriminative call features, while the textual tokens ‘free,’ ‘claim,’ and ‘reply’ were the most important textual featuresof the framework. The proposed study demonstrated an improvement over the performance of deep learning baselines, and while preserving interpretability for real-time mobile applications. LSTM, CNN, and BERT techniques were the baselines used.
关键词
暂无
报告人
ABDUL KHADAR SHAIK
Research Scholor B. S. Abdur Rahman Crescent Institute Of Science And Technology

稿件作者
V.A.S. Lakshmi V Narsaraopet Engineering College
Ramesh Babu Bolla ESWAR COLLEGE OF ENGINEERING
RESHMA SYED Vignan's Foundation for Science, Technology and Research
ABDUL KHADAR SHAIK B. S. Abdur Rahman Crescent Institute Of Science And Technology
发表评论
验证码 看不清楚,更换一张
全部评论
重要日期
  • 会议日期

    07月30日

    2026

    08月01日

    2026

  • 07月28日 2026

    初稿截稿日期

  • 08月03日 2026

    注册截止日期

主办单位
The United Societies of Science
承办单位
Kongunadu College of Engineering and Technology
协办单位
IEEE Section
IEEE Madras Section
历届会议
移动端
在手机上打开
小程序
打开微信小程序
客服
扫码或点此咨询