REVIEW 4 major objections 6 minor 255 references
From Past to Present: A Survey of Malicious URL Detection Techniques, Datasets and Code Repositories
T0 review · 4 major / 6 minor · reviewed 2026-08-16 · deepseek-v4-flash
Pith's one-line read A modality-based taxonomy reorganizes the malicious URL detection literature, and a curated registry of datasets and open-source code is offered as the field's first unified benchmark for reproducible evaluation.
desk verdict Useful modality-organized survey with recent LLM coverage, but the 'unified benchmark' claim is unsupported and the resource tables contain concrete misattributions. read the letter →
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
The reading
What carries the argument
The carrying mechanism is a hierarchical, modality-based taxonomy defined in Section 3: URL-based, HTML-based, JavaScript-based, visual-based, and hybrid-modality detection, each with its own technical principles and representative algorithms. This taxonomy is what turns a scattered literature into a structured map. Its companion machinery is the curated registry—Table 3 listing open-source implementations with links and Table 4 listing public datasets with sources, sizes, and access methods—together with the maintained GitHub repository that keeps the registry current. The taxonomy does the analytical work; the registry does the benchmarking work.
What would settle it
Pick a random sample of the repository URLs in Table 3 and the dataset access links in Table 4 and follow them: if a substantial fraction resolve to dead pages, relocated projects, or repositories containing code that does not match the cited algorithm, the 'unified benchmark for reproducible evaluation' claim is not met. A second check would take two papers the taxonomy assigns to different modalities and show that the same detection system could be described equally well in both categories, indicating the taxonomy's boundaries are not well-defined.
Extended reading notes
Core claim
On the paper's own terms, the central discovery is that the entire malicious-URL detection literature can be re-read through a modality lens that classifies each method by its primary data source: lexical URL features, HTML document features, JavaScript execution features, visual page features, and hybrid combinations. Within each category it surveys techniques from blacklists and heuristics through classical machine learning up to Transformers, GNNs, and LLM-based prompting and fine-tuning. It further claims to establish the first unified benchmark for reproducible evaluation by compiling 15 open-source code repositories and 14 public datasets, and it distills design principles for real-world deployment from systems like Monarch. The conclusion is that this consolidated resource framework enables meaningful cross-method performance comparisons and tracking of genuine methodological advancement.
Load-bearing premise
The paper's central benchmark claim assumes that the curated dataset and code-repository links are correct, complete, and stable, and that each cited work can be assigned unambiguously to exactly one modality; it verifies neither the links nor the assignment rules.
Editorial extensions
If this is right
- A researcher choosing a baseline for a new malicious-URL method can now start from the curated repository list in Table 3 instead of hunting through papers for code links.
- The modality taxonomy predicts which information channels—URL, HTML, JavaScript, visual—are already saturated and which combinations remain underexplored, guiding new multimodal fusion work.
- LLM- and Transformer-based defenses, which earlier reviews omitted, are folded into the same modality map, so their relationship to classical methods becomes comparable.
- The design principles in Section 6 give product teams a concrete checklist (accuracy, speed, scalability, adaptability, flexibility) for turning detection research into a service.
- If the curated registry stays current through the GitHub repository, cross-paper performance comparisons become reproducible rather than anecdotal.
Reading between the lines
- Editorial: the claim of being the 'first unified benchmark' depends on the registry staying alive; a survey can seed a benchmark, but only verified, archived copies of datasets and code would make the benchmark durable.
- Editorial: the taxonomy's five modalities could be extended to a sixth—network-traffic and user-behavior features—which the paper mentions but deliberately leaves out, suggesting a natural expansion for a follow-up survey.
- Editorial: a testable extension would be to run a single classifier from each modality category on a common dataset collection and report cross-modal accuracy, which would turn the curated registry into an actual leaderboard rather than a list of links.
Signed reviews
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. This survey reviews the field of malicious URL detection, organizing methods by data modality (URL, HTML, JavaScript, visual, and hybrid) and covering traditional machine learning, deep learning, Transformer/GNN, and LLM-based approaches. It also includes a short section on Arabic-language detection studies, tables of published algorithms with code repositories, tables of public datasets, definitions of evaluation metrics, design principles for deployed systems, and a discussion of challenges and future directions. The authors claim that the paper establishes 'the first unified benchmark for reproducible evaluation' by curating datasets and open-source implementations, and they maintain a GitHub repository of resources.
Significance. If its resource collection were accurate and its claims appropriately bounded, the survey would be a useful entry point for researchers: the modality-based taxonomy is a reasonable organizing principle, the coverage of Transformer/GNN/LLM-based detectors goes beyond several earlier surveys, and Tables 3 and 4 could help practitioners locate datasets and code. The paper does not present machine-checked proofs, experimental results, or a runnable evaluation harness; its main contribution is curation and classification of literature. However, the central 'unified benchmark' claim is substantially overreaching relative to the content, and the curation contains concrete errors (misattributed repository links, a nonexistent reference number, and a citation mismatch). These issues are fixable in revision, so the paper has potential as a survey after correction and recalibration of its claims.
major comments (4)
- [Section 8, Tables 3-4, companion repository] The conclusion asserts that the paper 'establish[es] the first unified benchmark for reproducible evaluation', but a reproducible benchmark requires fixed datasets with agreed train/test splits, a shared evaluation harness computing common metrics, and baseline results obtained by running that harness. The manuscript provides none of these: Section 5.1 lists repository URLs, Section 5.2 lists dataset sources, and the GitHub repository is described as 'ongoing curating datasets and open-source implementations' rather than as a benchmark harness. The claim should be softened to a curated resource collection unless the authors supply the actual harness, split definitions, metric implementations, and baseline results.
- [Table 3] Table 3 contains misattributed code-repository entries that undermine the curation. Entry [182] links to https://github.com/MjafarMashhadi/Haplophysh, but reference [182] is Saxe and Berlin's 'expose' paper, which is not the Haplophysh project. Entry [228] links to https://github.com/Microsoft/CNTK, a general deep-learning framework, while reference [228] is Wei et al.'s URL phishing detector with a convolutional neural network. Because the stated purpose of Table 3 is to provide usable baseline implementations, these misattributions need to be corrected and every entry should be verified before the resource can support benchmarking claims.
- [Section 7.1.1] Section 7.1.1 says that cloaking 'aims to evade the detection of phishing detection systems [404]', but the reference list contains no entry [404]; the subsequent discussion of server-side and client-side cloaking cites [243]-[249]. The in-text citation numbering appears corrupt. The authors must align the in-text citations with the final bibliography before resubmission.
- [Section 2.2] Section 2.2 attributes TfidfVectorizer to reference [14], which is Wang's 2022 comparison study 'Malicious url detection an evaluation of feature extraction and machine learning algorithm' rather than the documentation or original source of TfidfVectorizer. This citation mismatch, together with the errors in Table 3 and Section 7.1.1, indicates that the bibliography has not been systematically verified. Please re-check all reference-citation pairs.
minor comments (6)
- [Table 2] The Result column reports accuracy values obtained on different datasets and with different experimental protocols; without a dataset/column or a cautionary note, the numbers suggest direct comparability that does not exist.
- [Section 5.2] The sentence 'The specific creation time, creator, and link information are shown in Table 3' should refer to Table 4, since Table 4 lists the datasets.
- [Section 4] Several works reviewed in the Arabic section concern Arabic spam detection and fake news in social media rather than malicious URL detection (e.g., [213], [214], [216]); consider retitling the section or explicitly framing these as adjacent tasks that inform URL-centric detection.
- [Figures 2 and 3] Figures 2 and 3 appear without in-text references, even though Figure 3 presents the proposed taxonomy; add appropriate callouts.
- [Section 8] The phrase 'We rigorous coverage of Transformer, GNNs and LLMs' is an ungrammatical fragment, and the conclusion's claim of 'three high-impact directions' is not clearly aligned with the broader set of directions in Section 7.2.
- [Section 3.1.2, PA/CW/AROW paragraph] The sentence 'ROW enhances the robustness...' should be 'AROW enhances the robustness...', and the stacked-matrix notation in Equation (13) should be reconciled with the separate gate equations (14)-(16).
Circularity Check
No circularity: the survey derives no fitted quantity, and its self-citations are catalogue entries, not load-bearing evidence for a derived result.
full rationale
This manuscript is a survey and resource-curation paper rather than a derivation or empirical study, so the standard circularity patterns do not apply. There is no quantity fitted to data and then renamed as a prediction, no equation whose output is identical to its input by construction, and no uniqueness theorem imported from the authors' prior work to force a modeling choice. The paper's modality-based taxonomy is a classification scheme proposed by the authors; it organizes surveyed works by data modality, and calling it 'novel' is a novelty claim, not a derivational claim. The central conclusion that the paper 'establish[es] the first unified benchmark for reproducible evaluation' is unsupported, because Tables 3 and 4 provide curated links and the GitHub repository is described as 'ongoing curating datasets and open-source implementations' rather than a benchmark harness with splits, metrics, and baseline results. However, this is an overstatement or completeness/correctness weakness, not circularity: the claim does not reduce to its own inputs by construction, and the misattributed entries (e.g., Table 3 entry [182] pointing to a GitHub repository for Haplophysh while reference [182] is Saxe and Berlin's expose paper, and entry [228] pointing to Microsoft/CNTK while reference [228] is a CNN phishing detector) are curation errors, not circular reasoning. Similarly, the nonexistent reference [404] in Section 7.1.1 indicates a missing citation, not a self-referential derivation. The self-citations in the reference list (e.g., [74], [77], [78], [172], [234], [235] involving the corresponding author) are used only as surveyed/already-published works in the tables and narrative; none is invoked as the justification for a claimed result or as an external authority that forbids alternatives. They are therefore not load-bearing. Because the paper contains no fitted-input-as-prediction step and no self-citation chain supporting a central derived claim, the circularity score is 0.
Assumptions & free parameters
assumptions (2)
- domain assumption The selected references can be consistently classified into the proposed modality categories (URL, HTML, JavaScript, Visual, Hybrid).
- domain assumption The datasets and code repositories listed in Tables 3 and 4, including the GitHub repository URL, are stable, complete, and correctly attributed.
Cite this review
Pith. "Pith review of From Past to Present: A Survey of Malicious URL Detection Techniques, Datasets and Code Repositories." pith.science (2026). https://pith.science/paper/277BKY2Q
@misc{pith2026250416449,
author = {Pith},
title = {Pith review of: From Past to Present: A Survey of Malicious URL Detection Techniques, Datasets and Code Repositories},
year = {2026},
howpublished = {\url{https://pith.science/paper/277BKY2Q}},
note = {Machine review of arXiv:2504.16449}
}
read the original abstract
Malicious URLs persistently threaten the cybersecurity ecosystem, by either deceiving users into divulging private data or distributing harmful payloads to infiltrate host systems. Gaining timely insights into the current state of this ongoing battle holds significant importance. However, existing reviews exhibit 4 critical gaps: 1) Their reliance on algorithm-centric taxonomies obscures understanding of how detection approaches exploit specific modal information channels; 2) They fail to incorporate pivotal LLM/Transformer-based defenses; 3) No open-source implementations are collected to facilitate benchmarking; 4) Insufficient dataset coverage.This paper presents a comprehensive review of malicious URL detection technologies, systematically analyzing methods from traditional blacklisting to advanced deep learning approaches (e.g. Transformer, GNNs, and LLMs). Unlike prior surveys, we propose a novel modality-based taxonomy that categorizes existing works according to their primary data modalities (URL, HTML, Visual, etc.). This hierarchical classification enables both rigorous technical analysis and clear understanding of multimodal information utilization. Furthermore, to establish a profile of accessible datasets and address the lack of standardized benchmarking (where current studies often lack proper baseline comparisons), we curate and analyze: 1) publicly available datasets (2016-2024), and 2) open-source implementations from published works(2013-2025). Then, we outline essential design principles and architectural frameworks for product-level implementations. The review concludes by examining emerging challenges and proposing actionable directions for future research. We maintain a GitHub repository for ongoing curating datasets and open-source implementations: https://github.com/sevenolu7/Malicious-URL-Detection-Open-Source/tree/master.
Figures
Figures from the paper (4 more)
Reference graph
Works this paper leans on
-
[182]
Malicious url detection by dynamically mining patterns without pre-defined elements
Huang, D., Xu, K., Pei, J., 2014. Malicious url detection by dynamically mining patterns without pre-defined elements. World Wide Web 17, 1375–1394
2014
-
[228]
A deep learning for arabic sms phishing based on urls detection
Alsufyani, S., Alajmani, S., 2025. A deep learning for arabic sms phishing based on urls detection. International Journal of Advanced Computer Science & Applications 16
2025
-
[243]
Oest, A., Safaei, Y., Doupé, A., Ahn, G.J., Wardman, B., Tyers, K., 2019. Phishfarm: A scalable framework for measuring the effectiveness of evasion techniques against browser phishing blacklists, in: 2019 IEEE Symposium on Security and Privacy (SP), pp. 1344–1361. doi:10.1109/SP.2019.00049
arXiv 2019
-
[249]
Li, W., He, Y., Wang, Z., Alqahtani, S., Nanda, P., 2023. Uncovering flaws in anti-phishing blacklists for phishing websites using novel cloakingtechniques,in:Proceedingsofthe20thInternationalConferenceonSecurityandCryptography-Volume1:SECRYPT,INSTICC. SciTePress. pp. 813–821. doi:10.5220/0012135600003555
-
[14]
Malicious url detection an evaluation of feature extraction and machine learning algorithm
Wang, Y., 2022. Malicious url detection an evaluation of feature extraction and machine learning algorithm. Highlights in Science, Engineering and Technology 23, 117–123
2022
-
[1]
Malicious url detection using machine learning: A survey
Sahoo, D., Liu, C., Hoi, S.C., 2017. Malicious url detection using machine learning: A survey. arXiv preprint arXiv:1701.07179
arXiv 2017
-
[2]
Malicious url prediction using machine learning techniques
Aalla, H.V.S., Dumpala, N.R., Eliazer, M., 2021. Malicious url prediction using machine learning techniques. Annals of the Romanian Society for Cell Biology 25, 2170–2176
2021
-
[3]
blacklists
Sinha, S., Bailey, M., Jahanian, F., 2008. Shades of grey: On the effectiveness of reputation-based “blacklists”, in: 2008 3rd International Conference on Malicious and Unwanted Software (MALWARE), IEEE. pp. 57–64
2008
Show all 255 references
-
[4]
Apwg.https://apwg.org/
Anti-Phishing Working Group, 2023. Apwg.https://apwg.org/. [Online; accessed 10-April-2025]
2023
-
[5]
Detecting malicious urls using machine learning techniques: review and research directions
Aljabri,M.,Altamimi,H.S.,Albelali,S.A.,Al-Harbi,M.,Alhuraib,H.T.,Alotaibi,N.K.,Alahmadi,A.A.,Alhaidari,F.,Mohammad,R.M.A., Salah, K., 2022. Detecting malicious urls using machine learning techniques: review and research directions. IEEE Access 10, 121395– 121417
2022
-
[6]
Detectionofmaliciousurlsusingmachinelearning
Reyes-Dorta,N.,Caballero-Gil,P.,Rosa-Remedios,C.,2024. Detectionofmaliciousurlsusingmachinelearning. WirelessNetworks,1–18
2024
-
[7]
Malicious url detection: a survey, in: DEIM Forum F6–3
Aung, E.S., Yamana, H., 2020. Malicious url detection: a survey, in: DEIM Forum F6–3
2020
-
[8]
An effective detection approach for phishing websites using url and html features
Aljofey, A., Jiang, Q., Rasool, A., Chen, H., Liu, W., Qu, Q., Wang, Y., 2022. An effective detection approach for phishing websites using url and html features. Scientific Reports 12, 8842
2022
-
[9]
Asurveyofintelligentdetectiondesignsofhtmlurlphishingattacks
Asiri,S.,Xiao,Y.,Alzahrani,S.,Li,S.,Li,T.,2023. Asurveyofintelligentdetectiondesignsofhtmlurlphishingattacks. IEEEAccess11, 6421–6443. doi:10.1109/ACCESS.2023.3237798
2023
-
[10]
Towards detecting and classifying malicious urls using deep learning
Johnson, C., Khadka, B., Basnet, R.B., Doleck, T., 2020. Towards detecting and classifying malicious urls using deep learning. J. Wirel. Mob. Networks Ubiquitous Comput. Dependable Appl. 11, 31–48
2020
-
[11]
Detectionandanalysisofdrive-by-downloadattacksandmaliciousjavascriptcode,in:Proceedings of the 19th international conference on World wide web, pp
Cova,M.,Kruegel,C.,Vigna,G.,2010. Detectionandanalysisofdrive-by-downloadattacksandmaliciousjavascriptcode,in:Proceedings of the 19th international conference on World wide web, pp. 281–290
2010
-
[12]
Phishing url detection: A real-case scenario through login urls
Sánchez-Paniagua, M., Fernández, E.F., Alegre, E., Al-Nabki, W., González-Castro, V., 2022. Phishing url detection: A real-case scenario through login urls. IEEe Access 10, 42949–42960
2022
-
[13]
Phishing url detection using hybrid ensemble model
Pandey, A., Chadawar, J., 2022. Phishing url detection using hybrid ensemble model. INTERNATIONAL JOURNAL OF ENGINEERING RESEARCH & TECHNOLOGY (IJERT) 11
2022
-
[15]
Classification of malicious urls using machine learning
Abad, S., Gholamy, H., Aslani, M., 2023. Classification of malicious urls using machine learning. Sensors 23, 7760
2023
-
[16]
A url-based social semantic attacks detection with character-aware language model
Almousa, M., Anwar, M., 2023. A url-based social semantic attacks detection with character-aware language model. IEEE Access 11, 10654–10663
2023
-
[17]
An assessment of lexical, network, and content-based features for detecting malicious urls using machine learning and deep learning models
Aljabri, M., Alhaidari, F., Mohammad, R.M.A., Mirza, S., Alhamed, D.H., Altamimi, H.S., Chrouf, S.M.B., 2022. An assessment of lexical, network, and content-based features for detecting malicious urls using machine learning and deep learning models. Computational Intelligence ...
2022
-
[18]
An effective cost-sensitive xgboost method for malicious urls detection in imbalanced dataset
He, S., Li, B., Peng, H., Xin, J., Zhang, E., 2021. An effective cost-sensitive xgboost method for malicious urls detection in imbalanced dataset. IEEE Access 9, 93089–93096
2021
-
[19]
Highly predictive blacklisting., in: USENIX security symposium, pp
Zhang, J., Porras, P.A., Ullrich, J., 2008. Highly predictive blacklisting., in: USENIX security symposium, pp. 107–122
2008
-
[20]
An empirical analysis of phishing blacklists
Sheng, S., Wardman, B., Warner, G., Cranor, L., Hong, J., Zhang, C., 2009. An empirical analysis of phishing blacklists. Proceedings of Sixth Conference on Email and AntiSpam (CEAS)
2009
-
[21]
Protectsensitivesitesfromphishingattacksusingfeaturesextractablefrominaccessible phishing urls, in: 2013 IEEE international conference on communications (ICC), IEEE
Chu,W.,Zhu,B.B.,Xue,F.,Guan,X.,Cai,Z.,2013. Protectsensitivesitesfromphishingattacksusingfeaturesextractablefrominaccessible phishing urls, in: 2013 IEEE international conference on communications (ICC), IEEE. pp. 1990–1994
2013
-
[22]
A novel approach for phishing detection using url-based heuristic, in: 2014 international conference on computing, management and telecommunications (ComManTel), IEEE
Nguyen, L.A.T., To, B.L., Nguyen, H.K., Nguyen, M.H., 2014. A novel approach for phishing detection using url-based heuristic, in: 2014 international conference on computing, management and telecommunications (ComManTel), IEEE. pp. 298–303
2014
-
[23]
Detecting phishing web sites: A heuristic url-based approach, in: 2013 International Conference on Advanced Technologies for Communications (ATC 2013), IEEE
Nguyen, L.A.T., To, B.L., Nguyen, H.K., Nguyen, M.H., 2013. Detecting phishing web sites: A heuristic url-based approach, in: 2013 International Conference on Advanced Technologies for Communications (ATC 2013), IEEE. pp. 597–602
2013
-
[24]
Heuristic-based strategy for phishing prediction: A survey of url-based approach
da Silva, C.M.R., Feitosa, E.L., Garcia, V.C., 2020. Heuristic-based strategy for phishing prediction: A survey of url-based approach. Computers & Security 88, 101613
2020
-
[25]
Detection of phishing urls using heuristics-based approach, in: 2022 5th Information Technology for Education and Development (ITED), pp
Salihu, S.A., Oladipo, I.D., Wojuade, A.A., Abdulraheem, M., Babatunde, A.O., Ajiboye, A.R., Balogun, G.B., 2022. Detection of phishing urls using heuristics-based approach, in: 2022 5th Information Technology for Education and Development (ITED), pp. 1–7. doi:10.1109/ITED5663...
2022
-
[26]
Usinglexicalfeaturesformaliciousurldetection–amachinelearningapproach
Joshi,A.,Lloyd,L.,Westin,P.,Seethapathy,S.,2019. Usinglexicalfeaturesformaliciousurldetection–amachinelearningapproach. arXiv preprint arXiv:1910.06277
2019 arXiv
-
[27]
Machinelearningbasedphishingdetectionfromurls
Sahingoz,O.K.,Buber,E.,Demir,O.,Diri,B.,2019. Machinelearningbasedphishingdetectionfromurls. ExpertSystemswithApplications 117, 345–357
2019
-
[28]
Svms for the blogosphere: Blog identification and splog detection, in: AAAI spring symposium on computational approaches to analysing weblogs
Kolari, P., Finin, T., Joshi, A., et al., 2006. Svms for the blogosphere: Blog identification and splog detection, in: AAAI spring symposium on computational approaches to analysing weblogs
2006
-
[29]
Ma, J., Saul, L.K., Savage, S., Voelker, G.M., 2009a. Beyond blacklists: learning to detect malicious web sites from suspicious urls, in: Proceedings of the 15th ACM SIGKDD international conference on Knowledge discovery and data mining, pp. 1245–1254
-
[30]
Identifyingsuspiciousurls:anapplicationoflarge-scaleonlinelearning,in:Proceedings of the 26th annual international conference on machine learning, pp
Ma,J.,Saul,L.K.,Savage,S.,Voelker,G.M.,2009b. Identifyingsuspiciousurls:anapplicationoflarge-scaleonlinelearning,in:Proceedings of the 26th annual international conference on machine learning, pp. 681–688. :Preprint submitted to Elsevier Page 26 of 34
-
[31]
Malicious url detection based on kolmogorov complexity estimation, in: 2012 IEEE/WIC/ACM International Conferences on Web Intelligence and Intelligent Agent Technology, pp
Pao, H.K., Chou, Y.L., Lee, Y.J., 2012. Malicious url detection based on kolmogorov complexity estimation, in: 2012 IEEE/WIC/ACM International Conferences on Web Intelligence and Intelligent Agent Technology, pp. 380–387. doi:10.1109/WI-IAT.2012.258
2012 doi
-
[32]
Phishscore:Hackingphishers’minds,in:10thinternationalconferenceonnetworkand service management (CNSM) and workshop, IEEE
Marchal,S.,François,J.,State,R.,Engel,T.,2014a. Phishscore:Hackingphishers’minds,in:10thinternationalconferenceonnetworkand service management (CNSM) and workshop, IEEE. pp. 46–54
-
[33]
Phishstorm:Detectingphishingwithstreaminganalytics
Marchal,S.,François,J.,State,R.,Engel,T.,2014b. Phishstorm:Detectingphishingwithstreaminganalytics. IEEETransactionsonNetwork and Service Management 11, 458–471
-
[34]
A framework for detection and measurement of phishing attacks, in: Proceedings of the 2007 ACM workshop on Recurring malcode, pp
Garera, S., Provos, N., Chew, M., Rubin, A.D., 2007. A framework for detection and measurement of phishing attacks, in: Proceedings of the 2007 ACM workshop on Recurring malcode, pp. 1–8
2007
-
[35]
Prophiler:afastfilterforthelarge-scaledetectionofmaliciouswebpages,in:Proceedings of the 20th international conference on World wide web, pp
Canali,D.,Cova,M.,Vigna,G.,Kruegel,C.,2011. Prophiler:afastfilterforthelarge-scaledetectionofmaliciouswebpages,in:Proceedings of the 20th international conference on World wide web, pp. 197–206
2011
-
[36]
Judgingasitebyitscontent:learningthetextual,structural,andvisualfeaturesofmaliciousweb pages, in: Proceedings of the 4th ACM Workshop on Security and Artificial Intelligence, pp
Bannur,S.N.,Saul,L.K.,Savage,S.,2011. Judgingasitebyitscontent:learningthetextual,structural,andvisualfeaturesofmaliciousweb pages, in: Proceedings of the 4th ACM Workshop on Security and Artificial Intelligence, pp. 1–10
2011
-
[37]
Malicious url detection using logistic regression
Rayala, R., Kuppa, R., Pasumarthi, S., Karthik, S., 2023. Malicious url detection using logistic regression. Authorea Preprints
2023
-
[38]
Mutual information based logistic regression for phishing url detection
Vajrobol, V., Gupta, B.B., Gaurav, A., 2024. Mutual information based logistic regression for phishing url detection. Cyber Security and Applications 2, 100044
2024
-
[39]
Learning url embedding for malicious website detection
Yan, X., Xu, Y., Cui, B., Zhang, S., Guo, T., Li, C., 2020. Learning url embedding for malicious website detection. IEEE Transactions on Industrial Informatics 16, 6673–6681. doi:10.1109/TII.2020.2977886
2020
-
[40]
Malicious web content detection by machine learning
Hou, Y.T., Chang, Y., Chen, T., Laih, C.S., Chen, C.M., 2010. Malicious web content detection by machine learning. expert systems with applications 37, 55–60
2010
-
[41]
Cross-layer detection of malicious websites, in: Proceedings of the third ACM conference on Data and application security and privacy, pp
Xu, L., Zhan, Z., Xu, S., Ye, K., 2013. Cross-layer detection of malicious websites, in: Proceedings of the third ACM conference on Data and application security and privacy, pp. 141–152
2013
-
[42]
Phishari: Automatic realtime phishing detection on twitter, in: 2012 eCrime Researchers Summit, IEEE
Aggarwal, A., Rajadesingan, A., Kumaraguru, P., 2012. Phishari: Automatic realtime phishing detection on twitter, in: 2012 eCrime Researchers Summit, IEEE. pp. 1–12
2012
-
[43]
Detection of forwarding-based malicious urls in online social networks
Cao, J., Li, Q., Ji, Y., He, Y., Guo, D., 2016. Detection of forwarding-based malicious urls in online social networks. International Journal of Parallel Programming 44, 163–180
2016
-
[44]
Textual and visual content-based anti-phishing: a bayesian approach
Zhang, H., Liu, G., Chow, T.W., Liu, W., 2011. Textual and visual content-based anti-phishing: a bayesian approach. IEEE transactions on neural networks 22, 1532–1546
2011
-
[45]
Detect malicious web pages using naive bayesian algorithm to detect cyber threats
Magdacy Jerjes, A.Z.A., Dawod, A.Y., Abdulqader, M.F., 2023. Detect malicious web pages using naive bayesian algorithm to detect cyber threats. Wireless Personal Communications , 1–13
2023
-
[46]
Classification of malicious urls using naive bayes and genetic algorithm
Koca, M., Avcı, İ., Al-hayani, M.A.S., 2023. Classification of malicious urls using naive bayes and genetic algorithm. Sakarya University Journal of Computer and Information Sciences 6, 80–90
2023
-
[47]
Vundavalli,V.,Barsha,F.,Masum,M.,Shahriar,H.,Haddad,H.,2020.Maliciousurldetectionusingsupervisedmachinelearningtechniques, in: 13th International Conference on Security of Information and Networks, pp. 1–6
2020
-
[48]
Mankar, N.P., Sakunde, P.E., Zurange, S., Date, A., Borate, V., Mali, Y.K., 2024. Comparative evaluation of machine learning models for malicious url detection, in: 2024 MIT Art, Design and Technology School of Computing International Conference (MITADTSoCiCon), IEEE. pp. 1–7
2024
-
[49]
Onlinepassive-aggressivealgorithms
Crammer,K.,Dekel,O.,Keshet,J.,Shalev-Shwartz,S.,Singer,Y.,2006. Onlinepassive-aggressivealgorithms. JournalofMachineLearning Research 7, 551–585
2006
-
[50]
Confidence-weighted linear classification, in: Proceedings of the 25th international conference on Machine learning, pp
Dredze, M., Crammer, K., Pereira, F., 2008. Confidence-weighted linear classification, in: Proceedings of the 25th international conference on Machine learning, pp. 264–271
2008
-
[51]
Lexical feature based phishing url detection using online learning, in: Proceedings of the 3rd ACM Workshop on Artificial Intelligence and Security, pp
Blum, A., Wardman, B., Solorio, T., Warner, G., 2010. Lexical feature based phishing url detection using online learning, in: Proceedings of the 3rd ACM Workshop on Artificial Intelligence and Security, pp. 54–60
2010
-
[52]
Adaptiveregularizationofweightvectors
Crammer,K.,Kulesza,A.,Dredze,M.,2009. Adaptiveregularizationofweightvectors. Advancesinneuralinformationprocessingsystems 22
2009
-
[53]
Phishdef: Url names say it all, in: 2011 Proceedings IEEE INFOCOM, IEEE
Le, A., Markopoulou, A., Faloutsos, M., 2011. Phishdef: Url names say it all, in: 2011 Proceedings IEEE INFOCOM, IEEE. pp. 191–195
2011
-
[54]
Malicious url filtering—a big data application, in: 2013 IEEE international conference on big data, IEEE
Lin, M.S., Chiu, C.Y., Lee, Y.J., Pao, H.K., 2013. Malicious url filtering—a big data application, in: 2013 IEEE international conference on big data, IEEE. pp. 589–596
2013
-
[55]
Detecting malicious short urls on twitter
Alshboul, Y., Nepali, R., Wang, Y., 2015. Detecting malicious short urls on twitter. AMCIS 2015 Proceedings Search
2015
-
[56]
Anomaly based web phishing page detection, in: 2006 22nd Annual Computer Security Applications Conference (ACSAC’06), IEEE
Pan, Y., Ding, X., 2006. Anomaly based web phishing page detection, in: 2006 22nd Annual Computer Security Applications Conference (ACSAC’06), IEEE. pp. 381–392
2006
-
[57]
Click traffic analysis of short url spam on twitter, in: 9th IEEE international conference on collaborative computing: networking, applications and worksharing, IEEE
Wang, D., Navathe, S.B., Liu, L., Irani, D., Tamersoy, A., Pu, C., 2013. Click traffic analysis of short url spam on twitter, in: 9th IEEE international conference on collaborative computing: networking, applications and worksharing, IEEE. pp. 250–259
2013
-
[58]
Malicious url and intrusion detection using machine learning, in: 2024 International Conference on Information Networking (ICOIN), IEEE
Hamza, A., Hammam, F., Abouzeid, M., Ahmed, M.A., Dhou, S., Aloul, F., 2024. Malicious url and intrusion detection using machine learning, in: 2024 International Conference on Information Networking (ICOIN), IEEE. pp. 795–800
2024
-
[59]
Detecting malicious websites by learning ip address features, in: 2012 IEEE/IPSJ 12th International Symposium on Applications and the Internet, IEEE
Chiba, D., Tobe, K., Mori, T., Goto, S., 2012. Detecting malicious websites by learning ip address features, in: 2012 IEEE/IPSJ 12th International Symposium on Applications and the Internet, IEEE. pp. 29–39
2012
-
[60]
Modeling of ship fuel consumption based on multisource and heterogeneous data: Case study of passenger ship
Zhu, Y., Zuo, Y., Li, T., 2021. Modeling of ship fuel consumption based on multisource and heterogeneous data: Case study of passenger ship. Journal of Marine Science and Engineering 9, 273
2021
-
[61]
O’Reilly Media, Inc
Géron, A., 2022. Hands-on machine learning with Scikit-Learn, Keras, and TensorFlow: Concepts, tools, and techniques to build intelligent systems. " O’Reilly Media, Inc."
2022
-
[62]
Urlnet: Learning a url representation with deep learning for malicious url detection
Le, H., Pham, Q., Sahoo, D., Hoi, S.C., 2018. Urlnet: Learning a url representation with deep learning for malicious url detection. arXiv preprint arXiv:1802.03162 . :Preprint submitted to Elsevier Page 27 of 34
2018 arXiv
-
[63]
Texception:Acharacter/word-leveldeeplearningmodelforphishingurldetection,in: ICASSP2020-2020IEEEInternationalConferenceonAcoustics,SpeechandSignalProcessing(ICASSP),pp.2857–2861
Tajaddodianfar,F.,Stokes,J.W.,Gururajan,A.,2020. Texception:Acharacter/word-leveldeeplearningmodelforphishingurldetection,in: ICASSP2020-2020IEEEInternationalConferenceonAcoustics,SpeechandSignalProcessing(ICASSP),pp.2857–2861. doi:10.1109/ ICASSP40776.2020.9053670
2020
-
[64]
Deepcharacter-levelanomaly detectionbasedona convolutionalautoencoderfor zero-dayphishingurldetection
Bu,S.J., Cho,S.B.,2021. Deepcharacter-levelanomaly detectionbasedona convolutionalautoencoderfor zero-dayphishingurldetection. Electronics 10, 1492
2021
-
[65]
Grambeddings:anewneuralnetworkforurlbasedidentificationofphishingwebpagesthrough n-gram embeddings
Bozkir,A.S.,Dalgic,F.C.,Aydos,M.,2023. Grambeddings:anewneuralnetworkforurlbasedidentificationofphishingwebpagesthrough n-gram embeddings. Computers & Security 124, 102964
2023
-
[66]
Attentionisallyouneed
Vaswani,A.,Shazeer,N.,Parmar,N.,Uszkoreit,J.,Jones,L.,Gomez,A.N.,Kaiser,L.,Polosukhin,I.,2023. Attentionisallyouneed. URL: https://arxiv.org/abs/1706.03762,arXiv:1706.03762
2023 arXiv
-
[67]
Bert: Pre-training of deep bidirectional transformers for language understanding
Devlin, J., Chang, M.W., Lee, K., Toutanova, K., 2019. Bert: Pre-training of deep bidirectional transformers for language understanding. URL:https://arxiv.org/abs/1810.04805,arXiv:1810.04805
2019 arXiv
-
[68]
Detectingphishingurlsusingtheberttransformermodel,in:2023IEEEInternational Conference on Big Data (BigData), pp
Otieno,D.O.,Abri,F.,Namin,A.S.,Jones,K.S.,2023. Detectingphishingurlsusingtheberttransformermodel,in:2023IEEEInternational Conference on Big Data (BigData), pp. 2483–2492. doi:10.1109/BigData59044.2023.10386782
2023
-
[69]
Bert-based approaches to identifying malicious urls
Su, M.Y., Su, K.L., 2023. Bert-based approaches to identifying malicious urls. Sensors 23, 8499
2023
-
[70]
S,J.K.,B,A.,2023.Phishingurldetectionbyleveragingrobertaforfeatureextractionandlstmforclassification,in:2023SecondInternational Conference on Augmented Intelligence and Sustainable Systems (ICAISS), pp. 972–977. doi:10.1109/ICAISS58487.2023.10250684
2023
-
[71]
Asiri,S.,Xiao,Y.,Li,T.,2024.Phishtransformer:Anovelapproachtodetectphishingattacksusingurlcollectionandtransformer.Electronics
2024
-
[72]
URL:https://www.mdpi.com/2079-9292/13/1/30, doi:10.3390/electronics13010030
-
[73]
Tcurl: Exploring hybrid transformer and convolutional neural network on phishing url detection
Wang, C., Chen, Y., 2022. Tcurl: Exploring hybrid transformer and convolutional neural network on phishing url detection. Knowledge- Based Systems 258, 109955. URL:https://www.sciencedirect.com/science/article/pii/S0950705122010486, doi:https: //doi.org/10.1016/j.knosys.2022.109955
2022
-
[74]
A hybrid transformer ensemble approach for phishing website detection, in: 2023 International Conference on Self Sustainable Artificial Intelligence Systems (ICSSAS), pp
Mandapati, K.S., Meesala, S., Maddela, D., Ponnada, K., Neyyala, H., Shaik, E.A., 2023. A hybrid transformer ensemble approach for phishing website detection, in: 2023 International Conference on Self Sustainable Artificial Intelligence Systems (ICSSAS), pp. 1–8. doi:10.1109/I...
2023
-
[75]
Transurl:Improvingmaliciousurldetectionwithmulti-layertransformer encoding and multi-scale pyramid features
Liu,R.,Wang,Y.,Guo,Z.,Xu,H.,Qin,Z.,Ma,W.,Zhang,F.,2024. Transurl:Improvingmaliciousurldetectionwithmulti-layertransformer encoding and multi-scale pyramid features. Computer Networks 253, 110707. URL:https://www.sciencedirect.com/science/ article/pii/S1389128624005395, doi:htt...
2024
-
[76]
An integrated model based on deep learning classifiers and pre-trained transformer for phishing url detection
Do, N.Q., Selamat, A., Fujita, H., Krejcar, O., 2024. An integrated model based on deep learning classifiers and pre-trained transformer for phishing url detection. Future Generation Computer Systems 161, 269–285. URL:https://www.sciencedirect.com/science/ article/pii/S0167739...
2024 doi
-
[77]
Urdu text reuse detection at phrasal level using sentence transformer-based approach
Mehak, G., Muneer, I., Nawab, R.M.A., 2023. Urdu text reuse detection at phrasal level using sentence transformer-based approach. Expert Systems with Applications 234, 121063. URL:https://www.sciencedirect.com/science/article/pii/S0957417423015658, doi:https://doi.org/10.1016/...
2023
-
[78]
A large-scale pretrained deep model for phishing url detection, in: ICASSP 2023-2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), IEEE
Wang, Y., Zhu, W., Xu, H., Qin, Z., Ren, K., Ma, W., 2023a. A large-scale pretrained deep model for phishing url detection, in: ICASSP 2023-2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), IEEE. pp. 1–5
2023
-
[79]
Alightweightmulti-viewlearningapproachforphishingattackdetectionusingtransformer with mixture of experts
Wang,Y.,Ma,W.,Xu,H.,Liu,Y.,Yin,P.,2023b. Alightweightmulti-viewlearningapproachforphishingattackdetectionusingtransformer with mixture of experts. Applied Sciences 13, 7429
-
[80]
Exposure: Finding malicious domains using passive dns analysis., in: Ndss, pp
Bilge, L., Kirda, E., Kruegel, C., Balduzzi, M., 2011. Exposure: Finding malicious domains using passive dns analysis., in: Ndss, pp. 1–17
2011
-
[81]
Journal of Network and Computer Applications , 104170
Aljofey,A.,Bello,S.A.,Lu,J.,Xu,C.,2025.Comprehensivephishingdetection:Amulti-channelapproachwithvariantstcnfusionleveraging url and html features. Journal of Network and Computer Applications , 104170
2025
-
[82]
A deep learning-based phishing detection system using cnn, lstm, and lstm-cnn
Alshingiti, Z., Alaqel, R., Al-Muhtadi, J., Haq, Q.E.U., Saleem, K., Faheem, M.H., 2023. A deep learning-based phishing detection system using cnn, lstm, and lstm-cnn. Electronics 12, 232
2023
-
[83]
Researchonwebsitephishingdetectionbasedonlstmrnn,in:2020IEEE4thInformationTechnology,Networking,Electronic and Automation Control Conference (ITNEC), pp
Su,Y.,2020. Researchonwebsitephishingdetectionbasedonlstmrnn,in:2020IEEE4thInformationTechnology,Networking,Electronic and Automation Control Conference (ITNEC), pp. 284–288. doi:10.1109/ITNEC48623.2020.9084799
2020
-
[84]
Performanceanalysisofmaliciousurldetectionbyusingrnnandlstm,in:2020FourthInternational ConferenceonComputingMethodologiesandCommunication(ICCMC),pp.454–458
Arivukarasi,M.,Antonidoss,A.,2020. Performanceanalysisofmaliciousurldetectionbyusingrnnandlstm,in:2020FourthInternational ConferenceonComputingMethodologiesandCommunication(ICCMC),pp.454–458. doi:10.1109/ICCMC48092.2020.ICCMC-00085
2020
-
[85]
Antiphishstack: Lstm-based stacked generalization model for optimized phishing url detection
Aslam, S., Aslam, H., Manzoor, A., Chen, H., Rasool, A., 2024. Antiphishstack: Lstm-based stacked generalization model for optimized phishing url detection. Symmetry 16, 248
2024
-
[86]
Gupta, N., Thapliyal, S., Sharma, A., Sheladia, J., Wazid, M., Giri, D., 2024. Deep learning approach for malicious url detection using cnn, rnn, lstm and bi-lstm models, in: 2024 6th International Conference on Computational Intelligence and Networks (CINE), IEEE. pp. 1–5
2024
-
[87]
A hybrid dnn–lstm model for detecting phishing urls
Ozcan, A., Catal, C., Donmez, E., Senturk, B., 2023. A hybrid dnn–lstm model for detecting phishing urls. Neural Computing and Applications , 1–17
2023
-
[88]
A stacking model using url and html features for phishing webpage detection
Li, Y., Yang, Z., Chen, X., Yuan, H., Liu, W., 2019. A stacking model using url and html features for phishing webpage detection. Future Generation Computer Systems 94, 27–39
2019
-
[89]
Delta: automatic identification of unknown web-based infection campaigns, in: Proceedings of the 2013 ACM SIGSAC conference on Computer & communications security, pp
Borgolte, K., Kruegel, C., Vigna, G., 2013. Delta: automatic identification of unknown web-based infection campaigns, in: Proceedings of the 2013 ACM SIGSAC conference on Computer & communications security, pp. 109–120
2013
-
[90]
Identification of malicious web pages with static heuristics, in: 2008 Australasian Telecommunication Networks and Applications Conference, IEEE
Seifert, C., Welch, I., Komisarczuk, P., 2008. Identification of malicious web pages with static heuristics, in: 2008 Australasian Telecommunication Networks and Applications Conference, IEEE. pp. 91–96
2008
-
[91]
Phishing detection based associative classification data mining
Abdelhamid, N., Ayesh, A., Thabtah, F., 2014. Phishing detection based associative classification data mining. Expert Systems with Applications 41, 5948–5959
2014
-
[92]
Phishhaven—anefficientreal-timeaiphishingurlsdetectionsystem
Sameen,M.,Han,K.,Hwang,S.O.,2020. Phishhaven—anefficientreal-timeaiphishingurlsdetectionsystem. IeeeAccess8,83425–83443
2020
-
[93]
Analysis of phishing website detection using cnn and bidirectional lstm, in: 2020 4th International Conference on Electronics, Communication and Aerospace Technology (ICECA), pp
Pooja, A.S.S.V.L., Sridhar, M., 2020. Analysis of phishing website detection using cnn and bidirectional lstm, in: 2020 4th International Conference on Electronics, Communication and Aerospace Technology (ICECA), pp. 1620–1629. doi:10.1109/ICECA49313.2020. :Preprint submitted ...
2020
-
[94]
Pham,T.T.T.,Hoang,V.N.,Ha,T.N.,2018. Exploringefficiencyofcharacter-levelconvolutionneuronnetworkandlongshorttermmemory onmaliciousurldetection,in:Proceedingsofthe2018VIIInternationalConferenceonNetwork,CommunicationandComputing,pp.82–86
2018
-
[95]
Look before you leap: Detecting phishing web pages by exploiting raw url and html characteristics
Opara, C., Chen, Y., Wei, B., 2024. Look before you leap: Detecting phishing web pages by exploiting raw url and html characteristics. Expert Systems with Applications 236, 121183
2024
-
[96]
Web2vec: Phishing webpage detection method based on multidimensional features driven by deep learning
Feng, J., Zou, L., Ye, O., Han, J., 2020. Web2vec: Phishing webpage detection method based on multidimensional features driven by deep learning. IEEE Access 8, 221214–221224. doi:10.1109/ACCESS.2020.3043188
2020
-
[97]
Detecting phishing attacks using a combined model of lstm and cnn
Ariyadasa, S., Fernando, S., Fernando, S., 2020. Detecting phishing attacks using a combined model of lstm and cnn. International Journal of ADVANCED AND APPLIED SCIENCES 7, 56–67. doi:10.21833/ijaas.2020.07.007
2020 doi
-
[98]
Manjula, M., Venkatesh, Kenchamma, R.H., Basapur, S.B., 2024. Pd-uhd features: Phishing detection approach using uncooked url, html content and domain name features, in: 2024 Second International Conference on Networks, Multimedia and Information Technology (NMITCON), pp. 1–8....
2024
-
[99]
Htmlphish:Enablingphishingwebpagedetectionbyapplyingdeeplearningtechniquesonhtmlanalysis, in: 2020 International Joint Conference on Neural Networks (IJCNN), pp
Opara,C.,Wei,B.,Chen,Y.,2020. Htmlphish:Enablingphishingwebpagedetectionbyapplyingdeeplearningtechniquesonhtmlanalysis, in: 2020 International Joint Conference on Neural Networks (IJCNN), pp. 1–8. doi:10.1109/IJCNN48605.2020.9207707
2020
-
[100]
Combining long-term recurrent convolutional and graph convolutional networks to detect phishing sites using url and html
Ariyadasa, S., Fernando, S., Fernando, S., 2022. Combining long-term recurrent convolutional and graph convolutional networks to detect phishing sites using url and html. IEEE Access 10, 82355–82375
2022
-
[101]
Ouyang,L.,Zhang,Y.,2021.Phishingwebpagedetectionwithhtml-levelgraphneuralnetwork,in:2021IEEE20thInternationalConference onTrust,SecurityandPrivacyinComputingandCommunications(TrustCom),pp.952–958.doi:10.1109/TrustCom53373.2021.00133
2021
-
[102]
Reinforceddisentangledhtmlrepresentationlearningwithhard-sampleminingforphishingwebpage detection
Yoon,J.H.,Buu,S.J.,Kim,H.J.,2025. Reinforceddisentangledhtmlrepresentationlearningwithhard-sampleminingforphishingwebpage detection. Electronics 14, 1080
2025
-
[103]
Huang, Y., Yang, Q., Qin, J., Wen, W., 2019. Phishing url detection via cnn and attention-based hierarchical rnn, in: 2019 18th IEEE International Conference On Trust, Security And Privacy In Computing And Communications/13th IEEE International Conference On Big Data Science A...
2019
-
[104]
Phishingwebsitedetectionthroughmulti-modelanalysisofhtmlcontent, in: International Conference on Theoretical and Applied Computing, Springer
Çolhak,F.,Ecevit,M.İ.,Uçar,B.E.,Creutzburg,R.,Dağ,H.,2024. Phishingwebsitedetectionthroughmulti-modelanalysisofhtmlcontent, in: International Conference on Theoretical and Applied Computing, Springer. pp. 171–184
2024
-
[105]
Detecting malicious web links and identifying their attack types, in: 2nd USENIX Conference on Web Application Development (WebApps 11)
Choi, H., Zhu, B.B., Lee, H., 2011. Detecting malicious web links and identifying their attack types, in: 2nd USENIX Conference on Web Application Development (WebApps 11)
2011
-
[106]
A deep learning approach for detecting malicious javascript code
Wang, Y., Cai, W.d., Wei, P.c., 2016. A deep learning approach for detecting malicious javascript code. Security and Communication Networks 9, 1520–1534
2016
-
[107]
Laskov, P., Šrndić, N., 2011. Static detection of malicious javascript-bearing pdf documents, in: Proceedings of the 27th Annual Computer Security Applications Conference, Association for Computing Machinery, New York, NY, USA. p. 373–382. URL:https://doi.org/ 10.1145/2076732....
2011
-
[108]
Xu, W., Zhang, F., Zhu, S., 2013. Jstill: mostly static detection of obfuscated malicious javascript code, in: Proceedings of the Third ACM Conference on Data and Application Security and Privacy, Association for Computing Machinery, New York, NY, USA. p. 117–128. URL: https:/...
2013
-
[109]
Jsod: Javascript obfuscation detector
AL-Taharwa, I.A., Lee, H.M., Jeng, A.B., Wu, K.P., Ho, C.S., Chen, S.M., 2015. Jsod: Javascript obfuscation detector. Security and Communication Networks 8, 1092–1107. URL:https://doi.org/10.1002/sec.1064, doi:10.1002/sec.1064
2015 doi
-
[110]
Nicolay,J.,Spruyt,V.,DeRoover,C.,2016. Staticdetectionofuser-specifiedsecurityvulnerabilitiesinclient-sidejavascript,in:Proceedings of the 2016 ACM Workshop on Programming Languages and Analysis for Security, Association for Computing Machinery, New York, NY, USA. p. 3–13. URL...
2016
-
[111]
Gorji, A., Abadi, M., 2014. Detecting obfuscated javascript malware using sequences of internal function calls, in: Proceedings of the 2014 ACMSoutheastConference,AssociationforComputingMachinery,NewYork,NY,USA. URL:https://doi.org/10.1145/2638404. 2737181, doi:10.1145/2638404.2737181
2014
-
[112]
Detectionandmitigationofmaliciousjavascriptusinginformationflowcontrol,in:2014Twelfth Annual International Conference on Privacy, Security and Trust, pp
Sayed,B.,Traoré,I.,Abdelhalim,A.,2014. Detectionandmitigationofmaliciousjavascriptusinginformationflowcontrol,in:2014Twelfth Annual International Conference on Privacy, Security and Trust, pp. 264–273. doi:10.1109/PST.2014.6890948
2014
-
[113]
Xue,Y.,Wang,J.,Liu,Y.,Xiao,H.,Sun,J.,Chandramohan,M.,2015. Detectionandclassificationofmaliciousjavascriptviaattackbehavior modelling, in: Proceedings of the 2015 International Symposium on Software Testing and Analysis, Association for Computing Machinery, New York, NY, USA. ...
2015
-
[114]
Schütt, K., Kloft, M., Bikadorov, A., Rieck, K., 2012. Early detection of malicious behavior in javascript code, in: Proceedings of the 5th ACM Workshop on Security and Artificial Intelligence, Association for Computing Machinery, New York, NY, USA. p. 15–24. URL: https://doi....
2012
-
[115]
Corona, I., Maiorca, D., Ariu, D., Giacinto, G., 2014. Lux0r: Detection of malicious pdf-embedded javascript code through discriminant analysisofapireferences,in:Proceedingsofthe2014WorkshoponArtificialIntelligentandSecurityWorkshop,AssociationforComputing Machinery, New York,...
2014
-
[116]
Wang, J., Xue, Y., Liu, Y., Tan, T.H., 2015. Jsdc: A hybrid approach for javascript malware detection and classification, in: Proceedings of the 10th ACM Symposium on Information, Computer and Communications Security, Association for Computing Machinery, New York, NY, USA. p. ...
2015
-
[117]
Obfuscatedmaliciousjavascriptdetectionschemeusingthefeaturebasedondivided url, in: 2017 23rd Asia-Pacific Conference on Communications (APCC), pp
Morishige,S.,Haruta,S.,Asahina,H.,Sasase,I.,2017. Obfuscatedmaliciousjavascriptdetectionschemeusingthefeaturebasedondivided url, in: 2017 23rd Asia-Pacific Conference on Communications (APCC), pp. 1–6. doi:10.23919/APCC.2017.8303992
2017
-
[118]
Detectionofmaliciousjavascriptonanimbalanceddataset
Phung,N.M.,Mimura,M.,2021. Detectionofmaliciousjavascriptonanimbalanceddataset. InternetofThings13,100357. URL:https:// www.sciencedirect.com/science/article/pii/S2542660521000019, doi:https://doi.org/10.1016/j.iot.2021.100357
2021
-
[119]
A machine learning approach to malicious javascript detection using fixed length vector representation, in: 2018 International Joint Conference on Neural Networks (IJCNN), pp
Ndichu, S., Ozawa, S., Misu, T., Okada, K., 2018. A machine learning approach to malicious javascript detection using fixed length vector representation, in: 2018 International Joint Conference on Neural Networks (IJCNN), pp. 1–8. doi:10.1109/IJCNN.2018.8489414. :Preprint subm...
2018
-
[120]
Jodavi,M.,Abadi,M.,Parhizkar,E.,2015. Jsobfusdetector:Abinarypso-basedone-classclassifierensembletodetectobfuscatedjavascript code, in: 2015 The International Symposium on Artificial Intelligence and Signal Processing (AISP), pp. 322–327. doi:10.1109/AISP. 2015.7123508
2015
-
[121]
Research on malicious javascript detection technology based on lstm
Fang, Y., Huang, C., Liu, L., Xue, M., 2018. Research on malicious javascript detection technology based on lstm. IEEE Access 6, 59118– 59125. doi:10.1109/ACCESS.2018.2874098
2018
-
[122]
Malicious javascript detection based on bidirectional lstm model
Song, X., Chen, C., Cui, B., Fu, J., 2020. Malicious javascript detection based on bidirectional lstm model. Applied Sciences 10. URL: https://www.mdpi.com/2076-3417/10/10/3440, doi:10.3390/app10103440
2020 doi
-
[123]
Taylor–hho algorithm: A hybrid optimization algorithm with deep long short-term for malicious javascript detection
Alex, S., Dhiliphan Rajkumar, T., 2021. Taylor–hho algorithm: A hybrid optimization algorithm with deep long short-term for malicious javascript detection. International Journal of Intelligent Systems 36, 7153–7176
2021
-
[124]
Dabral, S., Agarwal, A., Mahajan, M., Kumar, S., 2017. Malicious pdf files detection using structural and javascript based features, in: Information,CommunicationandComputingTechnology:SecondInternationalConference,ICICCT2017,NewDelhi,India,May13,2017, Revised Selected Papers ...
2017
-
[125]
Wang, Q., Zhou, J., Chen, Y., Zhang, Y., Zhao, J., 2013. Extracting urls from javascript via program analysis, in: Proceedings of the 2013 9thJointMeetingonFoundationsofSoftwareEngineering,AssociationforComputingMachinery,NewYork,NY,USA.p.627–630. URL: https://doi.org/10.1145/...
2013
-
[126]
Maliciousjavascriptcodedetectionbasedonhybridanalysis,in:201825thAsia-PacificSoftwareEngineering Conference (APSEC), IEEE
He,X.,Xu,L.,Cha,C.,2018. Maliciousjavascriptcodedetectionbasedonhybridanalysis,in:201825thAsia-PacificSoftwareEngineering Conference (APSEC), IEEE. pp. 365–374
2018
-
[127]
Jstrong: Malicious javascript detection based on code semantic representation and graph neural network
Fang, Y., Huang, C., Zeng, M., Zhao, Z., Huang, C., 2022. Jstrong: Malicious javascript detection based on code semantic representation and graph neural network. Computers & Security 118, 102715. URL:https://www.sciencedirect.com/science/article/pii/ S0167404822001110, doi:htt...
2022
-
[128]
Detecting malicious javascript using structure-based analysis of graph representation
Rozi, M.F., Ban, T., Ozawa, S., Yamada, A., Takahashi, T., Kim, S., Inoue, D., 2023. Detecting malicious javascript using structure-based analysis of graph representation. IEEE Access 11, 102727–102745. doi:10.1109/ACCESS.2023.3317266
2023
-
[129]
A survey and classification of web phishing detection schemes
Varshney, G., Misra, M., Atrey, P.K., 2016. A survey and classification of web phishing detection schemes. Security and Communication Networks 9, 6266–6284
2016
-
[130]
A survey of url-based phishing detection, in: DEIM forum, pp
Aung, E.S., Zan, C.T., Yamana, H., 2019. A survey of url-based phishing detection, in: DEIM forum, pp. G2–3
2019
-
[131]
Detecting phishing web pages with visual similarity assessment based on earth mover’s distance (emd)
Fu, A.Y., Wenyin, L., Deng, X., 2006. Detecting phishing web pages with visual similarity assessment based on earth mover’s distance (emd). IEEE transactions on dependable and secure computing 3, 301–311
2006
-
[132]
Anantiphishingstrategybasedonvisualsimilarityassessment
Liu,W.,Deng,X.,Huang,G.,Fu,A.Y.,2006. Anantiphishingstrategybasedonvisualsimilarityassessment. IEEEInternetComputing10, 58–65
2006
-
[133]
Detectionofphishingwebpagesbasedonvisualsimilarity,in:Specialinterest tracks and posters of the 14th international conference on World Wide Web, pp
Wenyin,L.,Huang,G.,Xiaoyue,L.,Min,Z.,Deng,X.,2005. Detectionofphishingwebpagesbasedonvisualsimilarity,in:Specialinterest tracks and posters of the 14th international conference on World Wide Web, pp. 1060–1061
2005
-
[134]
Goldphish:Usingimagesforcontent-basedphishinganalysis,in:2010Fifthinternationalconference on internet monitoring and protection, IEEE
Dunlop,M.,Groat,S.,Shelly,D.,2010. Goldphish:Usingimagesforcontent-basedphishinganalysis,in:2010Fifthinternationalconference on internet monitoring and protection, IEEE. pp. 123–128
2010
-
[135]
Distinctive image features from scale-invariant keypoints
Lowe, D.G., 2004. Distinctive image features from scale-invariant keypoints. International journal of computer vision 60, 91–110
2004
-
[136]
Phishzoo: Detecting phishing websites by looking at them, in: 2011 IEEE fifth international conference on semantic computing, IEEE
Afroz, S., Greenstadt, R., 2011. Phishzoo: Detecting phishing websites by looking at them, in: 2011 IEEE fifth international conference on semantic computing, IEEE. pp. 368–375
2011
-
[137]
Mitigate web phishing using site signatures, in: TENCON 2010 - 2010 IEEE Region 10 Conference, pp
Huang, C.Y., Ma, S.P., Yeh, W.L., Lin, C.Y., Liu, C.T., 2010. Mitigate web phishing using site signatures, in: TENCON 2010 - 2010 IEEE Region 10 Conference, pp. 803–808. doi:10.1109/TENCON.2010.5686582
2010
-
[138]
Verilogo: Proactive phishing detection via logo recognition
Wang, G., Liu, H., Becerra, S., Wang, K., Belongie, S., Shacham, H., Savage, S., 2011. Verilogo: Proactive phishing detection via logo recognition. Department of Computer Science main & Engineering
2011
-
[139]
Intelligent visual similarity-based phishing websites detection
Chen, J.L., Ma, Y.W., Huang, K.L., 2020. Intelligent visual similarity-based phishing websites detection. Symmetry 12. URL:https: //www.mdpi.com/2073-8994/12/10/1681
2020
-
[140]
Speeded-up robust features (surf)
Bay, H., Ess, A., Tuytelaars, T., Van Gool, L., 2008. Speeded-up robust features (surf). Computer vision and image understanding 110, 346–359
2008
-
[141]
Acomputervisiontechniquetodetectphishingattacks,in:2015FifthInternationalConferenceonCommunication Systems and Network Technologies, pp
Rao,R.S.,Ali,S.T.,2015. Acomputervisiontechniquetodetectphishingattacks,in:2015FifthInternationalConferenceonCommunication Systems and Network Technologies, pp. 596–601. doi:10.1109/CSNT.2015.68
2015 doi
-
[142]
Comparisons of machine learning techniques for detecting malicious webpages
Kazemian, H., Ahmed, S., 2015. Comparisons of machine learning techniques for detecting malicious webpages. Expert Systems with Applications 42, 1166–1177. URL:https://www.sciencedirect.com/science/article/pii/S0957417414005284, doi:https: //doi.org/10.1016/j.eswa.2014.08.046
2015 doi
-
[143]
Contrast context histogram—an efficient discriminating local descriptor for object recognition and image matching
Huang, C.R., Chen, C.S., Chung, P.C., 2008. Contrast context histogram—an efficient discriminating local descriptor for object recognition and image matching. Pattern Recognition 41, 3071–3077. URL:https://www.sciencedirect.com/science/article/pii/ S0031320308000988, doi:https...
2008 doi
-
[144]
Fighting phishing with discriminative keypoint features
Chen, K.T., Chen, J.Y., Huang, C.R., Chen, C.S., 2009. Fighting phishing with discriminative keypoint features. IEEE Internet Computing 13, 56–63
2009
-
[145]
Antiphishing model with url & image based webpage matching
Arade, M.S., Bhaskar, P., Kamat, R., 2011. Antiphishing model with url & image based webpage matching. Int. J. Comput. Sci. Technol. IJCST 2, 282–286
2011
-
[146]
Mitigating online fraud by ant phishing model with url & image based webpage matching
Balamuralikrishna, T., Raghavendrasai, N., Sukumar, M.S., 2012. Mitigating online fraud by ant phishing model with url & image based webpage matching. International Journal of Scientific & Engineering Research 3, 1–6
2012
-
[147]
Adeeplearningtechniqueforwebphishingdetectioncombinedurlfeaturesandvisualsimilarity
Al-Ahmadi,S.,2020. Adeeplearningtechniqueforwebphishingdetectioncombinedurlfeaturesandvisualsimilarity. InternationalJournal of Computer Networks & Communications (IJCNC) Vol 12
2020
-
[148]
Abdelnabi, S., Krombholz, K., Fritz, M., 2020. Visualphishnet: Zero-day phishing website detection by visual similarity, in: Proceedings of the 2020 ACM SIGSAC Conference on Computer and Communications Security, Association for Computing Machinery, New York, NY, USA. p. 1681–1...
2020
-
[149]
Lam, I.F., Xiao, W.C., Wang, S.C., Chen, K.T., 2009. Counteracting phishing page polymorphism: An image layout analysis approach, in: Advances in Information Security and Assurance, Third International Conference and Workshops, ISA 2009, pp. 270–279. doi:10.1007/ 978-3-642-02617-1_28
2009
-
[150]
A threshold selection method from gray-level histograms
Otsu, N., 1979. A threshold selection method from gray-level histograms. IEEE Transactions on Systems, Man, and Cybernetics 9, 62–66. doi:10.1109/TSMC.1979.4310076
1979
-
[151]
Baitalarm: Detecting phishing sites using similarity in fundamental visual features, in: 2013 5th International Conference on Intelligent Networking and Collaborative Systems, pp
Mao, J., Li, P., Li, K., Wei, T., Liang, Z., 2013. Baitalarm: Detecting phishing sites using similarity in fundamental visual features, in: 2013 5th International Conference on Intelligent Networking and Collaborative Systems, pp. 790–795. doi:10.1109/INCoS.2013.151
2013 doi
-
[152]
Hybrid solution to detect and filter zero-day phishing attacks, in: International Conference on Emerging Research in Computing, Information, Communication and Applications
Gupta, B.B., Mishra, A., 2014. Hybrid solution to detect and filter zero-day phishing attacks, in: International Conference on Emerging Research in Computing, Information, Communication and Applications
2014
-
[153]
Visual similarity-based phishing detection without victim site information, in: 2009 IEEE Symposium on Computational Intelligence in Cyber Security, pp
Hara, M., Yamada, A., Miyake, Y., 2009. Visual similarity-based phishing detection without victim site information, in: 2009 IEEE Symposium on Computational Intelligence in Cyber Security, pp. 30–36. doi:10.1109/CICYBS.2009.4925087
2009
-
[154]
A svm-based technique to detect phishing urls
Huang, H., Qian, L., Wang, Y., 2012. A svm-based technique to detect phishing urls. Information Technology Journal 11, 921–925
2012
-
[155]
Detectionofhiddenfraudulenturlswithintrustedsitesusinglexicalfeatures,in:2013International Conference on Availability, Reliability and Security, IEEE
Sorio,E.,Bartoli,A.,Medvet,E.,2013. Detectionofhiddenfraudulenturlswithintrustedsitesusinglexicalfeatures,in:2013International Conference on Availability, Reliability and Security, IEEE. pp. 242–247
2013
-
[156]
Malicious urls detection using decision tree classifiers and majority voting technique
Patil, D.R., Patil, J.B., et al., 2018. Malicious urls detection using decision tree classifiers and majority voting technique. Cybernetics and Information Technologies 18, 11–29
2018
-
[157]
Machado,L.,Gadge,J.,2017.Phishingsitesdetectionbasedonc4.5decisiontreealgorithm,in:2017InternationalConferenceonComputing, Communication, Control and Automation (ICCUBEA), pp. 1–5. doi:10.1109/ICCUBEA.2017.8463818
2017
-
[158]
Malicious website detection using random forest and pearson correlation for effective feature selection
Sangra, E., Agrawal, R., Gundalwar, P.R., Sharma, K., Bangri, D., Nandi, D., 2024. Malicious website detection using random forest and pearson correlation for effective feature selection. International Journal of Advanced Computer Science and Applications 15. URL: http://dx.do...
2024
-
[159]
Dtof-ann:Anartificialneuralnetworkphishingdetectionmodelbasedondecisiontreeand optimal features
Zhu,E.,Ju,Y.,Chen,Z.,Liu,F.,Fang,X.,2020. Dtof-ann:Anartificialneuralnetworkphishingdetectionmodelbasedondecisiontreeand optimal features. Applied Soft Computing 95
2020
-
[160]
Longshort-termmemory
Hochreiter,S.,Schmidhuber,J.,1997. Longshort-termmemory. NeuralComputation9,1735–1780. doi:10.1162/neco.1997.9.8.1735
1997 doi
-
[161]
Comparative study of catboost, xgboost, and lightgbm for enhanced url phishing detection: a performance assessment
Odeh, A., Al-Haija, Q.A., Aref, A., Taleb, A.A., 2023. Comparative study of catboost, xgboost, and lightgbm for enhanced url phishing detection: a performance assessment. Journal of Internet Services and Information Security 13, 1–11
2023
-
[162]
Improving phishing website detection using a hybrid two-level framework for feature selection and xgboost tuning
Jovanovic, L., Jovanovic, D., Antonijevic, M., Nikolic, B., Bacanin, N., Zivkovic, M., Strumberger, I., 2023. Improving phishing website detection using a hybrid two-level framework for feature selection and xgboost tuning. Journal of Web Engineering 22, 543–574. doi:10.13052/...
2023
-
[163]
Malicious-url detection using logistic regression technique
Vanitha, N., Vinodhini, V., 2019. Malicious-url detection using logistic regression technique. International Journal of Engineering and Management Research (IJEMR) 9, 108–113
2019
-
[164]
Machine learning algorithms for detection of cyber threats using logistic regression
Gonaygunta, H., 2023. Machine learning algorithms for detection of cyber threats using logistic regression. Department of Information Technology, University of the Cumberlands
2023
-
[165]
Machine learning-based malicious website detection using logistic regression algorithm
Pastika, P.B., Alamsyah, A., 2024. Machine learning-based malicious website detection using logistic regression algorithm. Engineering, MAthematics and Computer Science Journal (EMACS) 6, 207–213
2024
-
[166]
Thakur,I.,Panda,K.,Kumar,S.,2022. Deeplearningmethodsformaliciousurldetectionusingembeddingtechniquesaslogisticregression with lasso penalty and random forest, in: 2022 Seventh International Conference on Parallel, Distributed and Grid Computing (PDGC), pp. 181–186. doi:10.110...
2022
-
[167]
Llms are one-shot url classifiers and explainers
Rashid, F., Ranaweera, N., Doyle, B., Seneviratne, S., 2025. Llms are one-shot url classifiers and explainers. Computer Networks 258, 111004
2025
-
[168]
Prompting large language models for malicious webpage detection, in: 2023 IEEE 4th International Conference on Pattern Recognition and Machine Learning (PRML), pp
Li, L., Gong, B., 2023. Prompting large language models for malicious webpage detection, in: 2023 IEEE 4th International Conference on Pattern Recognition and Machine Learning (PRML), pp. 393–400. doi:10.1109/PRML59573.2023.10348229
2023
-
[169]
Multimodal large language models for phishing webpage detection and identification
Lee, J., Lim, P., Hooi, B., Divakaran, D.M., 2024. Multimodal large language models for phishing webpage detection and identification. arXiv preprint arXiv:2408.05941
2024 arXiv
-
[170]
Chatphishdetector: Detecting phishing sites using large language models
Koide, T., Nakano, H., Chiba, D., 2024. Chatphishdetector: Detecting phishing sites using large language models. IEEE Access
2024
-
[171]
Apollo:Agpt-basedtooltodetectphishingemailsandgenerateexplanationsthatwarnusers
Desolda,G.,Greco,F.,Viganò,L.,2024. Apollo:Agpt-basedtooltodetectphishingemailsandgenerateexplanationsthatwarnusers. arXiv preprint arXiv:2410.07997
2024 arXiv
-
[172]
Prompt engineering or fine-tuning? a case study on phishing detection with large language models
Trad, F., Chehab, A., 2024. Prompt engineering or fine-tuning? a case study on phishing detection with large language models. Machine Learning and Knowledge Extraction 6, 367–384
2024
-
[173]
Pmanet:Maliciousurldetectionviapost-trainedlanguagemodelguided multi-level feature attention network
Liu,R.,Wang,Y.,Xu,H.,Qin,Z.,Zhang,F.,Liu,Y.,Cao,Z.,2025. Pmanet:Maliciousurldetectionviapost-trainedlanguagemodelguided multi-level feature attention network. Information Fusion 113, 102638. URL:https://www.sciencedirect.com/science/article/ pii/S1566253524004160, doi:https://...
2025
-
[174]
Memon, A., Manjotho, A.A., 2024. Apformer: Anti-phishing transformer for website-phishing detection via joint feature learning, in: 2024 International Conference on Engineering & Computing Technologies (ICECT), IEEE. pp. 1–5
2024
-
[175]
Life-long phishing attack detection using continual learning
Ejaz, A., Mian, A.N., Manzoor, S., 2023. Life-long phishing attack detection using continual learning. Scientific reports 13, 11488
2023
-
[176]
Enhancing unsupervised anomaly detection with score-guided network
Huang, Z., Zhang, B., Hu, G., Li, L., Xu, Y., Jin, Y., 2023. Enhancing unsupervised anomaly detection with score-guided network. IEEE Transactions on Neural Networks and Learning Systems
2023
-
[177]
Self-supervision-augmented deep autoencoder for unsupervised visual anomaly detection
Huang, C., Yang, Z., Wen, J., Xu, Y., Jiang, Q., Yang, J., Wang, Y., 2021. Self-supervision-augmented deep autoencoder for unsupervised visual anomaly detection. IEEE Transactions on Cybernetics 52, 13834–13847
2021
-
[178]
Learning to detect malicious urls
Ma, J., Saul, L.K., Savage, S., Voelker, G.M., 2011. Learning to detect malicious urls. ACM Transactions on Intelligent Systems and Technology (TIST) 2, 1–24
2011
-
[179]
Malicious web page detection based on on-line learning algorithm, in: 2011 International Conference on Machine Learning and Cybernetics, IEEE
Zhang, W., Ding, Y.X., Tang, Y., Zhao, B., 2011. Malicious web page detection based on on-line learning algorithm, in: 2011 International Conference on Machine Learning and Cybernetics, IEEE. pp. 1914–1919. :Preprint submitted to Elsevier Page 31 of 34
2011
-
[180]
Ma, J., Kulesza, A., Dredze, M., Crammer, K., Saul, L., Pereira, F., 2010. Exploiting feature covariance in high-dimensional online learning,in:ProceedingsoftheThirteenthInternationalConferenceonArtificialIntelligenceandStatistics,JMLRWorkshopandConference Proceedings. pp. 493–500
2010
-
[181]
Design and evaluation of a real-time url spam filtering service, in: 2011 IEEE symposium on security and privacy, IEEE
Thomas, K., Grier, C., Ma, J., Paxson, V., Song, D., 2011. Design and evaluation of a real-time url spam filtering service, in: 2011 IEEE symposium on security and privacy, IEEE. pp. 447–462
2011
-
[183]
expose: A character-level convolutional neural network with embeddings for detecting malicious urls, file paths and registry keys
Saxe, J., Berlin, K., 2017. expose: A character-level convolutional neural network with embeddings for detecting malicious urls, file paths and registry keys. arXiv preprint arXiv:1702.08568
2017 arXiv
-
[184]
Classification of url bitstreams using bag of bytes, in: 2018 21st Conference on Innovation in Clouds, Internet and Networks and Workshops (ICIN), IEEE
Shima, K., Miyamoto, D., Abe, H., Ishihara, T., Okada, K., Sekiya, Y., Asai, H., Doi, Y., 2018. Classification of url bitstreams using bag of bytes, in: 2018 21st Conference on Innovation in Clouds, Internet and Networks and Workshops (ICIN), IEEE. pp. 1–5
2018
-
[185]
Character level based detection of dga domain names, in: 2018 International Joint Conference on Neural Networks (IJCNN), IEEE
Yu, B., Pan, J., Hu, J., Nascimento, A., De Cock, M., 2018. Character level based detection of dga domain names, in: 2018 International Joint Conference on Neural Networks (IJCNN), IEEE. pp. 1–8
2018
-
[186]
Rajitha, K., Vijayalakshmi, D., 2018. Suspicious urls filtering using optimal rt-pfl: A novel feature selection based web url detection, in: Smart Computing and Informatics: Proceedings of the First International Conference on SCI 2016, Volume 2, Springer. pp. 227–235
2018
-
[187]
Towards detection of phishing websites on client-side using machine learning based approach
Jain, A.K., Gupta, B.B., 2018. Towards detection of phishing websites on client-side using machine learning based approach. Telecommu- nication Systems 68, 687–700
2018
-
[188]
What’sinaurl:Fastfeatureextractionandmaliciousurldetection,in:Proceedingsofthe3rdACMonInternational Workshop on Security and Privacy Analytics, pp
Verma,R.,Das,A.,2017. What’sinaurl:Fastfeatureextractionandmaliciousurldetection,in:Proceedingsofthe3rdACMonInternational Workshop on Security and Privacy Analytics, pp. 55–63
2017
-
[189]
Nlpbasedphishingattackdetectionfromurls,in:InternationalConferenceonIntelligentSystems Design and Applications, Springer
Buber,E.,Diri,B.,Sahingoz,O.K.,2017. Nlpbasedphishingattackdetectionfromurls,in:InternationalConferenceonIntelligentSystems Design and Applications, Springer. pp. 608–618
2017
-
[190]
Anefficientphishingwebpagedetector
He,M.,Horng,S.J.,Fan,P.,Khan,M.K.,Run,R.S.,Lai,J.L.,Chen,R.J.,Sutanto,A.,2011. Anefficientphishingwebpagedetector. Expert systems with applications 38, 12018–12027
2011
-
[191]
Two-phase malicious web page detection scheme using misuse and anomaly detection
Yoo, S., Kim, S., Choudhary, A., Roy, O., Tuithung, T., 2014. Two-phase malicious web page detection scheme using misuse and anomaly detection. International Journal of Reliable Information and Assurance 2, 1–9
2014
-
[192]
Context-sensitiveandkeyworddensity-basedsupervisedmachinelearningtechniquesformalicious webpage detection
Altay,B.,Dokeroglu,T.,Cosar,A.,2019. Context-sensitiveandkeyworddensity-basedsupervisedmachinelearningtechniquesformalicious webpage detection. Soft Computing 23, 4177–4191
2019
-
[193]
A machine learning based approach for phishing detection using hyperlinks information
Jain, A.K., Gupta, B.B., 2019. A machine learning based approach for phishing detection using hyperlinks information. Journal of Ambient Intelligence and Humanized Computing 10, 2015–2028
2019
-
[194]
Choi, Y., Kim, T., Choi, S., Lee, C., 2009. Automatic detection for javascript obfuscation attacks in web pages through string pattern analysis, in: Future Generation Information Technology: First International Conference, FGIT 2009, Jeju Island, Korea, December 10-12,
2009
-
[195]
Unbalanced web phishing classification through deep reinforcement learning
Maci, A., Santorsola, A., Coscia, A., Iannacone, A., 2023. Unbalanced web phishing classification through deep reinforcement learning. Computers 12, 118
2023
-
[196]
Anovelapproachformaliciousurldetectionbasedonthejointmodel
Yuan,J.,Liu,Y.,Yu,L.,2021. Anovelapproachformaliciousurldetectionbasedonthejointmodel. SecurityandCommunicationNetworks 2021, 4917016
2021
-
[197]
Malicious url detection based on machine learning
Do Xuan, C., Nguyen, H.D., Tisenko, V.N., 2020. Malicious url detection based on machine learning. International Journal of Advanced Computer Science and Applications 11
2020
-
[198]
DR, U.S., Patil, A., et al., 2023. Malicious url detection and classification analysis using machine learning models, in: 2023 International Conference on Intelligent Data Communication Technologies and Internet of Things (IDCIoT), IEEE. pp. 470–476
2023
-
[199]
Mm-convbert-lms:detectingmaliciouswebpagesviamulti-modallearningand pre-trained model
Tong,X.,Jin,B.,Wang,J.,Yang,Y.,Suo,Q.,Wu,Y.,2023. Mm-convbert-lms:detectingmaliciouswebpagesviamulti-modallearningand pre-trained model. Applied Sciences 13, 3327
2023
-
[200]
Patgiri, R., Katari, H., Kumar, R., Sharma, D., 2019. Empirical study on malicious url detection using machine learning, in: Distributed Computing and Internet Technology: 15th International Conference, ICDCIT 2019, Bhubaneswar, India, January 10–13, 2019, Proceedings 15, Spri...
2019
-
[201]
Cyber threat intelligence-based malicious url detection model using ensemble learning
Alsaedi, M., Ghaleb, F.A., Saeed, F., Ahmad, J., Alasli, M., 2022. Cyber threat intelligence-based malicious url detection model using ensemble learning. Sensors 22, 3373
2022
-
[202]
Phishing urls detection using sequential and parallel ml techniques: comparative analysis
Nagy, N., Aljabri, M., Shaahid, A., Ahmed, A.A., Alnasser, F., Almakramy, L., Alhadab, M., Alfaddagh, S., 2023. Phishing urls detection using sequential and parallel ml techniques: comparative analysis. Sensors 23, 3467
2023
-
[203]
Deep learning-based intrusion detection methods in cyber-physical systems: Challenges and future trends
Umer, M., Sadiq, S., Karamti, H., Alhebshi, R.M., Alnowaiser, K., Eshmawi, A., Song, H., Ashraf, I., 2022. Deep learning-based intrusion detection methods in cyber-physical systems: Challenges and future trends. Electronics 11, 3326
2022
-
[204]
Less is more: Robust and novel features for malicious domain detection
Hajaj, C., Hason, N., Dvir, A., 2022. Less is more: Robust and novel features for malicious domain detection. Electronics 11, 969
2022
-
[205]
Analysis of the performance impact of fine-tuned machine learning model for phishing url detection
Abdul Samad, S.R., Balasubaramanian, S., Al-Kaabi, A.S., Sharma, B., Chowdhury, S., Mehbodniya, A., Webber, J.L., Bostani, A., 2023. Analysis of the performance impact of fine-tuned machine learning model for phishing url detection. Electronics 12, 1642
2023
-
[206]
Intelligent deep machine learning cyber phishing url detection based on bert features extraction
Elsadig, M., Ibrahim, A.O., Basheer, S., Alohali, M.A., Alshunaifi, S., Alqahtani, H., Alharbi, N., Nagmeldin, W., 2022. Intelligent deep machine learning cyber phishing url detection based on bert features extraction. Electronics 11, 3647
2022
-
[207]
Multimodel phishing url detection using lstm, bidirectional lstm, and gru models
Roy, S.S., Awad, A.I., Amare, L.A., Erkihun, M.T., Anas, M., 2022. Multimodel phishing url detection using lstm, bidirectional lstm, and gru models. Future Internet 14, 340
2022
-
[208]
Malicious url detection based on associative classification
Kumi, S., Lim, C., Lee, S.G., 2021. Malicious url detection based on associative classification. Entropy 23, 182
2021
-
[209]
Olawsds: an online arabic web spam detection system
Al-Kabi, M.N., Wahsheh, H.A., Alsmadi, I.M., 2014. Olawsds: an online arabic web spam detection system. Int J Adv Comput Sci Appl 5, 105–110. :Preprint submitted to Elsevier Page 32 of 34
2014
-
[210]
Content-based analysis to detect arabic web spam
Al-Kabi, M., Wahsheh, H., Alsmadi, I., Al-Shawakfa, E., Wahbeh, A., Al-Hmoud, A., 2012. Content-based analysis to detect arabic web spam. Journal of Information Science 38, 284–296
2012
-
[211]
A link and content hybrid approach for arabic web spam detection
Wahsheh, H.A., Al-Kabi, M.N., Alsmadi, I.M., 2013. A link and content hybrid approach for arabic web spam detection. International Journal of Intelligent Systems and Applications (IJISA) 5, 30–43
2013
-
[212]
Web mining techniques to block spam web sites
El-Mohdy, E.M., El-Gamal, A., Elrefaey, H., 2018. Web mining techniques to block spam web sites. International Journal of Computer Applications 975, 8887
2018
-
[213]
Arabicspamdetectionintwitter,in:The2ndWorkshoponArabic Corpora and Processing Tools 2016 Theme: Social Media, p
AlTwairesh,N.,AlTuwaijri,M.,AlMoammar,A.,AlHumoud,S.,2016. Arabicspamdetectionintwitter,in:The2ndWorkshoponArabic Corpora and Processing Tools 2016 Theme: Social Media, p. 38
2016
-
[214]
Analysis of web spam for non-english content: toward more effective language-based classifiers
Alsaleh, M., Alarifi, A., 2016. Analysis of web spam for non-english content: toward more effective language-based classifiers. PloS one 11, e0164383
2016
-
[215]
A proposed spam detection approach for arabic social networks content, in: 2017 International Conference on Mathematics and Information Technology (ICMIT), IEEE
Mataoui, M., Zelmati, O., Boughaci, D., Chaouche, M., Lagoug, F., 2017. A proposed spam detection approach for arabic social networks content, in: 2017 International Conference on Mathematics and Information Technology (ICMIT), IEEE. pp. 222–226
2017
-
[216]
Alkhair, M., Meftouh, K., Smaïli, K., Othman, N., 2019. An arabic corpus of fake news: Collection, analysis and classification, in: Arabic Language Processing: From Theory to Practice: 7th International Conference, ICALP 2019, Nancy, France, October 16–17, 2019, Proceedings 7,...
2019
-
[217]
Spam detection on arabic twitter, in: Social Informatics: 12th International Conference, SocInfo 2020, Pisa, Italy, October 6–9, 2020, Proceedings 12, Springer
Mubarak, H., Abdelali, A., Hassan, S., Darwish, K., 2020. Spam detection on arabic twitter, in: Social Informatics: 12th International Conference, SocInfo 2020, Pisa, Italy, October 6–9, 2020, Proceedings 12, Springer. pp. 237–251
2020
-
[218]
Detecting arabic spam reviews in social networks based on classification algorithms
Najadat, H., Alzubaidi, M.A., Qarqaz, I., 2021. Detecting arabic spam reviews in social networks based on classification algorithms. Transactions on Asian and Low-Resource Language Information Processing 21, 1–13
2021
-
[219]
Enhancing detection of arabic social spam using data augmentation and machine learning
Alkadri, A.M., Elkorany, A., Ahmed, C., 2022. Enhancing detection of arabic social spam using data augmentation and machine learning. Applied Sciences 12. URL:https://www.mdpi.com/2076-3417/12/22/11388, doi:10.3390/app122211388
2022 doi
-
[220]
Sentifilter: A personalized filtering model for arabic semi-spam content based on sentimental and behavioral analysis
Alsulami, M.M., Al-Aama, A.Y., 2020. Sentifilter: A personalized filtering model for arabic semi-spam content based on sentimental and behavioral analysis. Int. J. Adv. Comput. Sci. Appl 11
2020
-
[221]
Efficientarabicandenglishsocialspamdetectionusingatransformerand2dconvolutionalneuralnetwork-based deep learning filter
Kihal,M.,Hamza,L.,2025. Efficientarabicandenglishsocialspamdetectionusingatransformerand2dconvolutionalneuralnetwork-based deep learning filter. International Journal of Information Security 24, 56. URL:https://doi.org/10.1007/s10207-024-00975-0, doi:10.1007/s10207-024-00975-0
2025 doi
-
[222]
A real-time framework for opinion spam detection in arabic social networks
Ezzat, C.A., Alkadri, A.M., Elkorany, A., 2025. A real-time framework for opinion spam detection in arabic social networks. Egyptian Informatics Journal 29, 100626
2025
-
[223]
TowardsMachineLearningforGulfDialecticalArabicMaliciousContentDetectioninSocialMedia
Alorini,D.,2018. TowardsMachineLearningforGulfDialecticalArabicMaliciousContentDetectioninSocialMedia. Ph.D.thesis.Howard University
2018
-
[224]
An ensemble approach for spam detection in arabic opinion texts
Saeed, R.M., Rady, S., Gharib, T.F., 2022. An ensemble approach for spam detection in arabic opinion texts. Journal of King Saud University - Computer and Information Sciences 34, 1407–1416. URL:https://www.sciencedirect.com/science/article/pii/ S1319157819307414, doi:https://...
2022 doi
-
[225]
Predicting rogue content and arabic spammers on twitter
Alharbi, A.R., Aljaedi, A., 2019. Predicting rogue content and arabic spammers on twitter. Future Internet 11, 229
2019
-
[226]
Evaluating arabic spam classifiers using link analysis, in: Proceedings of the 3rd international conference on information and communication systems, pp
Wahsheh, H.A., Al-kabi, M.N., Alsmadi, I.M., 2012. Evaluating arabic spam classifiers using link analysis, in: Proceedings of the 3rd international conference on information and communication systems, pp. 1–5
2012
-
[227]
Intelligent analysis of arabic tweets for detection of suspicious messages
AlGhamdi, M.A., Khan, M.A., 2020. Intelligent analysis of arabic tweets for detection of suspicious messages. Arabian Journal for Science and Engineering 45, 6021–6032
2020
-
[229]
Semi-supervised conditional gan for simultaneous generation and detection of phishing urls: A game theoretic perspective
Kamran, S.A., Sengupta, S., Tavakkoli, A., 2021. Semi-supervised conditional gan for simultaneous generation and detection of phishing urls: A game theoretic perspective. arXiv preprint arXiv:2108.01852
2021 arXiv
-
[230]
Accurateandfasturlphishingdetector:aconvolutionalneural network approach
Wei,W.,Ke,Q.,Nowak,J.,Korytkowski,M.,Scherer,R.,Woźniak,M.,2020. Accurateandfasturlphishingdetector:aconvolutionalneural network approach. Computer Networks 178, 107275
2020
-
[231]
Domurls_bert: Pre-trained bert-based model for malicious domains and urls detection and classification
Mahdaouy, A.E., Lamsiyah, S., Idrissi, M.J., Alami, H., Yartaoui, Z., Berrada, I., 2024. Domurls_bert: Pre-trained bert-based model for malicious domains and urls detection and classification. arXiv preprint arXiv:2409.09143
2024
-
[232]
An adversarial attack analysis on malicious advertisement url detection framework
Nowroozi, E., Abhishek, Mohammadi, M., Conti, M., 2023. An adversarial attack analysis on malicious advertisement url detection framework. IEEE Transactions on Network and Service Management 20, 1332–1344. doi:10.1109/TNSM.2022.3225217
2023
-
[233]
Tsgn: Transaction subgraph networks assisting phishing detection in ethereum
Wang, J., Chen, P., Xu, X., Wu, J., Shen, M., Xuan, Q., Yang, X., 2022. Tsgn: Transaction subgraph networks assisting phishing detection in ethereum. URL:https://arxiv.org/abs/2208.12938,arXiv:2208.12938
2022 arXiv
-
[234]
Efficient phishing url detection using graph-based machine learning and loopy belief propagation
Guo, W., Wang, Q., Yue, H., Sun, H., Hu, R.Q., 2025. Efficient phishing url detection using graph-based machine learning and loopy belief propagation. URL:https://arxiv.org/abs/2501.06912,arXiv:2501.06912
2025 arXiv
-
[235]
Fed-urlbert:Client-sidelightweightfederatedtransformersforurlthreat analysis
Li,Y.,Wang,Y.,Xu,H.,Guo,Z.,Zhang,F.,Liu,R.,Ma,W.,2023. Fed-urlbert:Client-sidelightweightfederatedtransformersforurlthreat analysis. arXiv preprint arXiv:2312.03636
2023 arXiv
-
[236]
Urlbert: A contrastive and adversarial pre-trained model for url classification
Li, Y., Wang, Y., Xu, H., Guo, Z., Cao, Z., Zhang, L., 2024. Urlbert: A contrastive and adversarial pre-trained model for url classification. arXiv preprint arXiv:2402.11495
2024 arXiv
-
[237]
Detectingspamurlsinsocialmediaviabehavioralanalysis,in:AdvancesinInformationRetrieval:37thEuropean Conference on IR Research, ECIR 2015, Vienna, Austria, March 29-April 2, 2015
Cao,C.,Caverlee,J.,2015. Detectingspamurlsinsocialmediaviabehavioralanalysis,in:AdvancesinInformationRetrieval:37thEuropean Conference on IR Research, ECIR 2015, Vienna, Austria, March 29-April 2, 2015. Proceedings 37, Springer. pp. 703–714
2015
-
[238]
Warningbird: Detecting suspicious urls in twitter stream., in: Ndss, pp
Lee, S., Kim, J., 2012. Warningbird: Detecting suspicious urls in twitter stream., in: Ndss, pp. 1–13
2012
-
[239]
Efficient learning using forward-backward splitting
Singer, Y., Duchi, J.C., 2009. Efficient learning using forward-backward splitting. Advances in Neural Information Processing Systems 22
2009
-
[240]
McDonald,R.,Hall,K.,Mann,G.,2010. Distributedtrainingstrategiesforthestructuredperceptron,in:Humanlanguagetechnologies:The 2010 annual conference of the North American chapter of the association for computational linguistics, pp. 456–464
2010
-
[241]
Liblinear: A library for large linear classification
Fan, R.E., Chang, K.W., Hsieh, C.J., Wang, X.R., Lin, C.J., 2008. Liblinear: A library for large linear classification. the Journal of machine Learning research 9, 1871–1874. :Preprint submitted to Elsevier Page 33 of 34
2008
-
[242]
Large-scale automatic classification of phishing pages., in: Ndss, p
Whittaker, C., Ryner, B., Nazif, M., 2010. Large-scale automatic classification of phishing pages., in: Ndss, p. 2010
2010
-
[244]
Eshete, B., Villafiorita, A., Weldemariam, K., 2013. Binspect: Holistic analysis and detection of malicious web pages, in: Security and Privacy in Communication Networks: 8th International ICST Conference, SecureComm 2012, Padua, Italy, September 3-5, 2012. Revised Selected Pa...
2013
-
[245]
Insideaphisher’smind:Understandingtheanti-phishingecosystem throughphishingkitanalysis,in:2018APWGSymposiumonElectronicCrimeResearch(eCrime),pp.1–12
Oest,A.,Safei,Y.,Doupé,A.,Ahn,G.J.,Wardman,B.,Warner,G.,2018. Insideaphisher’smind:Understandingtheanti-phishingecosystem throughphishingkitanalysis,in:2018APWGSymposiumonElectronicCrimeResearch(eCrime),pp.1–12. doi:10.1109/ECRIME.2018. 8376206
2018 doi
-
[246]
PhishTime: Continuous longitudinal measurement of the effectiveness of anti-phishing blacklists, in: 29th USENIX Security Symposium (USENIX Security 20), USENIX Association
Oest, A., Safaei, Y., Zhang, P., Wardman, B., Tyers, K., Shoshitaishvili, Y., Doupé, A., 2020. PhishTime: Continuous longitudinal measurement of the effectiveness of anti-phishing blacklists, in: 29th USENIX Security Symposium (USENIX Security 20), USENIX Association. pp. 379–...
2020
-
[247]
Sunrisetosunset:Analyzingthe end-to-endlifecycleandeffectivenessofphishingattacksatscale,in:29thUSENIXSecuritySymposium(USENIXSecurity20),USENIX Association
Oest,A.,Zhang,P.,Wardman,B.,Nunes,E.,Burgis,J.,Zand,A.,Thomas,K.,Doupé,A.,Ahn,G.J.,2020. Sunrisetosunset:Analyzingthe end-to-endlifecycleandeffectivenessofphishingattacksatscale,in:29thUSENIXSecuritySymposium(USENIXSecurity20),USENIX Association. pp. 361–377. URL:https://www.u...
2020
-
[248]
Crawlphish:Large-scaleanalysisofclient-sidecloakingtechniquesinphishing,in:2021IEEESymposiumonSecurity and Privacy (SP), pp
Zhang, P., Oest, A., Cho, H., Sun, Z., Johnson, R., Wardman, B., Sarker, S., Kapravelos, A., Bao, T., Wang, R., Shoshitaishvili, Y., Doupé, A.,Ahn,G.J.,2021. Crawlphish:Large-scaleanalysisofclient-sidecloakingtechniquesinphishing,in:2021IEEESymposiumonSecurity and Privacy (SP)...
2021
-
[250]
Maroofi, S., Korczyński, M., Duda, A., 2020. Are you human? resilience of phishing detection to evasion techniques based on human verification, in: Proceedings of the ACM Internet Measurement Conference, Association for Computing Machinery, New York, NY, USA. p. 78–86. URL:htt...
2020
-
[251]
Reinforcement learning: a survey
Kaelbling, L.P., Littman, M.L., Moore, A.W., 1996. Reinforcement learning: a survey. J. Artif. Int. Res. 4, 237–285
1996
-
[252]
Beyond the west: Revealing and bridging the gap between western and chinese phish- ing website detection
Yuan, Y., Apruzzese, G., Conti, M., 2025. Beyond the west: Revealing and bridging the gap between western and chinese phish- ing website detection. Computers & Security 148, 104115. URL:https://www.sciencedirect.com/science/article/pii/ S0167404824004206, doi:https://doi.org/1...
2025
-
[253]
Online learning: A comprehensive survey
Hoi, S.C., Sahoo, D., Lu, J., Zhao, P., 2021. Online learning: A comprehensive survey. Neurocomputing 459, 249–289. :Preprint submitted to Elsevier Page 34 of 34 Figure 1:Example of URL Figure 2:Count the number of papers published on the topic of malicious URL detection (Web ...
2021
-
[254]
A survey on unsupervised learning algorithms for detecting abnormal points in streaming data, in: 2022 International Joint Conference on Neural Networks (IJCNN), pp
Ngo Bibinbe, A.M.S., Mbouopda, M.F., Mbiadou Saleu, G.R., Mephu Nguifo, E., 2022. A survey on unsupervised learning algorithms for detecting abnormal points in streaming data, in: 2022 International Joint Conference on Neural Networks (IJCNN), pp. 1–8. doi:10.1109/ IJCNN55064....
2022
-
[2009]
Proceedings 1, Springer. pp. 160–172
Reviewed August 16, 2026 · model on record in the stance chip above.
Discussion (0). Continue with ORCID to comment.