مرکز اطلاعات علمی Scientific Information Database (SID) - Trusted Source for Research and Academic Resources

Persian Verion

مرکز اطلاعات علمی Scientific Information Database (SID) - Trusted Source for Research and Academic Resources

video

مرکز اطلاعات علمی Scientific Information Database (SID) - Trusted Source for Research and Academic Resources

sound

مرکز اطلاعات علمی Scientific Information Database (SID) - Trusted Source for Research and Academic Resources

Persian Version

مرکز اطلاعات علمی Scientific Information Database (SID) - Trusted Source for Research and Academic Resources

View:

1,731
مرکز اطلاعات علمی Scientific Information Database (SID) - Trusted Source for Research and Academic Resources

Download:

0
مرکز اطلاعات علمی Scientific Information Database (SID) - Trusted Source for Research and Academic Resources

Cites:

Information Journal Paper

Title

OPTIMIZATION OF MULTIPLE KERNELS IN TWIN SVM FOR DECREASING WEB SPAM PAGE DETECTION SEMANTIC GAP

Pages

  0-0

Keywords

TWIN SUPPORT VECTOR MACHINE (TSVM)Q2
GENETIC ALGORITHM (GA)Q2

Abstract

 Web pages are crawled and indexed by SEARCH ENGINEs for fast accessing data on the web. One of the challenges in the SEARCH ENGINEs is WEB SPAM PAGEs. There are many approaches to WEB SPAM PAGEs detection such as measurement of HTML code style similarity, pages linguistic pattern analysis and MACHINE LEARNING algorithm on page content features. One of the famous algorithms has been used in MACHINE LEARNING approach is Support Vector Machine (SVM) classifier. Unfortunately SVM could not achieve a reasonable accuracy in this scope. In order to classify non-linear data in a linear manner, the SVM needs to use the idea of the kernel, which leads to enhanced classification capabilities. A kernel, implicitly maps the data to a higher-dimensional space. Recently basic structure of SVM has been changed by new extensions called Twin SVM (TSVM) to increase robustness and classification accuracy using two separate hyperplanes. Because of using two separate hyperplanes in TSVM, it is better to use MULTIPLE KERNELS in it. Kernel functions are designed based on specific data sample. Therefore they cannot use for general purpose. In this paper we improved accuracy of web spam detection by using two nonlinear kernels into TSVM as an improved extension of SVM. These two kernels have been created based on genetic algorithm. The classifier ability to data separation has been increased by using two separated kernels for each class of data. Effectiveness of new proposed method has been experimented with two publicly used spam datasets called UK-2007 and UK-2006.

Cites

  • No record.
  • References

  • No record.
  • Cite

    APA: Copy

    ZARE CHAHOUKI, M.A., & MOHAMMADI, S.H.R.. (2017). OPTIMIZATION OF MULTIPLE KERNELS IN TWIN SVM FOR DECREASING WEB SPAM PAGE DETECTION SEMANTIC GAP. TABRIZ JOURNAL OF ELECTRICAL ENGINEERING, 46(4 (78)), 0-0. SID. https://sid.ir/paper/359155/en

    Vancouver: Copy

    ZARE CHAHOUKI M.A., MOHAMMADI S.H.R.. OPTIMIZATION OF MULTIPLE KERNELS IN TWIN SVM FOR DECREASING WEB SPAM PAGE DETECTION SEMANTIC GAP. TABRIZ JOURNAL OF ELECTRICAL ENGINEERING[Internet]. 2017;46(4 (78)):0-0. Available from: https://sid.ir/paper/359155/en

    IEEE: Copy

    M.A. ZARE CHAHOUKI, and S.H.R. MOHAMMADI, “OPTIMIZATION OF MULTIPLE KERNELS IN TWIN SVM FOR DECREASING WEB SPAM PAGE DETECTION SEMANTIC GAP,” TABRIZ JOURNAL OF ELECTRICAL ENGINEERING, vol. 46, no. 4 (78), pp. 0–0, 2017, [Online]. Available: https://sid.ir/paper/359155/en

    Related Journal Papers

    Related Seminar Papers

  • No record.
  • Related Plans

  • No record.
  • Recommended Workshops






    Move to top
    telegram sharing button
    whatsapp sharing button
    linkedin sharing button
    twitter sharing button
    email sharing button
    email sharing button
    email sharing button
    sharethis sharing button