期刊文献+

一种基于主题分类的文本过滤方法及其硬件实现

Text Filtering Algorithm Based on Topic Classification and Realization of DSP Hardware
下载PDF
导出
摘要 针对不良文本的过滤问题,提出一种基于主题分类的文本过滤方法,通过对文本信息进行向量化,引人文本特征抽取技术,筛选出针对文本内容的最优的特征项集合,利用SVM分类技术,来判断文本的态度和立场,达到内容审查过滤的目的.并利用DSP在硬件上加以实现,实验表明该方法同传统的过滤方法相比具有较高的准确率和召回率,且过滤时间大幅减少. Concerning Chinese vicious-topic information text filtering problems, this paper presents an improved text filtering method based on SVM classifier. Through replacing filtering method based on words feature items can be distinguished from different classes effectively. The experimental results indicate that this method is of both higher precision rate and recall rate compared with traditional methods.
出处 《湖南工程学院学报(自然科学版)》 2010年第2期49-52,共4页 Journal of Hunan Institute of Engineering(Natural Science Edition)
关键词 文本过滤 文本分类 支持向量机 DSP text filtering text classification support vector machine (SVM) DSP
  • 相关文献

参考文献3

二级参考文献12

  • 1Lewis D. D.. An evaluation of phrasal and clustered representalions on a text categorization task. In: Proceedings of SIGIR'92,the 15st ACM International Conference on Research and Development in Information Retrieval, Copenhagen, Denmark,1992, 37-50.
  • 2Sebastiani F,. Machine learning in automated text categorization. ACM Computing Surveys, 2002, 34(1): 1-47.
  • 3Lewis D.. Naive bayes at forty: The independence assumption in information retrieval. In: Proceedings of the 10th European Conference on Machine Learning, Chemnitz, Germany, 1998,4-15.
  • 4Salton G.. Automatic Text Processing: The Transformation,Analysis, and Retrieval of Information by Computer. Reading,MA: Addison Wesley, 1989.
  • 5Mitchell T. M.. Machine Learning. New York: McCraw Hill,1996.
  • 6Joachims T.. Text categorization with support vector machines: Learning with many relevant features. In: Proceedings of the 10th European Conference on Machine Learning,Chemnitz, Germany, 1998, 137-142.
  • 7Yang Y. , Liu X.. A Re-examination of text categorization methods. In: Proceedings of SIGIR'99, the 22nd ACM International Conference on Research and Development in Information Retrieval, Berkeley, CA, 1999, 42-49.
  • 8樊兴华.因果推理和文本分类.清华大学博士后出站报告,2004.
  • 9Larkey L. S.. Automatic essay grading using text categorization techniques.. In: Proceedings of SIGIR'98, the 21st ACM International Conference on Research and Development in Information Retrieval, Melbourne, Australia, 1998, 90-95.
  • 10Dumais S. T. , Platt J. , Hecherman D. , Sahami M.. Inductive learning algorithms and representation for text categorization.In: Proceedings of CIKM'98, the 7th ACM International Conference on Information and Knowledge Management, Bethesda, MD, 1998, 148-155.

共引文献159

相关作者

内容加载中请稍等...

相关机构

内容加载中请稍等...

相关主题

内容加载中请稍等...

浏览历史

内容加载中请稍等...
;
使用帮助 返回顶部