期刊文献+
共找到2篇文章
< 1 >
每页显示 20 50 100
A Semi-automatic Method Based on Statistic for Mandarin Semantic Structures Extraction in Specific Domains 被引量:1
1
作者 熊英 朱杰 孙静 《Journal of Shanghai Jiaotong university(Science)》 EI 2004年第4期25-29,共5页
This paper proposed a new method of semi-automatic extraction for semantic structures from unlabelled corpora in specific domains. The approach is statistical in nature. The extracted structures can be used for shallo... This paper proposed a new method of semi-automatic extraction for semantic structures from unlabelled corpora in specific domains. The approach is statistical in nature. The extracted structures can be used for shallow parsing and semantic labeling. By iteratively extracting new words and clustering words, we get an inital semantic lexicon that groups words of the same semantic meaning together as a class. After that, a bootstrapping algorithm is adopted to extract semantic structures. Then the semantic structures are used to extract new 展开更多
关键词 and augment the semantic lexicon. The resultant semantic structures are interpreted by persons and are amenable to hand-editing for refinement. In this experiment the semi-automatically extracted structures S SA provide recall rate of 84.
下载PDF
ML-Parser:An Eficient and Accurate Online Log Parser 被引量:1
2
作者 Yu-Qian Zhu Jia-Ying Deng +3 位作者 Jia-Chen Pu Peng Wang Shen Liang Wei Wang 《Journal of Computer Science & Technology》 SCIE EI CSCD 2022年第6期1412-1426,共15页
A log is a text message that is generated in various services,frameworks,and programs.The majority of log data mining tasks rely on log parsing as the first step,which transforms raw logs into formatted log templates.... A log is a text message that is generated in various services,frameworks,and programs.The majority of log data mining tasks rely on log parsing as the first step,which transforms raw logs into formatted log templates.Existing log parsing approaches often fail to effectively handle the trade-off between parsing quality and performance.In view of this,in this paper,we present Multi-Layer Parser(ML-Parser),an online log parser that runs in a streaming manner.Specifically,we present a multi-layer structure in log parsing to strike a balance between efficiency and effectiveness.Coarse-grained tokenization and a fast similarity measure are applied for efficiency while fine-grained tokenization and an accurate similarity measure are used for effectiveness.In experiments,we compare ML-Parser with two existing online log parsing approaches,Drain and Spell,on ten real-world datasets,five labeled and five unlabeled.On the five labeled datasets,we use the proportion of correctly parsed logs to measure the accuracy,and ML-Parser achieves the highest accuracy on four datasets.On the whole ten datasets,we use Loss metric to measure the parsing quality.ML-Parse achieves the highest quality on seven out of the ten datasets while maintaining relatively high efficiency. 展开更多
关键词 log parsing online approach structure extraction similarity measure
原文传递
上一页 1 下一页 到第
使用帮助 返回顶部