A Code Classification Method Based on TF-IDF

@inproceedings{Wang2018ACC,
  title={A Code Classification Method Based on TF-IDF},
  author={Ke Wang and Jian-Hong Jiang and MA Rui-Yun},
  year={2018}
}
The main purpose of the study is to find the code with similar possibilities to effectively avoid the adverse effects of code duplication. Through the clustering pretreatment of document feature information, to extract the relevant features of the document. Then the basic characteristics are used to cluster the document, to find out the best number of clusters. According to the reasonable number of clusters that have been found, using the vectors that generated through TF-IDF method, combined… CONTINUE READING

References

Publications referenced by this paper.
SHOWING 1-10 OF 15 REFERENCES

Android malicious code detection based on sensitive permissions and function call graphs[J

X Zhu, J Wang, Y Du, J. Bai
  • Journal of Sichuan University (Natural Science Edition),
  • 2016
VIEW 1 EXCERPT

Method for measuring the directional conflict between evidences based on improved cosine similarity[J

Y Mao, D Zhang, L. Wang
  • Systems Engineering and Electronics,
  • 2016
VIEW 1 EXCERPT

A Fast Detection Method of Paper Similarity Based on Spark[J

K Zhuo, G Tong, W. Yu
  • Library and Information Work,
  • 2015
VIEW 1 EXCERPT

Similar Papers

Loading similar papers…