duanmu1736 2011-05-17 20:16
浏览 44
已采纳

使用PHP在多个条目中查找剽窃的可能性

I am working on a web application that tracks helpdesk entries. We want to find a way to prevent people from copying and pasting their notes regarding common issues - we want original helpdesk entries to be written for every trouble-call.

In any case, we have thousands of entries and some of them are similar, I am trying to find a way of comparing them all to eachother and pointing out any entries that are very similar to others, i.e. 80% likely to be a direct copy, etc.

I've looked into similar_text() and a few other built-in PHP functions, but I am interested in hearing if anyone else has done something similar before. I don't believe I can use similar_text() efficiently since I need to compare multiple entries against each other, not two strings.

Any input is appreciated.

  • 写回答

3条回答 默认 最新

  • doulai2573 2011-05-18 04:15
    关注

    I do think similar_text() would do what you want. As long as your machine has enough memory to handle the comparisons, it should work fine. Also look at levenshtein() and soundex().

    本回答被题主选为最佳回答 , 对您是否有帮助呢?
    评论
查看更多回答(2条)

报告相同问题?

悬赏问题

  • ¥15 r语言神经网络自变量重要性分析
  • ¥15 基于双目测规则物体尺寸
  • ¥15 wegame打不开英雄联盟
  • ¥15 公司的电脑,win10系统自带远程协助,访问家里个人电脑,提示出现内部错误,各种常规的设置都已经尝试,感觉公司对此功能进行了限制(我们是集团公司)
  • ¥15 救!ENVI5.6深度学习初始化模型报错怎么办?
  • ¥30 eclipse开启服务后,网页无法打开
  • ¥30 雷达辐射源信号参考模型
  • ¥15 html+css+js如何实现这样子的效果?
  • ¥15 STM32单片机自主设计
  • ¥15 如何在node.js中或者java中给wav格式的音频编码成sil格式呢