Created: 2011-07-09 02:09
Updated: 2019-03-03 03:30
License: mit

This is a copyright violation detector running on Wikimedia Labs.

It can search the web for content similar to a given article, and graphically compare an article to a specific URL. Some technical details are expanded upon in a blog post.



  • If using Tool Labs, you should clone the repository to ~/www/python/src, or otherwise symlink it to that directory. A virtualenv should be created at ~/www/python/venv.

  • Install all dependencies listed above.

  • Create an SQL database with the cache and cache_data tables defined by earwigbot-plugins.

  • Create an earwigbot instance in .earwigbot (run earwigbot .earwigbot). In .earwigbot/config.yml, fill out the connection info for the database by adding the following to the wiki section:

          host: <hostname of database server>
          db:   <name of database>

    If additional arguments are needed by oursql.connect(), like usernames or passwords, they should be added to the _copyviosSQL section.

  • Run ./ to minify JS and CSS files.

  • Start the web server (on Tool Labs, webservice2 uwsgi-python start).

Cookies help us deliver our services. By using our services, you agree to our use of cookies Learn more