<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article
  PUBLIC "-//NLM//DTD Journal Publishing DTD v2.0 20040830//EN" "http://dtd.nlm.nih.gov/publishing/2.0/journalpublishing.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" article-type="research-article" dtd-version="2.0" xml:lang="EN">

  <front>

    <journal-meta>

      <journal-title>Information Technology Journal</journal-title>

      <issn pub-type="ppub">1812-5638</issn>

      <issn pub-type="epub">1812-5646</issn>

      <publisher>

        <publisher-name>Asian Network for Scientific Information</publisher-name>

      </publisher>

    </journal-meta>


    <article-meta>

      <article-id pub-id-type="doi">10.3923/itj.2013.2465.2469</article-id>


      <title-group>

        <article-title><![CDATA[A Chunk-based Copy Detection Approach for Multimedia Documents]]></article-title>

      </title-group>


      <contrib-group>

        <contrib contrib-type="author" xlink:type="simple">


          <name name-style="western">

            <surname>Guo</surname>

            <given-names>Li</given-names>

          </name>


          <name name-style="western">

            <surname>Jin</surname>

            <given-names>Bo</given-names>

          </name>


          <name name-style="western">

            <surname>Huang</surname>

            <given-names>Degen</given-names>

          </name>


        </contrib>

      </contrib-group>


      <pub-date pub-type="collection">




        <month>12</month>


        <year>2013</year>

      </pub-date>


      <volume>12</volume>

      <issue>12</issue>


      <abstract><![CDATA[<p>Copy detection is important to both intellectual property 
  protection and information retrieval. Prior researches on copy detection (e.g., 
  hash-based fingerprinting algorithm, etc.) concentrated on document-level copy 
  detection. These researches have difficulty to detect complicated multimedia 
  documents. In this study, a chunk-based copy detection approach is discussed. 
  And a Fingerprint-based Heuristic Algorithm (FHA) is proposed based on fixed-length 
  chunks and overlap of chunks. With experiments and results, the proposed approach 
  can deal with both total copy detection and partial copy detection of multimedia 
  documents.</p>]]></abstract>


    </article-meta>

  </front>


  <ref-list>
















      <ref id="13460">

        <label>1</label>

        <citation citation-type="journal" xlink:type="simple">

          <person-group person-group-type="author">

            <name name-style="western">

              <surname>Xiao, C., W. Wang, X. Lin, J.X. Yu and G. Wang,</surname>

              <given-names></given-names>

            </name>

          </person-group>

          <year>2011</year>

          <article-title><![CDATA[Efficient similarity joins for near-duplicate detection.]]></article-title>

          <source>ACM Trans. Database Syst., Vol. 36. </source>

          <volume>2011</volume>

        </citation>

      </ref>


















      <ref id="13461">

        <label>2</label>

        <citation citation-type="journal" xlink:type="simple">

          <person-group person-group-type="author">

            <name name-style="western">

              <surname>Urvoy, T., E. Chauveau, P. Filoche and T. Lavergne,</surname>

              <given-names></given-names>

            </name>

          </person-group>

          <year>2008</year>

          <article-title><![CDATA[Tracking Web spam with HTML style similarities.]]></article-title>

          <source>ACM Trans. Web., Vol.2.</source>

          <volume>2008</volume>

        </citation>

      </ref>






      <ref id="1151931">

        <label>3</label>

        <citation citation-type="journal" xlink:type="simple">

          <person-group person-group-type="author">

            <name name-style="western">

              <surname>Stajano, F. and P. Wilson,</surname>

              <given-names></given-names>

            </name>

          </person-group>

          <year>2011</year>

          <article-title><![CDATA[Understanding scam victims: Seven principles for systems security.]]></article-title>

          <source>Commun. ACM.,</source>

          <volume>54</volume>

          <fpage>70</fpage>

          <lpage>75</lpage>

        </citation>

      </ref>


















      <ref id="1051745">

        <label>4</label>

        <citation citation-type="journal" xlink:type="simple">

          <person-group person-group-type="author">

            <name name-style="western">

              <surname>Elmagarmid, A.K., P.G. Ipeirotis and V.S. Verykios,</surname>

              <given-names></given-names>

            </name>

          </person-group>

          <year>2007</year>

          <article-title><![CDATA[Duplicate record detection: A survey.]]></article-title>

          <source>IEEE Trans. Knowledge Data Eng.,</source>

          <volume>19</volume>

          <fpage>1</fpage>

          <lpage>16</lpage>

        </citation>

      </ref>






























      <ref id="13463">

        <label>5</label>

        <citation citation-type="journal" xlink:type="simple">

          <person-group person-group-type="author">

            <name name-style="western">

              <surname>Puppin, D., F. Silvestri, R. Perego and R.A. Baeza-Yates,</surname>

              <given-names></given-names>

            </name>

          </person-group>

          <year>2010</year>

          <article-title><![CDATA[Tuning the capacity of search engines: Load-driven routing and incremental caching to reduce and balance the load.]]></article-title>

          <source>ACM Trans. Inform. Syst.,</source>

          <volume>2010</volume>

        </citation>

      </ref>






      <ref id="1151939">

        <label>6</label>

        <citation citation-type="journal" xlink:type="simple">

          <person-group person-group-type="author">

            <name name-style="western">

              <surname>Potthast, M., A. Barron-Cedeno, B. Stein and P. Rosso,</surname>

              <given-names></given-names>

            </name>

          </person-group>

          <year>2011</year>

          <article-title><![CDATA[Cross-language plagiarism detection.]]></article-title>

          <source>Lang. Resour. Eval.,</source>

          <volume>45</volume>

          <fpage>45</fpage>

          <lpage>62</lpage>

        </citation>

      </ref>






























      <ref id="13464">

        <label>7</label>

        <citation citation-type="journal" xlink:type="simple">

          <person-group person-group-type="author">

            <name name-style="western">

              <surname>Bar-Yossef, Z., I. Keidar and U. Schonfeld,</surname>

              <given-names></given-names>

            </name>

          </person-group>

          <year>2009</year>

          <article-title><![CDATA[Do not Crawl in the DUST: Different URLs with similar text.]]></article-title>

          <source>ACM Trans. Web, Vol. 3.</source>

          <volume>2009</volume>

        </citation>

      </ref>




  </ref-list>

</article>

