{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,8,25]],"date-time":"2026-08-25T21:52:49Z","timestamp":1787694769675,"version":"build-2784847793"},"reference-count":12,"publisher":"MIT Press","issue":"3","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Computational Linguistics"],"published-print":{"date-parts":[[2013,9]]},"abstract":"<jats:p>Paraphrases are sentences or phrases that convey the same meaning using different wording. Although the logical definition of paraphrases requires strict semantic equivalence, linguistics accepts a broader, approximate, equivalence\u2014thereby allowing far more examples of \u201cquasi-paraphrase.\u201d But approximate equivalence is hard to define. Thus, the phenomenon of paraphrases, as understood in linguistics, is difficult to characterize. In this article, we list a set of 25 operations that generate quasi-paraphrases. We then empirically validate the scope and accuracy of this list by manually analyzing random samples of two publicly available paraphrase corpora. We provide the distribution of naturally occurring quasi-paraphrases in English text.<\/jats:p>","DOI":"10.1162\/coli_a_00166","type":"journal-article","created":{"date-parts":[[2013,5,17]],"date-time":"2013-05-17T16:12:51Z","timestamp":1368807171000},"page":"463-472","source":"Crossref","is-referenced-by-count":99,"title":["What Is a Paraphrase?"],"prefix":"10.1162","volume":"39","author":[{"given":"Rahul","family":"Bhagat","sequence":"first","affiliation":[{"name":"USC Information Sciences Institute"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Eduard","family":"Hovy","sequence":"additional","affiliation":[{"name":"USC Information Sciences Institute"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"281","reference":[{"key":"R1","first-page":"171","volume-title":"Frame, Fields, and Contrasts: New Essays in Semantic Lexical Organization.","author":"Clark E. V.","year":"1992"},{"key":"R2","doi-asserted-by":"publisher","DOI":"10.1162\/coli.08-003-R1-07-044"},{"key":"R3","doi-asserted-by":"publisher","DOI":"10.4324\/9781315835839"},{"key":"R4","doi-asserted-by":"publisher","DOI":"10.3115\/1220355.1220406"},{"key":"R5","doi-asserted-by":"publisher","DOI":"10.7551\/mitpress\/7287.001.0001"},{"key":"R6","doi-asserted-by":"publisher","DOI":"10.1007\/978-94-009-8467-7_8"},{"key":"R8","doi-asserted-by":"publisher","DOI":"10.1016\/S0022-5371(71)80035-X"},{"key":"R9","unstructured":"Huang, S., D. Graff, and G. Doddington. 2002. Multiple-translation Chinese corpus.Linguistic Data Consortium, Philadelphia, PA."},{"key":"R10","doi-asserted-by":"publisher","DOI":"10.1145\/502512.502559"},{"key":"R11","doi-asserted-by":"crossref","first-page":"37","DOI":"10.1075\/slcs.31.08mel","volume-title":"Lexical Functions in Lexicography and Natural Language Processing.","author":"Mel'cuk I.","year":"1996"},{"key":"R12","doi-asserted-by":"publisher","DOI":"10.1075\/slcs.129"},{"key":"R13","volume-title":"Nonparametric Statistics for the Behavioral Sciences.","author":"Siegal S.","year":"1988"}],"container-title":["Computational Linguistics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mitpressjournals.org\/doi\/pdf\/10.1162\/COLI_a_00166","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,4,30]],"date-time":"2025-04-30T06:32:45Z","timestamp":1745994765000},"score":1,"resource":{"primary":{"URL":"https:\/\/direct.mit.edu\/coli\/article\/39\/3\/463-472\/1434"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2013,9]]},"references-count":12,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2013,9]]}},"alternative-id":["10.1162\/COLI_a_00166"],"URL":"https:\/\/doi.org\/10.1162\/coli_a_00166","relation":{},"ISSN":["0891-2017","1530-9312"],"issn-type":[{"value":"0891-2017","type":"print"},{"value":"1530-9312","type":"electronic"}],"subject":[],"published":{"date-parts":[[2013,9]]}}}