Slashdot Mirror


Competition Seeks Best Approaches To Detecting Plagiarism

marpot writes "Does your school/university check your homeworks/theses for plagiarism? Nowadays, probably Yes, but are they doing it properly? Little is known about plagiarism detection accuracy, which is why we conduct a competition on plagiarism detection, sponsored by Yahoo! We have set up a corpus of artificial plagiarism which contains plagiarism with varying degrees of obfuscation, and translation plagiarism from Spanish or German source documents. A random plagiarist was employed who attempts to obfuscate his plagiarism with random sequences of text operations, e.g., shuffling, deleting, inserting, or replacing a word. Translated plagiarism is created using machine translation."

48 of 289 comments (clear)

  1. Insightful fact... by telchine · · Score: 3, Funny

    Here's an insightful fact related to this article:

    Little is known about plagiarism detection accuracy

    1. Re:Insightful fact... by gnick · · Score: 3, Informative

      But a lot of faith is put in it. I've got a friend that works at the University of Phoenix. We caught up not long ago and he was singing praises about how you just dump a paper into this tool he uses and it instantly tells you the exact percentage of plagiarism content in the student's paper. Too high == disciplinary action - Apparently without even bothering tracking sources or verifying specific plagiarized sections.

      Of course, this all came to me second hand - I've not used the tools myself.

      --
      He's getting rather old, but he's a good mouse.
    2. Re:Insightful fact... by Erwos · · Score: 4, Interesting

      The tools are fairly good, but, in my experience, they'll always report 3-7% or so of your paper as plagiarized, just because it's pretty difficult to write about _anything_ without unknowingly using previously written words. I would _hope_ that anyone who would pursue disciplinary action from such a tool's results would at least take a look to see if the sections being flagged are consequential.

      I have no idea how good they are with catching paraphrasing, though... it strikes me that the semi-intelligent plagiarizers would be doing that more than a straight copy and paste. There's also the "acceptable vs unacceptable" distinction to be made.

      --
      Plausible conjecture should not be misrepresented as proof positive.
    3. Re:Insightful fact... by eln · · Score: 3, Insightful

      That sort of thing is just unfair. In my opinion, plagiarism is indeed a heinous crime in an academic setting because it goes against everything the pursuit of academics is supposed to be about. Given that, the punishment should be severe.

      However, since the punishment for plagiarism should be severe, there should be great care to investigate it properly. If you can show a preponderance of evidence that not only is a paper plagiarized, but you can accurately identify the source(s) from which each plagiarized section of it was copied, then the student should be expelled after the first offense. If you can't come up with that evidence, though, you should not be punishing the student.

      I thought professors had legions of grad students to ferret this sort of thing out, why do they need these programs? Trusting a decision that could permanently impact a student's entire life to a computer program seems careless and dangerous.

    4. Re:Insightful fact... by mathx314 · · Score: 2, Interesting

      It can be substantially higher than that as well. In high school I wrote a five page paper about A Tale of Two Cities, with a few lengthy quotes, being a book by Dickens. Since it wasn't a terribly long paper and I had length quotes, I got somewhere around 20% plagiarized. Fortunately my teacher was smart enough to check before accusing me, but I remember hearing some talk from a later English teacher that the department was considering a 10% cutoff, above which you received disciplinary action regardless of the circumstances.

    5. Re:Insightful fact... by BillCable · · Score: 4, Interesting

      My wife teaches for Phoenix. Probably 90% of the plagiarism she sees is from students copying and pasting whole papers word-for-word from random cheat sites. Occasionally she'll get someone who fails to properly quote sources, but that's very much the minority. For the most part, the cheaters aren't all that bright, nor do they try to hide their cheating. They're just hoping they get away with it.

    6. Re:Insightful fact... by Deagol · · Score: 2, Insightful

      I think that the objection here comes from the lack of transparency of the product being used. You input a paper, and you get a percentage answer. You're not given a list of papers/sources that registered a match (it would seem, anyway -- I don't know), thus you cannot verify the claims of the machine. Of course, being proprietary systems, I highly doubt that the vendor will allow inspection of the methods of detection or the database.

      The point is, that 35% means *nothing* useful without the exact context it was generated in.

      As we've seen with black-box voting machines, block-box web filters, and black-box breathalyzers, I suspect we'll see many lawsuits about black-box plagiarism detectors. After all, such a program can adversely affect one's long-term future, so the system better damned well be transparent and close to infallible (at least as much as the human-based method of detection).

    7. Re:Insightful fact... by El_Muerte_TDS · · Score: 4, Insightful

      For the most part, the cheaters aren't all that bright, nor do they try to hide their cheating.

      How would you know? The best cheaters won't be caught, but that doesn't mean they're not cheaters.

    8. Re:Insightful fact... by bob.appleyard · · Score: 2, Informative

      When I was at university, one of the lecturers showed us the plagiarism detection tool. Sure, it gave you a percentage, but it also gave you some output showing the passages in the text vs. what the program thought those passages had been taken from. He showed that most of the things that the tool had detected there were inconsequential, on the paper he was using for the demonstration.

      --
      How dare you be so modest!! You conceited bastard!!
    9. Re:Insightful fact... by johnsonav · · Score: 3, Insightful

      The best cheaters won't be caught, but that doesn't mean they're not cheaters.

      Sufficiently advanced cheating is indistinguishable from original work.

      How can you know that everyone isn't cheating? Do you give up? Or, try and pick the low-hanging fruit?

      --
      ... and that's when the C.H.U.D.'s came at me.
    10. Re:Insightful fact... by bcrowell · · Score: 4, Insightful

      In my opinion, plagiarism is indeed a heinous crime in an academic setting because it goes against everything the pursuit of academics is supposed to be about. Given that, the punishment should be severe. [...] the student should be expelled after the first offense

      I teach physics at a community college, and although I don't assign the kind of term papers you'd see in an English course, I do grade homework, lab writeups, and exams, and plagiarism is an issue that comes up. My school's policy is that the only punishment the professor can give for cheating is to assign a zero on that particular assignment. This is, in my opinion, almost no punishment at all; typically the reason people cheat is because they know they're going to fail, so assigning an F isn't a punishment, it's more like assigning the grade that the student actually earned. The school's administration tells us that this policy is the way it is because of a recent legal decision in California. Before this rule was imposed on us, my policy had been to give the student an F in the course if it was a serious case of cheating. In any case, my school, like most community colleges, has an extremely late drop deadline (the 14th week of the semester), so, e.g., if I give a student an F on an exam for cheating on the exam, the student will typically just drop the course, resulting in no penalty on his transcript other than a W, which will not affect his GPA.

      My school does provide a process where the professor can file a form to report academic misconduct. The form is then supposed to be followed up on by the dean, filed somewhere, and referred to later if the student shows a repeating pattern of cheating. Theoretically the student can be expelled, but never on the first offense. My experience is that this process doesn't actually seem to work, because the administrators involved aren't interested in spending the time and meeting with angry students. The threat hanging over the heads of the profs and deans is always that the parents will sue. Avoiding lawsuits is always the administration's top priority, far higher than education.

      The long and the short of it is that when a student makes a calculated decision to risk cheating, he's usually doing it based on a realistic assessment that the consequences of getting caught are extremely mild.

      However, since the punishment for plagiarism should be severe, there should be great care to investigate it properly. If you can show a preponderance of evidence that not only is a paper plagiarized, but you can accurately identify the source(s) from which each plagiarized section of it was copied, then the student should be expelled after the first offense. If you can't come up with that evidence, though, you should not be punishing the student.

      There is absolutely no way, at least at my school, that a student would ever be expelled for plagiarism. To get expelled, you would have to physically attack someone. You seem to be imagining a situation in which the professor and/or the school punishes the student just because a particular piece of software flashes a message on the screen saying "plagiarized." I can't believe that anyone would ever do that. Of course you're going to look at the text that matched, and see whether you really believe that it looks like it was plagiarized.

      I thought professors had legions of grad students to ferret this sort of thing out, why do they need these programs?

      No, most professors do not have grad students to do this. I work at a community college. No grad students. My wife teaches at Cal State LA. They have grad students, but the grad students don't work as TAs or graders; the professors have to grade 100% of the written work.

      Trusting a decision that could permanently impact a student's entire life to a computer program seems careless and dangerous.

      I don't think anyone does trust such a decision to a program. They use the program as a first step.

    11. Re:Insightful fact... by SerpentMage · · Score: 2, Interesting

      I think that this is very dangerous...

      Let me tell you about a situation. I was a speaker until recently. And around 98 I was giving a talk on technology X. Another speaker who was from the company who created the technology also gave a talk on technology X. Me and this other speaker knew each other, but we did not converse.

      Oddly our two talks were VERY VERY similar. He in a private manner accused me of copying his slide deck. Since he was a more well known speaker and I a newbie it seemed all logical.

      It was only when a good friend of mine who also worked at the company jumped in and said, "Naa, he would not do that."

      Then when my good friend came later to talk to me he asked, "you did not copy, right?"

      Answer was a definite NO! I did not copy. We just happened to be thinking along the same lines and came up with a VERY VERY similar slide deck.

      In other words a fluke! And this is why I hate statistics and numbers without a thought behind it.

      --

      "You can't make a race horse of a pig"
      "No," said Samuel, "but you can make very fast pig"
    12. Re:Insightful fact... by samcan · · Score: 3, Informative

      Forget 20%, I had a rough draft with as high as 61%! The particular service we used in high school was Turnitin.com, and a research paper I wrote for high school had an appendix with a copy of the 1805 Treaty of Tripoli (as a help for the teacher)...the website flagged that as 18% plagiarized, from some random Bell Atlantic user's website.

      Excluding that, the site would flag random sentences, and would flag part of a sentence as plagiarized, skip a word or two, and then say the rest of the sentence was plagiarized from the same source!

      An example is shown below (words in bold are supposedly plagiarized from one source, words in italics from another):

      Thus, the Founding Fathers wanted to create a government that was stable, and protected the rights of the people.

      Another example from a paper on the Russo-German war of 1941:

      They propose that German troops push all the way to the outskirts of Moscow, causing Joseph Stalin to abandon the city. While escaping, his train is destroyed by German planes, removing all signiïcant leadership to the Red Army.

      In another paper, when I quoted an article, I listed the title of the article in-text. Turnitin reported that the title of the article was plagiarism...of the article I was citing!

      Turnitin.com has "features" for excluding the quoted text, and excluding the bibliography, but as I use LaTeX, and like to use block quotes, the usefulness of these features are questionable.

      In my opinion, Turnitin.com is a joke.

    13. Re:Insightful fact... by mh1997 · · Score: 3, Funny

      I said "for the most part." It'd actually be a lot more effort to cheat and do enough to get away with it, than it would to just write the paper correctly. The people who are cheating seem to be doing it out of laziness or desperation. They run out of time to complete the assignment, so they Google something and use whatever pops up.

      I completely agree that It'd actually be a lot more effort to cheat and do enough to get away with it, than it would to just write the paper correctly. The people who are cheating seem to be doing it out of laziness or desperation. They run out of time to complete the assignment, so they Google something and use whatever pops up.

    14. Re:Insightful fact... by Urza9814 · · Score: 2, Interesting

      Here's the interesting thing: I have a professor that uses such tools (specifically TurnItIn.com), and I submitted a paper not too long ago and was told it was 6% plagiarized. No big deal. My prof said 20-30% would be allowable, as like you said, it's hard to write anything without seemingly plagiarizing. But the problem is this: I didn't cheat, I never saw anyone else's papers, and nobody ever saw mine. Yet a week later, after everyone else had sumbitted, suddenly I was at 23% plagiarized. Now, my professor didn't make any mention of it, but this raises a question about such services - they provide no way to see what the percentage was when submitted, only what it is now. And depending on how specific the prompt was that is being submitted, your percentage plagiarized can increase dramatically from other students submitting their own responses.

      Oh, and of course the reason nobody should act on such tools alone - I have yet to see one that can will determine if a source has been cited or not. That doesn't mean there aren't ones that do that out there - I would be surprised if there weren't - but with TurnItIn as my example again, if I made heavy use of attributed quotes in my paper, I may start off with 20%+ plagiarized. And after everyone else submits I may even break 50%. Even without plagiarizing a single sentence. Anyone who is stupid enough to rely entirely on the score some program gives has no place in education.

    15. Re:Insightful fact... by Idiomatick · · Score: 2, Insightful

      That is the goal. Culling the crappy cheaters is the same as culling the crappy students. So long as you are failing a high enough quota you will ensure a high enough quality of students make it through. This isn't new at all. Just make it so that to cheat and succeed it requires you to be as smart or smarter than someone doing the work legitimately.

    16. Re:Insightful fact... by DangerFace · · Score: 2, Insightful

      I think this kind of hits the nail on the head. The problem with plagiarism detection is that if you're writing a paper on the Russo-German war of 1941, or classical conditioning, or yaddah yaddah yaddah, is that unless you have found some significant new information, which is highly doubtful, everything you write will have been written before. The purpose of writing these papers - in general, at least - isn't in order to educate the entire field but to show that you have the ability to put together a coherent piece of work.

      In this day and age plagiarism is a bit like cheatbot.exe. When you can subcontract your work out to Indian PhDs, and Turnitin.com make every piece of work handed in to them, ever, available for download for a small fee, the only possible defence against plagiarism is decent teachers working decent hours and getting to know their pupils well enough to recognise their work. Admittedly, that isn't exactly ironclad but it's the best method for teaching anyway and it's the best way to avoid false positives, which is a priority for me since I don't plagiarise.

    17. Re:Insightful fact... by severoon · · Score: 3, Funny

      Sufficiently advanced cheating is called learning.

      Here's the proper way to cheat—it never failed me in university philosophy courses. Let's say you're supposed to read two or three of Nietzsche's works and write on the topic of Nietzsche: Feminist or Misogynist? You could try to read his books, but you won't understand them. Even if you do understand them, you will need to research his life in order to interpret his works in the proper context. And, after all that, when you finally do all this legwork, you'll only learn that he specifically designed his writings and behavior to lead you into a black hole. No normal human has a chance.

      So here's what you do. You put down the primary sources and go to the library. Read papers published by Ph.D. students that interpret Nietzsche's works and struggle to answer the question before you. Make notes on the general points of the argument and the supporting quotes across several of these papers (they're generally pretty short, and way easier to understand that the primary text). You can even read some Nietzsche if you're feeling adventurous, but I don't recommend it.

      Once you've formulated your own fervently held beliefs about Nietzsche in this way (by ripping them off of original thoughts by people that actually cared), you can leave with only your general notes outlining the arguments and citing the supporting quotes. If there's a good amount of material to choose from, make sure you choose an interpretation that is controversial (but well-supported)...don't turn in just another paper that will make the TA's eye's glaze or the professor want to put a gun in his mouth—liven things up a bit for those poor saps, they're stuck studying philosophy their entire lives! Look at all of the material you've collected and turn it over in your brain...try to synthesize your own controversial conclusion drawn from the points that others have worked so hard to create. Now go party for a couple of days to let it all sink in. The more beer you drink during this time, the less likely that some random quote you read will bubble up from the depths verbatim and get you busted. Once the requisite few days have been partied away, sit down and write the sentence or two that ties together all of the supporting material that you have decided to randomly & provocatively tie together. Include the points and supporting quotes to "prove" what you're saying.

      Instant A. Takes an hour, maybe two at the uni libe, and maybe another couple of hours to draft a typical 5-8 pager. This takes other students in the class weeks of devotion to achieve, leaving you plenty of time to study for your other courses, or study the local bar scene, or interact with the student body (as it were -winkwink-). The best part is, when you sit down to write after a couple of days of partying, it could be a paper, or it could be an in-class midterm or final. Whatever...either way, you're covered.

      --
      but have you considered the following argument: shut up.
    18. Re:Insightful fact... by bcrowell · · Score: 2, Informative

      Take, for instance, JPLs bullying to prevent the publication of the fact that its Titan lander photos -- which contain smoke plumes, amazingly enough -- actually are left-right reversed photos of Pearl Harbor, taken from a Japanese plane.

      I find this hard to believe. If you're going to make this statement, I would suggest that each time you say it, you also provide a URL or some other information that would allow the reader to verify your claims. Otherwise it just comes off like a kooky conspiracy theory.

    19. Re:Insightful fact... by DudeTheMath · · Score: 2, Interesting

      The important point when "incorporating research done by someone else into your own" (as BrokenHalo mentions--see, I'm citing my quotation!) is to cite (not necessarily quote) the other someone. This is how I got my A's in college and HS. Failing to do so is plagiarism.

      If you use someone else's idea, you cite it ("Hey, someone else thought of this before me."). That's it. If you use someone else's words, you quote it ("Someone else said 'exactly this'.").

      If you don't use someone's exact words, it makes it harder to spot and/or prove plagiarism, but it doesn't mean you can't be caught. And the brain is an amazing thing: You'd be surprised how often that clever phrase you write two or three days later, regardless of intervening beer, is a nearly exact quote.

      --
      You save only 59 seconds over 8 miles by going 75 instead of 65. Do you really have to pass that guy? Do the Math!
    20. Re:Insightful fact... by Zerth · · Score: 2, Interesting

      I think it might be an april fool's day prank, but I found this: http://csma31.csm.jmu.edu/physics/rudmin/titan/titan.htm

  2. Here is my perosnal take on the article... by svendsen · · Score: 4, Funny

    Does your school/university check your homeworks/theses for plagiarism? Nowadays, probably Yes, but are they doing it properly? Little is known about plagiarism detection accuracy, which is why we conduct a competition on plagiarism detection, sponsored by Yahoo! We have set up a corpus of artificial plagiarism which contains plagiarism with varying degrees of obfuscation, and translation plagiarism from Spanish or German source documents. A random plagiarist was employed who attempts to obfuscate his plagiarism with random sequences of text operations, e.g., shuffling, deleting, inserting, or replacing a word. Translated plagiarism is created using machine translation

    1. Re:Here is my perosnal take on the article... by MadKeithV · · Score: 2, Funny

      It's encrypted with double ROT-13.

  3. Plausible test? by fuzzyfuzzyfungus · · Score: 4, Insightful

    Now, I understand that plagiarism is common among the weakest of undergrad writers; but "machine translation from Spanish or German source documents" and "random text operations" seem like unrealistic experimental stimuli.

    In order to be a success, a plagiarized paper has to survive scrutiny by automated systems, if any are deployed, and human graders, if any are paying attention. Machine translation and text mangling should trivially defeat automated systems, at least any that aren't cranked well into World o' false positives territory; but would they pass human scrutiny? Even if they did, handing in something produced by machine translation and text mangling would probably earn you a referral to "Remedial English 101 For Life".

  4. Re:Defeat Plagerism by gnick · · Score: 5, Funny

    Simply using words would not constitute plagiarism. You just can't allow students to use words that somebody else has used before.

    For more information of this technique, please read my recent paper, Clickous Verandim Redundo Berata Quizzomandus.

    --
    He's getting rather old, but he's a good mouse.
  5. Irony by Shadow+Wrought · · Score: 4, Funny

    Just imagine everyone's surprise when all the entrants turn in the exact same process.

    --
    If brevity is the soul of wit, then how does one explain Twitter?
  6. Plagiarism detection is easy by DingerX · · Score: 4, Insightful

    A plagiarised paper just smells bad, and is characterized by shifts in voices and writing styles, sudden ignorance of the the critical points raised earlier. The same author who can't write a grammatically correct sentence one moment is throwing down complex constructions the next The harder part is identifying the source of the plagiarism. For undergraduate papers, even the harder part is trivial. After all, the point of plagiarism is that the author is too lazy to write anything original.

    For academics (professors), the situation isn't all that different. Plagiarism is usually a mix of stupidity, laziness and pressure to get stuff done. It usually happens where big, popularizing authors try to rip off the obscure ones (go back twenty years a la Mr. Ambrose, or pick something in a different language, preferably Italian), or when someone needs a book in an obscure field, and tries to pirate something really obscure.

    Even so, if a plagiarist has enemies who give a damn, they can find the source fairly fast. So why construct a test for the most obfuscated cases, when a plagiarist clever enough to obfuscate could simply write something original and sufficiently clever?

  7. We could ... by PPH · · Score: 4, Funny

    ... use the same system the US Patent Office uses for finding prior art.

    On second thought, scratch that idea.

    --
    Have gnu, will travel.
  8. detecting it is easy by Anonymous Coward · · Score: 2, Funny

    Calculate an md5 hash of the paper, if it matches the md5 of another, it's plagiarized.

    1. Re:detecting it is easy by fuzzyfuzzyfungus · · Score: 4, Funny

      And that is why I always change the font and margins on papers that I plagiarize...

  9. i don't get why you would do this by llamapater · · Score: 2, Interesting

    It's a monkeys on a typewriter thing. these companies add papers to there database as they compare them. If you feed enough papers into a database eventually they will all come back plagiarized there are not an infinite number of possible term papers there are only so many things that could be written for a topic that make sense, and most English teachers recycle topics. why English departments buy into this I don't understand let it go for long enough(it would only take another decade or two at most) and you will start getting people who didn't even know they were plagiarizing getting kicked out of college, I'm not talking about improper citations I'm talking about guy in Washington has the same idea as a guy in New York 20 years later. I'm not a lawyer, so i don't know if this is possible, but couldn't they copyright these databases in some form or render them proprietary. If they did that there business model could change to just collecting royalties.

  10. Too many false positives by russotto · · Score: 3, Interesting

    I once was on a Fido forum with someone who would often write responses nearly word-for-word identical to mine. It was uncanny; I'd see his post and recognize my own writing, only to realize it wasn't mine. Timestamps would sometimes show my post was written first, sometimes his. I imagine some others on the forum thought at least one of us was a sock puppet, but neither of us was.

    (If he's on slashdot, he's probably composing a post just like this one)

    That probably happens rarely. But build a big enough database, and it will happen often. Particularly given the restricted problem domains in undergraduate papers. It's not just a computer problem; even humans will think "plagiarism" when they see two papers with similar ideas and similar turns of phrase. Which I think demonstrates that plagiarism cannot be established satisfactorily merely by showing similarity between papers.

  11. Require submission of drafts; meet with students by cpu_fusion · · Score: 5, Interesting

    Plagiarism is a symptom of professors only being involved in the last step: reviewing the final product.

    Require the students to submit multiple drafts. Meet with them for 15 minutes each and discuss their thought processes on the ongoing paper. You'll get better final products, teach people not to procrastinate, and smoke-out people who have no involvement in their "own work."

    What, can't do that because you have 60 students in a class? Well, there's part of the problem too.

    We're trying to find a technology solution to a problem with less student-teacher interaction. Typical!

  12. The fingerprint model. by MarkvW · · Score: 2, Insightful

    Law enforcement uses automated fingerprint detection to identify possible matches. It never claims a match based on the computer.

    Using a program as the sole plagiarism judge and jury is profoundly unfair. If a university wants to discipline a student for a plagiarism hit, then it needs to obtain the source document--and pay the source document's creator if necessary to obtain it.

    Confronting the student with the alleged source gives the student a fair chance to defend himself/herself.

  13. The humanities are in trouble. by Areyoukiddingme · · Score: 5, Interesting

    Seriously, the humanities are in trouble. With over 6 billion people on the planet, it's extremely difficult to have an original thought. This sets the stage for endless repetition. Add to that the fact that the very process of teaching the humanities usually means imparting a teacher's single interpretation of the source material to the students who then do the natural thing when it comes to writing a paper and parrot back to the teacher what they've heard, knowing that's the only way to get a good grade, and the resulting combination is deadly.

    The papers are all going to be similar from the beginning, because it's a rare instructor who actually encourages dissenting opinions (and that fault in teaching is a whole other discussion of its own). Then the papers are going to be similar because there really are only so many ways to interpret the source material that are defensible. And finally, the papers are heavily likely to be similar to at least one other paper written about the subject, when every paper ever written on the subject is considered (exactly what the plagiarism sites attempt to do).

    I think the problem this competition is trying to solve is intractable in the face of the current educational system. It's gotten to the point where, if the software considers a large enough number of sources, even the instructor's own papers are going to look like plagiarism.

    Hell, look at the Slashdot comment system. A million people read the front page, but only a few thousand post comments. Thousands more are content to simply moderate the comments, and face it, comments they agree with are more likely to be modded up, one way or another. Then compare the modded comments. We get a lot of duplicate or near duplicate thought, and hence near duplicate comments on every article. Why? Because when you get enough people together in one place, discussing the same subject in writing, there are only so many viewpoints and only so many comments that won't get modded down for being of the "cubic what?" variety.

    Time to go back to grading on spelling and grammar. We've reached the end of the grading on ideas road. Coherency of presentation is all we have left. (One could argue it's all we ever had.)

    1. Re:The humanities are in trouble. by Areyoukiddingme · · Score: 2, Funny

      Shit, by the time I came back to the keyboard after writing this post and not hitting submit, there were 30 other posts that said the same thing. I must be a plagiarist.... Damnit.

    2. Re:The humanities are in trouble. by Archangel+Michael · · Score: 2, Interesting

      In my experience many professors (too many) basically ask for Plagiarized papers. There are a variety of reasons why they ask for plagiarized papers, mostly having to do with either laziness or wanting a particular view point regurgitated.

      People in my generation had Cliff Notes which one could spew forth a regurgitated version from Cliff Notes and get an A, while those with original thoughts would be graded much more harshly.

      I once researched a paper where it was fully documented with sources and such, all my own research and writing and got an D- on it. I finished the class basically plagiarizing my roommates papers from the year before and scoring A's and B's on all the other papers.

      What is the point of using your brain when it is punished?

      Not all Professors are like this, but too many are.

      --
      Agent K: A *person* is smart. People are dumb, stupid, panicky animals, and you know it.
    3. Re:The humanities are in trouble. by CodeBuster · · Score: 2, Insightful

      Perhaps I'm letting my engineering background run away with me

      I don't think so. I received my undergraduate degree in CS and I remember the ongoing feuds between the Humanities and the Sciences and especially the engineering disciplines (which includes CS at many universities) which tend to be the more pragmatic and practical ones among the scientists. The humanists would always dismiss us and our profession(s) as merely a necessary evil of modern society, considering their own bullshit musings to be the highest form of human development, the epitome of achievement, and what we would all be doing if we were somehow freed from the concerns of daily living and left with unlimited time to devote to philosophical discussion of the human condition. Scientists, and to a lesser extent engineers, tend to view the entire history of this planet and humans in general as merely temporary concerns in a temporary society and seek instead to understand the universe itself which existed before us and will probably be around long after our demise. This necessarily leads them into the study of mathematics which is really the antithesis of the humanities and causes some rather spectacular misunderstandings as the two opposite world views clash; but enough of this bullshit, I am beginning to sound like a humanist rather than an engineer.

  14. Uh, Use Google? by chainLynx · · Score: 2, Interesting

    Here's a good article explaining how Google makes plagiarism detection easy: http://questioncopyright.org/node/4 There was a story a couple years ago about one of these plagiarism detection services, Turnitin, getting sued for copyright infringement... does anyone know if that went anywhere? http://education.zdnet.com/?p=953

  15. Re:Require submission of drafts; meet with student by Colonel+Korn · · Score: 2, Insightful

    Plagiarism is a symptom of professors only being involved in the last step: reviewing the final product.

    Require the students to submit multiple drafts. Meet with them for 15 minutes each and discuss their thought processes on the ongoing paper. You'll get better final products, teach people not to procrastinate, and smoke-out people who have no involvement in their "own work."

    What, can't do that because you have 60 students in a class? Well, there's part of the problem too.

    We're trying to find a technology solution to a problem with less student-teacher interaction. Typical!

    I never taught a class involving humanities paper writing (in the science classes I taught, I could detect borrowed work by asking our kids to explain the calculations in their presentations and reports), but my wife meets with students several at least once after they turn in a required outline and bibliography to her. The bibliography, meeting, and my wife's extensive knowledge of scholarship in her field have made plagiarism rare and very obvious. Also, they make the students write vastly better papers and learn a lot more. Even having students meet with a TA to discuss paper ideas and progress is a huge help, and required outlines, drafts, and (especially) bibliographies should be part of the writing process in every lower level undergrad class. In upper level classes, the meeting is sufficient.

    --
    "I zero-index my hamsters" - Willtor (147206)
  16. This could be useful for blog unduplication by Animats · · Score: 2, Insightful

    This is a useful mechanism for search engines, which need to distinguish original content from hundreds or thousands of blogs echoing it. Imagine the Web with all the duplicate, repetitive material ignored. No wonder Yahoo is supporting this. Someone over there is thinking.

  17. Plagiarism vs. Ghostrwriting by jcohen · · Score: 3, Interesting

    I realize that plagiarism detection represents an interesting problem in computer science, and that it goes some distance toweard solving a serious problem. However, I read an article in the Chronicle of Higher Education, behind a paywall, alas, which leads me to believe that it is only a partial solution to academic dishonesty. The article suggested that, thanks to the Internet, the costs of human capital are now so low that hiring a ghostwriter to compose one's papers, sidestepping the problem of plagiarism to begin with, is far more expedient than plagiarism itself. It described a Russian-"businessman"-headed network of Filipino paper-writers, most paid between $1 and $3 a page, who are able to market their services to the West through a web site and remote call centers. At $20/page to the end-user, with no possibility of plagiarism detection, I think that most desperate students would find this a good deal. In my opinion, ghostwriting will supplant plagiarism as time goes on.

    What is a teacher to do? In-class writing samples would seem to be the only hope of detecting ghostwriting. Students could, of course, argue that at home, they can "polish" their papers, and that therefore they will not resemble the in-class samples. Moreover, checking samples against papers is a thankless and time-consuming task which is only a preliminary to actually evaluating the work. Perhaps there is a computer-based solution to this, but, in the meantime, perhaps potential ghostwriting customers could take their desires to their logical conclusion, and simply buy their degrees on the Internet directly.

    --
    "Imaginary solutions to real problems."
  18. My Dissertation by Kryis · · Score: 3, Interesting

    The Computer Science department at my uni routinely scans final year dissertations using automated software. Mine was flagged up as "possibly plagiarised"; a significant amount of content could be found elsewhere on the web (can't remember the exact percentage).

    My project supervisor said when he got the email from the system saying it came back positive he was very surprised - given the small amount of research in the area (there are only 5 or 6 papers on the same topic that I am aware of), and no other research on that exact method of solving the problem .

    When I found this out I was more than a little worried - I wasn't aware of copying any other work . It turns out that it had picked up on stupid stuff, like the boilerplate at the beginning of the dissertation, or phrases like "In conclusion,", and nothing longer than 3 or 4 words in any paragraph.

    This sort of plagiarism detection that detects word shuffling is fine for people that REALLY don't have a clue (i.e. the ones that forget to change the @author javadoc tag when copying their friends Java coursework), but it would still be relatively trivial to change enough words in a sentence to fool the system.

  19. You said it: Plagiarism detection is easy by kcdoodle · · Score: 2, Interesting

    If you have graded more than 2 assignments in your life, and really read each and every paper, and provided good critical feedback, then it is really easy to spot a plagiarized paper.

    Also, a grader usually knows the subject matter and has read many other good and bad works on the subject. You can get a feel for a person's writing style and depth of knowledge on a subject in just a few sentences. Then when you "smell something fishy", then it usually is.

    So far, whenever I "smell something fishy" I try to find the best sentence near the fishiness and paste it into Google. Plagiarists are not going to rewrite every sentence, if they do, then they probably learned something anyway. No, plagiarists are just lazy and in a hurry and deep down they know they deserve to be caught.

    --

    - I live the greatest adventure anyone could possibly desire. - Tosk the Hunted
  20. Wrong Problem by green1 · · Score: 2, Interesting

    They are trying to invalidate plagarism detection software by proving that you can still manage to plagarise in a way it won't detect (false negative). The thing is, this isn't the problem with plagarism software, the real problem is where it detects plagarism when none in fact took place (false positive). This will happen in a few ways:

    1) There have been several highly publicized incidents where students have been in big trouble for plagarising their own work. This is ludicrous, they wrote it in the first place!

    2) A large enough database of phrases, paragraphs, etc. will eventually encompass the majority of ways of phrasing a particular idea, therefore when discussing an existing idea the odds of saying something that has been said before will eventually approach certainty.
    Now this wouldn't necessarilly apply if you were inventing a whole new concept, but in most classes that is not what you are being asked to do, instead you are asked to research how something has already been done. There is bound to be duplication here, especially as the database grows. This doesn't mean you plagarised something, merely that someone else has worded something similarily in the past. (For it to be plagarism you would have had to have seen and copied that earlier work, in this case you may not even know about it.)

  21. Who needs plagiarism? by Ralph+Spoilsport · · Score: 5, Insightful
    When you've got Markov Generators?

    And the Postmodernism Generator?

    You don't have to write much of anything at all. Would you get a good grade? Fuck no. Would they FLUNK YOU FOR IT? Fuck no. Because its graded by untenured faculty who have to curry favour with students, or its graded by Grad Assistants who don't give a shit, and why should they.

    Oh, look, a paper by Cindy Bleethstain. She's a fucking idiot. Let's see. Hmmmm. Yup. Incomprehensible bullshit, as usual. Give her a C+ because some of it is intelligible and kind of funny.

    Oh, look another paper by Guido LeDouchebag. Bottlecaps are smarter than this turnip. Hmmm. Yup. More incomprehensible bullshit. C+. At least he finally discovered the spellchecker.

    THAT'S what it is often like, unfortunately.

    I read the paper, and if there is a passage that is noticeably different in tone, I'll copy past a section into Google and see where they pulled it. 9 times out of 10, it's a direct lift from a web page, unattributed. I send it back, and tell them "Footnotes, please. Also, automatic single grade loss. right off the top."

    If it comes back still broken, then I nail 'em for plagiarism. It's a big deal, and requires paperwork I don't like to fill out...

    So far I've only had one student have the cajones to not bother fixing their attributions, and he got crucified by the Ethics board. He was an arrogant little prick, too.

    RS

    --
    Shoes for Industry. Shoes for the Dead.
    1. Re:Who needs plagiarism? by HikingStick · · Score: 2, Interesting

      My problem with automatic checks is that there is always a chance that someone's seemingly original thought may actually reflect thoughts someone else already may have put down on paper. I remember being accused of copying someone else's work once. It was in the early '80s when the Internet as we know it was not part of general public awareness. When the instructor interrogated me on the sentence (one sentence in a paper at least five pages long), he insisted I copied it from some specific book or article. I had absolutely no clue what the guy was talking about. At the time, all I ever read was Fred Saberhagen, Tolkein, Piers Anthony, and Terry Brooks. After what seemed like forever (it was probably no more than ten minutes), he finally realized that I had no clue about his source, and I'm guessing he realized that whatever I wrote matched the way I wrote the rest of the paper and the way I used the spoken word. Yech! I haven't thought about that situation in a long time. I must be getting old.

      --
      I use irony whenever I can, but my shirts are still wrinkled...
  22. But if the teacher cares about the students... by AliasMarlowe · · Score: 3, Insightful

    The students cannot fake it, if the teacher cares about them learning.

    Many many many moons ago, I was a Chem. Eng. grad student. This was before the internet existed, and before my beard had turned gray. One of my duties to pay my way was supervising a lab course for undergrads, and marking the students' lab reports (they were expected to produce about 20 pages per week just on this one lab course). I insisted on interviewing them individually on their reports, where they had to explain their results and conclusions. Nobody tried faking anything twice, because it was caught immediately; they had to read up and understand the background, or they were in deep shit. That class got the highest average mark ever in the year-end exam on the associated theory (the professor was pleasantly surprised).

    --
    Those who can make you believe absurdities can make you commit atrocities. - Voltaire