Slashdot Mirror


95% of User-Generated Content Is Bogus

coomaria writes "The HoneyGrid scans 40 million Web sites and 10 million emails, so it was bound to find something interesting. Among the things it found was that a staggering 95% of User Generated Content is either malicious in nature or spam." Here is the report's front door; to read the actual report you'll have to give up name, rank, and serial number.

12 of 192 comments (clear)

  1. Want to get ripped? by Anonymous Coward · · Score: 5, Funny

    I got ripped in 2 weeks. learn how with secret juice formula.

  2. Let me be the first to post that this is BS. by nicknamenotavailable · · Score: 5, Funny

    That is so untrue. There is value in what I write.

  3. It might be true, but it's also irrelevent. by onion2k · · Score: 5, Insightful

    95% of user-generated posts on Web sites are spam or malicious.

    The fact is that there are millions of old blogs, unused forums, ancient guestbooks, etc that are easy to spam automatically. While it might very well be true that 95% of comments on the internet are spam of some sort, they're probably read by a tiny fraction of internet users. People tend to stick to about a dozen big sites that get very little rubbish posted on them at all.

    Car analogy: 95% of cars are rusty old heaps of crap that can't move. Thankfully they're in scrapyards and not on the roads.

    1. Re:It might be true, but it's also irrelevent. by mwvdlee · · Score: 5, Funny

      95% of humans are over 100 years old. Most of them are dead.

      --
      Slashdot social media options: AIM, ICQ, Yahoo, Jabber and Mobile Text. Why no MySpace?
    2. Re:It might be true, but it's also irrelevent. by Anonymous Coward · · Score: 5, Funny

      That should be on Fox News.

      "Number of dead people reaches all time high!"

    3. Re:It might be true, but it's also irrelevent. by Trepidity · · Score: 5, Informative

      It seems that at least as well as anyone can estimate, the current population really is about 5% of the total humans who've ever lived.

    4. Re:It might be true, but it's also irrelevent. by CAIMLAS · · Score: 5, Insightful

      A lot of forum software works well, until it gets "behind the curve", and then the site maintainer pulls the site*.

      By "behind the curve" I mean any of the following can/does happen:
      1) Forum software gets out of date and user fails to upgrade due to modifications or similar, resulting in spam.
      2) Forum software gets popular without having a good security model and/or update cycle, resulting in exploits.
      3) Gets inundated with comment approvals and the forum (or blog) gets ignored or set to auto-allow out of frustration.

      * By "pulls the site" I mean "abandons it but doesn't take it down". That's typically the end result.

      It's a lot of work to maintain your own forum and/or blog: managing spam can and will take hours+ from your day if you've not got a good automated and/or textual way to deal with it: web interfaces are clumsy.

      Car analogy: 95% of cars are rusty old heaps of crap that can't move. Thankfully they're in scrapyards and not on the roads.

      Yet, unlike most of those cars, the actual blog content is not necessarily useless. I have seen quite a few abandoned blogs and/or forums which have 3-10 year old information on them which is by no means useless; it's just getting buried.

      Digital archeologists of the future will probably have to figure out an automated way to prune back the spam to find the actual Internet, the way things are going.

      Consider: if spam accounts for 95% of all user-generated content, and said user-generated content is actually a non-trivial percentage of all actual content online (believable), consider how much bandwidth gets wasted by these spammers. (Thankfully, I suspect most of the 'user generated content spam' doesn't show up on the first couple search page results so it's not going to likely be perused with regularity - unless it's more heavily seeded on topics common folks search.)

      --
      ~/ssh slashdot.org ssh: connect to host slashdot.org port 22: too many beers
    5. Re:It might be true, but it's also irrelevent. by Anonymous Coward · · Score: 5, Funny

      Well, then MSNBC would just rip into Fox for inferring these unfortunate individuals should no longer vote. CNN would chime in and blame the lack of universal health care for the deaths.

  4. Was going to RTFA but it's probably bogus by syousef · · Score: 5, Funny

    ...95% probability actually. So I didn't bother.

    --
    These posts express my own personal views, not those of my employer
  5. So Sturgeon was right by Aussie · · Score: 5, Interesting

    "Ninety percent of everything is crud."

    http://en.wikipedia.org/wiki/Sturgeon's_Law

  6. Re:This just in by Smegly · · Score: 5, Funny

    a staggering 95% of User Generated Content is... ...spam. Here is the report's front door; to read the actual report you'll have to give up name, rank, and serial number.

    Give up your Name, rank, email... so we can enlighten you with valuable information from our partners.

  7. Re:This just in by VoltageX · · Score: 5, Informative

    Sorry to hijack this, but http://securitylabs.websense.com/content/Assets/WSL_ReportQ3Q4FNL.PDF seems to be the direct link to the paper.

    --
    "Anonymous could not immediately be reached for further comment." - International Business Times