Forum Moderators: Robert Charlton & goodroi
Both these DCs all, and I mean ALL of my sites that have 301's set-up got hit with www and non-www duplicate. These 301s have been up since February.
Hope these DCs are not the end result.
Anyone else experience this?
I am really at a loss.
When someone using a free webhosting site copies my 4 year old domain that doesn't expire for 5 years, it should be obvious to Google engineers who is the original website, and who is copying the content - without even a human looking at the page.I'm not suggesting that free hosting sites/pages get dropped or punished.
Quite the contrary. It's quite easy to get scraper sites removed if they are hosted on free hosting companies. Gather up all your proof, get the very strict copyright rules off the host site and send all this info to the hosting company. It's been my experience they drop the scrapers instantly without question. For more info on how to contact these hosts search for Stop 302 Redirect Hijacking.
Janiss
My site, which was decimated in May from Bourbon, came back in July, has now been decimated again as of September 21. Some I know who has researched these things extensively says it's because of my pages being framed by others and because of being hijacked - all this adds up to a glitch on Google's part that makes it refuse to acknowledge content sites like mine and leaves them low in the SERPs.
Just add a pop out of frames script on all affected pages. However check to make sure it works in all browsers.
If you believe your site has been hijacked do the search mentioned above to see actions you can take to stop it.
For all those who have seen their sites index quadrupled and drop in the SERPs--are you using session IDs? If so you might want to read the following:
A Comment from googleguy:So what's the problem with a session id, and why doesn't Googlebot crawl them? Well, we don't just have one machine for crawling. Instead, there are lots of bot machines fetching pages in parallel. For a really large site, it's easily possible to have many different machines at Google fetch a page from that site. The problem is that the web server would serve up a different session-id to each machine! That means that you'd get the exact same page multiple times--only the url would be different. It's things like that which keep some search engines from crawling dynamic pages, and especially pages with session-ids.
Google can do some smart stuff looking for duplicates, and sometimes inferring about the url parameters, but in general it's best to play it safe and avoid session-ids whenever you can.
Google's Webmaster Technical Guidelines:
*Use a text browser such as Lynx to examine your site, because most search engine spiders see your site much as Lynx would. If fancy features such as JavaScript, cookies, session IDs, frames, DHTML, or Flash keep you from seeing all of your site in a text browser, then search engine spiders may have trouble crawling your site.
*Allow search bots to crawl your sites without session IDs or arguments that track their path through the site. These techniques are useful for tracking individual user behavior, but the access pattern of bots is entirely different. Using these techniques may result in incomplete indexing of your site, as bots may not be able to eliminate URLs that look different but actually point to the
Thanks.
- My provider moved and so my IP adress changed
But I cannot see a reason why my site is banned completely.
If your provider placed your site on a shared IP this could be the problem. Possibly someone else on that IP was banned which means all sites on that IP get banned.
Solution: Request a dedicated IP address. Most reputable hosts only charge $1.00 more per month. If your host doesn't provide this then find a new host.
In my case I think, I have found the source of my problem.After Allegra I used robots.txt and URL removal console to remove duplicate content. This was in March. After that I continously had a robots.txt with
User-agent: *
Disallow: dup1.php
...
If your .php scripts are using Session IDs see my message above re Google not being able to process them correctly.
Do you see any new canonical url issues?
We do.
We have had 301s in place since February. How are we getting hit with the canonical issue now?
I can not understand how Google got so dumb!
My biggest concern is, like sites in the past, we are dropping in rankings now, but are going to disappear completely once the effects of the cononical url issue fully kicks in.
thank you for your help. Fortunately my provider is a good friend of mine and my ip is very dedicated ;-)
I was just wondering if moving to another ip could cause a problem.
Meanwhile I found out what caused my problem: it is in fact duplicate content which came up again - I stated that somewhere in this thread...
Concerning the Session-IDs: That was my mistake three years ago when I kicked out myself from google the first time.
It was someone like you who helped me out then. That's why I love being here: People help each other.
Does anybody remember engines like Excite, Alta Vista (remember black monday?, or InfoSeek? Are we seeing another slow death?
Google does not care. Why?
I think Google is trying to monetize their adwords. Quality search is not a priority. Priority is developing all of their 'new' ideas.
Majority of general public is so hooked on 'Googling' it, they don't even realize that their are being chalked up crap results. They scam free listings and jump to adwords quicker than you can 'giddiup!'.
For Mr. Cutts to claim no knowledge of 301 issues is ridiculous!
I brought a domain anysite.co.uk on 05-Sep-2005
Added content to site made manually before 17-Sep-2005
Added first link on the Internet on 19-Septem-2005
Indexed on 20-September, appear in Google all pages indexed ~60 pages.
25-Sept-2005 last visit of Googlebot.
26-Sep-2005 dissapeared from Google.
Can someone explain me what is going on?
It is about my IP? .htacces for 301 redirect from non-www to www?
Regards,
Dan
It has become painfully obvious (IMHO) that engines are years away from being able to use a technical approach only to discerning the good sites from the bad. Without human intervention, far too many quality sites will be killed in an effort to clear out the spam. What good will a technical approach be (however advanced) if in the end the serps you are left with are only 50-60% as "good" as they could be due to so amny quality sites being lef tout in the cold?
Having said that, the google gravy train has been nice, but it won't make or break my business. Making a few grand a day is all too easy without google, you just have to find the traffic elsewhere.
Here are few interesting threads about the sandbox.
[webmasterworld.com...]
[webmasterworld.com...]
[webmasterworld.com...]
I hope this helps.
aff_dan, sounds like classic sandbox to me. Your pages won't be seen in Google for an extended period of time now.
I might disagree with that. I mean, he could very well be in the sandbox, and I hope he is not. But, he could just be in the freshy index, not fully indexed yet.
Not much time has passed since he loaded his site.
I would wait a few more weeks before diagnosing him with a sandbox.
[copyscape.com...] and see if your content has been duped.