콘텐츠로 건너뛰기
메뉴
커뮤니티에 참여하려면 회원 가입을 하시기 바랍니다.
신고된 질문입니다
1 회신
3987 화면

Hello,


Google can't for some reason fetch generated robots.txt, although it should be able to do so. I can access robots.txt using curl with following output:

curl http://www.mydomain.com/robots.txt

User-agent: *
Disallow: /web/login
Allow: *

User-Agent: Googlebot
Disallow: /web/login


Google is complaining:


Failed: Robots.txt unreachable

Any idea what is wrong?

Also, because of that Google can't access sitemap.xml. 

Another problem is about sitemap.xml. I contains URL's with http, not https prefix. They are valid, as we have http->https redirection rule, but I would prefer to have it correctly in sitemap in the first place. Any help with that?


Many thanks in advance.


Lumir

아바타
취소
작성자 베스트 답변

I have found a problem and fixed it. In our case problem was, that we had issue with Nginx proxy settings. When we were accessing the our domain webpages using curl (command line) or Safari, everything seems to be working. But when we tried access website using Firefox, we received an SSL error:

SSL_ERROR_RX_UNEXPECTED_NEW_SESSION_TICKET

We had to move from all sites handled by the Nginx line 

ssl_session_tickets off;

to the /etc/nginx/nginx.conf, section http {} and restart the Nginx.

More info here https://serverfault.com/questions/1021041/browsers-reported-ssl-error-when-one-of-the-server-blocks-in-nginx-configur

This was preventing Google accessing URL's of the domain. Now it's fixed.


아바타
취소
관련 게시물 답글 화면 활동
1
1월 23
3666
1
6월 17
6126
2
7월 15
9228
1
2월 25
1094
2
12월 24
6121