I want to create a crawler with C#. The problem is that some websites

Question

0

Asked: June 1, 20262026-06-01T11:58:57+00:00 2026-06-01T11:58:57+00:00

I want to create a crawler with C#. The problem is that some websites

0

I want to create a crawler with C#. The problem is that some websites have disabled black listed crawlers in their robots.txt file, using:

User-agent: *
Disallow: /

Is there a way I could fake my request to show that I’m for instance Googlebot?

Report

Leave an answer
Cancel reply

You must login to add an answer.

Need An Account,

1 Answer

Editorial Team · Answer 1 · 2026-06-01T11:58:58+00:00

HttpWebRequest has .UserAgent, however – I would simply say: don’t.

Of course, your point re robots.txt is rather moot; that is for you to follow. If you write a badly behaved tool that ignores robots.txt regardless of what you claim as your user-agent, then you should expect to be blacklisted fairly quickly.

In particular, trying to impersonate any of the major players is very dubious. Frankly I’d expect most major sites to also check the incoming IP range.

Sign Up

Sign In

Forgot Password

The Archive Base Latest Questions

I want to create a crawler with C#. The problem is that some websites

Leave an answerCancel reply

1 Answer

Leave an answer
Cancel reply