I am trying to code a crawler based on PHP with curl. I have

Question

0

Asked: June 15, 20262026-06-15T13:46:18+00:00 2026-06-15T13:46:18+00:00

I am trying to code a crawler based on PHP with curl. I have

0

I am trying to code a crawler based on PHP with curl. I have database of 20,000-30,000 URLs that I have to crawl. Each call to curl to fetch a webpage takes around 4-5 seconds.

How can I optimize this and reduce the time required to fetch a page?

Report

Leave an answer
Cancel reply

You must login to add an answer.

Need An Account,

1 Answer

Editorial Team · Answer 1 · 2026-06-15T13:46:19+00:00

You can use curl_multi_* for that. The amount of curl resources you append to one multi handle is the amount of parallel requests it will do. I usually start with 20-30 threads, depending on the size of returned content (make sure your script won’t terminate on memory limit).

Note, that it will run as long as it takes to run the slowest request. So if a request times out, you might wait for very long. To avoid that, it can be a good idea to set timeout to some acceptable value.

You can see the code example at my answer in another thread here.

Sign Up

Sign In

Forgot Password

The Archive Base Latest Questions

I am trying to code a crawler based on PHP with curl. I have

Leave an answerCancel reply

1 Answer

Leave an answer
Cancel reply