Get stucked when use firecrawl to crawl (choose a website as knowledge base) #11531

Closed
opened 2026-02-21 19:00:50 -05:00 by yindo · 0 comments
Owner

Originally created by @ninekirin on GitHub (Mar 18, 2025).

Self Checks

  • This is only for bug report, if you would like to ask a question, please head to Discussions.
  • I have searched for existing issues search for existing issues, including closed ones.
  • I confirm that I am using English to submit this report (我已阅读并同意 Language Policy).
  • [FOR CHINESE USERS] 请务必使用英文提交 Issue,否则会被关闭。谢谢!:)
  • Please do not modify this template :) and fill in all the required fields.

Dify version

1.0.1

Cloud or Self Hosted

Self Hosted (Docker)

Steps to reproduce

  1. Create a new knowledge base
  2. Choose sync from website and choose firecrawl
  3. Input url
  4. Set options: crawl child page/limit: 200/subpage limit: 5
  5. Run

✔️ Expected Behavior

Job succeeded and I can choose what pages to import

Actual Behavior

Stucked at 55/141, the logs of firecrawl indicated that Job succeeded but in dify the status was always scraping.

Image

And new requests produced when the status was scraping, infinite loops and never stop.

Image

Originally created by @ninekirin on GitHub (Mar 18, 2025). ### Self Checks - [x] This is only for bug report, if you would like to ask a question, please head to [Discussions](https://github.com/langgenius/dify/discussions/categories/general). - [x] I have searched for existing issues [search for existing issues](https://github.com/langgenius/dify/issues), including closed ones. - [x] I confirm that I am using English to submit this report (我已阅读并同意 [Language Policy](https://github.com/langgenius/dify/issues/1542)). - [x] [FOR CHINESE USERS] 请务必使用英文提交 Issue,否则会被关闭。谢谢!:) - [x] Please do not modify this template :) and fill in all the required fields. ### Dify version 1.0.1 ### Cloud or Self Hosted Self Hosted (Docker) ### Steps to reproduce 1. Create a new knowledge base 2. Choose sync from website and choose firecrawl 3. Input url 4. Set options: crawl child page/limit: 200/subpage limit: 5 5. Run ### ✔️ Expected Behavior Job succeeded and I can choose what pages to import ### ❌ Actual Behavior Stucked at 55/141, the logs of firecrawl indicated that `Job succeeded` but in dify the status was always scraping. ![Image](https://github.com/user-attachments/assets/9857595f-d249-4f5a-b1ed-dcd2a8e3d0d9) And new requests produced when the status was `scraping`, infinite loops and never stop. ![Image](https://github.com/user-attachments/assets/4774df11-030c-4d3c-ab78-4d09a4d31d5a)
yindo added the 🐞 bug label 2026-02-21 19:00:50 -05:00
yindo closed this issue 2026-02-21 19:00:50 -05:00
Sign in to join this conversation.
1 Participants
Notifications
Due Date
No due date set.
Dependencies

No dependencies set.

Reference: langgenius/dify#11531