proposing r.jina.ai as an alternative to FireCrawl to extract web content #7803
JOduMonT
started this conversation in
Suggestion
Replies: 2 comments 1 reply
|
side comment; as a non Chinese user I don't understand why I'm force to click this
|
1 reply
|
Another option worth considering: Purify (https://github.com/Easonliuliang/purify) Single Go binary, no Redis/Playwright dependencies. Focuses on token reduction — tested 52-99% savings depending on the site. Also has a built-in MCP server if you are using Claude or Cursor. Self-hosting is just one binary or docker compose up. Might be a good fit for Dify knowledge base ingestion where you want clean content without the overhead. |
0 replies
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Uh oh!
There was an error while loading. Please reload this page.
Self Checks
1. Is this request related to a challenge you're experiencing? Tell me about your story.
I'm using r.jina.ai to extract web content and it would be nice if it was proposed in DiFy.
Since r.jina.ai is already in the tools, I believe it would not be so complex to do.
2. Additional context or comments
I'm using r.jina.ai to extract web content and it would be nice if it was proposed in DiFy.
Since r.jina.ai is already in the tools, I believe it would be not so complex to do.
3. Can you help us with this feature?
All reactions