- Firecrawl does not provide Pay-As-You-Go Pricing, you can use Scrappey to get HTML page source without getting blocked and then pass it to the llm-reader
get_processed_textfunction to get LLM Ready Text. - You can use any API to get HTML page source or your own set up as well. I have found Scrappey to be most cost effective and performant. Feel free to share any other similar available option.
Reasons to choose Scrappey and llm-reader repo
Start for as little as $0.001 per URL, scale as needed, and never lose credits. Larger deposits automatically unlock discounts and your balance never expires.
Apart from the Pay-As-You-Go Option if you prefer predictable monthly billing then Scrappey plans give you more value, and unused credits roll over so you never waste money.
€15/month → 16,500 URLs ~84% cheaper than Firecrawl's $19 plan
€49/month → 59,300 URLs
€99/month → 125,200 URLs ~9% cheaper than Firecrawl's $99 plan
All with unused credits roll over
llm-reader output is optimized to give fewer tokens - meaning you pay less when sending data to LLMs.
If your websites do not require JavaScript rendering or need for browser then you can use the faster and cheaper (5 times cheaper) HTTP mode.
Our code for converting webpages into clean, LLM-ready text is open source. You can self-host, modify, or integrate it into your existing workflows.
Steps to Use Scrappey with llm-reader repo
- Install llm-reader as per README. You can skip installing Playwright dependencies and browser as we will use scrappey to get HTML page source.
- To get HTML page source without getting blocked, we use Scrappey API. This is what you pay for. The conversion to LLM-ready text is 100% free because it's open source.
- Get your API Key from Scrappey.
- Example Code:
import requests from url_to_llm_text.get_llm_input_text import get_processed_text # pass html source text to get llm ready text # Get HTML Page Source without getting blocked headers = { 'Content-Type': 'application/json', } params = { 'key': 'your_scrappey_api_key', } # visit scrappey request builder for more parameters json_data = { 'cmd': 'request.get', 'url': 'your_url', # 'requestType': 'request', # if using HTTP request mode 'browserActions': [ { 'type': 'wait', 'wait': 3, }, ], } response = requests.post('https://publisher.scrappey.com/api/v1', params=params, headers=headers, json=json_data) # Get LLM Ready Text res = response.json().get('solution','') if res != '': page_source = res.get('response','') llm_text = await get_processed_text(page_source,'https://parseextract.com') print(llm_text)
Please make sure to use the Scrappey API link provided above. It offers an affordable and high-performance way to obtain LLM-ready text, and using this link also helps support my open-source work through affiliation, at no additional cost to you. A win-win for everyone involved.