Here's the comparison table: https://github.com/apify/mcpc?tab=readme-ov-file#related-wor...
748 karma · joined October 31, 2011
Software engineer | Founder & CEO @ Apify
https://x.com/jancurn
https://apify.com/jancurn
https://www.linkedin.com/in/jancurn
Here's the comparison table: https://github.com/apify/mcpc?tab=readme-ov-file#related-wor...
- https://github.com/apify/mcpc
- https://github.com/chrishayuk/mcp-cli
- https://github.com/wong2/mcp-cli
- https://github.com/f/mcptools
- https://github.com/adhikasp/mcp-client-cli
- https://github.com/thellimist/clihub
- https://github.com/EstebanForge/mcp-cli-ent
- https://github.com/knowsuchagency/mcp2cli
- https://github.com/philschmid/mcp-cli
- https://github.com/steipete/mcporter
- https://github.com/mattzcarey/cloudflare-mcp
- https://github.com/assimelha/cmcpWe’re running a marketplace of 8,000+ tools called Actors for all kinds of web data extraction and automation use cases (see https://apify.com/store). Just last month, we paid out more than $500k to community developers who publish these Actors on the Apify platform.
The unit economics work for niche tools: scrapers for specific platforms, packaged open-source tools, MCP servers, or API wrappers. Too small for building SaaS around it, but developers earn a few thousand dollars per month as passive income.
We believe there can be many more such Actors. So we're putting $1M in prizes on the table to motivate developers to build new, useful Actors. Our bet is that 10,000 new specific tools can widely expand the capabilities of many AI agents and unlock a lot of value.
This is Jan, the founder of Apify (https://apify.com/) — a full-stack web scraping platform.
With the help of Python community and the early adopters feedback, after an year of building Crawlee for Python in beta mode, we are launching Crawlee for Python v1.0.0.
The main features are:
- Unified storage client system: less duplication, better extensibility, and a cleaner developer experience. It also opens the door for the community to build and share their own storage client implementations.
- Adaptive Playwright crawler: makes your crawls faster and cheaper, while still allowing you to reliably handle complex, dynamic websites. In practice, you get the best of both worlds: speed on simple pages and robustness on modern, JavaScript-heavy sites.
- New default HTTP client `ImpitHttpClient` (https://crawlee.dev/python/api/class/ImpitHttpClient), powered by the Impit (https://github.com/apify/impit) library): fewer false positives, more resilient crawls, and less need for complicated workarounds. Impit is also developed as an open-source project by Apify, so you can dive into the internals or contribute improvements yourself: you can also create your own instance, configure it to your needs (e.g., enable HTTP/3 or choose a specific browser profile), and pass it into your crawler.
- Sitemap request loader: easier to start large-scale crawls where sitemaps already provide full coverage of the site
- Robots exclusion standard: not only helps you build ethical crawlers, but can also save time and bandwidth by skipping disallowed or irrelevant pages
- Fingerprinting: each crawler run looks like a real browser on a real device. Using fingerprinting in Crawlee is straightforward: create a fingerprint generator with your desired options and pass it to the crawler.
- Open telemetry: monitor real-time dashboards or analyze traces to understand crawler performance. easier to integrate Crawlee into existing monitoring pipelines
For details, you can read the announcement blog post: https://crawlee.dev/blog/crawlee-for-python-v1
Our team and I will be happy to answer here any questions you might have.
we’re publishing this whitepaper that describes a new concept for building serverless microapps called Actors, which are easy to develop, share, integrate, and build upon. Actors are a reincarnation of the UNIX philosophy for programs running in the cloud.
Our goal is to make Actors an open web standard. We’d love to hear your thoughts.
Here’s the corresponding GitHub repo: https://github.com/apify/actor-whitepaper
BTW we continuously update this exhaustive post covering all legal aspects of web scraping: https://blog.apify.com/is-web-scraping-legal/
Please note that this is the first release, and we'll keep adding many more features as we go, including anti-blocking, adaptive crawling, etc. To see where this might go, check https://github.com/apify/crawlee
The real motivation for this project comes from a simple frustration that it's not easy to make living out of developing your hobby open source software, even though it's being used by thousands of companies and brings them significant value. The only accessible way to make living from your software is by selling it as SaaS, which in turn is still overly complicated. We believe there is a better way and today we'd like to show it to you. I'm looking forward to hear what you think.
The real motivation for this project comes from a simple frustration that it's not easy to make living out of developing your hobby open source software, even though it's being used by thousands of companies and brings them significant value. The only accessible way to make living from your software is by selling it as SaaS, which in turn is still overly complicated. We believe there is a better way and today we'd like to show it to you.
I'm looking forward to hear what you think.
Apify builds software technology and infrastructure that helps anyone leverage the full potential of the largest source of information ever created in the history of humankind - the Web. Our mission is to let people automate mundane tasks on the web and spend their time on things that matter. We strive to keep the web open as a public good and a basic right for everyone, regardless of the way you want to use it, as its creators intended.
Join Apify and help us make the web more programmable!
It’s pretty funny to see that in 2006 Elon Musk didn’t even deserve a mention of his name. How things have changed...
Please give it a try, it's free. We'd love to hear what you think - both good or bad things :)
Disclaimer: I'm a co-founder of Apify :)
https://www.apify.com/docs/actor https://www.apify.com/docs/scheduler
Disclaimer: I'm a co-founder of Apify
Anyway, we hope you’ll give it a shot and we’re really looking forward to hear what you think about it. All feedback welcome!
For example, a simple actor to convert HTML to PDF looks like this:
https://www.apify.com/jancurn/url-to-pdf
More info:
https://www.apify.com/docs/actor
https://www.apify.com/docs/sdk/apify-runtime-js/latest
https://www.apify.com/library?type=acts
Disclaimer: I'm a co-founder of Apify