Apify
Hosted actors, proxies, storage, scheduling, and web automation
The visible web scraping platform loop is buildable, but a credible replacement needs more than the first screen. Apify earns its keep through connectors, auth, reliability, so expect a weekend or multi-day build and a narrower personal scope.
Build verification: not recorded. How we judge buildability
What you give up
- OAuth app verification
- durable execution at scale
- schema drift handling and enterprise controls
- calendar-provider edge cases and timezone correctness
- hundreds of maintained connectors
Why people still pay
Apify: One automation is easy. Subscribers pay for maintained integrations, credentials, retries, observability, and someone else owning breakage.
Your build guide
The stack, security requirements, and agent rules for a focused replacement.
Before you start
- Runtime and tools: Python, HTTP parsing, SQLite and optional Playwright for explicitly permitted rendered pages.
- Before starting: A small authorized URL set, reviewed selectors or field schema and a user-approved crawl frequency; no proxy-evasion machinery.
Use these project rules and optional skill references alongside the prompt. Review each skill before adding it to your agent; the AGENTS.md export includes the same guidance.
Project rule — domain: Store ExtractorVersion, Run, FrontierURL, PageResult and DatasetRow; deduplicate by canonical product ID and persist the next page only after current rows commit.
Project rule — scope and recovery: Ship one trusted extractor, not arbitrary hosted actors. Constrain browser networking and resources; blocked access is a stopped run, not a trigger for fingerprint or account evasion.
Project rule — acceptance: Stop on page three, restart the run and introduce a changed selector on page four; previous products remain unique and the selector failure is shown instead of an empty successful dataset.
Project rule — delivery: document real setup commands and permissions; do not claim a build, accuracy level, performance result or security certification that has not been demonstrated.
Recommended skill: modern-python — structure the Python worker or explicitly optional read-only utility with pinned dependencies, typed boundaries and clear failure handling. Follow the maintainer's installation instructions and match its requirements to the chosen runtime.
Recommended skill: agent-browser — inspect the approved browser workflow and reproduce permitted page interactions; the CLI/browser runtime is a separate prerequisite. Follow the maintainer's installation instructions and match its requirements to the chosen runtime.
Implementation plan
Phase 1
Pin the working slice and create its example input: Run one owned-site catalogue extractor from a reviewed URL list, inspect each request, checkpoint pagination and export a structured dataset. Confirm setup: A small authorized URL set, reviewed selectors or field schema and a user-approved crawl frequency; no proxy-evasion machinery.
Phase 2
Implement persistence and write-time invariants before decorating the UI: Store ExtractorVersion, Run, FrontierURL, PageResult and DatasetRow; deduplicate by canonical product ID and persist the next page only after current rows commit.
Phase 3
Connect the working view to real saved state. Bound hosts, redirects, page count, body size and timeout. Reject private-network destinations throughout redirects, respect access restrictions and backoff, and distinguish extraction failure from a missing field.
Phase 4
Expose the app-specific limits and recovery path in context: Ship one trusted extractor, not arbitrary hosted actors. Constrain browser networking and resources; blocked access is a stopped run, not a trigger for fingerprint or account evasion.
Phase 5
Walk through this concrete acceptance case and preserve its exported evidence: Stop on page three, restart the run and introduce a changed selector on page four; previous products remain unique and the selector failure is shown instead of an empty successful dataset. Finish the README and backup/restore instructions; report unfinished capabilities explicitly.
Build the following focused alternative to Apify. This is a deliberately limited personal or small-team substitute, not parity with the paid service. WORKING SLICE Run one owned-site catalogue extractor from a reviewed URL list, inspect each request, checkpoint pagination and export a structured dataset. SETUP AND ARCHITECTURE Use Python, HTTP parsing, SQLite and optional Playwright for explicitly permitted rendered pages. Prerequisites: A small authorized URL set, reviewed selectors or field schema and a user-approved crawl frequency; no proxy-evasion machinery. Before integrating anything, record actual versions and permissions, plus model files or provider limits only where used, in the README; make unavailable dependencies visible rather than simulating success. DOMAIN MODEL AND INVARIANTS Store ExtractorVersion, Run, FrontierURL, PageResult and DatasetRow; deduplicate by canonical product ID and persist the next page only after current rows commit. IMPLEMENTATION CONTRACT Bound hosts, redirects, page count, body size and timeout. Reject private-network destinations throughout redirects, respect access restrictions and backoff, and distinguish extraction failure from a missing field. Provide an input/setup view, the main work view, and a review/export view appropriate to this workflow. Preserve the last saved state if a job or save fails. Include empty, loading, permission-denied, partial and retryable-error states. Log identifiers and error categories without secret values or unnecessary private content. APP-SPECIFIC BOUNDARY AND RECOVERY Ship one trusted extractor, not arbitrary hosted actors. Constrain browser networking and resources; blocked access is a stopped run, not a trigger for fingerprint or account evasion. ACCEPTANCE SCENARIO Stop on page three, restart the run and introduce a changed selector on page four; previous products remain unique and the selector failure is shown instead of an empty successful dataset. Also reopen the app after an interrupted operation, confirm the saved record/export remains inspectable, and document the recovery action. These are implementation acceptance requirements, not a claim that this guide has been tested. DELIVERY Deliver a runnable repository with migrations or project-format versioning, a non-sensitive example, environment/permission setup, the exact manual acceptance steps, and a backup/export-and-restore walkthrough. Implement the working slice before optional integrations; list any deferred paid-product capabilities honestly. Do not add capabilities outside the working slice just to resemble the original product. PROJECT RULES FOR AGENTS.md Keep the domain invariants above executable at the write boundary. Propose scope changes before adding providers or permissions. Never fabricate source evidence, publish results, identity matches or successful delivery. Preserve user originals and require an explicit confirmation for destructive changes or external publication.
$ open in your agent (prompt prefilled, you press enter), copy the prompt or copy or download AGENTS.md
prompt copied. want to know what dies next week?
new verdicts + top votes, weekly. free. one-click out.
Apify pricing
| plan | monthly | annual (per mo) | what you get |
|---|---|---|---|
| free | $0/workspace | $0/workspace | $5 prepaid platform usage/month; $0.20 per compute unit; 16 GB RAM; 25 concurrent runs; 5 datacenter proxy IPsHard stop when the monthly prepaid usage is exhausted. |
| creator | $1/workspace | — | $500 one-time platform-usage bonus valid for 6 months; 64 GB RAM; 32 concurrent runs; own and Universal Actors only; 10 GB residential proxy and 10,000 SERP monthly maximumsSix-month prepaid promotional plan costing $6 total; the $500 bonus is one-time and expires after 6 months. |
| starter | $29/workspace | $26.10/workspace | $29 prepaid usage/month; $0.20 per compute unit; 64 GB RAM; 32 concurrent runs; 30 datacenter IPs then $1/IP |
| scale | $199/workspace | $179.10/workspace | $199 prepaid usage/month; $0.16 per compute unit; 256 GB RAM; 128 concurrent runs; 200 datacenter IPs then $0.80/IP |
| business | $999/workspace | $899.10/workspace | $999 prepaid usage/month; $0.13 per compute unit; 512 GB RAM; 256 concurrent runs; 500 datacenter IPs then $0.60/IP |
| enterprise | — | — | Custom prepaid usage, compute, proxy, concurrency, security, SLA and support allowancesCustom quote. |
free tier$5 platform usage/month, 16 GB RAM, 25 concurrent runs and 5 datacenter proxy IPs; services stop until reset when the $5 is exhausted
billingfree + monthly or annual paid plans; annual subscriptions save 10%; usage beyond prepaid credits is billed pay-as-you-go
hidden costsThe subscription is prepaid usage, not unlimited service: compute, residential/SERP/datacenter proxies, storage, transfer and paid Store Actor event fees draw down credits or overrun; unused monthly credits expire; extra concurrent runs cost $5/run/month, RAM $1/GB/month and training $150/hour
pricing sources checked 2026-08-14 · pricing source ↗
Questions about Apify
Can you build your own Apify with AI?
Partly. The visible web scraping platform loop is buildable, but a credible replacement needs more than the first screen. Apify earns its keep through connectors, auth, reliability, so expect a weekend or multi-day build and a narrower personal scope.
What does the Apify build prompt cover?
The prompt starts with this scope: Run one owned-site catalogue extractor from a reviewed URL list, inspect each request, checkpoint pagination and export a structured dataset. Full-product capabilities excluded from the comparison include: OAuth app verification; durable execution at scale; schema drift handling and enterprise controls. Follow the implementation plan and its prerequisites before expanding the build.
How do I use the prompt, AGENTS.md and agent skills?
Start with the Apify prerequisites and stack, then copy the prompt into your coding agent. Save the project rules as AGENTS.md in the project root. Linked skills are optional packages or source instructions for specific tasks; review their current contents and install only those matching the chosen stack. A skill does not supply API credentials or verify the finished app.
How long will this Apify project take?
The catalogue estimate is weekend to multi-day for the limited scope. Setup, integration approvals, debugging, deployment and ongoing maintenance can add time. This is an estimate, not a delivery guarantee.
What would I give up by replacing Apify?
OAuth app verification; durable execution at scale; schema drift handling and enterprise controls; calendar-provider edge cases and timezone correctness; hundreds of maintained connectors. Apify: One automation is easy. Subscribers pay for maintained integrations, credentials, retries, observability, and someone else owning breakage.
What price is this guide comparing against?
The recorded Starter plan is $29 reference price (monthly), checked 2026-08-14. Check the linked pricing source before buying. Building your own also has hosting, API and maintenance costs; the recorded amount is not a guaranteed saving.
What can I use instead of building Apify?
The prior-art section lists Activepieces, n8n as starting points. Review their current scope, license and maintenance before adopting one.