You will get Data Extraction from Public Websites Using Python


Project details
I will build a Python script to extract publicly available data from a website and deliver it in CSV, Excel, or JSON format.
✅ Only publicly accessible pages
✅ No login, passwords, or private accounts
✅ Respect robots.txt and website terms
✅ One website per project
Typical use cases:
Product listings
Job postings
Articles or blog data
Business directories
You will receive:
Clean, structured dataset
Python script (optional)
Clear explanation of extracted fields
⚠️ Client must confirm the website allows scraping before ordering.
✅ Only publicly accessible pages
✅ No login, passwords, or private accounts
✅ Respect robots.txt and website terms
✅ One website per project
Typical use cases:
Product listings
Job postings
Articles or blog data
Business directories
You will receive:
Clean, structured dataset
Python script (optional)
Clear explanation of extracted fields
⚠️ Client must confirm the website allows scraping before ordering.
Data Tool
PythonWhat's included
| Service Tiers |
Starter
$45
|
Standard
$85
|
Advanced
$150
|
|---|---|---|---|
| Delivery Time | 2 days | 3 days | 4 days |
Number of Pages Mined/Scraped | 1 | 4 | 10 |
Number of Sources Mined/Scraped | 1 | 1 | 1 |
Number of Revisions | 2 | 3 | 5 |
Optional add-ons
You can add these on the next page.
Fast Delivery
+$20 - $30
Additional Page Mined/Scraped
+$10
Additional Source Mined/Scraped
+$20
Additional Revision
+$10
Deliver Python Script
+$25Frequently asked questions
About Trojan
Google Workspace Automation Engineer | Python, APIs, AI Workflows
Tirana, Albania - 2:35 am local time
If your team lives in Gmail / Sheets / Drive and things are messy (copy-paste, CSV chaos, missed follow-ups, no audit trail), I turn that into an automated workflow with logs, alerts, and clean handoff docs. 📩📊📁🧾🔔
What I deliver:
* 🤖 Gmail / Drive / Sheets automation using Python + Google APIs
* 🔄 Intake → processing → database/dashboard (Sheets or Postgres/BigQuery)
* ✅ Approvals + 🔔 notifications (email/Slack) + 🧩 run tracking
* 🧼 Data cleaning, 🧹 dedupe, ✅ validation, and 🚨 exception reporting
* 🧠 “AI add-ons” when useful: 📝 summarization, 🏷️ classification, 🧭 routing (with guardrails 🛡️)
How I work (so projects don’t drift):
1. 🗺️ Map your workflow + success criteria (30–60 min)
2. 🚀 Ship a working MVP fast
3. 🧱 Harden it: 🔁 retries, 🪵 logs, 📈 monitoring, and 📚 documentation
Send me your current process (a Loom 🎥, screenshots 🖼️, or a sample Sheet 📄), and I’ll propose the simplest automation that saves time immediately. ⏱️💡
Steps for completing your project
After purchasing the project, send requirements so Trojan can start the project.
Delivery time starts when Trojan receives requirements from you.
Trojan works on your project following the steps below.
Revisions may occur after the delivery date.
I review your requirements
Once you provide the target website, data fields, and output format, I confirm feasibility and ask for clarifications if needed.
I build and test your custom scraper
I create the scraper, extract the required data, clean/format it, and prepare the final dataset.