Data Engineer: Customer and Prospect Database (CAPDB) Build
Only freelancers located in the U.S. may apply.U.S. located freelancers only
Project Overview We are launching a CAPDB (Customer and Prospect Database) program: a single, scored database of every current customer and every target prospect in our market. The CAPDB consolidates CRM, billing, product usage, and third party firmographic data into one master account database, applies AI assisted enrichment to fill and validate attributes, and scores every account against the profile of our best customers. You will be a dedicated hands-on builder for this program, working alongside our internal business systems and data team and reporting to the Director of IT and Business Systems. Scope of Work • Data model design. Design and build the CAPDB schema (master account, contact, and activity layers) in our Azure SQL / Microsoft Fabric environment. • Ingestion and consolidation. Build and maintain pipelines that consolidate data from our internal systems. • Entity resolution. Match, merge, and deduplicate account records across sources into a clean master record with clear survivorship rules. • AI enrichment. Build LLM assisted enrichment workflows to append, standardize, and validate firmographic attributes such as firm size, practice areas, geography, and tech stack. • Account scoring. Partner with GTM stakeholders to implement a scoring and tiering model (for example Ideal / Emerging / Acceptable / Avoid) based on fit and historical performance signals such as retention, expansion, and deal velocity. • Salesforce activation. Push scores, tiers, and enriched attributes back into Salesforce so sales and marketing can act on them in their daily workflows. • Documentation and handoff. Document the data model, pipelines, and scoring logic so our internal team can own the platform long term. Required Qualifications • 5+ years in data engineering or analytics engineering roles • Expert SQL and hands-on experience with Azure SQL Database • Microsoft Fabric experience (Data Factory pipelines, Lakehouse or Warehouse, semantic models), or deep experience with the equivalent Azure Synapse / Power BI stack • Working knowledge of the Salesforce data model and Salesforce data integration • Python for pipeline and enrichment work • Proven experience with data consolidation, master data management, or entity resolution across multiple systems • Strong written communication; able to work independently with a distributed team Nice to Have • Experience calling LLM APIs (Anthropic Claude, Azure OpenAI) for data enrichment or classification at scale • Familiarity with any of our surrounding stack: Fivetran, Chargebee, Pendo, Gong, ZoomInfo • Power BI report development and DAX • B2B SaaS RevOps or GTM analytics background; ICP definition, account scoring, or CAPDB style projects
- More than 30 hrs/weekHourly
- 6+ monthsDuration
- IntermediateExperience Level
$45.00
-
$65.00
Hourly- Remote Job
- Ongoing projectProject Type
Skills and Expertise
Activity on this job
- Proposals:50+
- Interviewing:1
- Invites sent:1
- Unanswered invites:0
About the client
- United StatesEast Quogue6:17 PM
- $177K total spent24 hires, 10 active
- 2,494 hours
Explore similar jobs on Upwork
How it works
Create your free profileHighlight your skills and experience, show your portfolio, and set your ideal pay rate.
Work the way you wantApply for jobs, create easy-to-by projects, or access exclusive opportunities that come to you.
Get paid securelyFrom contract to payment, we help you work safely and get paid securely.
About Upwork
- 4.9/5(Average rating of clients by professionals)
- G2 2021#1 freelance platform
- 49,000+Signed contract every week
- $2.3BFreelancers earned on Upwork in 2020
Find the best freelance jobs
Growing your career is as easy as creating a free profile and finding work like this that fits your skills.
Trusted by