
11 - 50 employees
💼 Consulting
📦 Logistics
🏛️ Government
Consulting • Logistics • Government
Civic Marketplace is on a mission to unlock local agency procurement through innovation, by building a marketplace that saves time and taxpayer dollars. Their platform provides government procurement offices with the ability to make purchases with ease, simplicity, and full legal compliance. By using cutting-edge technology, Civic Marketplace enhances public service delivery with transparency and efficiency, ensuring access to a network of reliable, pre-approved vendors. They focus on supporting diverse suppliers to foster local economic development, addressing the challenges faced by procurement officers in outdated systems and inefficient processes.
🔥 0 minutes ago
🌐 United Kingdom, United States – Remote
⏰ Full Time
🟡 Mid-level
🟠 Senior
🚰 Data Engineer
👻 Ghost score 13%
Improve your chances of getting an interview by checking your resume score before you apply.

11 - 50 employees
💼 Consulting
📦 Logistics
🏛️ Government
Consulting • Logistics • Government
Civic Marketplace is on a mission to unlock local agency procurement through innovation, by building a marketplace that saves time and taxpayer dollars. Their platform provides government procurement offices with the ability to make purchases with ease, simplicity, and full legal compliance. By using cutting-edge technology, Civic Marketplace enhances public service delivery with transparency and efficiency, ensuring access to a network of reliable, pre-approved vendors. They focus on supporting diverse suppliers to foster local economic development, addressing the challenges faced by procurement officers in outdated systems and inefficient processes.
• Build and own a trusted canonical procurement dataset for analytics and agentic procurement • Diagnose and fix data quality issues across supplier, agency, solicitation, contract, and awarded-quote data • Design and implement reliable, observable, cost-effective ingestion pipelines from APIs, flat files, scraped pages, PDFs, and other external sources • Resolve supplier entities across legal entities, spellings, trading names, subsidiaries, registries, and identifiers • Define shared canonical models for suppliers, contracts, commodities, and agencies • Build testing, lineage, freshness, provenance, and self-serve data capabilities • Design ingestion and storage strategies for agentic workflows and optimize vector pipelines such as Pinecone for RAG • Collaborate with the Head of Engineering to shape architecture • Write pipelines, carry the pager, and investigate raw sources when data appears incorrect • Collaborate with product engineering, customer success, and the Growth Lead • Serve as the first dedicated data engineering hire and function lead alongside the Head of Engineering
• Data engineering experience in a startup or scale-up, where you've built the platform rather than inherited one • Strong SQL and Python, and the judgement to know when the warehouse is the wrong place to solve a problem • A data model you designed that other people had to live with, including the parts you would do differently now • Real experience of messy external sources: half-documented APIs, flat-file drops, scraped pages, PDFs, and data you neither control nor can correct • Entity resolution or record linkage experience, or clear evidence you'd be good at it • Commercial judgement to distinguish revenue-blocking data problems from merely interesting ones • Instinct to diagnose why a number is wrong before proposing a fix • Instinct to build systems that scale rather than pipelines that run once • Genuine AI fluency • Determination to stay positive through the ups and downs of a fast-moving startup • Comfort operating remotely across timezones with real autonomy and not much oversight • Curiosity about public-sector data • Govtech, public sector, civic tech, public records, or open-data experience is advantageous • Familiarity with supplier and procurement data, including SAM.gov, UEI, DUNS, NAICS, UNSPSC, cooperative purchasing, COG, or NIGP is advantageous • Experience with OpenSearch, Elasticsearch, or vector pipelines such as Pinecone or Milvus feeding an LLM product is advantageous • Bilingual English and Spanish is especially valuable • Experience being the first data hire is advantageous
• Competitive salary and early-stage equity • Comprehensive medical, dental, and vision insurance • Flexible PTO • Remote-first, with real flexibility across timezones • Full AI tool stack: Claude Pro, HubSpot, Make, Notion, and more • Regular team offsites, including international meet-ups • Direct access to the founding team and a front-row seat to building something that matters
Apply Now🕒 4 days ago
Senior Data Engineer building secure data pipelines and reusable AI platform infrastructure for Oyster’s global employment platform. Enabling LLM, RAG, vector search, and production AI capabilities.
🕒 5 days ago
Senior Data Engineer managing PostgreSQL, SQL Server, and Oracle databases for M3’s global healthcare market research business. Applying AI and automation to improve data reliability, security, and delivery efficiency.
🕒 6 days ago
Lead Data Engineer building semantic, MCP, and data-quality infrastructure for Mojo Mortgages, a fintech mortgage brokerage. Enabling governed self-service analytics in an FCA-regulated environment.
🕒 6 days ago
Lead Data Engineer directing Microsoft Fabric, Python, and PySpark data engagements for CreateFuture, an AI-native consulting partner. Leading teams, architecture, pipelines, and client delivery.
🕒 September 28
Geospatial Data Engineer building scalable GIS pipelines and data products for Arden University. Supporting evidence-based decisions across its expanding digital higher-education ecosystem.