Data Engineer, Legal AI Tech Scale-Up

Job not on LinkedIn

🔥 1 hour ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Xayn

Xayn

11 - 50 employees

Founded 2017

🤖 Artificial Intelligence

⚖️ Legal

☁️ SaaS

Artificial Intelligence • Legal • SaaS

Xayn is a pioneering legal tech company that offers Noxtua, Europe’s first sovereign Legal AI, designed to enhance legal workflows for lawyers. Noxtua provides advanced tools such as legal research capabilities and document drafting assistance through its AI models that are trained on high-quality legal data, ensuring compliance with European data protection laws. This innovative solution aims to streamline legal processes and improve efficiency in the legal industry.

📋 Description

• Design, build, and optimize end-to-end ETL pipelines for legal data from multiple jurisdictions • Work extensively with XML-based legal data feeds: parse, validate, normalize, and transform XML structures into scalable internal schemas and unified document formats • Develop and maintain data models and storage schemas that support continuously updated datasets while ensuring consistency, scalability, and accuracy across diverse datasets • Coordinate data handover and integration from multiple internal and external data providers, ensuring reliable and timely updates • Implement and continuously refine metadata enrichment strategies to maximize searchability, ranking quality, and relevance of legal information in vector databases. • Build and maintain a high-performance search and retrieval infrastructure enabling agent-based systems to retrieve the most relevant legal information efficiently • Collaborate with product, AI, and legal domain experts to deliver high-quality, reliable data solutions • Own the data integration of one jurisdiction end-to-end

🎯 Requirements

• at least 2 years of professional experience in data engineering • Strong Python skills with experience in designing robust data pipelines • Experience in building and maintaining reliable ET and RAG pipelines • Solid understanding of data modeling, quality, filtering, validation, and consistency • Familiarity with containerization (Docker), CI/CD pipelines, and version control (Git) • Strong grasp of data structures, algorithms, system design principles, and software engineering best practices • Expertise in working with graph databases • Familiarity with developing and deploying NLP models is a bonus • English proficiency at the C2 level

🏖️ Benefits

• 100% remote work possible (given a German residence), other countries upon request • Flexible working hours • 26 days vacation + December 24th & 31st off, + 1 additional vacation day per year of employment (up to 30 days) • Urban Sports Club Membership discounts, depending on location • Laptop (Lenovo or Mac) plus €1,000 net home office setup budget (paid with your first salary)

Apply Now

Similar Jobs

🕒 4 days ago

AWISEE

11 - 50

📣 Marketing

☁️ SaaS

🛍️ eCommerce

AI Engineer designing and building data ingestion pipelines for digital marketing solutions. Collaborating with teams to develop AI-ready datasets and implement data quality frameworks.

AWS

Docker

ETL

Kubernetes

Python

SQL