Data Engineer – Complex Data Pipelines

🕒 vor 27 Tagen

🇫🇷 Frankreich – Remote

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

🚰 Dateningenieur

👻 Geisterscore 12%

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

Jetzt Bewerben
Ähnliche Remote-Jobs finden

📊 Überprüfen Sie Ihre Lebenslauf-Bewertung für diese Stelle

Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

Logo of MARSS Group

MARSS Group

51 - 200 Mitarbeiter

📦 Logistik

💼 Beratung

🎖️ Verteidigung

Logistics • Consulting • Defense

Die MARSS Group ist ein Technologieunternehmen, das sich auf die Entwicklung fortschrittlicher Sicherheits- und Überwachungssysteme zur Stärkung der nationalen Sicherheit spezialisiert hat. Das 2005 gegründete Unternehmen verfügt über mehr als 15 Jahre Erfahrung in Forschung und Zusammenarbeit mit der EU, der NATO und verschiedenen Verteidigungsbehörden. Zu den technologischen Innovationen von MARSS gehören integrierte sensorbasierte Überwachung, künstliche Intelligenz und Open-Source Intelligence (OSINT). Diese Technologien schützen weltweit kritische Infrastrukturen, maritime Anlagen, Spezialeinheiten und hochrangige Personen. Das Produktportfolio umfasst Systeme wie NiDAR™ zur Weiterentwicklung von Command-and-Control-Funktionen sowie verschiedene Lösungen zur Abwehr unbemannter Luftfahrtsysteme und zur Verbesserung des Lagebilds in unterschiedlichen Sicherheitskontexten.

Beschreibung

• Design and build data pipelines from scratch, from data ingestion through processing, transformation, storage and consumption • Design ingestion for distributed recording nodes that are offline most of the time, including local buffering, resumable transfer, and reconciliation of late-arriving or out-of-order data on reconnection • Define, together with the ML team, which data is prioritised during short and bandwidth-limited connection windows • Design pipelines capable of handling structured/tabular data, images, video and temporal/time-series data • Work with sensor-generated sequential data and handle clock drift across nodes to ensure trustworthy downstream timing • Develop robust and scalable data processing solutions using Python and SQL • Design data models and storage approaches, including capacity planning and retention for large volumes of image and video data on self-managed storage • Own workflow definitions in the orchestration layer, including ordering, retry, idempotency and backfill behaviour • Build processes for data ingestion, transformation, validation, quality control and traceability • Develop tools supporting data preparation and availability for machine learning and AI applications • Collaborate with Machine Learning Engineers, DevOps and Software Engineers to understand data requirements and provide solutions • Ensure pipelines are reliable, maintainable and scalable as data volumes and use cases increase • Identify data-quality issues, including gaps and duplicates caused by node outages and retries, and define data-quality and pipeline-health monitoring • Define the overall architecture and technical standards for the in-house data platform • Document pipeline architecture, data flows and technical solutions

🎯 Anforderungen

• Strong professional experience in a closely related data engineering role • Demonstrated experience designing and implementing data pipelines from zero (not cloud-based), including architectural and technical decisions • Good programming skills in Python & SQL • Professional experience working with several different types of data • Experience building pipelines involving at least some of the following: images, video, sensor data, time-series or other sequential data • Experience with systems that must tolerate unreliable or absent network connectivity and recover gracefully, offline-first, store-and-forward, edge collection or similar architectures • Experience running data infrastructure on bare metal or self-managed servers, rather than exclusively on managed cloud services • Good understanding of data ingestion, transformation, storage, validation and data-quality principles • Experience working with large or complex datasets • Strong Linux skills, including comfort with filesystems, storage, services and network troubleshooting • Good software engineering practices, including Git, testing, code review and documentation • Ability to independently investigate technical problems and propose appropriate architecture and solutions • Fluent English, written and spoken.

🏖️ Vorteile

• Full-time employment (CDI)

Jetzt Bewerben

Ähnliche Jobs

🕒 vor 28 Tagen

Free2move

201 - 500

📦 Logistik

🚘 Automobilindustrie

🏥 Gesundheitswesen

Data Engineer building real-time pricing pipelines and ML infrastructure for Free2move’s global mobility platform. Owning production data products across France, Spain, or Italy.

🇫🇷 Frankreich – Remote

💵 €50.000 - €60.000 / Jahr

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

🚰 Dateningenieur

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

BeReal.

51 - 200

👥 B2C

📱 Medien

Senior Data Engineer constructing and managing data pipelines for BeReal's growth. Collaborating with engineers and product teams to enhance data-driven solutions.

🇫🇷 Frankreich – Remote

💰 €60.000.000 Series B - BeReal. im 2022-10

⏰ Vollzeit

🟠 Senior

🚰 Dateningenieur

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 2 Monaten

Shotgun

11 - 50

📣 Marketing

✈️ Reisen

Revenue Data & Operations Engineer optimizing data infrastructure at Shotgun. Focused on data pipelines, analytics, and custom app development to support commercial operations and revenue growth.

🇫🇷 Frankreich – Remote

💰 Venture Round im 2020-02

⏰ Vollzeit

🟢 Junior

🟡 Mittelstufe

🚰 Dateningenieur

🗣️🇫🇷 Französisch erforderlich

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 2 Monaten

Actian

201 - 500

💼 Beratung

🏥 Gesundheitswesen

📦 Logistik

Software Architect designing resilient, cost-efficient cloud data platforms for Actian, a data management company. Improving developer tooling, security, observability, scalability, and cloud infrastructure standards.

🇫🇷 Frankreich – Remote

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

🚰 Dateningenieur

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 6 Monaten

SPARTEO

51 - 200

📣 Marketing

🤖 Künstliche Intelligenz

Lead Data Engineer architecting high-volume batch and streaming data systems for Sparteo’s AI-powered advertising technologies. Leading data engineers, reliability, governance, and Generative AI integration.

🇫🇷 Frankreich – Remote

⏰ Vollzeit

🟠 Senior

🚰 Dateningenieur

🗣️🇺🇸🇬🇧 Englisch erforderlich