Principal Data Engineer, LLM/AI Platforms

🕒 June 12

🏄 California – Remote

info

💵 $195k - $290k / year

⏰ Full Time

🔴 Lead

🚰 Data Engineer

🦅 H1B Visa Sponsor

info
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of CrowdStrike

CrowdStrike

5001 - 10000 employees

Founded 2011

🔒 Cybersecurity

☁️ SaaS

🤖 Artificial Intelligence

Cybersecurity • SaaS • Artificial Intelligence

CrowdStrike is a cybersecurity company that provides cloud-based security services to stop breaches. It is recognized as a leader in endpoint protection, identity and cloud security, and managed detection and response. CrowdStrike's platform, Falcon, integrates artificial intelligence to offer real-time visibility, detection, and protection against sophisticated cyber threats. The company is lauded for its effectiveness in securing networks and data, making it a trusted partner for businesses worldwide.

📋 Description

• Architect, implement, and optimize data platforms and pipelines specifically designed to support LLMs, Retrieval-Augmented Generation (RAG), and sophisticated AI agentic systems at Exabyte scale. • Drive the adoption and deployment of agentic workflows and agent harnessing techniques to create autonomous, data-driven security features. • Design and implement highly scalable, fault-tolerant, and cost-effective data solutions, emphasizing rapid iteration and high-quality deployment. • Write elegant, production-ready code with a focus on performance, maintainability, and testing rigor, ensuring the ability to ship fast without compromising quality. • Provide technical leadership and deep expertise in data modeling, normalization, and semantic cataloging for AI/ML workloads. • Establish best practices for MLOps/DataOps surrounding LLMs, including monitoring, observability, and zero-touch recovery mechanisms for AI services. • Actively mentor engineers, conducting technical workshops, leading design reviews, and strengthening the team's knowledge in cutting-edge AI platform technologies. • Collaborate across the organization with Data Scientists, Product Managers, and other engineering teams to transform research prototypes into robust, production-grade services. • Own the end-to-end lifecycle of critical data services: development, testing, deployment, and monitoring.

🎯 Requirements

• Master’s degree or PhD in Computer Science, Data Engineering, or a related STEM field, or equivalent practical experience • 10+ years of progressive experience in Data Engineering/Platform Engineering, with at least 3 years focused on architecting and building platforms for AI/ML or Data Science at massive scale • Demonstrable hands-on experience in LLM engineering (fine-tuning, prompt engineering, deployment), RAG, and developing agentic workflows • Proven track record of designing and delivering large-scale distributed systems (sharding, partitioning, concurrency) • Exceptional ability to write clean, elegant, performant, and well-tested code, coupled with a proactive mindset for delivering results quickly • A thorough understanding of engineering practices, including effective peer code reviews, resilient architecture design, and comprehensive testing paradigms • Prior experience in a Principal or Staff level engineering role, demonstrating technical leadership and mentorship capabilities.

🏖️ Benefits

• Market leader in compensation and equity awards • Comprehensive physical and mental wellness programs • Competitive vacation and holidays for recharge • Paid parental and adoption leaves • Professional development opportunities for all employees regardless of level or role • Employee Networks, geographic neighborhood groups, and volunteer opportunities to build connections • Vibrant office culture with world class amenities • Great Place to Work Certified™ across the globe

Apply Now

Similar Jobs

🕒 June 12

Pivotal Health

51 - 200

⚕️ Healthcare Insurance

☁️ SaaS

🤝 B2B

Staff Software Engineer developing and architecting data warehouse solutions for healthcare reimbursement. Focused on building scalable infrastructure and data governance frameworks.

🇺🇸 United States – Remote

💵 $210k - $240k / year

💰 Pre seed on 2024-01

⏰ Full Time

🔴 Lead

🚰 Data Engineer

🕒 June 12

Pivotal Health

51 - 200

⚕️ Healthcare Insurance

☁️ SaaS

🤝 B2B

Staff Data Engineer building data pipelines and warehouse infrastructure for healthcare reimbursement. Collaborating with software engineers and stakeholders to optimize data usage and quality.

🇺🇸 United States – Remote

💵 $200k - $230k / year

💰 Pre seed on 2024-01

⏰ Full Time

🔴 Lead

🚰 Data Engineer

🕒 June 12

Gainwell Technologies

10,000+ employees

⚕️ Healthcare Insurance

Advisor Data Architect leveraging technology to improve health outcomes managing Medicaid and Major Health Care Payer data projects. Collaborating with clients to ensure data integrity and governance.

🇺🇸 United States – Remote

💵 $99.2k - $141.7k / year

💰 Grant on 2023-06

⏰ Full Time

🟠 Senior

🔴 Lead

🚰 Data Engineer

🦅 H1B Visa Sponsor

info

AWS

Azure

BigQuery

Cloud

NoSQL

Tableau

🕒 June 10

Texas Windstorm Insurance Association

201 - 500

🤝 Non-profit

🏛️ Government

Data Warehouse Developer leveraging expertise in Data Warehousing and Business Intelligence to design resilient data solutions for TWIA/TFPA. Collaborating across teams to transform data into meaningful insights for decision-making.

🇺🇸 United States – Remote

⏰ Full Time

🟠 Senior

🔴 Lead

🚰 Data Engineer

ETL

Guidewire

Informatica

SDLC

SSIS

Tableau

🕒 June 10

Teamworks

501 - 1000

⚽ Sports

☁️ SaaS

🤖 Artificial Intelligence

Staff Data Engineer co-defining AWS lakehouse architecture and leading data maturity initiatives for innovative sports tech platform. Collaborating across teams to integrate performance data into analytics and ML insights.

🇺🇸 United States – Remote

💵 $216k / year

💰 $235M Series F - Teamworks on 2025-06

⏰ Full Time

🔴 Lead

🚰 Data Engineer