DevOps Lead – Architect

Job not on LinkedIn

🔥 5 minutes ago

🦀 Maryland – Remote

infoinfo

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

infoinfo

👻 Ghost score 10%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Dynanet Corporation

Dynanet Corporation

51 - 200 employees

💼 Consulting

🏥 Healthcare

📦 Logistics

Consulting • Healthcare • Logistics

Dynanet Corporation is a company providing technology solutions including application services, cloud enablement, cybersecurity, DevSecOps, legacy system modernization, and robotic process automation. It employs a 4D methodology for rapid development that aligns IT investments to mission outcomes to improve business performance. Dynanet focuses on IT modernization and human-centered design to address core business needs and meet organizational objectives. The company serves diverse clients, including federal agencies, by customizing systems to meet specific mandates and modernizing aging systems to perform in modern environments. Dynanet emphasizes a strong process methodology and the power of people and connections in delivering successful IT solutions.

📋 Description

• Serve as designated Key Personnel for an upcoming federal health agency program providing enterprise DevSecOps platform services • Own the technical architecture and security posture of managed DevSecOps platforms and supporting AWS infrastructure across development, test, implementation, and production accounts • Establish uniform operating patterns for provisioning, identity, backup and restore, observability, and security • Define and maintain Terraform Infrastructure as Code standards, versioned module catalogs, drift detection, and remediation • Implement policy-as-code preventive controls and GitOps workflows for Kubernetes-based services • Maintain continuous Authority to Operate and security documentation in the client’s compliance repository • Implement compliance-as-code and documentation-as-code practices for continuous evidence, scan results, baselines, and SBOM generation • Support security control assessments, security impact analyses, penetration testing, and maintenance of required security plans • Operate artifact repository and binary analysis platforms, including segregation, retention, quota, scanning, and remediation workflows • Define observability architecture for application performance monitoring, centralized logging, OpenTelemetry, metrics, and dashboards • Establish Site Reliability Engineering practices, including SLI/SLO definition, error budgets, toil reduction, and capacity planning • Design and maintain disaster recovery playbooks, failover/failback procedures, recovery drills, and RTO/RPO validation • Implement Zero Trust practices including network segmentation, identity-aware access, least-privilege roles, and secrets rotation • Lead FinOps and cost optimization, including tagging, rightsizing, discounted pricing coverage, scheduling, anomaly detection, and cost reporting • Maintain architecture documentation as code, including ADRs, C4 diagrams, and API contracts • Serve as technical interface to the independent quality assurance contractor and client security officers • Participate in a 24x7x365 on-call rotation and lead quarterly resilience and security exercises

🎯 Requirements

• Deep AWS architecture experience in a multi-account environment, including VPC and network design, IAM boundary design, and account segmentation across environments • Expert-level Terraform experience, including modular design, state governance, remote backends, and drift detection at an organizational scale • Strong Kubernetes architecture experience, preferably AWS EKS, including workload design, Helm, and container security practice • Demonstrated ownership of security compliance for a federal system boundary, including maintaining or supporting an Authority to Operate and the associated documentation set • Hands-on experience with artifact management and binary analysis platforms, preferably JFrog Artifactory and Xray • Practical observability architecture experience across metrics, logs, and traces, including OpenTelemetry and at least one enterprise application performance monitoring platform • Demonstrated Site Reliability Engineering practice, including defining service level indicators and objectives and operating against error budgets • Experience designing and testing disaster recovery and backup strategies with validated recovery time and recovery point objectives • Experience with cloud cost optimization practice, including tagging governance, rightsizing, and utilization-driven decision making • Bachelor's degree in Computer Science, Engineering, Information Systems, or a related technical discipline • Additional directly relevant experience may be considered in lieu of a degree • Minimum of ten (10) years of infrastructure, DevOps, or cloud engineering experience, including at least four (4) years in an architecture or technical lead capacity • Minimum of three (3) years supporting federal systems subject to FISMA and Authority to Operate requirements • Candidates must be able to provide a resume suitable for submission as designated Key Personnel • Strong written and verbal communication skills • Highly organized with the ability to prioritize, balance, and effectively advance multiple competing priorities in a high-volume, fast-paced environment • Ability to interact in a professional and collaborative manner with fellow Dynanet Teammates and the clients, and business partners • Ability and desire to challenge and educate yourself to support and advance IT services delivery in the Federal agencies served • Excellent judgment and creative problem-solving skills • Respond to team member and client requests via email, MS teams, or other communication means during core business hours • Active listening and collaboration skills • Preferred: prior architecture ownership within a federal health environment • Preferred: familiarity with agency compliance documentation repositories, automated inventory tooling, and federal security control baselines • Preferred: experience with compliance automation frameworks such as InSpec, OSCAL-based control authoring, or continuous controls monitoring • Preferred: familiarity with NIST SP 800-218, NIST SP 800-53, and Zero Trust reference architecture under OMB M-22-09 • Preferred: experience with policy-as-code tooling including Open Policy Agent and Conftest • Preferred: experience operating Grafana, Prometheus, SonarQube, TestRail, or Apache JMeter • Preferred: relevant certifications such as AWS Certified Solutions Architect - Professional, Certified Kubernetes Administrator, CISSP, or HashiCorp Certified: Terraform Associate

🏖️ Benefits

• Industry Competitive Compensation • Medical and Dental Insurance • Paid Time Off/Holidays • 401(k) Retirement Plans with Matching • Remote Work • Paid Training • Employee Referral Program • Employee Development Program

Apply Now

Similar Jobs

🔥 9 hours ago

NeuroFlow

51 - 200

🏥 Healthcare

☁️ SaaS

🤝 B2B

Senior Site Reliability Engineer owning reliable, compliant infrastructure for NeuroFlow’s behavioral health platform. Leading SLOs, incident response, cloud operations, and secure delivery across AWS and Azure.

🔥 11 hours ago

RTX

10,000+ employees

🏭 Manufacturing

💼 Consulting

📦 Logistics

Reliability Engineer improving equipment reliability and maintenance strategies across RTX aerospace manufacturing sites. Applying RCA, FMEA, CMMS/EAM analytics and predictive-maintenance practices.

🔥 12 hours ago

Ad Hoc LLC

501 - 1000

💼 Consulting

🏥 Healthcare

📦 Logistics

DevOps Engineer III building AWS and Kubernetes infrastructure for Ad Hoc’s government digital services. Improving CI/CD, security, reliability, and developer experience for Veterans Affairs products.

🔥 13 hours ago

Circle

501 - 1000

💳 Fintech

₿ Crypto

🌐 Web 3

Senior SRE operating secure Kubernetes and Terraform infrastructure for Circle’s digital-asset financial platform. Improving reliability, observability, automation, and production resilience.

🔥 17 hours ago

RELX

10,000+ employees

💼 Consulting

🏥 Healthcare

🛡️ Insurance

SRE Engineering Lead managing teams and reliability initiatives for LexisNexis Risk Solutions' cloud-based risk platforms. Driving Kubernetes, Terraform, Azure, automation, incident response, and operational excellence.