Site Reliability Engineer / Full Stack Engineer
Lockheed Martin
- Location
- Dartmouth
- Work model
- On-Site
- Level
- Mid
- Posted
- 4d ago
Skills
About this role
Standard Job Description This position will be a core member of our digital delivery team. It will cover a wide range of activities for the design, implementation, and support of systems used across the Lockheed Martin Canada Inc. (Lockheed Martin) user community. It will fulfil a key role in enabling cloud services for Lockheed Martin’s International community, migrating existing systems to the cloud and supporting the creating of new cloud-based services. The candidate will be proficient in infrastructure, security, software development, database management systems and systems integration. They will have experience using DevOPs best practices using agile process and common technology stacks. The ideal candidate will be an individual contributor capable of working with multiple stakeholder groups in support of international Information Technology (IT) services and the broader Lockheed Martin organization. They will use their technical skills to contribute to the completion of milestones associated with specific projects.
Key Responsibilities
Troubleshoot production issues and coordinate with the development team to streamline code deployment Development and Operations of Amazon Web Services (AWS) Cloud based services that include Tenant, Subscription, vNET, Virtual Machines/EC2, Storage Accounts, KeyVaults, Role-Based Access Control (RBAC), Policy and basic Infrastructure as a Service (IaaS) use cases. This also includes Landing Zone and Active Directory deployments and configurations System Maintenance: Manage, monitor, and maintain Linux (RHEL) and Windows Server environments across both physical hardware and virtualized infrastructure (VMware). Network & Security: Oversee network infrastructure, including firewalls, switches, and routers. Implement security patches, manage access controls (Active Directory/FreeIPA), and ensure robust backup and disaster recovery protocols. Automation & Scripting: Reduce manual overhead by writing and maintaining automation scripts using Ansible, Terraform, PowerShell, or Python. Support & Troubleshooting: Diagnose and resolve complex hardware, software, and network anomalies. Documentation: Maintain clear, up-to-date documentation for system configurations, operational procedures, and network topology. Conduct systems tests for security, performance, and availability Familiarity with industry cloud systems, processes, and toolsets that include GitHub/GitLab and Visual Studio Code Advance foundational processes to drive process and cost efficiencies within Cloud platforms and service offerings Implement automation tools and frameworks (Continuous Integration/Continuous Delivery (CI/CD) pipelines) Conduct systems tests for security, performance, and availability Develop and maintain design and troubleshooting documentation Analyze code and communicate detailed reviews to development teams to ensure a marked improvement in applications and the timely completion of projects Translate functional needs and desired capabilities into value-add technical solutions underpinned by short and long-term technical roadmaps that are aligned with Lockheed Martin strategy and initiatives. Basic Qualifications Experience with designing/developing/integrating systems with IT infrastructure technologies (networking, telecommunications, computing, storage, virtualization including containers) Hands-on experience with AWS such as Virtual Private Cloud (VPC), Elastic Compute Cloud (EC2), Containers, Simple Storage Service (S3), Elastic Load Balancing (ELB), Relational Database Service (RDS), Route53, Cloud Formation, Cloud Watch, Identity and Access Management (IAM), Code Commit, Lambda, Cloud Trail, Application Programming Interface (API) Gateways and others Experience with Managing Cloud Networks (Aviatrix) Experience of infrastructure configuration using Ansible Knowledge of Infrastructure as code principles Experience of virtual infrastructure (VMware) Experience of coding with