Top Skills • Amazon ECS (Elastic Container Service) • Target Groups • Task Definitions • Infrastructure as Code (IaC) • Setting up a Git repository in the cloud using Terraform • CI/CD Pipelines • Connecting a new application to existing pipelines aws and datadog design and implementation experience Description We are looking for a Site Reliability Engineer to work as part of a lean, product focused engineering organization. This role is about building and operating reliable cloud based systems by writing code, automating infrastructure and delivery workflows, and reducing friction for developers and users. You will work closely with product and application engineers to design, deploy, and operate systems with clear ownership and practical engineering judgment. We expect you to use modern tooling, including AI assisted tools where appropriate, to speed up automation, troubleshooting, and operations while remaining accountable for correctness, security, and reliability. This role favors simple, effective solutions, hands on ownership, and continuous improvement within small Agile teams. What you will do: • Design, build, and operate cloud infrastructure for critical production and non production applications with reliability and simplicity as primary goals • Architect and evolve multi-account AWS foundations (organizations, accounts, IAM boundaries, guardrails, and environment separation) to enable secure, scalable delivery • Design and operate cloud networking architecture (VPCs, routing, segmentation, ingress/egress, connectivity patterns) to support reliability, security, and compliance requirements • Treat reliability, security, and compliance as first class design concerns throughout the system lifecycle • Build tooling and automation that reduces errors, shortens recovery time, and improves day to day operations • Implement monitoring, logging, and alerting that make system behavior observable and actionable • Use AI assisted tools to accelerate infrastructure delivery, automation, troubleshooting, and root cause analysis, applying engineering judgment to validate outcomes • Implement reliability guardrails for releases (progressive delivery, safe rollbacks, change risk controls) and provide production support during deployments. • Participate in incident response, perform root cause analysis, and drive durable improvements that prevent recurrence • Work closely with application engineers to co own system design, operation, and continuous improvement • Maintain clear, lightweight documentation that supports shared ownership and effective on call operations What we need from you: • BA/BS, in a related technical field; or the equivalent in education and work experience • 8+ years of experience in DevOps, SRE, platform engineering, or similar roles supporting application teams running production services • Strong CI/CD experience (Jenkins and Git-based workflows preferred), including building secure, reliable pipelines and enabling teams to ship safely • Experience implementing and operating observability platforms (logging/metrics/alerting); Elastic Stack/OpenSearch experience is a plus • Hands-on, demonstrable experience designing and operating AWS environments, and enabling application teams to adopt AWS correctly (networking, IAM, security, reliability, and cost awareness) • Infrastructure as Code experience (Terraform preferred; CloudFormation acceptable), including building reusable modules/patterns and managing changes through review and automation • Experience supporting CI/CD builds and deployment patterns for common application stacks (for example Java, NodeJS, and .NET) • Experience scripting in Bash, Python, or PowerShell • Experience working on large scale cloud-based web applications What we would like from you: • Ability to clearly communicate both verbally and in writing with client and team members, including experience documenting and presenting findings • Excellent analytical skills, organizational abilities, and problem-solving skills • Familiarity with AI agentic development tools (e.g., Claude Code, GitHub Copilot, Windsurf) and practical experience applying them to infrastructure and operations workflows • Self-starter who works efficiently in a fast-paced environment with changing priorities and a geographically distributed team • Ability to think creatively and seek optimum solutions • Ability to grasp loosely defined concepts and transform them into tangible results and key deliverables • Diagnostic skills with the ability to analyze technical, business and financial issues and options • Ability to infer from previous examples, willingness to understand how an application is put together • Action-oriented, with the ability to quickly deal with change • Someone who will embody our values of courage, integrity, collaboration, inclusion, connection and fun. Additional Skills & Qualifications What we would like from you: • Ability to clearly communicate both verbally and in writing with client and team members, including experience documenting and presenting findings • Excellent analytical skills, organizational abilities, and problem-solving skills • Familiarity with AI agentic development tools (e.g., Claude Code, GitHub Copilot, Windsurf) and practical experience applying them to infrastructure and operations workflows • Self-starter who works efficiently in a fast-paced environment with changing priorities and a geographically distributed team • Ability to think creatively and seek optimum solutions • Ability to grasp loosely defined concepts and transform them into tangible results and key deliverables • Diagnostic skills with the ability to analyze technical, business and financial issues and options • Ability to infer from previous examples, willingness to understand how an application is put together • Action-oriented, with the ability to quickly deal with change • Someone who will embody our values of courage, integrity, collaboration, inclusion, connection and fun. Experience Level Expert Level Job Type & Location This is a Contract position based out of Chicago, IL. Pay and Benefits The pay range for this position is $75.00 - $80.00/hr. Individual compensation offered for this position within this range will depend on many factors, including qualifications, skills, relevant experience, job knowledge, geographic location, internal equity, and other pertinent job-related factors. Eligibility requirements apply to some benefits and may depend on your job classification and length of employment. Benefits are subject to change and may be subject to specific elections, plan, or program terms. If eligible, the benefits available for this temporary role may include the following: • Medical, dental & vision • Critical Illness, Accident, and Hospital • 401(k) Retirement Plan – Pre-tax and Roth post-tax contributions available • Life Insurance (Voluntary Life & AD&D for the employee and dependents) • Short and long-term disability • Health Spending Account (HSA) • Transportation benefits • Employee Assistance Program • Time Off/Leave (PTO, Vacation or Sick Leave) Workplace Type This is a hybrid position in Chicago,IL. Application Deadline This position is anticipated to close on Aug 31, 2026. About TEKsystems We're partners in transformation. We help clients activate ideas and solutions to take advantage of a new world of opportunity. We are a team of 80,000 strong, working with over 6,000 clients, including 80% of the Fortune 500, across North America, Europe and Asia. As an industry leader in Full-Stack Technology Services, Talent Services, and real-world application, we work with progressive leaders to drive change. That's the power of true partnership. TEKsystems is an Allegis Group company. The company is an equal opportunity employer and will consider all applications without regards to race, sex, age, color, religion, national origin, veteran status, disability, sexual orientation, gender identity, genetic information or any characteristic protected by law. About TEKsystems and TEKsystems Global Services We’re a leading provider of business and technology services. We accelerate business transformation for our customers. Our expertise in strategy, design, execution and operations unlocks business value through a range of solutions. We’re a team of 80,000 strong, working with over 6,000 customers, including 80% of the Fortune 500 across North America, Europe and Asia, who partner with us for our scale, full-stack capabilities and speed. We’re strategic thinkers, hands-on collaborators, helping customers capitalize on change and master the momentum of technology. We’re building tomorrow by delivering business outcomes and making positive impacts in our global communities. TEKsystems and TEKsystems Global Services are Allegis Group companies. Learn more at TEKsystems.com. The company is an equal opportunity employer and will consider all applications without regard to race, sex, age, color, religion, national origin, veteran status, disability, sexual orientation, gender identity, genetic information or any characteristic protected by law. San Francisco Fair Chance Ordinance: Pursuant to the San Francisco Fair Chance Ordinance, for all positions located in the city and county of San Francisco, we will consider for employment qualified applicants with arrest and conviction records. Massachusetts Lie Detector: It is unlawful in Massachusetts to require or administer a lie detector test as a condition of employment or continued employment. An employer who violates this law shall be subject to criminal penalties and civil liability. Use of Artificial Intelligence (AI): We may use Artificial Intelligence (AI) to support parts of our hiring process, including sourcing, screening, and evaluating candidates. AI helps assess applications and qualifications, but final decisions are made by our hiring team. By applying, you acknowledge and agree that your application may be reviewed using AI tools.
Windows DevOps Cloud Engineer
Baesystems
DevOps Cloud Engineer
Baesystems
Senior DevOps Engineer
Level99 Entertainment
Senior Manager of Data Science Production Engineering, DevOps
Empleos en Natera
Senior DevOps Automation Engineer
VetsEZ
Associate, DevOps Engineer
Blackrock