Joining Amex Tech means discovering and shaping your contribution to something big. Here, you can work alongside talented tech teams and build a unique career with the Powerful Backing of American Express. With a range of opportunities to work with the latest technologies, and a commitment to back the broader engineering community through open source, our mission is to power your success. Because Amex Tech is powered by our technology, our culture, and our colleagues. The Technology organization enables and accelerates the company’s growth strategies, delivering global capabilities and services in support of Amex’s customers and colleagues, while maintaining 24/7 servicing and availability to ensure an uninterrupted, high-quality customer experience. Technology provides the foundation for everything we do in the company while driving differentiation through building and leveraging innovative technology and data insights. How will you make an impact in this role? As a Software Engineer specializing in Distributed Platforms, you will be responsible for ensuring the reliability, scalability, performance, and resiliency of enterprise applications running in distributed and cloud environments. You will leverage your expertise in software engineering, production support, automation, and cloud technologies to resolve complex technical issues and enhance platform stability. You will collaborate with engineering, product, infrastructure, and operations teams to build, automate, monitor, and support highly available distributed systems while driving modernization and operational excellence Software Engineering & Platform Development Provide hands-on support for complex enterprise applications and distributed platforms. Possess 5+ years of experience in Distributed Application Support and Platform Engineering. Design, develop, test, debug, and deploy scalable software solutions using Java/Python, hands-on coding & debugging, SQL/DB basics, API and Cloud exposure. Work with SQL databases, REST APIs, and cloud technologies. Participate in change management, root cause analysis (RCA), and troubleshooting of production issues. Build automation solutions to improve operational efficiency, platform resiliency, and system reliability. Follow ITIL best practices for Incident, Problem, and Change Management. Exposure to APM tool experience, SRE/Production Support Experience, ITIL, Exposure to AIOps/ML, Dashboard/Alerts creation through observability tools Runtime Engineering, Reliability & Operations Ensure platform availability, scalability, reliability, and performance. Monitor and improve operational metrics such as SLOs, SLAs, MTTR, and MTBF. Diagnose and resolve production incidents, application outages, performance issues, and infrastructure-related problems. Apply Site Reliability Engineering (SRE) principles to improve operational excellence. Support disaster recovery, high availability, capacity planning, and business continuity initiatives Distributed Platform Engineering Strong experience supporting Linux/Unix-based distributed applications. Experience with Java-based microservices and REST APIs. Knowledge of cloud platforms such as AWS or GCP or Azure Familiarity with containers and distributed application architectures. Understanding of hybrid cloud environments and enterprise integration. Data, Integration & Automation Experience with REST APIs, MQ, Connect:Direct, and event-driven architectures. Knowledge of relational and NoSQL databases such as PostgreSQL, Redis, Couchbase, and DB2. Strong scripting and automation skills using Python, Bash, Ansible, and Jenkins. Experience building deployment, monitoring, and operational automation. Observability & Monitoring Experience with monitoring and observability tools such as Splunk, Dynatrace, AppDynamics, ELK, Grafana, or Prometheus. Ability to analyze system metrics, logs, and application performance to identify optimization opportunities. Build dashboards, alerts, and monitoring solutions to improve platform health. DevOps & Security Experience with Git, Jenkins, Maven, and CI/CD pipelines. Understanding of enterprise security practices, access management, and compliance requirements. Ensure applications meet non-functional requirements including availability, scalability, performance, security, and maintainability. Depending on factors such as business unit requirements, the nature of the position, cost and applicable laws, American Express may provide visa sponsorship for certain positions.
Senior Operations Engineering Engineer
New York ISO
Operations Engineering Engineer
New York ISO
Engineer I - Systems Operations Engineering (SOE)
Dukeenergy
Legal and Operations Engineering Lead
Perplexity
SIEM Engineering & Operations Lead - Vice President
Db
Director, Model Engineering & Operations
Caresource