Senior Platform Engineer (Remote US)
Summary
Senior Platform Engineer provides US-hours production support for a C#/.NET-based Product Catalog running on AWS, Apigee, Rancher/Kubernetes, diagnosing latency, API failures, and stability issues.
Senior Platform Engineer (Remote US)
We are seeking a Senior Platform Engineer to provide US-hours production support and platform engineering for the Product Catalog application. This role focuses on diagnosing and resolving connection, latency, and stability issues across the AWS-based Product Catalog stack, including Apigee, Rancher/Kubernetes, and the C# services.
The engineer will work closely with existing offshore resources and key stakeholders to ensure application reliability and performance during US business hours.
Requirements
AWS Cloud
- Hands-on experience troubleshooting production workloads in AWS
- Familiarity with core services backing an enterprise app (e.g., compute, database, networking, load balancing, monitoring)
- Experience with Apigee / Apigee X (or similar API gateway) in a production environment
- Ability to debug API failures, timeouts, policy issues, and routing problems
- Strong working knowledge of Kubernetes concepts (pods, deployments, services, scaling)
- Experience using Rancher (or comparable tooling) for cluster and workload management
- Ability to diagnose pod instability, restarts, resource constraints, and related performance issues
- Proficiency in C# and .NET, especially for API and service‑oriented architectures
- Strong debugging and troubleshooting skills in distributed systems
- Experience with rules engines or rule-based business logic
- Strong SQL skills for performance analysis and data troubleshooting
- Willingness and aptitude to learn Product Catalog's rules model and support rules work similar to our Senior Engineers over time
- Proven track record providing production support for complex, multi‑tier applications
- Strong analytical and problem‑solving skills; comfortable owning issues end‑to‑end
- Good communication skills for working with both technical teams and business stakeholders during active incidents with the team.
Responsibilities
- Provide US-hours production support for Product Catalog, focusing on connection, latency, and general performance issues.
- Triage and troubleshoot issues across the AWS, Apigee, Rancher/Kubernetes, C# services, and database layers.
- Investigate and resolve problems with API calls, including timeouts, missing logs, and routing/connection failures.
- Collaborate with offshore teams who implemented the AWS migration to stabilize and optimize the current environment.
- Work with Product Catalog stakeholders to identify root causes and drive sustainable fixes, not just workarounds.
- Support and optimize the C# codebase that powers Product Catalog and its API integrations.
- Develop proficiency in Product Catalog's rules engine and SQL-based rules storage, gradually taking on non‑production rules work to offload Sr Support Engineers.
- Participate in knowledge sharing and documentation to reduce single‑point‑of‑failure risk within the Product Catalog team.
- Cover US Business Hours: 9:00 AM – 6:00 PM EST
Our Culture
- Open communication and a focus on psychological safety;
- Recognition of achievements and contributions;
- Collaboration across multicultural teams;
- Less bureaucracy, more focus on results.
- Intro call
- Technical Interview
- Manager Interview
- Client Interview
- Pre-offer stage + Reference Check (if requested)
- Official Offer