Senior Platform Engineer
rew
📍 Remote🌐 Remote🕐 27d ago🔗 himalayas
Job Description
We are seeking a **Senior Platform Engineer** to provide **US-hours production support and platform engineering** for the Product Catalog application. This role focuses on diagnosing and resolving connection, latency, and stability issues across the AWS-based Product Catalog stack, including Apigee, Rancher/Kubernetes, and the C# services.
The engineer will work closely with existing offshore resources and key stakeholders to ensure application reliability and performance during US business hours.
**Requirements**
**AWS Cloud**
* Hands-on experience troubleshooting production workloads in AWS
* Familiarity with core services backing an enterprise app (e.g., compute, database, networking, load balancing, monitoring)
**Apigee / API Gateway**
* Experience with Apigee / Apigee X (or similar API gateway) in a production environment
* Ability to debug API failures, timeouts, policy issues, and routing problems
**Rancher / Kubernetes**
* Strong working knowledge of Kubernetes concepts (pods, deployments, services, scaling)
* Experience using Rancher (or comparable tooling) for cluster and workload management
* Ability to diagnose pod instability, restarts, resource constraints, and related performance issues
**C# / Application Services**
* Proficiency in C# and .NET, especially for API and service‑oriented architectures
* Strong debugging and troubleshooting skills in distributed systems
**Rules / SQL (Nice-to-Have but Important)**
* Experience with rules engines or rule-based business logic
* Strong SQL skills for performance analysis and data troubleshooting
* Willingness and aptitude to learn Product Catalog's rules model and support rules work similar to our Senior Engineers over time
**General**
* Proven track record providing production support for complex, multi‑tier applications
* Strong analytical and problem‑solving skills; comfortable owning issues end‑to‑end
* Good communication skills for working with both technical teams and business stakeholders during active incidents with the team.
**Responsibilities**
* Provide **US-hours production support** for Product Catalog, focusing on connection, latency, and general performance issues.
* Triage and troubleshoot issues across the **AWS, Apigee, Rancher/Kubernetes, C# services, and database** layers.
* Investigate and resolve problems with **API calls**, including timeouts, missing logs, and routing/connection failures.
* Collaborate with offshore teams who implemented the AWS migration to **stabilize and optimize** the current environment.
* Work with Product Catalog stakeholders to **identify root causes** and drive sustainable fixes, not just workarounds.
* Support and optimize the **C# codebase** that powers Product Catalog and its API integrations.
* Develop proficiency in Product Catalog's **rules engine and SQL-based rules storage**, gradually taking on non‑production rules work to offload Sr Support Engineers.
* Participate in**knowledge sharing and documentation** to reduce single‑point‑of‑failure risk within the Product Catalog team.
**Working conditions**
* Cover US Business Hours: 9:00 AM – 6:00 PM EST
**Why work with us?**
**Our Culture
* Open communication and a focus on psychological safety;
* Recognition of achievements and contributions;
* Collaboration across multicultural teams;
* Less bureaucracy, more focus on results.
**Hiring Process**
* Intro call
* Technical Interview
* Manager Interview
* Client Interview
* Pre-offer stage + Reference Check (if requested)
* Official Offer
**
Originally posted on [Himalayas](https://himalayas.app)