Full-Stack Engineer
1. Application Observability Setup under AWS CloudWatch based on client requirements
- Design, configure and maintain AWS CloudWatch Synthetics Canaries and related monitoring capabilities for client applications, in accordance with application observability and monitoring requirements.
- Write and refine canary test scripts per application to accurately validate functionality/availability; iterate on scripts in UAT before promoting to PRD, and refine again in PRD to account for environment-specific differences (network paths, auth, endpoint configs).
- Implement application health checks, availability monitoring, performance monitoring, dashboards, metrics, logs and alerts to provide proactive visibility of application health.
- Enable CloudWatch dashboard/metrics visibility in observability PRD account so canary results are viewable by stakeholders.
- Troubleshoot and resolve any canary failures or false positives during rollout.
- Work with application teams to troubleshoot performance and availability issues identified through monitoring and observability tools.
- Support continuous improvement of application monitoring, alerting and operational dashboards.
- Plan and execute the migration of the existing managed CI/CD runner to the managed runner for applications.
- Assess existing build and deployment configurations, dependencies, environment variables, credentials and integration requirements prior to migration.
- Perform configuration, testing, troubleshooting and validation to ensure applications can be successfully built and deployed through the managed environment.
- Work with relevant platform and application teams to resolve deployment pipeline and runner-related issues.
- Design, develop, configure and maintain backend integration components and services required by client applications.
- Set up and maintain Model Context Protocol (MCP) servers and associated APIs/services to enable AI-enabled applications to securely access authorized knowledge sources.
- Develop integrations with client systems and knowledge repositories.
- Implement appropriate authentication, authorization, API management, logging, error handling and data transformation mechanisms for system-to-system integrations.
- Work with application and data owners to understand source-system interfaces and develop reusable integration services where appropriate.
- Work closely with product and application teams to establish, configure and manage CI/CD deployment pipelines for client applications.
- Support the onboarding and migration of existing applications onto deployment processes.
- Review existing application deployment architecture and progressively migrate manually configured infrastructure and deployment components towards Infrastructure-as-Code (IaC).
- Develop and maintain IaC templates/scripts for consistent and repeatable provisioning and deployment across development, testing and production environments.
- Troubleshoot CI/CD pipeline, deployment, environment and infrastructure configuration issues.
- Promote reusable deployment patterns and automation to reduce manual deployment effort and configuration inconsistencies.
The candidate should have hands-on knowledge or experience in several of the following areas:
- Cloud infrastructure, preferably AWS and/or GCC
- GitHub, Git or equivalent source code management platforms
- CI/CD pipelines and deployment automation
- Enterprise DevSecOps platforms
- CI/CD runners and build/deployment agents
- Application observability and APM
- Logging, metrics, tracing, dashboards, and alerting
- Observability platforms
- Docker and containerized applications
- Kubernetes or managed container platforms
- Infrastructure-as-Code and configuration automation
- Linux and/or Windows server administration
- SFTP and managed file-transfer solutions
- REST APIs and system integration
- Networking fundamentals including DNS, routing, firewall rules, proxies and load balancers
- TLS/SSL certificates, secrets, and identity/access management
- Scripting using PowerShell, Bash, Python, or equivalent
- Application incident, problem, and change management
- Diploma/Degree in Computer Science, Information Technology, Computer Engineering, or a related discipline preferred.
- Preferably 3–5 years of relevant experience in application infrastructure, DevOps, cloud engineering, platform engineering, or application operations.
- Hands-on experience supporting production applications and troubleshooting across application and infrastructure layers.
- Experience with CI/CD, source code management, and deployment automation.
- Experience with cloud-hosted applications and modern application architectures.
- Good understanding of application observability, performance monitoring, and operational support.
- Experience working in Singapore Government ICT/GCC environments would be advantageous.
- Experience with enterprise DevSecOps, observability or other Singapore Government Tech Stack platforms would be advantageous.
- Strong troubleshooting and analytical skills with the ability to diagnose issues across application, infrastructure, and integration layers.
- Able to work effectively across application developers, infrastructure teams, cybersecurity teams, vendors, and WOG service providers.
- Strong operational mindset with emphasis on reliability, automation, standardization, and maintainability.
- Able to translate application requirements into practical infrastructure and deployment solutions.
- Comfortable managing multiple applications and technical stakeholders.
- Good documentation and communication skills.
- Proactive in identifying operational risks and opportunities for engineering improvement.