freehire launches on Product Hunt on 26 August.

Follow →

Senior Software Engineer - Cloud

Open 27d

You will design, develop, and maintain core cloud platform services including compute orchestration, resource management, multi-tenancy, and API gateway components. You'll build and optimize RESTful/gRPC APIs and microservices supporting cloud resource provisioning, lifecycle management, and monitoring. You'll develop scalable, fault-tolerant distributed systems handling high-throughput workloads across multi-region deployments, and implement service mesh, rate limiting, authentication, and authorization mechanisms. You'll write clean, well-tested, well-documented code following best practices like CI/CD, code review, and TDD. On the infrastructure side, you'll design cloud-native architectures using Kubernetes and Docker, build Infrastructure-as-Code solutions, optimize CI/CD pipelines, and implement service discovery, configuration, and secrets management. You'll design data access layers and storage solutions across relational, NoSQL, and caching systems, optimizing queries and implementing migration and backup strategies. You'll implement observability solutions including distributed tracing, structured logging, and metrics collection, build alerting and SLI/SLO frameworks, and participate in on-call rotations and incident response. You'll collaborate across teams to define technical requirements, participate in architecture and code reviews, mentor junior engineers, and evaluate new technologies to improve the platform.

Responsibilities

  • Design, develop, and maintain core cloud platform services including compute orchestration, resource management, multi-tenancy, and API gateway components
  • Build and optimize RESTful/gRPC APIs and microservices for cloud resource provisioning, lifecycle management, and monitoring
  • Develop scalable, fault-tolerant distributed systems handling high-throughput workloads across multi-region deployments
  • Implement service mesh, rate limiting, authentication, and authorization mechanisms
  • Write clean, well-tested, well-documented code following CI/CD, code review, and TDD practices
  • Design and implement cloud-native architectures using Kubernetes, Docker, and container orchestration platforms
  • Develop and maintain Infrastructure-as-Code solutions using Terraform or Pulumi
  • Build and optimize CI/CD pipelines for automated testing, building, and deployment
  • Implement service discovery, configuration management, and secrets management
  • Collaborate with SRE and infrastructure teams to ensure high availability, disaster recovery, and capacity planning
  • Design and implement data access layers and storage solutions including relational, NoSQL, and caching systems
  • Optimize database queries, indexing strategies, and connection pooling
  • Implement data migration, backup, and recovery strategies
  • Implement observability solutions including distributed tracing, structured logging, and metrics collection
  • Design and build alerting mechanisms and SLI/SLO frameworks
  • Conduct performance profiling, load testing, and capacity planning
  • Participate in on-call rotations and incident response, driving root cause analysis and post-mortem improvements
  • Collaborate with AI platform, product, and infrastructure teams to define technical requirements
  • Participate in architecture design reviews, code reviews, and technical discussions
  • Mentor junior engineers and contribute to team knowledge sharing
  • Evaluate and adopt new technologies, frameworks, and tools

Requirements

  • Bachelor's or Master's degree in Computer Science, Software Engineering, Electrical Engineering, or related field
  • Minimum 5 years of professional experience in software engineering, with at least 3 years focused on cloud platform development, distributed systems, or backend services at scale
  • Strong proficiency in one or more backend programming languages: Go, Java, Python, or Rust
  • Deep experience with cloud-native technologies including Kubernetes, Docker, microservices architecture, and service mesh (Istio, Envoy, Linkerd)
  • Solid understanding of distributed systems concepts: consensus algorithms, eventual consistency, sharding, replication, and fault tolerance
  • Experience with major cloud platforms (AWS, GCP, or Azure)
  • Strong knowledge of database technologies (PostgreSQL, MySQL, MongoDB, Redis) and experience with data modeling, query optimization, and storage system design
  • Proficiency in CI/CD tools and practices (GitHub Actions, Jenkins, ArgoCD) and Infrastructure-as-Code (Terraform, Pulumi)
  • Experience with observability and monitoring tools (Prometheus, Grafana, Jaeger, ELK) and SRE practices
  • Knowledge of cloud security best practices, including authentication/authorization (OAuth, JWT, RBAC), encryption, and compliance standards (SOC 2, ISO 27001)
  • Familiarity with AI/ML infrastructure and workload management is a plus (GPU scheduling, model serving, training pipelines)
  • Fluent in English; proficiency in Chinese is a plus

Benefits

  • Attractive welfare benefits and developmental opportunities such as training and mentoring

See also