Software Engineer – Platform
per year (also listed as $81.73–$144.23/hr)
Job4Pak tracks remote platform-engineering roles like this one because demand for senior infrastructure talent in the USA keeps climbing as more companies shift to AI-first, cloud-native engineering. This particular opening sits at the center of that shift — building the internal tooling and infrastructure that lets product teams ship faster and safer.
About the Role
This is a Software Engineer opportunity focused on platform infrastructure, internal developer tooling, shared libraries, automation, and reliability. The role exists to help product engineers ship features quickly while maintaining strong performance, uptime, security, and scalability — building the systems, tools, and “paved paths” that make development and deployment safer and more reliable.
Platform Infrastructure & Developer Tooling
- Own and evolve AWS infrastructure using Terraform and Pulumi Cloud
- Treat infrastructure as a product engineering teams depend on
- Design and maintain internal developer tooling, shared libraries, SDKs, code generation, data access patterns, and service scaffolding
- Create golden paths for common workflows such as new service setup, background jobs, event streams, and APIs
- Build platform defaults that make security, observability, and consistency easier to adopt
- Reduce engineering toil through automation and self-service tooling
CI/CD, Environments & Delivery
- Design and build CI/CD systems that let engineers deploy dozens of times a day with confidence
- Maintain Buildkite pipelines and TypeScript pipeline-as-code workflows
- Build and support per-PR ephemeral environments
- Improve local development tooling, templates, and CI-integrated workflows
Reliability, Observability & Security
- Drive reliability through SLOs, autoscaling, incident response, and postmortems
- Build observability tooling and shared instrumentation libraries
- Improve alerting so signals are useful rather than noisy
- Enforce security best practices across IAM, secrets management, encryption, and audit logging
Data Platform, Scaling & AI-First Engineering
- Own reliability and performance of Aurora PostgreSQL — provisioning, backups, failover, tuning
- Solve scaling problems across Aurora PostgreSQL, Kafka throughput, and cost-efficient compute autoscaling
- Build tooling and best practices for AI-first software engineering, including support for autonomous code-change agents
- Contribute to architecture decisions, documentation, and shared engineering standards
What You’ll Bring
- Experience building internal platforms or developer tooling — code generation, CLIs, templates, shared SDKs, frameworks
- Strong TypeScript skills and strong API design judgment
- Deep AWS experience: ECS, Lambda, VPC, ALB, IAM, RDS, ElastiCache, MSK, OpenSearch, S3, CloudWatch, CloudTrail, GuardDuty
- Strong Terraform and/or Pulumi experience, including modules, workspaces, CI-driven plan/apply workflows
- Experience with production observability stacks — Datadog, CloudWatch, Sentry, distributed tracing, SLOs
- Experience operating Aurora PostgreSQL at scale, including read replicas and query tuning
- Reliability engineering mindset: SLOs, error budgets, incident response
- Strong written communication skills
Tech Environment
- AWS infrastructure managed with Terraform and Pulumi Cloud
- Docker on ECS (EC2/Fargate)
- Aurora PostgreSQL, ElastiCache Redis, MSK Kafka, OpenSearch
- Buildkite CI/CD with TypeScript pipeline-as-code
- TypeScript monorepo — Node/Express, React, GraphQL/Apollo
- Datadog, CloudWatch, Sentry, Cloudflare, GitHub
Benefits
- 100% of employee and family medical premiums covered
- Vision and dental coverage
- 401(k), HSA and FSA options
- Employee loan program through a lender partner
- Remote-first culture with flexible time off