- Location
- shanghai, CN
- Type
- Full-time
- Department
- Engineering
- Seniority
- Senior
- Experience
- 5+ years
- Source
- Breezy HR
Description
The role is based in the GMT+8 Timezone. We can engage you in Singapore, China or remotely if applicable.
About Us
DeGate is a self-custodial crypto wallet built to make earning on blockchains simple. Users deposit USD Coin (USDC) once and swap directly into assets on any supported chain — over 10,000,000 tokens across 17 blockchains — with no bridging and no need to hold gas tokens on each chain. Alongside swapping, the platform offers Turbo Range, which earns fees from liquidity provision on assets such as Bitcoin (BTC) and Tesla (TSLA); Simple Earn, which puts USDC to work through curated yield vaults; and Stocks, which gives access to 100+ onchain stocks, indices, gold and real-world assets, traded around the clock and settled in USDC. Funds stay under the user's own control, backed by audited contracts and 24/7 customer support.
The easy way to earn on blockchains: https://app.degate.com/
We Want You
We are looking for individuals who are keen to join our group of talented project contributors. You are someone who:
- has an immense interest in blockchain-related technology
- well on your way to being a Crypto-Native
- has an insatiable thirst for growth and learning,
- is keen to build and grow together with the industry
About the Role
Contribute to the design and evolution of our cloud infrastructure and internal developer platform while owning the delivery and continuous improvement of key platform components. This is a hands-on senior individual contributor (IC) role in which you will balance reliability, developer experience, and cost while collaborating with the team to drive platform adoption and cross-team delivery.
We are looking for an engineer with solid fundamentals and production experience who can solve complex problems independently and is ready to take on broader technical responsibility. You do not need to be an expert in every platform domain on day one; we value learning agility, ownership, and consistent delivery. This role offers hands-on opportunities to solve complex platform challenges and grow into end-to-end ownership of a platform domain.
What you'll be doing
Architecture and Platform Engineering
- Contribute to the target architecture, technical direction, and roadmap for cloud infrastructure and the internal developer platform, and own the delivery of key components or milestones.
- Identify architectural bottlenecks and technical debt, propose actionable improvements, and use RFCs, ADRs, and architecture reviews to build alignment and drive implementation.
- Design, build, and operate infrastructure on AWS and Kubernetes, including Amazon EKS.
- Manage modular, multi-account, and multi-environment infrastructure as code with Terraform.
- Build and maintain CI/CD infrastructure.
- Standardize application delivery on a GitOps workflow built with Helm and Argo CD, and make it self-service through templates, golden paths, CLIs, APIs, or portals.
- Build delivery performance measurement into the platform, collecting and analyzing DORA metrics to drive continuous improvement in delivery speed and change safety.
- Build an Observability 2.0 platform centered on wide events, with OLAP infrastructure (e.g., ClickHouse) as its unified data foundation that also serves other analytical workloads, improving system visibility and reliability.
- Help operate and improve shared services such as PostgreSQL, Redis, and Apache Kafka, progressively taking independent ownership of the reliability, capacity planning, and performance of one or more of them.
- Work with the security team to integrate IAM, KMS, Secrets Manager, SOPS, Sealed Secrets, and related controls into secure-by-default platform workflows.
- Participate in an on-call rotation and major incident response, perform root cause analysis, and drive follow-up improvements.
Technical Impact and Collaboration
- Work with engineering, security, and business teams to deliver cross-team projects with clear objectives and scope, and own the outcomes.
- Contribute to and continuously improve engineering standards for automation, testing, observability, security, and documentation.
- Participate in design and post-incident reviews, share practical knowledge, and help the team improve system design and troubleshooting.
What you'll need
- 5+ years of experience in platform engineering, cloud infrastructure, SRE, DevOps, or a related software engineering role, with hands-on experience building and operating production systems and independently delivering a key component or service.
- Experience contributing to architectural design, system migrations, or cross-team delivery, with the ability to own a solution from design through implementation once the objectives and scope are clear.
- Strong software engineering skills, with the ability to design, build, and operate production platform services, APIs, CLIs, plugins, or internal developer tools.
- Strong English-language technical research skills, with a habit of working from primary sources — official documentation, RFCs and design docs, research papers and technical reports, GitHub issues and pull requests, upstream open-source communities, international conferences such as KubeCon and AWS re:Invent, and engineering blogs from leading technology teams — to stay current with the state of the art.
- Able to independently research technologies that are not yet widely adopted locally and have little secondary coverage: analyze source code, build prototypes, make informed technology choices, and clearly explain fit and boundaries, adoption costs, risks, and alternatives.
- A product mindset for internal developer platforms and developer tools, with experience improving developer experience through user feedback, pilots, documentation, or metrics.
- Ability to write clear RFCs, ADRs, architecture diagrams, and technical documentation, communicate decisions, tradeoffs, and risks, and contribute effectively to cross-team technical discussions.
- Strong foundations in Linux, DNS, TLS, HTTP/TCP, and load balancing, as well as identity and access management, secrets management, and infrastructure security.
- Experience operating production workloads on AWS, with hands-on experience in several areas such as EKS, IAM, VPC, load balancing, object storage, and managed databases, plus the ability to learn unfamiliar areas quickly.
- Production Kubernetes experience, with an understanding of cluster lifecycle, networking, storage, and RBAC, plus hands-on experience with one or more of the following: CRDs, controllers/operators, or add-ons.
- Production operations or performance-tuning experience with at least one of PostgreSQL, Redis, or Apache Kafka, and a willingness to expand into other shared services.
- Strong Terraform skills, including reusable module design and experience managing multi-account or multi-environment infrastructure.
- Experience with Helm, GitOps, and CI/CD, including hands-on use of Argo CD, Flux, or a similar GitOps tool.
- Hands-on experience with one or more observability tools such as Prometheus, Grafana, Loki, or Tempo.
Bonus Points
- Experience with multi-region or high-availability architectures, backup and recovery, disaster recovery, or large-scale distributed systems.
- Experience with financial or trading systems, or other high-reliability environments.
- Experience with data modeling, schema design, and query optimization on ClickHouse or other OLAP databases, or building observability on wide events.
- Experience researching, piloting, or taking an AI developer platform from concept to production to improve software development, delivery, and operations; this scope does not include model training, inference platforms, or GPU infrastructure.
- Familiarity with LLM APIs, MCP, tool use, agent workflows, RAG, and evaluation, including experience with coding agents, MCP servers, knowledge connectors, or AI developer toolchains.
- Understanding of agent identity and authorization, sandboxing, auditing, human approval, and rollback, along with experience evaluating AI workflows for quality, latency, cost, and human-intervention requirements.
- Experience with DevSecOps, cloud security, or infrastructure security, including IAM, least-privilege access, KMS, secrets management, software supply chain security, or vulnerability management.
- Experience with cloud cost governance, capacity management, or FinOps.
- Open-source contributions, technical writing, or public speaking.
- Experience building software or automation tooling in Rust.
- Experience administering GitHub Enterprise at the organization level, including permissions, rulesets, SSO/SCIM, and OIDC federation.
How We'll Measure Success
- Understand the platform architecture and phased roadmap, and independently deliver assigned key components and milestones.
- Improve platform adoption, availability, and self-service task success rates while reducing the change failure rate.
- Reduce lead times for infrastructure provisioning, environment setup, application deployment, and incident recovery.
- Reduce manual operations and support tickets while improving DORA metrics and infrastructure cost efficiency.
- Demonstrate sound technical judgment, clear communication, and consistent execution while growing into end-to-end ownership of a platform domain.