- Type
- Full-time
- Department
- Engineering
- Experience
- 4+ years
- Closing date
- Today
- Source
- Vincere
Description
Key Responsibilities
-
Assist in implementing a robust, scalable and high‑performance data storage and management platform for genomic data.
-
Support the design and documentation of system architecture, data models, data flow diagrams and system interfaces.
-
Help identify and resolve performance bottlenecks across data storage, ETL and retrieval pipelines.
-
Contribute to the design and implementation of scalability plans to accommodate rapid growth of genomic data.
-
Prepare technical documentation and participate in testing activities across the SDLC (unit test, SIT, UAT, load and regression tests) to ensure functional and operational requirements are met.
-
Support routine operational activities to maintain system stability, performance and reliability.
-
Perform other duties as assigned by senior officers.
Requirements
-
Bachelor’s degree in Computer Science or related discipline, with at least 4 years of relevant working experience.
-
Solid understanding of data warehouse concepts, data modelling and ETL processes.
-
Hands‑on experience with storage technologies such as object storage (e.g. S3), parallel file systems, and databases (e.g. ClickHouse, MySQL, PostgreSQL, AWS RDS).
-
Experience with cloud‑based data technologies, including data security and access control mechanisms.
-
Familiarity with Linux system administration, shell scripting, system monitoring and performance tuning.
-
Knowledge of containerisation using Docker and deployment on Kubernetes.
-
Knowledge of CI/CD pipelines and related DevOps tools and practices.