- Salary
- $93k – $180k
- Location
- (USA) Trust Building AR Bentonville Home Office, United States of America
- Type
- Full-time
- Department
- Engineering
- Seniority
- Senior
- Experience
- 3+ years
- Closing date
- Today
- Source
- Workday
Description
What you'll do...
Position: Senior Data Engineer
Job Location: 811 Excellence Dr, Bentonville, AR 72716
Duties: Data Strategy: understand, articulate, and apply principles of the defined strategy to routine business problems that involve a single function. Data Source Identification: support the understanding of the priority order of requirements and service level agreements. Helps identify the most suitable source for data that is fit for purpose. Performs initial data quality checks on extracted data. Data Transformation and Integration: extract data from identified databases. Creates data pipelines and transform data to a structure that is relevant to the problem by selecting appropriate techniques. Develops knowledge of current data science and analytics trends. Tech. Problem Formulation: translate/ co-own business problems within one's discipline to data related or mathematical solutions. Identifies appropriate methods/tools to be leveraged to provide a solution for the problem. Shares use cases and gives examples to demonstrate how the method would solve the business problem. Understanding Business Context: provide recommendations to business stakeholders to solve complex business issues. Develops business cases s for projects with a projected return on investment or cost savings. Translates business requirements into projects, activities, and tasks and aligns to overall business strategy and develops domain specific artifact. Serves as an interpreter and conduit to connect business needs with tangible solutions and results. Identify and recommend relevant business insights pertaining to their area of work. Data Modeling: analyze complex data elements, systems, data flows, dependencies, and relationships to contribute to conceptual, physical, and logical data models. Develops the Logical Data Model and Physical Data Models including data warehouse and data mart designs. Defines relational tables, primary and foreign keys, and stored procedures to create a data model structure. Evaluates existing data models and physical databases for variances and discrepancies. Develops efficient data flows. Analyzes data-related system integration challenges and proposes appropriate solutions. Creates training documentation and trains end-users on data modeling. Oversees the tasks of less experienced programmers and stipulates system troubleshooting supports. Code Development and Testing: write code to develop the required solution and application features by determining the appropriate programming language and leveraging business, technical, and data requirements. Creates test cases to review and validate the proposed solution design. Creates proofs of concept. Tests the code using the appropriate testing approach. Deploys software to production servers. Contributes code documentation, maintains playbooks, and provides timely progress updates. Data Governance: establish, modify, and document data governance projects and recommendations. Implements data governance practices in partnership with business stakeholders and peers. Interprets company and regulatory policies on data. Educates others on data governance processes, practices, policies, and guidelines. Provides recommendations on needed updates or inputs into data governance policies, practices, or guidelines.
Minimum education and experience required: Bachelor’s degree or the equivalent in Computer Science or a related field plus 3 years of experience in software engineering or related experience OR Master’s degree or the equivalent in Computer Science or a related field plus 1 year of experience in software engineering or related experience
Skills required: Must have experience with: Designing and implementing large-scale data pipelines and ETL processes using Python, Scala, PySpark, Spark and SQL for batch and streaming workloads; designing and building production streaming and event-driven pipelines (Kafka, Google Pub/Sub) including real-time ingestion and processing; Architecting and managing enterprise cloud data solutions on GCP (BigQuery, Dataproc, Cloud Storage, Pub/Sub, Cloud SQL, Secret Manager, Compute Engine); open‑source data technologies (Hadoop, Spark, Kafka, Hive); Designing data models that translate business requirements into scalable, query-efficient schemas for analytics using BigQuery and Hive; Developing, optimizing and maintaining complex analytical queries across BigQuery, Hive and relational databases; Managing scheduling, orchestration and monitoring of ETL and ELT workflows using Apache Airflow while implementing comprehensive data warehousing concepts, dimensional modeling and ETL best practices; Performing performance tuning and cost optimization for Spark jobs, Dataproc clusters and cloud storage systems by analyzing execution metrics and resource utilization; Implementing and managing comprehensive CI/CD pipelines for data applications using git-based version control systems and advanced cloud infrastructure provisioning tools such as Terraform; Establishing and maintaining data quality frameworks, security protocols and governance standards while applying standard procedures for managing data, monitoring compliance requirements, and ensuring efficient operations across all distributed data processes. Employer will accept any amount of experience with the required skills.
Wal-Mart is an Equal Opportunity Employer.
Rate of pay: $92,934.00 - 180,000.00/year
Walmart and its subsidiaries are committed to maintaining a drug-free workplace and has a no tolerance policy regarding the use of illegal drugs and alcohol on the job. This policy applies to all employees and aims to create a safe and productive work environment.