Associate Data Engineer
Posted 5hrs ago
Employment Information
Report this job
Job expired or something wrong with this job?
Job Description
Associate Data Engineer building data pipelines, warehouses, and BI datasets for TaskUs’s outsourced digital services. Supporting ETL automation, data quality, modeling, testing, and monitoring.
Responsibilities:
- Design, develop, and maintain data pipelines and data warehouses for analytics infrastructure
- Configure and set up solutions using pre-approved patterns to move data from operational and external environments to the business intelligence environment
- Manage data integration and ETL jobs, including consolidation, aggregation, filtering, validation, data quality routines, and dimension creation
- Configure and deploy automated jobs using pre-approved reusable patterns
- Execute production-level jobs while considering data dependencies, schedules, validation routines, and data stream updates
- Design, develop, verify, and continuously optimize data models according to business requirements
- Collaborate with technical users to define data requirements and establish business rules
- Collaborate with business stakeholders to deliver and enhance the BI platform and provide relevant, accurate insights through new datasets
- Perform unit and system integration testing
- Conduct peer reviews, assist with testing, and maintain current documentation
- Monitor error-handling and logging tools and processes
- Troubleshoot configuration and data errors
- Monitor performance and raise appropriate flags
- Identify opportunities to optimize the ETL environment
- Implement monitoring, quality, and validation processes to ensure data accuracy and integrity
Requirements:
- At least 2 years of experience in Data Engineering
- Knowledge of Python, Apache Kafka, AWS Redshift, AWS Glue, AWS S3, and Pentaho Data Integration
- Strong knowledge of data warehousing concepts, including traditional and MPP database designs, star and snowflake schemas
- At least 2 years of data modeling experience
- At least 2 years of hands-on development experience using ETL tools such as Pentaho, SSIS, Informatica, Talend, Fivetran, or Airflow
- Knowledge of the architecture, design, and implementation of MPP databases such as Teradata, Snowflake, or Redshift
- 3-year development using cloud-based analytics solutions, preferably AWS or GCP
- Knowledge of designing and implementing streaming pipelines using Apache Kafka, Apache Spark, or Segment
- Knowledge of database tuning and ETL tuning
- Experience using Python in a cloud-based environment preferred
- Knowledge of NoSQL databases preferred
- Knowledge, experience, and exposure to AI tools preferred but not required
- Ability to work effectively across internal functional areas in ambiguous situations
- Structured thinking and effective communication
- Bachelor's/College Degree in Computer Science, Information Technology, Engineering (Computer/Telecommunication), or 4-8 years of experience in place of a degree
- Ability to work 8:00 PM to 5:00 AM PHT, subject to business needs
Benefits:
- Competitive industry salaries
- Comprehensive benefits packages
- Wellness support
- Inclusive environment
- Internal mobility opportunities
- Professional growth opportunities
- Work from home



















