

Search by job, company or skills

About Trustify Technology
Trustify Technology is a fast-growing software engineering and digital transformation company providing AI-driven software development, verification & validation, and automation solutions for global clients across the US, UK/EU, Japan, and Singapore.
Trustify aims to be the top-tier technology outsourcing partner from Vietnam, combining human expertise with AI power to deliver high-quality, cost-effective solutions.
About Data Engineer Position
We are seeking an experienced and highly skilled Data Engineer (Senior/Principal/Lead) with knowledge of Data Science to join our team. The ideal candidate will have a strong background in data engineering, as well as a good understanding of data science and machine learning concepts. This role will focus on building and maintaining data pipelines, managing data storage, and providing data solutions to support our organization's data-driven initiatives.
Responsibilities
o Design, build and maintain data pipelines that are scalable, reliable and efficient
o Develop and manage data storage solutions, including data warehouses and data lakes
o Implement and maintain ETL processes for various data sources
o Collaborate with data scientists to develop and implement machine learning models
o Ensure data quality, accuracy, and consistency across all data sources
o Collaborate with cross-functional teams to identify and implement data-driven solutions to business problems
o Stay up to date with emerging technologies, trends, and best practices in data engineering, data science, and machine learning
o Develop and maintain documentation of data processes, data models, and metadata
o Collaborate with data scientists to ensure their machine learning models can be integrated into production systems
o Ensure data privacy and security are maintained throughout all data processes
Requirements
o From 5 years of experience as a Data Engineer working in the IT industry.
o Bachelor's degree in Computer Science/ Software Engineering, Mathematics
o Must have experiences/ knowledge in Linux, ETL pipelines, Python, SQL
o Experience in working with Cloud services: AWS/ Azure/ GCP
o Nice to have experiences in Java, Airflow, GCP Git, Databricks
o Familiar with CI/CD pipelines, Docker, new Data tech and tools (Tableau, PowerBI)
o Strong experience with data manipulation tools and frameworks such as Apache Spark, or Apache Kafka
o Good understanding of data science and machine learning concepts, with experience implementing Data Models, Data Warehouse
o Analytical thinking, Teamwork, Automation mindset
o Good English communication skills
Job ID: 151559893
Skills:
Azure Synapse, Power Bi, Spark, Azure Data Lake, Databricks, Sparksql, Python, ML Pipelines
Skills:
Hadoop, Informatica, Apache Nifi, Apache Airflow, Hive, Gcp, Spark, Etl Tools, Azure, Talend, Python, AWS, HDFS
Skills:
data engineering , Machine Learning, Stl, Python, Deep Learning, Data Science Pipelines, Object-Oriented Programming, AI LLM tools, Jupyter Notebook
Skills:
snowflake , Web Services, Csv, Pyspark, Kafka, Postgres, Json, Sql, Devops, Spark Streaming, Xml, Databricks, Python, AWS, Parquet, web API frameworks, Spark structured streaming, Text, CI CD, Agentic AI, Delta, MLFlow, git-based version control
Skills:
data engineering , snowflake , Sql, ETL/ELT, AI/ML