Big Data Engineer

JOB DESCRIPTION

Selecting and integrating any Big Data tools and frameworks required to provide requested capabilities.
Implementing ETL process to transform data from OLTP databases to OLAP DB and Data Lake using event streaming platforms such as Kafka.
Develop, transform large datasets and maintain robust data pipelines that can support various use cases with high performance.
Monitoring performance and advising any necessary infrastructure changes.
Defining data retention, data governance policies and framework

JOB REQUIREMENT

At least 5 years experience in Java programming languages.
At least 5 years experience with Big Data, Java Spring, Kafka Streams, Spark Streams frameworks.
Experience in large scale deployment and performance tuning.
Experience with schema design and dimensional data modeling
Experience with non-relational and relational databases (MySQL, MongoDB)
Experience building and optimizing ‘big data’ data pipelines, architectures and data sets.
Experience with data pipeline and workflow management tool
Fluent written and spoken English.
Strong analytical and problem-solving skills.
Bonus Points if You
Have experience with Delta Lake technology.
Good Docker/Kubernetes knowledge is a plus.
Good Kibana, Elasticsearch ELK stack knowledge is a plus.

WHAT'S ON OFFER

We have a track record of success and a vision and a plan for a promising future. Our company has closed to 100% market share for player location regulatory compliance in the US gaming space. And we have fuelled that momentum with the expansion into new markets - media & entertainment and fintech.
We are proud of our values and we live them in all of our actions, conversations, and work: There’s always a way; Together we can do more; Aim higher. Then higher; Act with integrity; For the greater good.
We are proud to be part of a global team that develops award-winning solutions for some of the world’s largest and most innovative companies.
We will support you on your learning journey. We invest in employee career growth and development. Our learning & development commitment includes leadership and technical development, a substantial budget for education and training, as well as dedicated work hours for self-study.
We care about our team. Our team is talented, has a bias for action, and is known for their positive attitude and energy. Team members are generously rewarded with competitive salaries, incentives, and a comprehensive benefits package.
We care about giving back to the communities in which we live and work. We supports a  broad range of community initiatives through donations and employee volunteer activities.
We know that work can be fun. We take the time to create employee events and experiences where everyone can connect and celebrate.

CONTACT

PEGASI – IT Recruitment Consultancy | Email: recruit@pegasi.com.vn | Tel: +84 28 3622 8666
We are PEGASI – IT Recruitment Consultancy in Vietnam. If you are looking for new opportunity for your career path, kindly visit our website www.pegasi.com.vn for your reference. Thank you!

Job Summary

Company Type:

Product

Technical Skills:

Data Engineering, Java

Location:

Ho Chi Minh - Viet Nam

Working Policy:

Job ID:

J01078

Status:

Close

Related Job:

Senior Bioinformatics Engineer

Ho Chi Minh - Viet Nam


Product

Serve as the primary bioinformatics subject matter expert for engineering teams developing cloud-native bioinformatics software. Collaborate with software architects to translate scientific workflows into scalable distributed computing architectures. Help engineers understand the computational characteristics, assumptions, and limitations of existing bioinformatics tools. Validate that modernized applications preserve scientific correctness and produce reproducible results. Define biological data models, metadata standards, controlled vocabularies, and best practices for data harmonization. Guide engineering teams in designing scalable approaches for processing large genomic and multi-omics datasets. Evaluate open-source bioinformatics software (e.g. PLINK, Regenie, BCFtools, GATK, Nextflow workflows, etc.) and identify opportunities for cloud-native modernization. Work with distributed computing specialists to determine how algorithms can be parallelized using Spark and other large-scale execution frameworks. Develop validation datasets, benchmarking methodologies, and acceptance criteria for transformed applications. Review engineering designs to ensure biological accuracy and scientific integrity. Collaborate with AI engineering teams on using AI-assisted software transformation while ensuring scientific correctness. Stay current with advances in bioinformatics, computational biology, distributed computing, and cloud-based scientific software.

Negotiation

View details

Senior Software Engineer (Distributed Computing)

Ho Chi Minh - Viet Nam


Product

  • Python
  • AWS
  • Spark

Design and provide input on system architectures, contribute to coding standards, and mentor junior engineers. Work with various teams to ensure software solutions meet business requirements, international standards, and objectives. Tackle complex software development and integration challenges, optimizing system performance. Create high-quality Python code and integrate diverse software components into cohesive solutions, with a focus on cloud computing and life sciences applications. Keep up-to-date with new technologies and Python frameworks, particularly in cloud computing. Oversee testing, deployment, and comprehensive documentation of integrated systems. Actively participate in all phases of the software development lifecycle. This includes creating user stories and engaging in sprint planning to align development efforts with business objectives. Engage with multinational companies, demonstrating flexibility to occasionally adapt to US and EU time zones.

Negotiation

View details

Data and Insights Manager

Ho Chi Minh - Viet Nam


Product

  • Data Analyst
  • Business Intelligence
  • Data Science
  • Machine Learning

Oversee the data and information system to ensure adequacy, accuracy, and legitimacy of data products such as reports, insights, analytics tools, and models. Implement and oversee data governance and policies. Communicate data visions and strategies across functions and business units. Anticipate upcoming trends and business demands to propose impactful data deliverables. Oversee the pipeline to deliver high-quality, timely, and useful data to enable data usability, availability, and efficiency on the Information System (BI, Reportings, Dashboard). Ensure excellence in performing data ad-hoc requests. Identify analytics opportunities to support the growth of the organization. Develop analytics capacities through data and data science techniques, mastering knowledge about customers, products, and the business. Drive data-empowered innovations through ML/AI to improve the efficiency and quality of business and products. Raise awareness of the potential application of ML/AI. Provide data, ML/AI learnings, and strategic input to business visions and projects. Design the framework of AI/ML building, testing, and deployment at a large scale. Ensure the quality and efficiency of ML/AI deliverables.

Negotiation

View details