We are seeking an Entry-Level Data Engineer to work with Hadoop ecosystem technologies including HDFS, Hive, and MapReduce. The role involves writing SQL queries, creating and maintaining Hive tables, loading and cleaning large datasets, and assisting with ETL workflows. Candidates must have basic SQL, Hadoop, Hive, and Linux knowledge, with good problem-solving skills.
Key Highlights
Key Responsibilities
Technical Skills Required
Benefits & Perks
Nice to Have
Job Description
Entry Level Data Engineer
Full-Time
Candidate must be open to relocate
Responsibilities:
- Work with Hadoop ecosystem technologies such as HDFS, Hive, and MapReduce.
- Write and execute SQL queries to extract, transform, and analyze data.
- Create and maintain Hive tables and perform data processing using HiveQL.
- Load, clean, and validate large datasets stored in Hadoop.
- Assist senior developers with ETL/data processing workflows.
- Troubleshoot data-related issues and investigate query or processing errors.
- Perform basic data analysis and generate reports based on business requirements.
- Follow data quality, security, and documentation standards.
- Learn and work with tools such as Linux, Git, Python, or Spark as required.
Looking to advance your Data Science career with relocation support? Explore Data Science Jobs with Relocation Packages that include comprehensive packages to help you move and settle in your new role.
Discover our full range of relocation jobs with comprehensive support packages to help you relocate and settle in your new location.
Required Skills:
- Basic knowledge of SQL: SELECT, JOIN, GROUP BY, subqueries, aggregate functions.
- Understanding of Hadoop and HDFS concepts.
- Basic knowledge of Hive/HiveQL.
- Familiarity with Linux commands.
- Understanding of databases and data warehousing concepts.
- Good problem-solving and analytical skills.
Interested in relocating to United State? Check out our comprehensive Relocation Jobs in United State page with detailed relocation packages and benefits.
Good to Have:
- Apache Spark
- Python
- ETL concepts
- Data warehousing
- Git
- Cloud platforms such as AWS/Azure/GCP
Thanks
Similar Jobs
Explore other opportunities that match your interests