Data Engineer
4 days ago
Plano, TX, United States
Innova Solutions, Inc
Full-time
Free with email or Google
Save this job and keep your search organized
Create a free account to save jobs, create alerts and return to this listing from your dashboard.
Free with email or Google
By continuing, you agree to our Terms & Privacy Policy.
A client of Innova Solutions is immediately hiring for the Data Engineer (Python/S3 Storage Developer with AI driven platform Starburst)
Position Type: Full time, Hybrid Onsite Contract --- (*Hybrid Position – 3 days onsite and 2 days remote work in a week)
Duration: 12-18 Months contract
Location(s): Either Plano, TX / Jersey City, NJ / Charlotte, NC
LOCAL OR NEARBY CANDIDATES HIGHLY PREFERRED….SECOND ONSITE/IN-PERSON INTERVIEW REQUIRED
CLIENT INTERVIEW PROCESS:
• 1st Round Interview - Phone Screening
• 2nd Round Interview - In-Person/ Onsite Panel
*Note: While resources should know how to leverage AI for efficiency, they should not need AI to demonstrate technical knowledge and skills during interview.
Required Experience and Skills:
• Minimum 8-10 Years of experience working as a Data Engineer.
• Should have 3-5 years of hands-on experience developing data engineering solutions using Python.
• Should have 1-3 Years of Data Governance experience.
• Should have 3-5 years of experience in Python.
• Should have 3-5 years of experience in Hadoop
• 2+ years experience in AI Driven Platform Starburst.
• 2+ years of experience developing and supporting Apache Airflow DAGs in production environments.
• Experience designing, developing, and supporting data pipelines and automated workflow solutions.
• Experience working with S3-compatible object storage platforms or cloud object storage technologies.
• Experience creating, managing, and troubleshooting Apache Iceberg tables or similar modern table formats.
• Experience connecting to and integrating data from Oracle, SQL Server, Hive, Teradata, PostgreSQL, or similar enterprise platforms.
• Proficiency and good hands-on experience working in Linux/Unix environments, including shell scripting.
• Experience with source control systems such as GIT and collaborative software development practices.
• Understanding of Data Lakehouse and modern data platform concepts.
• Working knowledge of SQL and database design concepts.
• Understanding of Agile software development methodology.
• Strong analytical, troubleshooting, and problem-solving skills.
• Excellent written and verbal communication skills.
• Should have a Banking domain experience.
• Bachelor's degree in Computer Science, Engineering, Mathematics, Information Systems, or related STEM discipline.
Preferred Skills:
• Experience with Starburst, Trino, or Presto.
• Familiarity with Hadoop ecosystem technologies including Hive and HDFS.
• Experience with CI/CD pipelines and automation tools such as Jenkins, GitHub Actions, Bitbucket Pipelines, or Ansible.
• Experience implementing data quality, monitoring, and observability solutions.
• Knowledge of data governance, metadata management, and data lineage concepts.
• Experience supporting large-scale analytical data platforms.
• Exposure to container-based platforms such as Kubernetes or OpenShift.
• Understanding of enterprise security controls, authentication mechanisms, encryption standards, and access management practices.
• Exposure to reporting, analytics, AI, or Machine Learning (ML) data platforms.
Job Description:
Oure direct client’s Enterprise Finance Technology's AI, Data & Platform Services Technology Team is seeking a motivated Data Engineer to help develop and support a modern enterprise data platform built on Python, Apache Airflow, S3-compatible object storage, Apache Iceberg, and Starburst (AI Driven Platform).
In this role, the candidate will design, develop, and maintain data ingestion and transformation pipelines that create governed, scalable, and reusable data products. The engineer will build Airflow DAGs, onboard new data sources, establish and maintain data connectivity, and create and manage Iceberg tables that make curated datasets available for enterprise analytics and reporting.
The successful candidate will contribute to strategic modernization initiativ
Position Type: Full time, Hybrid Onsite Contract --- (*Hybrid Position – 3 days onsite and 2 days remote work in a week)
Duration: 12-18 Months contract
Location(s): Either Plano, TX / Jersey City, NJ / Charlotte, NC
LOCAL OR NEARBY CANDIDATES HIGHLY PREFERRED….SECOND ONSITE/IN-PERSON INTERVIEW REQUIRED
CLIENT INTERVIEW PROCESS:
• 1st Round Interview - Phone Screening
• 2nd Round Interview - In-Person/ Onsite Panel
*Note: While resources should know how to leverage AI for efficiency, they should not need AI to demonstrate technical knowledge and skills during interview.
Required Experience and Skills:
• Minimum 8-10 Years of experience working as a Data Engineer.
• Should have 3-5 years of hands-on experience developing data engineering solutions using Python.
• Should have 1-3 Years of Data Governance experience.
• Should have 3-5 years of experience in Python.
• Should have 3-5 years of experience in Hadoop
• 2+ years experience in AI Driven Platform Starburst.
• 2+ years of experience developing and supporting Apache Airflow DAGs in production environments.
• Experience designing, developing, and supporting data pipelines and automated workflow solutions.
• Experience working with S3-compatible object storage platforms or cloud object storage technologies.
• Experience creating, managing, and troubleshooting Apache Iceberg tables or similar modern table formats.
• Experience connecting to and integrating data from Oracle, SQL Server, Hive, Teradata, PostgreSQL, or similar enterprise platforms.
• Proficiency and good hands-on experience working in Linux/Unix environments, including shell scripting.
• Experience with source control systems such as GIT and collaborative software development practices.
• Understanding of Data Lakehouse and modern data platform concepts.
• Working knowledge of SQL and database design concepts.
• Understanding of Agile software development methodology.
• Strong analytical, troubleshooting, and problem-solving skills.
• Excellent written and verbal communication skills.
• Should have a Banking domain experience.
• Bachelor's degree in Computer Science, Engineering, Mathematics, Information Systems, or related STEM discipline.
Preferred Skills:
• Experience with Starburst, Trino, or Presto.
• Familiarity with Hadoop ecosystem technologies including Hive and HDFS.
• Experience with CI/CD pipelines and automation tools such as Jenkins, GitHub Actions, Bitbucket Pipelines, or Ansible.
• Experience implementing data quality, monitoring, and observability solutions.
• Knowledge of data governance, metadata management, and data lineage concepts.
• Experience supporting large-scale analytical data platforms.
• Exposure to container-based platforms such as Kubernetes or OpenShift.
• Understanding of enterprise security controls, authentication mechanisms, encryption standards, and access management practices.
• Exposure to reporting, analytics, AI, or Machine Learning (ML) data platforms.
Job Description:
Oure direct client’s Enterprise Finance Technology's AI, Data & Platform Services Technology Team is seeking a motivated Data Engineer to help develop and support a modern enterprise data platform built on Python, Apache Airflow, S3-compatible object storage, Apache Iceberg, and Starburst (AI Driven Platform).
In this role, the candidate will design, develop, and maintain data ingestion and transformation pipelines that create governed, scalable, and reusable data products. The engineer will build Airflow DAGs, onboard new data sources, establish and maintain data connectivity, and create and manage Iceberg tables that make curated datasets available for enterprise analytics and reporting.
The successful candidate will contribute to strategic modernization initiativ