Data Engineer
We are looking for a skilled Data Engineer to join our team. The ideal candidate will have a strong background in data engineering and be proficient in JAVA/Kotlin, Apache Spark, Apache Hive, Apache Airflow, and SQL Server. Responsible for designing, developing, and maintaining our data infrastructure and pipelines.
Primary Skills:
Java/Kotlin : Proficient in writing clean, efficient, and maintainable code using Java/Kotlin, or other modern programming languages.
Apache Spark, Hadoop , Kafka : Experience with Spark for large-scale data processing and Proven experience in distributed data engineering environments.
Apache Hive: Knowledge of Hive for data warehousing solutions.
Apache Airflow: Proficiency in using Airflow for orchestrating complex data workflows.
SQL Server: Experience with SQL Server for database management and querying.
Software Testing Principles : Developed unit and integration tests for data transformation logic usingPyTest and JUnit, ensuring data accuracy and reliability across multiple pipeline stages.
Docker, Kubernetes - Hands-on experience with Docker, Kubernetes, and cloud-native development to deploy and manage data services.
CI/CD tools - Familiarity with CI/CD tools (e.g., Jenkins, GitHub Actions, GitLab CI)
Responsibilities:
Design, build, and maintain scalable data pipelines across distributed systems using modern orchestration tools.
Implement and manage data governance frameworks to ensure data quality, security, and compliance.
Optimize and troubleshoot data processing jobs to ensure performance, reliability, and fault tolerance.
Apply software testing principles (unit and integration testing) to validate data workflows and transformations.
Develop and maintain data warehousing solutions with Apache Hive.
Orchestrate data workflows using Apache Airflow.
Manage and query databases using SQL Server.
Collaborate with data scientists and analysts to understand data requirements and deliver solutions.
Leverage containerization technologies (Docker, Kubernetes) to deploy and manage data services.
Implement CI/CD pipelines to automate testing, deployment, and monitoring of data applications.
Ensure data quality and integrity across all data pipelines.
Qualifications:
Proven experience as a Data Engineer or in a similar role.
Strong programming skills in Kotlin/Java.
Hands-on experience with Apache Spark, Apache Hive, Apache Airflow, and SQL Server.
Familiarity with cloud platforms (e.g., AWS, GCP, Azure) is a plus.
Excellent problem-solving skills and attention to detail.
Strong communication and teamwork abilities.
关于高知特 (Cognizant)
高知特(Cognizant)(纳斯达克代码:CTSH)作为一家AI Builder和相关技术服务提供商,致力于通过打造全栈AI解决方案,帮助企业将人工智能投资转化为实际价值。公司凭借深厚的行业经验、流程优化和工程技术专长,将企业独特的业务场景融入科技系统,赋能组织释放人才潜能,推动切实成果,并帮助全球企业在瞬息万变的环境中保持领先。如需了解更多详情,敬请访问 cognizant.ai 或关注@cognizant。
补充雇佣信息
薪酬信息截至本职位发布之日为准。Cognizant 保留在适用法律允许的范围内随时修改该信息的权利。
申请人可能需要通过现场面试或视频会议的方式参加面试。此外,候选人在每次面试时可能需要出示其当前所在州或政府签发的有效身份证件。
Cognizant 是一家提供平等就业机会的雇主。在招聘过程中,您的申请和候选资格不会因种族、肤色、性别、宗教、信仰、性取向、性别认同、国籍、残疾、遗传信息、怀孕、退伍军人身份或任何其他受联邦、州或地方法律保护的特征而受到影响。







