CTG, a Cegeka company, is at the forefront of digital transformation, providing IT and business solutions that accelerate project momentum and deliver desired value. Over nearly 60 years, we have earned a reputation as a faster and more reliable, results-driven partner. Our vision is to be an indispensable partner to our clients and the preferred career destination for digital and technology experts. CTG leverages the expertise of over 9,000 team members in 19 countries to provide innovative solutions. Together, we operate across the Americas, Europe, and India, working in close cooperation with over 3,000 clients in many of today's highest-growth industries. For more information, visit www.ctg.com . Our culture is a direct result of the people who work at CTG, the values we hold, and the actions we take. In other words, our people define our culture. It's a living, breathing thing that is renewed every day through the ways we engage with each other, our clients, and our communities. Part of our mission is to cultivate a workplace that attracts and develops the best people. CTG will consider for employment all qualified applicants including those with criminal histories in a manner consistent with the requirements of all applicable local, state, and federal laws. CTG is an Equal Opportunity Employer. CTG will assure equal opportunity and consideration to all applicants and employees in recruitment, selection, placement, training, benefits, compensation, promotion, transfer, and release of individuals without regard to race, creed, religion, color, national origin, sex, sexual orientation, gender identity and gender expression, age, disability, marital or veteran status, citizenship status, or any other discriminatory factors as required by law. CTG is fully committed to promoting employment opportunities for members of protected classes.
Data Engineer – AI & Analytics
Location
United States
Posted
3 days ago
Salary
$70 - $75 / hour
Seniority
Mid Level
Job Description
Data Engineer – AI & Analytics
Computer Task Group, Inc
Role Description CTG is seeking to fill a Data Engineer – AI & Analytics position for our client. Join a high-impact team helping organizations modernize their data ecosystems to power next-generation AI, machine learning, and analytics solutions. This role is ideal for an experienced Data Engineer who thrives on designing scalable data platforms, building modern data pipelines, and transforming complex data into trusted, AI-ready assets. Location: Remote Duration: 5 months Duties: - Assess customer data environments to evaluate data quality, structure, lineage, and AI/ML readiness. - Design, build, and optimize scalable ETL/ELT pipelines across APIs, relational databases, cloud storage, files, and streaming platforms. - Develop cloud-native data infrastructure supporting enterprise AI, analytics, and machine learning initiatives. - Build high-performance batch and real-time data processing pipelines. - Create reusable, production-ready data assets for analytics, business intelligence, and machine learning teams. - Implement best practices for data governance, security, privacy, metadata management, and regulatory compliance. - Optimize data models, storage formats, and pipeline performance for large-scale processing. - Develop and maintain technical documentation, architecture diagrams, and operational procedures. - Collaborate with data scientists, software engineers, analysts, and business stakeholders to deliver scalable data solutions. - Troubleshoot and resolve complex data integration and performance challenges across modern and legacy environments. Qualifications - Advanced proficiency in Python and SQL for enterprise data engineering. - Experience with PySpark or comparable distributed processing frameworks. - Hands-on experience with Databricks, Apache Spark, or similar cloud data platforms. - Experience using Apache Airflow or equivalent workflow orchestration tools. - Strong knowledge of Parquet, Delta Lake, and modern analytical storage formats. - Experience building streaming solutions using Kafka or equivalent messaging platforms. - Proficiency with Docker, Git, CI/CD pipelines, and DevOps practices. - Strong understanding of cloud data architectures, API development, and distributed systems. - Knowledge of data modeling, performance tuning, and scalable data architecture design. - Familiarity with Master Data Management (MDM), data governance, data lineage, PII compliance, and responsible AI data practices. - Exposure to analytics libraries, statistical computing frameworks, and Natural Language Processing (NLP) technologies is a plus. - Excellent analytical, troubleshooting, and collaboration skills. Requirements - 5+ years of experience in data engineering, cloud data platforms, or big data development. - Demonstrated success designing enterprise-scale data platforms supporting AI, analytics, or machine learning workloads. - Experience developing robust ETL/ELT pipelines across structured, semi-structured, and streaming data sources. - Strong background in application development, API development, debugging, and performance optimization. - Experience working with modern cloud technologies and distributed computing environments. - Ability to translate complex business and technical requirements into scalable data architecture solutions. Education - Bachelor's degree in Computer Science, Information Systems, Data Science, Engineering, or a related technical discipline. - Equivalent professional experience will also be considered. - Excellent verbal and written English communication skills and the ability to interact professionally with a diverse group are required. Benefits - The expected base salary for this position ranges from $70.00 to $75.00/hour. - Salary offers are based on a wide range of factors including relevant skills, training, experience, education, market factors, and where applicable, licensure or certifications obtained. - In addition to salary, a competitive benefit package is also offered. To Apply To be considered, please apply directly to this requisition using the link provided. Kindly forward this to any other interested parties. Thank you!
Related Guides
Related Categories
Related Job Pages
More Data Engineer Jobs
Lead Data Engineer - Python/PySpark/Databricks/AWS/AI
JPMorganChaseJPMorgan Chase & Co. (NYSE: JPM) is a leading global financial services firm with assets of $3.7 trillion and operations worldwide. The firm is a leader in investment banking, financial services for consumers and small businesses, commercial banking, financial transaction processing, and asset management. A component of the Dow Jones Industrial Average, JPMorgan Chase & Co. serves millions of consumers in the United States and many of the world’s most prominent corporate, institutional and government clients under its J.P. Morgan and Chase brands. Technology fuels every aspect of our company and is at the heart of everything we do. With over 50,000 technologists globally and an annual tech spend of $12 billion, we are dedicated to improving the design, analytics, development, coding, testing and application programming that goes into creating high quality software and new products. Learn more about technology at our firm, explore resources from our Distinguished Engineers, AI & ML researchers, and other experts; access the latest episode of our TechTrends podcast, and more at www.jpmorgan.com/technology. Information about JPMorgan Chase & Co. is available at www.jpmorganchase.com. ©2023 JPMorgan Chase & Co. All rights reserved. JPMorgan Chase is an Equal Opportunity Employer, including Disability/Veterans.
Join us as we embark on a journey of collaboration and innovation, where your unique skills and talents will be valued and celebrated. Together we will create a brighter future and make a meaningful difference. As a Lead Data Engineer - Python/PySpark/Databricks/AWS/AI at JPMorganChase within the Consumer & Community Banking, you are an integral part of an agile team that works to enhance, build, and deliver data collection, storage, access, and analytics solutions in a secure, stable, and scalable way. As a core technical contributor, you are responsible for maintaining critical data pipelines and architectures across multiple technical areas within various business functions in support of the firm’s business objectives. Job responsibilities - Generates data models for their team using firmwide tooling, linear algebra, statistics, and geometrical algorithms - Delivers data collection, storage, access, and analytics data platform solutions in a secure, stable, and scalable way - Implements database back-up, recovery, and archiving strategy - Evaluates and reports on access control processes to determine effectiveness of data asset security with minimal supervision - Uses enterprise-authorized AI capabilities within the work environment to accelerate data platform and model design analysis and documentation, validating outputs and handling data according to sensitivity and security requirements. - Applies reuse-first, AI-assisted practices within delivery and operational routines (e.g., backup/recovery validation and access control review support), ensuring traceability/auditability and alignment to resiliency and security expectations. Required qualifications, capabilities, and skills - Formal training or certification on Data Science engineering concepts and 5+ years applied experience - Expertise with Python, PySpark, Databricks, Snowflake, AWS and AI - Working experience with both relational and NoSQL databases - Experience and proficiency across the data lifecycle - Experience with database back-up, recovery, and archiving strategy - Proficient knowledge of linear algebra, statistics, and geometrical algorithms - Demonstrated experience using enterprise-authorized AI capabilities within the work environment to support data engineering workflows with strong validation habits and awareness of data sensitivity. - Ability to review and validate AI-assisted outputs (e.g., model/design summaries or operational checklists) before use, escalating when uncertain and following data handling requirements. Preferred qualifications, capabilities, and skills - Exposure to cloud technologies - Hands-on experience on Kafka or any streaming technology - Hands-on experience in Splunk, Dynatrace tools - Exposure to AI Driven development About UsChase is a leading financial services firm, helping nearly half of America’s households and small businesses achieve their financial goals through a broad range of financial products. Our mission is to create engaged, lifelong relationships and put our customers at the heart of everything we do. We also help small businesses, nonprofits and cities grow, delivering solutions to solve all their financial needs. We offer a competitive total rewards package including base salary determined based on the role, experience, skill set and location. Those in eligible roles may receive commission-based pay and/or discretionary incentive compensation, paid in the form of cash and/or forfeitable equity, awarded in recognition of individual achievements and contributions. We also offer a range of benefits and programs to meet employee needs, based on eligibility. These benefits include comprehensive health care coverage, on-site health and wellness centers, a retirement savings plan, backup childcare, tuition reimbursement, mental health support, financial coaching and more. Additional details about total compensation and benefits will be provided during the hiring process. We recognize that our people are our strength and the diverse talents they bring to our global workforce are directly linked to our success. We are an equal opportunity employer and place a high value on diversity and inclusion at our company. We do not discriminate on the basis of any protected attribute, including race, religion, color, national origin, gender, sexual orientation, gender identity, gender expression, age, marital or veteran status, pregnancy or disability, or any other basis protected under applicable law. We also make reasonable accommodations for applicants’ and employees’ religious practices and beliefs, as well as mental health or physical disability needs. Visit our FAQs for more information about requesting an accommodation. Equal Opportunity Employer/Disability/Veterans About the TeamOur Consumer & Community Banking division serves our Chase customers through a range of financial services, including personal banking, credit cards, mortgages, auto financing, investment advice, small business loans and payment processing. We’re proud to lead the U.S. in credit card sales and deposit growth and have the most-used digital solutions – all while ranking first in customer satisfaction.
Principal Data Engineer
RealTime eClinical SolutionsBetter research. Better business. Better outcomes.
• Define and drive the data science and AI roadmap, aligning model development priorities with business objectives and product strategy. • Lead end-to-end delivery of ML and AI solutions — from problem framing, data discovery, and model design through validation, deployment, and performance monitoring. • Translate ambiguous business problems into well-scoped data science workstreams, identifying quick wins alongside longer-term strategic initiatives. • Champion best practices in model development, including versioning, documentation, validation, and observability. • Design and implement NLP pipelines for use cases such as entity extraction, semantic mapping, classification, and retrieval-augmented generation (RAG). • Build and maintain forecasting and predictive models to support operational and strategic decision-making. • Apply statistical and machine learning methods to identify root causes of process inefficiencies and data quality issues. • Develop reusable data pipelines, crosswalk tables, and transformation workflows that support scalable, cross-functional data products. • Conduct current-state assessments of data architecture, sources, and quality; define future-state data models and governance standards. • Develop and maintain KPI reporting frameworks and dashboards that enable performance monitoring and data-driven decision-making. • Apply process optimization methodologies (e.g., Lean Six Sigma) to identify bottlenecks, reduce cycle time, and improve data accuracy. • Ensure analytical outputs are accurate, auditable, and aligned with regulatory and compliance requirements (e.g., HIPAA, GDPR). • Partner closely with Product, Engineering, and business stakeholders to clarify requirements, validate feasibility, and define measurable success criteria. • Communicate complex analytical findings and model outputs clearly to both technical and non-technical audiences, including executive stakeholders. • Define value-realization strategies for data and AI investments, ensuring ROI is tracked through improved search, reporting, and operational insight. • Mentor data analysts and junior data scientists through pairing, design reviews, and structured technical guidance. • Lead knowledge transfer of owned models, pipelines, and analytical frameworks to ensure team resilience and continuity. • Drive a culture of continuous learning, analytical rigor, and responsible AI within the data science function.
Azure Data Engineer
Legrand SlovenijaLegrand is the global specialist in electrical and digital building infrastructures. Our comprehensive offering of solutions for residential, commercial, and data center markets makes us a benchmark for customers worldwide. We harness technological and societal trends with lasting impacts on buildings with the purpose of improving life by transforming the spaces where people live, work, and meet with electrical and digital infrastructures and connected solutions that are simple, innovative, and sustainable. Legrand is a global, publicly traded company listed on the Euronext (Legrand SA EPA: LR). For more information, visit www.legrandgroup.com/en . Legrand, North & Central America (LNCA) is a leader in the AV, Lighting & Controls, Electrical, and Data Center markets. For more information, visit legrand.us . The industry-leading brands of Approved Networks, Ortronics, Raritan, Server Technology, and Starline empower Legrand’s Data, Power & Control to produce innovative solutions for data centers, building networks, and facility infrastructures. For more information, visit www.legrand.us Equal Opportunity Employer
Role Description Legrand has an exciting opportunity for an Azure Data Engineer to join our Information Technology Team. This is a remote position. - Lead the design, development, and optimization of enterprise data platforms on Azure. - Advance our data modernization strategy, including Azure Fabric adoption, data lakehouse architecture, and integration of enterprise systems such as ERP, MDM, and analytics platforms. - Work closely with business, analytics, and governance teams to deliver scalable, high-performance data solutions that enable trusted, data-driven decision-making. Benefits - Comprehensive medical, dental, and vision coverage. - High employer 401K match. - Paid time off (PTO) and holiday pay. - Short-term and long-term disability benefit plans. - Above-benchmark paid maternity and parental leave. - Bonus opportunities in accordance with the Company’s incentive plans. - Paid time off to volunteer. - Active/growing Employee Resource Group network. Company Description Legrand is the global specialist in electrical and digital building infrastructures. Our comprehensive offering of solutions for residential, commercial, and data center markets makes us a benchmark for customers worldwide. We harness technological and societal trends with lasting impacts on buildings with the purpose of improving life by transforming the spaces where people live, work, and meet with electrical and digital infrastructures and connected solutions that are simple, innovative, and sustainable. Legrand is a global, publicly traded company listed on the Euronext (Legrand SA EPA: LR). For more information, visit www.legrandgroup.com/en .
Senior Data Engineer
TripadvisorTripadvisor, founded in 2000, is an award-winning network for travel information that features real advice from global travelers. The world’s largest travel s
• Providing the organization’s data consumers high quality data sets by data curation, consolidation, and manipulation from a wide variety of large scale (terabyte and growing) sources. • Building high quality data pipelines and ETL processes that interact with terabytes of data on leading platforms such as Snowflake and BigQuery. • Developing and improving our enterprise data marts by creating efficient and scalable data models to be used across the organization. • Partnering with our analytics, data science, crm, and machine learning teams for data solutions. • Responsible for an enterprise data mart integrity, validation, and documentation. • Responsible for the data pipelines’ SLA and dependency management. • Writing technical documentation for data solutions, and presenting at design reviews • Solving data pipeline failure events and implementing sound anomaly detection • Working with various teams from analytics to product owners and front end developers on tracking solutions and solving technical challenges • Leading medium to large size projects in terms of writing technical specs and project planning • Mentoring junior members of the team.


