Job Closed
This listing is no longer active.
Inteligência, Inovação e Tecnologia.
Data Engineer – GCP, Databricks
Location
Brazil
Posted
66 days ago
Salary
0
Seniority
Senior
Job Description
Data Engineer – GCP, Databricks
Leega
• Company focused on delivering efficient and innovative service to its clients • Innovative professionals who are driven by challenges • Technically lead the evolution of the data platform • Guide the team and ensure the delivery of robust solutions in cloud environments
Job Requirements
- Lead the design and evolution of the data architecture (Data Lake, Data Warehouse, Lakehouse)
- Design, develop and optimize high-scale data pipelines (ETL/ELT)
- Serve as the technical reference for the data engineering team
- Work with distributed processing using Databricks (Spark/PySpark)
- Define best practices for development, testing and deployment (DataOps)
- Ensure data governance, quality, security and reliability
- Optimize performance and costs in GCP environments
- Integrate multiple data sources (batch and streaming)
- Support BI, Analytics and Data Science teams in making data available
- Participate in strategic decisions related to data and technology
Benefits
- Mandatory requirements**
- Specialist across multiple projects and technical team leadership.
- Minimum of 5 years' experience as a Data Engineer
- Solid experience with Google Cloud Platform (BigQuery, Dataflow, Cloud Storage, Pub/Sub)
- Advanced experience with Databricks and Apache Spark (PySpark)
- Proficiency in Python and SQL
- Experience with data modeling (relational, dimensional and Lakehouse)
- Experience with pipeline orchestration (Airflow/Cloud Composer or similar)
- Familiarity with code versioning and CI/CD practices
- Experience with large-scale data environments
- Desired requirements**
- Experience with Delta Lake and Lakehouse architecture
- Knowledge of infrastructure as code (Terraform)
- Experience with Docker/Kubernetes
- Knowledge of data streaming (Kafka, Pub/Sub)
- Certifications (GCP Professional Data Engineer, Databricks)
- Experience with observability and monitoring tools
- Differentials**
- Experience leading technical teams or data projects
- Strong communication skills with business stakeholders
- Knowledge of data governance, LGPD and information security
- Experience with BI tools (Looker, Power BI, Tableau)
- Advanced English
Related Guides
Related Categories
Related Job Pages
More Data Engineer Jobs
Senior Software Engineer, Data Centralization & Storage
AttentiveThe most comprehensive text message marketing solution.
Role Description Our engineering department consists of 200+ people across multiple teams, such as application development, infrastructure, data platform, machine learning, and security. We believe our company will win in the long run through product innovation. To get there, we obsess over iteratively delivering customer value through rapid prototyping and data-driven decision-making. We’re looking for a self-motivated, highly driven Software Engineer with a strong understanding of low-level distributed systems concepts - not just tool users, but systems thinkers. We want engineers who are comfortable reasoning about how infrastructure behaves under load, how services communicate over the network, and how data flows through complex architectures. Our team is focused on building the platform to support data driven products and strategic outcomes for our customers. We do this by providing an extensible and flexible compute platform to access highly available, mission critical data sets. What You’ll Accomplish - Architect high-throughput solutions that power our most critical operations, ensuring scalability and efficiency. - Expand and enhance our self-service platform, collaborating with cross-functional teams to fuel our AI, ML, and analytics goals. - Tackle complex distributed data challenges, streamline system integrations, and uphold high standards of quality and governance. - Champion cutting-edge technologies, keeping our platform at the forefront of industry advancements and enabling strategic outcomes. - Unify data from diverse systems, paving the way for experimentation and innovation while empowering teams with intuitive tools and frameworks. Your Expertise - Track record of debugging issues across layers—from application logic to infrastructure bottlenecks—and understand the tradeoffs in system design, not just the settings in a managed UI. - You don’t just know how to run jobs in tools, you understand how they operate under the hood and how connected storage impacts performance like: - How simple storage API semantics (e.g. consistency, eventual visibility, multi-part uploads) affect job execution. - The impact of network I/O and data locality on performance and cost. - The way resource scheduling and JVM tuning influence distributed job behavior. - Proven experience as a Software Engineer with a focus on high throughput scalable systems. - In-depth knowledge of high throughput processing technologies such as Spark, Flink, and/or Kafka. - Proficiency in Java and strong understanding of object-oriented design, data structures, algorithms, and optimization. - You have development experience working with data warehouse tools like Snowflake, Clickhouse, Trino, etc. - You’ve experience with open source data storage formats such as Apache Iceberg, Parquet, Arrow, or Hudi. - You are knowledgeable about data modeling, data access, and data replication techniques, such as CDC. - You have a proven track record of architecting applications at scale and maintaining infrastructure as code via Terraform. - You are excited by new technologies but are conscious of choosing them for the right reasons. What we use - Our backend is Java / Spring Boot microservices, built with Gradle, coupled with things like DynamoDB, Pulsar, Flink, Spark, Postgres, Planetscale, and Redis, hosted via AWS. - Our frontend is built with React and TypeScript, and uses best practices like GraphQL, Storybook, Radix UI, Vite, esbuild, and Playwright. Salary and Benefits - The US base salary range for this full-time position is $178,400 - $205,500 annually + equity + benefits. - Our salary ranges are determined by role, level and location. Company Values - Default to Action - Move swiftly and with purpose. - Be One Unstoppable Team - Rally as each other’s champions. - Champion the Customer - Our success is defined by our customers' success. - Act Like an Owner - Take responsibility for Attentive’s success.
Senior Data Architect, m/f/x
Cara CareCara Care is a medical device company on a mission to empower people “to live a healthier and happier life.” The company is made up of a diverse, friendly,
• We’re looking for a Senior Data Architect to join us immediately and help shape the backbone of our data-driven future. • If you love designing scalable data systems, making architectural decisions that matter, and turning complex data ecosystems into clean, reliable platforms - this is your role. • You’ll own the evolution of our data architecture end-to-end and play a key role in turning data into a real strategic asset across the company. • You’ll work closely with Product, Engineering, and Leadership, and your decisions will directly shape how data is stored, accessed, trusted, and used.
• Architect and own the governed Snowflake layer • Define connection standards for AI agents and applications • Encode PII scrubbing and compliance as code • Document the warehouse • Modernize the analytics codebase • Support the Tableau and Mixpanel layers • Set technical direction and raise the bar
• Develop and evolve data pipelines in Azure, GCP and Databricks environments, ensuring efficiency, performance and reliability; • Design solutions for large-scale data ingestion, transformation and provisioning; • Integrate data from multiple sources and systems, including Firebase; • Use Terraform for provisioning and standardizing infrastructure as code; • Identify opportunities to improve data architecture, proposing scalable and sustainable solutions; • Proactively troubleshoot issues and maintain platform stability; • Participate in technical definitions, contributing best practices and knowledge sharing; • Collaborate with multidisciplinary teams to ensure technical and business alignment; • Support initiatives to modernize and consolidate data in cloud environments.



