Please submit your CV in English.
Senior Data Engineer
Location
Worldwide
Posted
14 hours ago
Salary
0
Seniority
Senior
Job Description
Senior Data Engineer
Loka, Inc
Role Description We are seeking a Senior Data Engineer to join our growing team at Loka. In this role, you will: - Build scalable production-grade data pipelines and real-time data streams. - Write and design scalable, cloud-ready applications. - Lead technical projects through architecture, design, and implementation phases. - Collaborate with Machine Learning, Data Science, Design, Software Engineering, and Business teams to triage data or ETL issues. - Build tests and health checks to maintain code and data quality. - Monitor and analyze data flowing through various systems by adding appropriate visualizations and dashboards. - Provide updates and offer guidance to clients. Qualifications - Advanced Python and SQL. - Experience in ETL design, implementation, and maintenance. - Experience in AWS, GCP, or Azure delivering data-centric product experiences. - 5+ years experience in a related role. - Experience with in-memory and disk-based databases, relational and non-relational databases, full-text search engines, database design, development, and maintenance (e.g., MySQL, MongoDB, OpenSearch, DynamoDB, with bonus for Graph Databases like Neo4j). - Experience with data warehousing and multidimensional data models (columnar data modeling). - Working knowledge in Data Lake, Data warehouse, and massive parallel processing. - Strong problem-solving ability and ability to work through ambiguity and incomplete specifications. Requirements - Excellent English - Being a global team, we work entirely in English for meetings, customer calls, and business communications. - CV written in English. Preferred but Not Required - Working knowledge in IAM, Federated Authentication, SSO, SAML, Encryption, Security, APIs, Disaster Recovery, or Backup. - Strong ability in distributed systems for large-scale data processing. - Experience with Spark and Pandas. - Experience with Open Tables (Hudi, Delta Lake, DataBricks, Iceberg as related to Data Lakehouse). - Experience with Infrastructure as Code and CI/CD Pipelines. - Experience with QuickSight and DataViz. Personality Profile - Curious: You strive to learn and grow into different industries with a modern tech stack. - Autonomous: You thrive in a fully remote, globally distributed team. - Collaborative: You enjoy working across departments. - Adaptable: You operate with a startup mindset and move at startup speed. Benefits - Every other Friday off (26 extra days off a year). - Remote-first culture. - Explore and Relocation programs (three months work abroad or full international relocation). - Paid sick days and local holidays. - Business English classes program. - Continuous Learning Support. - Fitness and/or Mental Health Subscriptions. - Access to LokaLabs™, our internal research and development program. - Defined career path.
Related Guides
Related Categories
Related Job Pages
More Data Engineer Jobs
Role Description Build and optimize ETL pipelines for large-scale AI applications, integrating APIs, web scraping, and real-time data processing. Develop and maintain scalable AI data infrastructure using PySpark, Pandas, SQL, and cloud services (Azure, AWS, GCP). Implement retrieval-augmented generation (RAG) pipelines using FAISS, ChromaDB, and Pinecone for AI-driven insights. Deploy and monitor AI applications using FastAPI, Streamlit, Docker, for real-time performance. Work with cross-functional teams to ensure data security, compliance, and AI model reliability in production environments. Qualifications - Bachelor's or Master's degree in Computer Science, Data Engineering, Artificial Intelligence, or a related field. - At least 2-4 years of experience in data engineering, AI/ML, or cloud-based AI infrastructure. - Expertise in Python, PySpark, and SQL for data transformation and large-scale processing. - Experience with cloud platforms (AWS, Azure, GCP) for AI model deployment and data pipeline automation. - Hands-on experience with vector databases (FAISS, ChromaDB, Pinecone) for efficient data retrieval. - Proficiency in containerization and orchestration tools like Docker and Kubernetes. - Strong understanding of retrieval-augmented generation (RAG) and real-time AI model deployment. - Knowledge of API development and AI service integration using FastAPI and Streamlit/Dash. - Ability to optimize AI-driven automation processes and ensure model efficiency. - Strong analytical and problem-solving skills. - Excellent communication and collaboration abilities. Requirements - Expertise in Python, PySpark, and SQL for data transformation and large-scale processing. - Experience with cloud platforms (AWS, Azure, GCP) for AI model deployment and data pipeline automation. - Hands-on experience with vector databases (FAISS, ChromaDB, Pinecone) for efficient data retrieval. - Proficiency in containerization and orchestration tools like Docker and Kubernetes. - Strong understanding of retrieval-augmented generation (RAG) and real-time AI model deployment. - Knowledge of API development and AI service integration using FastAPI and Streamlit/Dash. - Ability to optimize AI-driven automation processes and ensure model efficiency. - Strong analytical and problem-solving skills. - Excellent communication and collaboration abilities. Company Description
Data Engineer
Shawmut Services LLCShawmut Services is a full-service strategic planning, organizing, and grassroots mobilization firm.
• Design, develop, and maintain scalable ETL/ELT data pipelines. • Build and optimize data architectures and warehouses (e.g., Snowflake, BigQuery Etc). • Integrate data from a wide variety of sources using APIs and web browser automations. • Collaborate with company leadership to understand data needs and implement ongoing improvements in system processes. • Monitor and ensure data quality, reliability, and performance. • Automate data workflows and implement data validation and testing processes. • Maintain up-to-date documentation for data pipelines, systems, and processes.
Doctoral Fellow – Mathematics, Engineering, Statistics, Data Science
Sistema FibraPelo Futuro da Indústria | Pelo Futuro do Trabalho
• Fine-tuning of SLMs tailored to specific needs • Implementation of agents and integrations as required • Implementation of multi-LLM routing • Preparation of technical and business documentation • Knowledge transfer • Conduct cost–benefit analysis.
Role Description Fuse Energy is a forward-thinking renewable energy startup on a mission to deliver a terawatt of renewable energy – fast. We're combining first-principles thinking with cutting-edge technology to build a radically better energy system. - Lead data center from early-stage opportunity through Ready-to-Build and energisation - Assess sites from technical, planning, power, and infrastructure perspectives - Develop grid connection and power strategies with National Grid, DNOs, and utility providers - Support planning, permitting, and stakeholder engagement processes - Identify and mitigate technical and development risks - Ensure projects are deliverable, scalable, and aligned with construction requirements - Experience developing data centers or large-scale infrastructure projects - Strong technical understanding of data center infrastructure, including power, cooling, and utility systems, with the ability to evaluate engineering decisions and challenge technical assumptions - Strong understanding of AUS planning and grid connection processes - Experience taking projects from early-stage development through design, permitting, and delivery readiness - Highly analytical, self-directed, and comfortable with significant ownership Qualifications - Experience developing data centers or large-scale infrastructure projects - Strong technical understanding of data center infrastructure, including power, cooling, and utility systems - Strong understanding of AUS planning and grid connection processes - Experience taking projects from early-stage development through design, permitting, and delivery readiness - Highly analytical, self-directed, and comfortable with significant ownership Requirements - Lead data center from early-stage opportunity through Ready-to-Build and energisation - Assess sites from technical, planning, power, and infrastructure perspectives - Develop grid connection and power strategies with National Grid, DNOs, and utility providers - Support planning, permitting, and stakeholder engagement processes - Identify and mitigate technical and development risks - Ensure projects are deliverable, scalable, and aligned with construction requirements Benefits - Competitive salary - Biannual bonus scheme - Fully expensed tech to match your needs


