Junior Machine Learning Engineer
Location
Brazil
Posted
7 days ago
Salary
0
Seniority
Junior
Job Description
Junior Machine Learning Engineer
Docket
• Support the development of IDP (Intelligent Document Processing) solutions: Contribute to creating, training, tuning models and fine-tuning for extracting unstructured data from complex documents. • Build the future of agentic AI: Assist in orchestrating workflows using frameworks such as LangGraph, n8n, and others, connecting document extraction to approval and formalization pipelines. • Transform data into products: Apply statistical techniques and data analysis to the structured datasets produced by our IDPs, helping identify patterns and insights that will become new AI products for our clients. • Collaborate with technical, product and commercial teams: Work side by side with Technical Leadership, Data Engineering, Product Managers and the Front-end team, evaluating, supporting and improving the transition of proofs of concept (PoCs) to production solutions. • Ensure quality and monitoring: Help evaluate model outputs (LLMs, RAGs and classical models), test prompts and document processes, ensuring solutions are accurate and scalable. • Learn and innovate: Stay up to date with the state of the art in Generative AI, bringing curiosity, ideas and proactivity to solve bottlenecks in our document pipeline.
Job Requirements
- Completed degree or near completion in Computer Science, Engineering, Statistics, Mathematics, Physics, Data Science or related STEM fields.
- Clear understanding of statistics and programming logic (know why to use a technique, not just how to import a library).
- Proficiency in Python focused on data manipulation and API consumption.
- Practical knowledge of SQL for extracting, joining and analyzing data in relational databases.
- Hunger to learn, resilience to handle errors during experiments and excellent communication to ask for help, clarify doubts and present ideas clearly.
- Fluency with AI tools: Daily use of tools such as ChatGPT, Claude, Cursor or GitHub Copilot to accelerate your development and study flow.
- Programming and analysis: Python (Pandas, NumPy, Scikit-learn, TensorFlow, PyTorch) and SQL (complex queries and data modeling).
- Generative AI and LLMs: Solid fundamentals in Prompt Engineering and a theoretical/practical understanding of how LLMs (OpenAI, Anthropic) and RAG (Retrieval-Augmented Generation) architectures work.
- Natural Language Processing (NLP) and Computer Vision: Basic understanding of concepts behind text extraction and document classification.
- Applied statistics: Ability to analyze large volumes of structured data and apply multivariate analyses or hypothesis testing to extract business value.
- Orchestration logic: Understanding of how APIs communicate and the logic of graph- or node-based flows (preparing to work with Agents).
Benefits
- Meal and food allowance via Flash for when hunger strikes.
- Health and dental insurance to take care of you.
- Life insurance for added peace of mind.
- Petlove benefits, because at Docket we understand your furry family matters too.
- Conexa Saúde, psychological support at your fingertips.
- Wellhub and TotalPass to keep your body moving.
- Galena, because learning and development are essential.
- Partnership with Sesc, access to leisure and cultural activities.
- Childcare assistance for parents with children up to 5 years old.
- Baby Cash when the family grows.
- Birthday month day off to celebrate as you deserve.
- Working hours: 44 hours per week
Related Guides
Related Job Pages
More Machine Learning Engineer Jobs
Machine Learning Engineer
AllCloudAllCloud is a leader in amplifying organizations’ cloud potential through AI. With a track record of hundreds of successful migrations and implementations across AWS and Salesforce, AllCloud has developed strategies and solutions that enable businesses of all sizes to remain at the forefront of innovation. AllCloud serves clients across the globe with offices in EMEA and North America.
Role Description We are looking for a savvy Machine Learning/Data Engineer to join our growing team of data experts. The hire will be primarily responsible for AI/ML projects on AWS, leveraging native services as well as custom-built models to deliver predictive insights to our customers. In addition, this hire will also support migrating to the cloud, optimizing our customers’ databases and data flows, and enriching our operational and functional data flow with AI/ML algorithms. The ideal candidate is confident in data in any form or scale and happy to learn and teach new data tools. The candidate enjoys optimizing data systems and building them from the ground up. The Machine Learning Engineer will support new system designs and migrate existing ones, working closely with solutions architects, project managers, and data scientists. They must be self-directed and comfortable supporting the data needs of multiple teams, systems, and products. The right candidate will be excited by the prospect of optimizing or re-designing our customers’ data architecture to support our next generation of products and data initiatives, and machine learning systems. Responsibilities - Keep our customers’ data separated and secure to meet compliance and regulations requirements. - Design, Build and Operate the infrastructure required for optimal extraction, transformation, and loading of data from a wide variety of data sources using SQL and cloud (mainly AWS) migration and ‘big data’ technologies. - Optimize various RDBMS engines in the cloud and solve customers' security, performance, and operation problems. - Design, Build and Operate large, complex data lakes that meet functional / non-functional business requirements. - Optimize various data types ingestion, storage, processing, and retrieval from near real-time events and IoT to unstructured data as images, audio, video and documents, and in between. - Use Jupyter Notebooks to build and deploy ML models. - Leverage AWS AI/ML pre built solutions to accelerate work for customers. - Work with customers and internal stakeholders, including the Executive, Product, Data, Software Development, and Design teams, to assist with data-related technical issues and support their data infrastructure and business needs. Qualifications - 3+ years of experience in a Data Scientist/Machine Learning Engineer role. - Bachelor's (Graduate preferred) degree in Computer Science, Mathematics, Informatics, Information Systems, or another quantitative field. - Experience with big data tools: Spark, ElasticSearch, Hadoop, Kafka, Kinesis etc. - Experience with relational SQL and NoSQL databases, such as MySQL or Postgres and DynamoDB or Cassandra. - Experience with AWS cloud services: EC2, RDS, EMR, Redshift etc. - Experience with functional and scripting languages: Python, Java, Scala, etc. - Experience with various ML models for classification, scoring and more. - Experience with Deep Learning Neural Networks (Convolution, NLP etc.) - Experience with AWS AI/ML Services. - Experience with Python coding. - Advanced working SQL knowledge and experience working with relational databases, query authoring (SQL) as well as working familiarity with a variety of databases. - Experience building and optimizing ‘big data’ data pipelines, architectures and data sets. - Strong analytic skills related to working with unstructured datasets. - Build processes supporting data transformation, data structures, metadata, dependency and workload management. - Working knowledge of message queuing, stream processing, and highly scalable ‘big data’ data stores. - Experience supporting and working with external customers in a dynamic environment. Certifications - AWS Machine Learning Specialty (Strongly Preferred) - AWS Solutions Architect - Associate (Strongly Preferred) Benefits - Our team inspires progress in each other and in our customers through our relentless pursuit of excellence; you will work with leaders who promote learning and personal development.
Senior Machine Learning Engineer
ChattermillTurn customer feedback from every channel into insights that drive better products, greater retention, and deep loyalty.
• Train, evaluate, and iterate on ML models and agentic systems for customer feedback, including owning our custom fine-tuning pipelines. Run experiments end-to-end, track results rigorously, and make clear recommendations on what to ship, iterate, or retire. • Build and maintain LLM-powered features: retrieval pipelines, reranking systems, insight agents, data mining agents, and automated taxonomy generation. • Design and run robust evaluation frameworks: build test sets, define metrics, evaluate non-deterministic systems, handle class imbalance, and automate checkpoint comparisons. • Improve and extend semantic search and retrieval, evolving from embedding-based approaches toward more advanced methods. • Write production-quality code and collaborate closely with Engineering on productionisation, model serving, data pipelines, and monitoring. • Work with Product and Commercial teams to translate business needs into practical ML solutions, and support client evaluations and accuracy benchmarking. • Mentor team members, review code and research, and bring relevant advances from the literature into the product.
Senior Machine Learning Engineer, Developer Advocacy
Grafana LabsGrafana Labs supports organizations’ monitoring, visualization and observability goals. 950,000+ active installations
• Lead the evolution of the Interactive Learning system's recommendation engine • Build and operate applied models for continuous improvement • Define measures of recommendation quality and partner across various teams for integration • Ship incremental improvements and enhance existing recommender features
Senior Machine Learning Engineer, Developer Advocacy
Grafana LabsGrafana Labs supports organizations’ monitoring, visualization and observability goals. 950,000+ active installations
• Lead the evolution of the recommendation system for Grafana's Interactive Learning system • Build, deploy, and operate recommendation models • Partner with software engineers and data analysts • Develop, validate, and monitor models • Ship incremental improvements and integrate them into the existing recommender system



