Application Site Reliability Engineer (SRE)

Location

Americas + 1 moreAll locations: Americas | Latin America (LATAM)

Posted

3 days ago

Salary

0

Seniority

Mid Level

Job Description

Application Site Reliability Engineer (SRE)

CXM Direct LLC

Role Description Join our Platform & Production Reliability team and help ensure the reliability, performance, and availability of our mission-critical trading systems. As an Application Site Reliability Engineer (SRE), you will own the day-to-day reliability of our .NET/C# services running on Windows, starting with our in-house liquidity bridge that connects MetaTrader trading servers to external liquidity providers. Over time, you will expand your impact across related trading and back-office services. This is a hands-on role for an engineer who enjoys solving production challenges, improving observability, automating operations, and building resilient systems where uptime directly impacts customer experience. Qualifications - Mid-Level (3–5 years) experience - Strong experience debugging and supporting .NET/C# applications in production - Hands-on experience with Windows Server environments - Strong PowerShell scripting skills - Experience with Python or Bash - Experience with Grafana, Prometheus, and Loki (or equivalent monitoring and observability tools) - Experience with modern CI/CD pipelines - Experience working with AWS - Hands-on experience with Terraform or other Infrastructure as Code (IaC) tools - Experience troubleshooting and supporting Aurora PostgreSQL or other relational database platforms - Practical experience with SLIs & SLOs, Error Budgets, Incident Response, Root Cause Analysis (RCA), Alert Design, Production Operations Requirements - Participate in the on-call rotation for production trading systems and lead incident response during service disruptions - Investigate production incidents, perform root cause analysis, and implement preventive actions to eliminate recurring issues - Build and maintain Grafana dashboards, Prometheus alerts, and operational health views across applications, infrastructure, and databases - Instrument .NET services to improve telemetry, metrics, logging, and visibility into service health and customer impact - Define, implement, and monitor Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets - Troubleshoot issues across .NET/C# applications, Windows Server, Aurora PostgreSQL databases, AWS infrastructure, CI/CD pipelines and deployments - Improve deployment safety, release automation, and rollback strategies - Partner with developers to improve application operability, resilience, and fault isolation - Automate operational tasks through scripting and infrastructure automation - Create and maintain runbooks, operational documentation, and incident response procedures - Continuously improve monitoring, alert quality, automation, and platform reliability Benefits - Work on mission-critical trading infrastructure that directly impacts customers - Solve challenging reliability and scalability problems in a real-time environment - Build world-class observability, automation, and deployment practices - Collaborate with experienced engineers in a modern engineering culture - Influence reliability strategy and engineering best practices across the platform

Related Categories

Related Job Pages

More Application Engineer Jobs

CXM Direct logo

Application Site Reliability Engineer, SRE

CXM Direct

Your Innovative and Reliable STP Broker

Full TimeRemoteTeam 51-200Since 2015H1B No Sponsor

• Own the day-to-day reliability of .NET/C# services running on Windows. • Participate in the on-call rotation for production trading systems and lead incident response during service disruptions. • Investigate production incidents, perform root cause analysis, and implement preventive actions to eliminate recurring issues. • Build and maintain Grafana dashboards, Prometheus alerts, and operational health views across applications, infrastructure, and databases. • Instrument .NET services to improve telemetry, metrics, logging, and visibility into service health and customer impact. • Define, implement, and monitor Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets. • Troubleshoot issues across .NET/C# applications, Windows Server, Aurora PostgreSQL databases, AWS infrastructure, CI/CD pipelines and deployments. • Improve deployment safety, release automation, and rollback strategies. • Partner with developers to improve application operability, resilience, and fault isolation. • Automate operational tasks through scripting and infrastructure automation. • Create and maintain runbooks, operational documentation, and incident response procedures. • Continuously improve monitoring, alert quality, automation, and platform reliability.

Argentina
Prolific logo

Senior Application Security Engineer

Prolific

Building a better world with better data.

Full TimeRemoteTeam 51-200Since 2014H1B Sponsor

Role Description Security at Prolific isn't an afterthought, it's foundational to how we build. As a company trusted by world-leading research institutions and AI labs to handle sensitive data at scale, the security of our application layer is critical. We handle participant data, researcher credentials, payment flows, and API integrations that demand rigorous protection at the code level. As a Senior Security Engineer, you'll be the technical authority on application security at Prolific. You'll work hands-on with our engineering teams to: - Find and fix vulnerabilities in our codebase - Perform security testing - Build security tooling - Embed secure development practices into how we ship software This isn't a governance or policy role; you'll be in the code, reviewing pull requests, threat modeling new features, and building the automation that keeps our platform secure as we scale. You'll report to the Head of Engineering/Platform and work cross-functionally with product engineering, platform, data, and TechOps teams. Qualifications - Several years in application/product security and a background in software engineering - Strong knowledge of OWASP Top 10 (Web & API) and modern attack paths (e.g. auth flaws, SSRF, injection, business logic abuse, supply chain) - Experience working with complex, large-scale systems and modern architectures - Hands-on security testing experience (especially Burp Suite) across web apps and APIs - Python for security tooling, automation, or custom detection (Django a plus) - Experience implementing and tuning SAST, SCA, DAST, and secret scanning in CI/CD - Practical threat modelling experience, including leading lightweight sessions - Strong collaboration skills, able to clearly explain issues and drive remediation - Builder mindset, you automate wherever possible Requirements - Experience with Django, Vue.js, MongoDB, GCP - Security champions or bug bounty programmes - Supply chain security (SCA, SBOMs, dependency review) - IaC security (e.g. Terraform, policy-as-code) - Hands-on certifications (OSCP, GWAPT, BSCP) - Experience in scaling environments building out security practices Benefits - Competitive salary - Benefits - Remote working - Impactful, mission-driven culture

United Kingdom

Title: Enterprise - Senior Application Engineer - JavaScript, Angular, REST Location: Annapolis Junction, MD Job Description: Erias Ventures was founded to serve its customers with an entrepreneurial mindset. We value creative problem-solving, open communication, and empowering our employees to make decisions and put forth new ideas. Our staff includes technical experts working across multiple disciplines, bringing diverse perspectives to every project. We are seeking engineers who wish to grow their careers and want to become part of a technically strong and growth-oriented company focused on bringing innovative solutions to the difficult mission problems facing our customers. Description We are seeking an Application Engineer to be part of the Secure the Enterprise initiative, develop capabilities to shift from the current manual system security evaluation and authorization process to a new model that emphasizes automation, streamlined processes and approvals, continuous monitoring and assessment, and network data gathering across the entire life cycle of a project. - Will provide the development, testing, deploying, and sustainment of various web-based capabilities utilizing Angular 14 or higher which will interact with ReST end points to make requests, receive formatted responses, and visualize, in various details, the data to UI front ends - Will focus on creating single page application dashboards and provide mockups for new development This position may allow for partial telework. Clearance A current Top-Secret/SCI with polygraph security clearance is required. Candidates cannot be sponsored or nominated for a government security clearance under this position. Experience Twelve (12) years minimum experience and a High School Diploma/GED. Ten (10) years minimum experience and an Associate's Degree. Eight (8) years minimum experience and a Bachelor's Degree. Six (6) years minimum experience and a Master's Degree. Four (4) years minimum experience and a Doctorate's Degree. ​ Required skills: - Experienced with using TypeScript JavaScript language - UI Technical Leadership required - Experienced in using Angular Framework to develop user interfaces - Experienced with analyzing json data structure when working with the ReST - Experienced with Cascading Style Sheets (CSS) to enhance the look and feel of user interfaces Desired skills: - Jira - Confluence - Agile Framework / SAFe - AWS</li> - Balsamiq - MongoDB Benefits Erias Ventures provides a complete package of wealth, health, and happiness benefits. The expected salary range for this position, depending on education and years of experience is $237,000 - $262,000. Wealth Benefits: - Above Market Hourly Pay - 11% Roth or Traditional 401k with Immediate Vesting and Deposit - Spot Bonuses for Assisting with Business Development and Company Growth - Professional Development Bonuses for Certificates and Degrees Health Benefits: - Company subsidized Medical Coverage - 100% Company Paid Vision and Dental Coverage - 100% Company Paid Long Term Disability, Short Term Disability, and Group Life Insurance - Monthly Wellness Reimbursement Happiness Benefits: - Paid Time Off with Flexible Work Schedules and Birthday Off - Amazon Prime Membership and Monthly Internet Reimbursement - Technology and Productivity Allowance for Equipment and Supplies - Morale Building and Company Events to Celebrate our Successes and Build our Community - Onboarding and Annual Swag - Company Paid Professional Development and Training At Erias Ventures, we are dedicated to fostering a diverse and inclusive workplace. As an equal opportunity employer, we ensure that all qualified applicants are considered for employment based on merit, without discrimination. We welcome individuals regardless of race, color, religion, gender, gender identity or expression, sexual orientation, national origin, genetics, disability, age, or veteran status. Referrals & Inquiries Do you know a cleared professional seeking to advance their career? Interested in earning some extra cash? If so, refer them to us with their name and contact details, and you could be eligible for a referral bonus of up to $10,000 for each successful hire. Not seeing the right position right now? Reach out to us, and we&rsquo;ll notify you as new contracts and opportunities become available!

Maryland
$237K - $262K / year
ContractRemoteTeam 201-500Since 2015H1B No Sponsor

• Own the stability and performance of mission-critical systems supporting business users. • Investigate system anomalies by correlating logs, metrics, and traces to identify root causes and determine remediation approaches. • Lead rapid incident response and drive root cause resolution. • Participate in an on-call rotation to support scheduled deployments and respond to production issues outside business hours. • Respond to and resolve customer-reported issues within SLA, communicating status and resolution to business stakeholders. • Proactively identify problems, performance bottlenecks, and areas for improvement before they impact customers. • Improve observability across monitoring, metrics, logging, and alerting. • Troubleshoot and resolve software integration issues, data integrity problems, and application errors. • Identify opportunities to improve system scalability, fault tolerance, and performance.

Brazil