Personalized, expert AI coaching for every manager—at 2% of the traditional cost
Staff Software Engineer – Platform
Location
New York
Posted
2 days ago
Salary
0
Seniority
Lead
Job Description
Staff Software Engineer – Platform
Valence
• Build the paved roads. Design and operate shared platform primitives used by multiple teams (authorization and routing guardrails, secure logging, rollout controls, observability, queue infrastructure, and the centralized LLM gateway) so teams build on standards instead of reinventing them. • Make change safe by default. Own the reliability and change-safety systems that let the org move fast without breaking trust: CI/CD gates, migration safety patterns, progressive delivery, rollback automation, and kill switches. • Enforce security at the platform layer. Ship secure defaults, policy enforcement frameworks, auditability, and data-access guard libraries so the safe path is the easy path. • Run global delivery controls. Implement single-domain strategy, region pinning and routing, residency-aware behavior, and cross-region config consistency checks for enterprise customers worldwide. • Raise the operational bar. Establish SLOs, synthetic checks, alerting standards, and incident runbooks and drills that drive faster detection and recovery. • Build the internal developer platform. Create service templates, blessed internal-tool paths, and standard deploy, monitoring, and auth patterns that make the right thing the default thing. • Operate an AI-native engineering plane. Build issue-to-PR automation infrastructure with governance, human-in-control workflows, and audit logs that convert signals into auditable, human-approved execution outcomes.
Job Requirements
- 8+ years of experience in software engineering, distributed systems, infrastructure, or platform/SRE roles, with a track record of building systems other teams depend on
- You build reusable primitives, not one-off features. You think in terms of paved roads, secure defaults, and standards that scale across many teams.
- Hands-on experience with CI/CD, progressive delivery, rollback and migration safety, observability, SLOs, and incident response.
- Comfortable developing and operating services in cloud environments (AWS, GCP, Azure) and working with containerization and orchestration (Docker, Kubernetes).
- Familiarity with implementing platform-level guardrails (authorization, auditability, data-access controls, and policy enforcement) in a privacy- and residency-conscious way.
- Strong software engineering skills, including writing maintainable code, debugging distributed systems, and collaborating in cross-functional teams.
- Eagerness to tackle unfamiliar problems, learn new technologies, and shape our platform and culture.
- Ability to explain technical ideas clearly and work effectively with both technical and non-technical stakeholders.
Benefits
- Comprehensive health coverage (medical, dental, vision) from day one
- Generous PTO, company-wide R&R shutdowns, and paid parental leave
- Retirement plan support for US and global employees
- Outsized missions from day one, with direct responsibility for company-defining projects
- Work alongside the executive team with transparency into strategy and decision-making
- Influence on direction through real-time customer feedback and market insights
- Meaningful ownership in a venture-backed company at a growth inflection point
- Financial upside that comes from scaling fast
- Top-up grants as we scale and you deliver exceptional performance
- A culture built for top talent: intensity to win, growth without limits, and a team that solves hard problems and celebrates big wins together
Related Guides
Related Categories
Related Job Pages
More Platform Engineer Jobs
• Manage and evolve our API Gateway integration platforms at enterprise scale • Manage the tooling and support for intelligent automated integration with flexibility and reliability as your mission • Analyze and refine API definition standards and drive support for innovation at scale in our API ecosystem • Guide and mentor junior engineers while setting the bar for operational excellence • Managing and advancing the API integration platforms and related infrastructure components, including software upgrades, performance monitoring, and application development team collaboration • Helping lead platform enhancement efforts that enable flexible and secure integration between internal and external services • Providing guidance and expertise for application service design, data and transaction workflows for scalability, reliability and extensibility • Following infrastructure-as-code best practices (Terraform, Helm, Ansible) for repeatable deployments and environment consistency • Driving operational excellence through on-call rotation support, thorough post-mortem investigation, and implementing solutions that improve resilience • Performing unit testing and complex debugging, and identifying opportunities for automated testing where appropriate • Performing analysis of technical feasibility and solution design, including estimation of work efforts, planning, and backlog organization • Ensuring upgrades, patching, and platform updates are proactively planned and executed without business disruption • Helping set reliability targets and defining operational metrics (availability, latency, error budgets) in line with SRE methodologies • Writing detailed technical documentation for systems, including architectural diagrams, and operational procedures for standardized actions used for maintenance and support
Senior Software Developer – Real-Time Processing Platform
Arctic WolfAt Arctic Wolf, we foster a collaborative and inclusive work environment that thrives on diversity of thought, background, and culture. Our commitment to bold growth and shaping the future of security operations is matched by our dedication to customer satisfaction, with over 10,000 customers worldwide and more than 2,000 channel partners globally. As we continue to expand globally and enhance our technology, Arctic Wolf remains the most trusted name in the industry. Recognized by Top Workplace USA (2021-2025). Best Places to Work – USA (2021-2025). Great Place to Work – Canada (2021-2024). Great Place to Work – UK (2024-2026). Kununu Top Company – Germany (2024-2026).
• Design, develop, and enhance core components of Arctic Wolf’s real-time detection and event processing platform. • Build and evolve large-scale distributed systems that process and analyze trillions of events per day in near real-time. • Develop capabilities across event processing systems, detection frameworks, event correlation platforms, and stream-processing infrastructure. • Solve complex engineering challenges related to scalability, reliability, performance, latency, throughput, and cost efficiency in cloud-native environments. • Contribute to the architecture and technical direction of next-generation platform capabilities. • Deliver high-quality, production-ready software and operate systems at massive scale. • Collaborate with cross-functional teams to continuously improve platform effectiveness and operational excellence. • Mentor and support other developers while fostering a culture of ownership, quality, and continuous improvement.
• Analyze business requirements to identify SharePoint-based solutions • Design and develop SharePoint solutions including custom workflows and applications • Collaborate with business stakeholders to ensure solutions meet their needs • Develop SharePoint solutions using development frameworks like .NET and PowerShell • Test and debug SharePoint solutions to ensure quality standards • Provide ongoing maintenance and support for SharePoint solutions • Create and maintain technical documentation for SharePoint solutions
• Implement the specialized pharmacovigilance agents: write system prompts, configure model parameters, build tool-use definitions, and define agent boundaries to ensure precise adverse event processing • Build and iterate prompt chains for each processing step: source document parsing, field extraction, MedDRA coding suggestions, causality assessment logic, narrative drafting, and E2B(R3) output generation • Develop the deterministic rule engine layer: implement ICH E2B field validation checks, MedDRA hierarchy verification, and regulatory logic constraints that operate alongside LLM outputs • Create and maintain evaluation datasets in collaboration with the pharmacovigilance domain team: annotated ground-truth cases, edge case libraries, and regression test suites • Develop and maintain Model Context Protocol (MCP) servers to expose enterprise applications, APIs, databases, and services as standardized tools for AI agents • Implement secure MCP integrations, tool definitions, authentication, and testing to enable reliable agent interaction with internal and external systems • Run accuracy benchmarks, analyze failure modes, and iterate on prompts and agent configurations to improve performance against defined thresholds • Implement the quality control agent's cross-verification logic: configure separate Claude instances, build comparison algorithms, and calibrate confidence scoring • Build human-in-the-loop feedback mechanisms: reviewer interfaces for accept/modify/reject decisions, structured feedback capture, and feedback-to-prompt-improvement pipelines




