
Mongodb · Dublin
We are seeking a Staff engineer to design, build, and operate the internal and external Observability stack for the MongoDB platform. Tens of thousands of custo...
We are seeking a Staff engineer to design, build, and operate the internal and external Observability stack for the MongoDB
platform. Tens of thousands of customers depend on our Observability stack to monitor their database clusters and to generate
actionable alerts to safeguard critical workloads. This is an opportunity to join a team that is responsible for all Observability
systems that support metrics, metric visualization, logs, traces, and alerts for MongoDB. We are looking for engineers with high
standards, and experience in setting direction and technical leadership for large engineering teams in designing and operating
complex distributed systems, with strict SLO on security, durability, availability and performance.
As MongoDB Atlas and its supporting infrastructure continue to experience rapid growth, the demand for high-cardinality
observability data for internal and external use cases means we need to continually innovate and scale our systems to the next
level. For example, MongoDB Observability systems need to handle 10’s of billions of metrics time series, all whilst processing
petabytes of logs, traces, and events. Our stack includes VictoriaMetrics, Splunk, Flink, WarpStream/Kafka, Java, Golang
Fluentbit. In addition to owning critical components of our observability infrastructure, as a Staff engineer on the team, you’ll
also work closely with other SWE, Product and SRE teams to promote and implement best practices in instrumenting and monitoring
their services. This is a highly collaborative role, and you will get to own some of the most relied upon internal infrastructure
at Mongo.
Our team champions a strong culture of inclusivity, diversity, and collaboration. If you want to be a deeply technical leader on a
collaborative team that applies low-level systems expertise to build the foundational infrastructure of a popular database, join
us! Let’s build a faster, more reliable, and exceptionally observable database system together.
We are looking to speak to candidates who are based in Dublin for our hybrid working model.
C/C++/Java/Rust mission critical software systems
engineers
of components that drive performance, scalability, cost-efficiency, and resiliency
the root cause of production issues
durability, availability, and performance) and maintainability
diagnosed and fixed a few customer or testing-reported issues
and using your experience to drive the long-term technical roadmap of the Observability Team
MongoDB is built for change, empowering our customers and our people to innovate at the speed of the market. We have redefined the
data platform for the AI era, enabling builders to create, transform, and disrupt industries with software. MongoDB’s unified data
platform, the most widely available, globally distributed data platform on the market, helps organizations modernize legacy
workloads, embrace innovation, and unleash AI. Our cloud-native platform, MongoDB Atlas, is the only globally distributed,
multi-cloud data platform and is available across AWS, Google Cloud, and Microsoft Azure.
With offices worldwide and over 67,000 customers, including 75% of the Fortune 100 and AI-native startups, relying on MongoDB for
their most important applications, we’re powering the next era of software.
Our compass at MongoDB is our Leadership Commitment, guiding how and why we make decisions, show up for each other, and win. It’s
what makes us MongoDB.
To drive the personal growth and business impact of our employees, we’re committed to developing a supportive and enriching
culture for everyone. From employee affinity groups, to fertility assistance and a generous parental leave policy, we value our
employees’ wellbeing and want to support them along every step of their professional and personal journeys. Learn more about what
it’s like to work at MongoDB, and help us make an impact on the world!
MongoDB is committed to providing any necessary accommodations for individuals with disabilities within our application and
interview process. To request an accommodation due to a disability, please inform your recruiter.
MongoDB is an equal opportunities employer.
Req ID: 1273361935
MongoDB is seeking a Staff Software Engineer to join the Atlas Clusters Organization. The organization is responsible for building MongoDB Atlas, our database as a service offering and fastest growing product. Atlas allows users to deploy fault-tolerant, secure, globally distributed MongoDB clusters in just minutes. This includes developing software to interface with the three major cloud providers (AWS, Azure, and GCP) in order to bring security, durability, availability, and performance to all deployments of MongoDB. This engineer will also work on our Atlas Data Federation & Archiving product. Atlas Data Federation & Archiving allows customers to move data from hot to cold storage and run federated queries over that data. We are forming a new Atlas Clusters team in the Dublin area. We are looking to speak to candidates who are based in Dublin and would like a hybrid or in-office working model. WHAT YOU’LL DO * Build and design new features for MongoDB Atlas and Atlas Data Federation & Archiving * Contribute to and lead complex technical projects * Work with stakeholders throughout MongoDB to build our roadmap and product offerings * Work with customers and support engineers to fix issues and become part of our on-call rotation * Collaborate with team members to develop our codebase, best practices, and design principles * Foster an inclusive and respectful work environment according to MongoDB's Core Values WE’RE LOOKING FOR SOMEONE WHO * Has at least 10+ years of professional software development experience * Is skilled at writing large-scale, distributed backend systems in a compiled language (Go, Java, C#, etc) * Has experience with at least one major cloud provider technology (AWS, Azure, GCP) * Has led the launch of a new module and maintained it in production * Is eager to solve tough problems * Has excellent communication skills * Is curious, collaborative, and motivated SUCCESS MEASURES * In 3 months, you'll have shipped code into production and collaborated with the team to solve tough problems * In 6 months, you'll have contributed to a large project and joined our on-call rotation * In 12 months, you'll have designed new features, led development work, and become a go-to expert on parts of the system ABOUT MONGODB MongoDB is built for change, empowering our customers and our people to innovate at the speed of the market. We have redefined the data platform for the AI era, enabling builders to create, transform, and disrupt industries with software. MongoDB’s unified data platform, the most widely available, globally distributed data platform on the market, helps organizations modernize legacy workloads, embrace innovation, and unleash AI. Our cloud-native platform, MongoDB Atlas, is the only globally distributed, multi-cloud data platform and is available across AWS, Google Cloud, and Microsoft Azure. With offices worldwide and over 67,000 customers, including 75% of the Fortune 100 and AI-native startups, relying on MongoDB for their most important applications, we’re powering the next era of software. Our compass at MongoDB is our Leadership Commitment, guiding how and why we make decisions, show up for each other, and win. It’s what makes us MongoDB. To drive the personal growth and business impact of our employees, we’re committed to developing a supportive and enriching culture for everyone. From employee affinity groups, to fertility assistance and a generous parental leave policy, we value our employees’ wellbeing and want to support them along every step of their professional and personal journeys. Learn more about what it’s like to work at MongoDB, and help us make an impact on the world! MongoDB is committed to providing any necessary accommodations for individuals with disabilities within our application and interview process. To request an accommodation due to a disability, please inform your recruiter. MongoDB is an equal opportunities employer. Req ID: 2273470263
MongoDB is built for change, empowering our customers and our people to innovate at the speed of the market. We have redefined the data platform for the AI era, enabling builders to create, transform, and disrupt industries with software. MongoDB’s unified data platform, the most widely available, globally distributed data platform on the market, helps organizations modernize legacy workloads, embrace innovation, and unleash AI. Our cloud-native platform, MongoDB Atlas, is the only globally distributed, multi-cloud data platform and is available across AWS, Google Cloud, and Microsoft Azure. With offices worldwide and over 67,000 customers, including 75% of the Fortune 100 and AI-native startups, relying on MongoDB for their most important applications, we’re powering the next era of software. Our compass at MongoDB is our Leadership Commitment, guiding how and why we make decisions, show up for each other, and win. It’s what makes us MongoDB. To drive the personal growth and business impact of our employees, we’re committed to developing a supportive and enriching culture for everyone. From employee affinity groups, to fertility assistance and a generous parental leave policy, we value our employees’ wellbeing and want to support them along every step of their professional and personal journeys. Learn more about what it’s like to work at MongoDB, and help us make an impact on the world! MongoDB is committed to providing any necessary accommodations for individuals with disabilities within our application and interview process. To request an accommodation due to a disability, please inform your recruiter. MongoDB is an equal opportunities employer. Req ID 1273361945
Work Mode: Flex 1+ days per week in our Dublin office Department: Engineering ABOUT THE COMPANY LearnUpon partners with over 1,600 organisations globally to unlock the potential of employees, customers & members through learning that’s easy, scalable and focused on results. Read more about life at LearnUpon here. ABOUT THE TEAM Our Engineering organization is dedicated to building robust, scalable infrastructure that handles world-scale platform demands. As part of the Site Reliability Engineering (SRE) team, we focus on system architecture, absolute performance, and technical innovation. Operating with high ownership and technical expertise, we are responsible for the scale-out of the LearnUpon infrastructure, championing internal self-service tooling, and embedding a culture of observability and shared operational responsibility across all engineering squads. ABOUT THE OPPORTUNITY As a Staff Site Reliability Engineer, you will be a principal technical leader and a key catalyst for our infrastructure's evolution. In this role, you will take ownership of our core platform resilience, driving the strategy to build out an advanced, cost-effective observability function spanning metrics, logs, and transaction tracking. This opportunity requires a strategic thinker who can design cross-team SLO/SLI frameworks, navigate complex distributed system requirements, and mentor talent to ensure LearnUpon scales efficiently to support our ambitious global goals. In addition, you’ll be responsible for: * Infrastructure Optimization: Identify opportunities to improve and scale our infrastructure for performance, observability, maintainability, and cost, by creating innovative solutions. * Observability Function Strategy: Lead our efforts to build an observability function that incorporates application metrics, application transaction tracking, and event log management. * Resilience & Scaling: Drive the processes to maintain resilient, scalable, and cost-effective infrastructure while working with other Engineering teams to provide solutions that meet their ongoing requirements. * Tooling & Self-Service: Build tools focused on measuring, monitoring, and alerting, with an eye towards self-service in order to promote Engineers’ ownership of observability. * Operational Agility & Support: React quickly to changing customer and business needs and actively participate in the team's on-call rota. Team Up-Leveling: Mentor junior talent and effectively communicate complex technical ideas to both technical and non-technical peers. SKILLS & EXPERIENCE Must-Haves * 7+ years of experience in a software or Ops role. * 5+ years of cloud engineering experience, with at least 2 years of experience with AWS. * Experience deploying Microservice environments using containerisation technologies such as Kubernetes and Docker. * Experience designing and implementing Observability tech stacks, championing its benefits to Engineering teams, and managing the associated cost analysis of metrics gathering, effort, and tooling. * Ability to architect the design of SLO/SLI implementations that balance the needs of different teams. * Experience building and supporting large-scale distributed systems that back a consumer app or website with associated requirements of performance, security, and disaster recovery. * Experience with implementing IaC (e.g., CloudFormation, Terraform, etc.), automation tooling (e.g., Puppet, Ansible etc.), and CI/CD (e.g., Jenkins, Travis CI, GitLab, etc.). * Experience using AI tools to streamline tasks and improve efficiencies. Nice-to-Haves * Experience with database scaling would be a strong plus. * Certification in AWS, any PaaS, and/or related technologies. *If you don’t tick every box but believe this role is a mutually good fit, please don’t hesitate to apply. We’d love to hear from you. WHY CHOOSE LEARNUPON? From comprehensive rewards and generous time off to meaningful investment in your growth and development, LearnUpon gives you the support, trust, and opportunity to do the most impactful work of your career. Learn more here. HIRING PROCESS * Qualified applicants may be invited to an initial screening call with a member of our TA Team. * Successful candidates will be invited to a series of practical interviews. * Finally, candidates will have an interview with our CTO. * Successful candidates will be contacted with an offer to join our team. Note: At LearnUpon, we utilise AI to enhance the speed and quality of our screening and assessment practices, but our hiring decisions are always human. LearnUpon is an Equal Opportunities Employer. We do not discriminate on the basis of gender, marital status, family status, age disability, sexual orientation, race, religion, membership of the Traveller community, or any other legally protected status. Check out our Careers site and Instagram to learn more about working at LearnUpon. By submitting your application, you agree to LearnUpon's Privacy Policy.