
BlaBlaCar · Paris
About BlaBlaCar BlaBlaCar is the world’s leading community-based travel app enabling 27 million members a year to carpool or travel by bus in 21 countries. Our...
About BlaBlaCar
BlaBlaCar is the world’s leading community-based travel app enabling 27 million members a year to carpool or travel by bus in 21 countries. Our team of 800 employees counts over 50 nationalities and is spread across our 5 global offices, 30% working fully remotely.
Your Mission
By joining our Foundations department, you will be working alongside talented individuals grouped in small agile teams that each have strong ownership on their piece of these goals. Foundations is composed of seven teams which “provide consistent, easy to use, infrastructures, services, and expertise to support BlaBlaCar’s growth and evolution”.
The Site Reliability Engineering team (SRE) is responsible to provide best in class Observability, Alerting and Incident management tools and processes to service teams. As an enabling team, we help BlaBlacar engineers to efficiently improve their service reliability. Empowering developers and bringing them our reliability expertise are at the core of our daily work.
Core Infrastructure: Kubernetes, Google Cloud Platform
GitOps/Delivery: GitHub, Terraform, Flux, Helm, Jenkins
Observability/Incident Management: Datadog, Opentelemetry, Grafana IRM,
In house Synthetic Tests platform: Playwright, Qualcium, SauceLabs
Languages: Go / Python for Tooling, Typescripts/JS for the testing platform
Your responsibilities
Support software engineers by creating, maintaining, and improving observability and alerting tools and frameworks. You embrace the use of AI, leveraging agentic to eliminate toil and streamline your daily tasks
Own the Service Level Objectives (SLOs) framework, assist in the design and maintenance of indicators (SLI) and objectives to ensure service reliability.
Owning the incident management process by defining best practices, standards, and ensuring continuous improvement through post-mortems and chaos engineering. While developers handle incidents within their scope, you could step in as Incident Commander during high-severity incidents, leading coordination efforts .
Develop and maintain tools, such as Terraform modules or Go apps, to help automate and enhance reliability across services.
Build and promote reporting on operational metrics and incidents to drive distributed and continuous improvement.
Your qualifications
1 to 5 years of experience in SRE, DevOps, or Software Engineering roles
Working in a multidisciplinary environment will request strong communication skills : you'll need to adapt your communication level to other teams expertise and be able to understand their needs
Strong knowledge of observability tools (e.g., Datadog) and understanding of metrics, logging, and tracing.
Troubleshooting/oncall experience in production environments, diagnosing and resolving technical issues effectively (experience with Kubernetes is a plus).
Full working proficiency in English
Fit with our BlaBlaPrinciples
Thriving in a collaborative, fast-growing and innovative environment
Ability to take ownership, aligned with business priorities and navigating in different contexts
Familiarity with incident management platforms (e.g., Grafana IRM) is a bonus
Experience working with Service Level Objectives (SLOs) and Service Level Indicators (SLIs)
Exposure to programming in Go or a strong interest in learning it.
Experience in integrating Opentelemetry
Backend services are built using multiple programming languages: while development skills aren't required, familiarity with object-oriented programming and scripting languages is an advantage.
Familiarity with web/mobile testing tools or a strong curiosity to understand how software is tested at scale.
What we have to offer
Hybrid status for this role : 2-3 days at the Office
4 additional weeks on top of legal maternity/paternity leaves
50% healthcare coverage (Alan)
Financial support for home office equipment
Minimum 25 days holiday per year
Local meal plan policy (Swile card)
50% transportation paid (Forfait Mobilité Durable)
Free unlimited carpooling & bus rides
Personal growth via trainings, mentorship, and internal mobility opportunities
Employee Stock ownership plan
Regular team building events
1 day off per year to test our product
Interested in joining the ride?
a 45-min video-call with Maxime, Talent Acquisition Manager, to get to know you, understand your career expectations and answer your questions
a 60-min video-call with Damien Bertau, Hiring Manager, to discuss your experience and share more details about the team
a 90-min system design interview with 2 team members to discuss about your technical expertise
a 45-min video-call with Maxime Fouilleul, Head of Foundations, to get a wider vision of the department and its strategy
Our hiring process lasts on average 25-30 days, offers usually come within 48 hours.
Please note that one of these interviews will be onsite.
About BlaBlaCar BlaBlaCar is the world’s leading community-based travel app enabling 27 million members a year to carpool or travel by bus in 21 countries. Our team of 800 employees counts over 50 nationalities and is spread across our 5 global offices, 30% working fully remotely. Your Mission SRE in the Engineering Experience Team, part of the Foundations Department, are responsible for designing, building and maintaining the Software Delivery platform, tools and standards that enable teams to confidently release changes up to production. We aim to accelerate delivery, simplify the Engineering experience, guarantee reliable workflow and satisfy Engineering needs at scale. By joining our Foundations Department, you will be working alongside talented individuals grouped in small agile teams that each have strong ownership on their stack and roadmaps. Foundations is composed of four teams (Engineering Experience, Cloud Infrastructure, Site Reliability Engineering & Quality Assurance) which “provide consistent, easy to use, secured infrastructure, services, and expertise to support BlaBlaCar’s growth and evolution”. The Engineering Experience Team has four main objectives, driving its roadmap: Reduce BlaBlaCar product’s time to market by designing, building and maintaining state of the art CI/CD and associated tooling to streamline day-to-day delivery from development teams Improve developers efficiency in providing AI tooling and infrastructure, ensuring the compliance with internal policies while keeping enough flexibility for experimentation Drive development teams towards autonomy through the provision of comprehensive training and support, clear guidelines, and effective tooling Leverage our existing FinOps framework to enable precise cost control for the Software-as-a-Service (SaaS) we use and manage, strategically balancing this with the need to support innovation and the adoption of new functionalities The role requires a global vision of the Engineering perimeter.You will champion the adoption and sharing of best practices among Engineering. Your approach should be that of an enabler, not a gatekeeper. You embrace the use of AI, leveraging code generators and assistants to eliminate toil and streamline your daily tasks and make development teams life easier. Crucially, your strong communication skills will be essential for ensuring a clear understanding of our users' needs. To fulfill the mission, you will be working with several stakeholders : The Product & Engineering teams, working with service team to ensure best understanding and usage of our Software Factory components. Developer Experience Engineers, to ensure that the best-in-class user experience is prioritized from the start. External SaaS providers, to deliver cutting-edge support for our internal users, as well as analyzing and recommending subscription adjustments to maximize the value BlaBlaCar derives from these services. Technical stack: Core Infrastructure: Google Cloud Platform, Kubernetes GitOps/Delivery: GitHub, Github Copilot, Github Actions & Jenkins, Terraform, Flux, Helm Datastores: Postgres, Cassandra, Elasticsearch, Kafka Observability: Datadog, Grafana Languages: Go for Infra/Security Tooling, Java for backend services, Python for data Your ResponsibilitiesIn cooperation with your Engineering Manager and the Engineering Experience team: Design, build and improve parts of our Software Factory, specifically but not only Continuous Integration and Continuous Delivery, to address scaling and resiliency needs on our cloud platforms; Implement tools and services to ease the work of developers and automate problem resolution; Collaborate with engineers and help them improve software development lifecycle and processes. Investigate and fix service issues; Your Qualifications You can demonstrate a strong experience with large scale continuous integration/delivery systems (e.g. GitHub Actions or Jenkins); You can demonstrate an experience with Cloud platforms, container and process isolation technologies, especially Docker and Kubernetes; You can demonstrate an experience with an SRE/DevOps oriented language (e.g. Go); You can demonstrate a good knowledge of Linux/Unix fundamentals; You embrace change, prioritize high-value tasks, and are results-driven and impact-oriented; You are a humble, collaborative, and communicative team player, focused on enabling developer empowerment and autonomy, eager to share knowledge and learn from others; You are at ease with English speaking. If you don’t meet 100% of the qualifications outlined above, tell us why you’d still be a great fit for this role in your application! What we have to offer Hybrid status for this role : 2-3 days at the Office 4 additional weeks on top of legal maternity/paternity leaves 50% healthcare coverage (Alan) Financial support for home office equipment Minimum 25 days holiday per year Local meal plan policy (Swile card) 50% transportation paid (Forfait Mobilité Durable) Free unlimited carpooling & bus rides Personal growth via trainings, mentorship, and internal mobility programs Employee Stock ownership plan Regular team building events 1 day off per year to test our product Interested in joining the ride? Here’s what your hiring journey will look like: a 45-min video-call with Maxime, Talent Acquisition Managers to get to know you, understand your career expectations, and answer your first questions a 60-min video-call with your future manager, Jean-Baptiste Favre, Engineering Manager, to get to know you, present you the team, and discuss your technical fit for the role. a technical assignment to evaluate your technical skills followed by a 60-min video-call with two Engineers. a 30-min video-call with Maxime Fouilleul, Head of Engineering, for vision fit and rounding off the process
We’re Capital on Tap 👋 💳 Capital on Tap started because small businesses were underserved. Big banks were slow, their products weren't fit for purpose, and small business owners often couldn't access what they needed. We set out to fix that. Today we're a financial platform - not just a credit card company. We offer a best-in-class business credit card, SME-focused spend management platform, a savings product that hit £1 billion in funds within its first year, and a growing suite of tools and financial products that make running a small business easier. 1,000+ employees, £20bn in annual card spend, 200,000+ customers, 17,000+ Trustpilot reviews averaging 4.7 stars, and we're profitable. We’ve done a pretty good job so far, but we’re just getting started! 📍London, Old Street | 🏢 2 Days in Office SRE at Capital On Tap 🌞 At Capital On Tap, we run a hybrid embedded SRE model. We aim to work closely with the teams within Capital On Tap to provide them the best support. Our main objective currently is to gain as much visibility into our platform's health while offering scalable solutions. What You’ll be doing: As a Site Reliability Engineer (SRE) you will help ensure our platforms are fast, reliable, and scalable. You’ll design, build, and monitor systems, prevent issues before they happen. Using SLAs, SLIs, and SLOs, you’ll guide feature launches while maintaining services that everyone can depend on. * Manage and automate Azure, Datadog, NGINX & Cloudflare * Develop and monitor Kubernetes and Serverless resources * Maintain infrastructure code with Terraform & CRDs / Crossplane * Improve systems, processes, and technologies; consult stakeholders to enhance platform performance * Getting involved in new application architecture & design processes * Design solutions to reduce toil, automate repetitive tasks and streamline workflows to reduce manual work and boost team productivity. * Create SLIs and SLOs; increase application visibility * Align with the Product team on SLAs and core service objectives * Collaborate with Platform Engineers for automated solutions and pipelines * Enhance user experience with infrastructure and pipeline optimisation * Support CI/CD tools such as Azure Devops, Octopus Deploy and Flux to streamline software delivery * Lead incident troubleshooting to safeguard customer experience We’re Looking For 🔎 Required skills: * Experience in managing a public cloud (Azure advantageous) * Experience in Azure DevOps, Octopus, Flux or other CI/CD tools * Experience with Linux and Microsoft Systems * Excellent communication skills and ability to collaborate with multiple teams in an agile environment * Proficient in contributing to IaC technologies involving expertise in writing, managing, and optimising infrastructure with tools such as Terraform and Pulumi * Experience working with a cloud monitoring solution (advantageous to have DataDog) * Experience with Kubernetes and Docker * Experience in at least one scripting language (Python, PowerShell, Go) Interview Process 🤝 * First stage: 30-minute intro, CV review, and values with Talent Partner * Second stage: 60 minute “Tech Chat” with Team Manager * Final stage: 75-minute Technical Task + 30 minute Interview with Head of Platform Engineering Diversity & Inclusion 🌈 We welcome, consider and encourage applications from anyone who shares our commitment to inclusivity. Join us in creating a space where authenticity thrives, and everyone can do their best work. Great Work Deserves Great Perks We try not to take ourselves too seriously (all the time) so we make sure our office is decked out with a pool table, arcade machine, beer tap, and a couple of office dogs thrown in for good measure. Check out our benefits: 🏥 Private Healthcare including dental and opticians services through Vitality ✈️ Worldwide travel insurance through Vitality 🎁 Anniversary Rewards (£250, £500, £750, 4-week fully paid sabbatical) 👛 Salary Sacrifice Pension Scheme up to 7% match 🏖️ 28 days holiday (plus bank holidays) 📖 Annual Learning and Wellbeing Budget 👪 Enhanced Parental Leave 🚲 Cycle to Work Scheme 🚂 Season Ticket Loan 💬 6 free therapy sessions per year 🐶 Dog Friendly Offices 🍫 Free drinks and snacks in our offices Check out more of our benefits, values and mission here. Other Info 👍Check out our ‘Top Tips’ for interviewing. ✔️Keep updated on new job opportunities by following us on Linkedin. 📧Email careers@capitalontap.com if you have any questions. Excited to work here? Apply! If you’d like to progress your career within our fast growing, profitable fintech then click apply and we will aim to get back to you within 3 working days (during busy periods this could take up to 5 working days.)
What do we do? Paddle offers digital product companies a completely different approach to their payment infrastructure. Instead of assembling and maintaining a complex stack of payments-related apps and services, we’re a Merchant of Record for our customers. That means we take away 100% of the pain of payment fragmentation. It’s faster, safer, cheaper, and, above all, way better. We’re backed by investors including KKR, FTV Capital, Kindred, Notion, and 83North and serve over 6000 software sellers in 245 territories globally. The Role: As a Site Reliability Engineer, you’ll be helping to drive our product and engineering department forward, ensuring reliability on different parts of the Paddle platform and helping our Engineers to work better and more efficiently. Paddle SRE team’s role is “Everything SRE”, with a focus on infrastructure, reliability standards, and practices. The SRE team is part of the Platform function. By following this model: * It’s easy to spot patterns and draw similarities between services and projects. * We act as a glue between disparate product teams, creating solutions out of distinct pieces of software. * Enable product engineers to use DevOps practices to maintain user-facing products without divergence in practice across the business. * Define production standards as code and work to smooth out any sharp edges to greatly simplify things for the product engineers running their services. You are empowered to use the right tech for the job. You’ll have the freedom to input into what technology and tooling are used and educate the rest of your colleagues accordingly. As an SRE, we want you to be a driving force of improving and automating how our product teams develop software at all stages of its lifecycle, which we strive to achieve with strong collaboration and communication with our fellow engineers. Tech Stack: * Go for our new services * PHP Laravel for our Classic system * Aurora MySQL and PostgreSQL for persistent data storage * Docker in production and local development * AWS ECS Fargate for our runtime * AWS SQS for asynchronous message queues * AWS EventBridge for our event bus * Redis for key/value store * Terraform for resource management * Honeycomb, SLOs and OpenTelemetry for our observability needs * Cloudflare for our firewall and DNS server What you'll do: * Develop and maintain tools to maximise engineering efficiency; such as but not limited to automating deployment infrastructure and database upgrades * Seek out processes that can be improved with automation and have internal Developer Experience as a main driver. Collaborate and enable engineers to do their jobs more efficiently, working with other engineers on a regular basis * Create, maintain and test our system disaster recovery process, including tooling to automate the process * You’ll be able to choose from a selection of AI tools to support day-to-day work (e.g. code generation, investigation, automation, and documentation), and we’ll back sensible experimentation with the right guardrails. * Handle production incidents, author blameless postmortems and enrich operational playbooks and runbooks * Monitoring, alerting, and SLO tracking; hands-on SRE work, not just DevOps-style monitoring * Run performance investigations (load testing, bottleneck analysis) and drive tuning across apps, data stores and AWS. * Own cost optimisation workstreams: right-sizing, autoscaling policies, workload scheduling, storage tiering, and identifying waste across ECS/Fargate, RDS/Aurora, SQS and observability. * Be an advocate of the GitOps methodology We'd love to hear from you if you: * A software development background, with experience shipping and operating production services, plus strong fundamentals in testing, code review, CI/CD, and debugging. * Have experience working across the AWS ecosystem, partnering closely with AWS Solution Architects and subject-matter experts to design, review, and operate production systems * A curiosity about AI and how it’s reshaping software development. * Collaborative, security-minded, and detail-oriented. We move quickly, so you’ll thrive if you enjoy a fast-paced environment and take pride in doing things the right way. * Knowledge of platform and ops concepts such as networking and Linux administration * Experience working with microservices and distributed systems at scale * Experience with monitoring tools: we use Opentelemetry, Honeycomb, Grafana, Pingdom and Incident.io Everyone is welcome at Paddle At Paddle, we’re committed to removing invisible barriers, both for our customers and within our own teams. We recognise and celebrate that every Paddler is unique and we welcome every individual perspective. As an inclusive employer we don’t care if, or where, you studied, what you look like or where you’re from. We’re more interested in your craft, curiosity, passion for learning and what you’ll add to our culture. We encourage you to apply even if you don’t match every part of the job ad, especially if you’re part of an underrepresented group. Please let us know if there’s anything we can do to better support you through the application process and in the workplace. We will do everything we can to support any accommodations needed. We’re committed to building a diverse team where everyone feels safe to be their authentic self. Let’s grow together. Our Values * Paddle Together - “None of us, is as smart as all of us” * Paddle Simply - “Simple can be harder than complex: you have to get your thinking clean to make it simple” * Paddle for others - “We can realise our wildest dreams, so long as we help enough other people to realise theirs” Why you’ll love working at Paddle We are a diverse, growing group of Paddlers across the globe who pride ourselves on our transparent, collaborative and respectful culture. We are a ‘digital-first’ company, which means you can work remotely, from one of our stylish hubs, or even a bit of both! We offer all team members unlimited holidays and 4 months paid family leave regardless of gender. We invest in learning and will help you with your personal development via constant exposure to new challenges, an annual learning fund, and regular internal and external training.