
dida · Berlin
WARUM DU DICH FÜR UNS ENTSCHEIDEN SOLLTEST? Du arbeitest in einem interdisziplinären Team von Mitarbeitern mit einem soliden Hintergrund in Mathematik und Stat...
Du arbeitest in einem interdisziplinären Team von Mitarbeitern mit einem soliden Hintergrund in Mathematik und Statistik. Wir
bieten flexible Arbeitszeiten (Voll- und Teilzeit) und haben ein schönes Büro mit gutem Kaffee in Berlin Schöneberg. Wir
bevorzugen ein hybrides Arbeitsmodell, sind aber auch offen für mobiles Arbeiten. Wir glauben an die Wissenschaft und unterstützen
Sie bei der Veröffentlichung Ihrer Forschungsergebnisse.
Hier sind kurze Beschreibungen von einigen Projekten, an denen wir gearbeitet haben.
Schätzung der Anzahl von Sonnenkollektoren, die auf ein Dach passen (Computer Vision):
Anhand eines Satellitenbildes und eines Bodenbildes eines Hauses sollen bestimmte Elemente eines Daches (einschließlich
Hindernisse, Dachgauben usw.) automatisch erkannt werden, um herauszufinden, wie viele Sonnenkollektoren darauf passen. Dies
beinhaltet die Ableitung von 3D-Informationen aus 2D-Bildern, um die Dachneigung zu ermitteln.
Erkennen, Klassifizieren und Vorschlagen der rechtlichen Wirksamkeit von Textabschnitten (NLP):
Automatisches Durchgehen von Tausenden von Rechtsdokumenten mit dem Ziel, bestimmte Absätze zu klassifizieren und ihre rechtliche
Wirksamkeit zu überprüfen. Dies beinhaltet die Konvertierung von Scans in Text, die Entwicklung eines Etikettierungsschemas
(Problemmodellierung) und die automatische Erkennung verschiedener Absätze, bevor die Inferenzaufgabe in Angriff genommen wird.
UNSERE MISSION dida ist eine der führenden Machine Learning Agenturen mit Aufträgen und Projekten in der Wissenschaft, der Industrie und dem öffentlichen Sektor. Unser Team, das überwiegend aus Machine Learning Scientists besteht, trägt zu neuen Erkenntnissen und Innovationen in der KI-Forschung bei. Wir setzen uns für eine bewusste Entwicklung und Integration von Machine Learning in allen Gesellschaftsbereichen ein und begreifen uns als Brücke zwischen Forschung und Industrie. Werde Teil unseres Teams und unterstütze uns dabei! Deine Aufgaben: * Linux-Administration: Patch-Management, Monitoring, Benutzerverwaltung sowie Analyse und Troubleshooting unserer Linux-Systeme. * Service-Operations: Planung, Bereitstellung und Konfiguration interner Dienste inklusive sauberer technischer Dokumentation. * Hardware-Betreuung: Unterstützung bei der physischen Betreuung von Server- und Client-Hardware, inklusive Komponententausch. * CI/CD & Automatisierung: Mitwirkung beim Aufbau und der Pflege von CI/CD-Pipelines sowie bei der Automatisierung wiederkehrender Ops-Tätigkeiten. * Frontend für interne Tools: Punktuelle Weiterentwicklung kleiner Frontend-Anwendungen für interne Prozess-Tools DEIN PROFIL * Du studierst ein technisches Fach (Mathe, Informatik, Wirtschaftsinformatik o. Ä.). * Du hast solide Kenntnisse im Umgang mit Linux (Shell, Paketmanagement, Systemadministration-Grundlagen). * Du hast erste Erfahrung mit Frontend-Entwicklung (HTML, CSS, JavaScript). * Du besitzt Grundverständnis für Netzwerke und gängige Protokolle. * Du hast Lust, dich in neue Themen einzuarbeiten und eigenständig Probleme zu lösen. * Du lebst in Berlin und hast noch mindestens 1,5 Jahre Studienzeit vor dir. WARUM DU DICH FÜR UNS ENTSCHEIDEN SOLLTEST? Werde Teil eines hochqualifizierten Teams, das Machine Learning Modelle für die KI- Forschung und Anwendung entwickelt: * Wir haben gemeinsam mit dem GFZ (Helmholtz-Zentrum für Geoforschung) und der ESA (European Space Agency) ein Modell entwickelt, das Satellitendaten auswertet, um die moderne Landwirtschaft zu fördern. * Wir haben die Erkennung bestimmter Wolkenstrukturen für den Deutschen Wetterdienst (DWD) automatisiert. * Wir unterstützen die Deutsche Bahn bei der Erkennung von ungewöhnlichen Objekten auf Gleisen. * Und wir haben den „AI for Earth Award“ gewonnen – für eine unserer Lösung, die illegale Minen auf Satellitenbildern im Regenwald erkennt. Was dich sonst noch erwartet: * Deep Dive in ML: Tauche ein in die Welt des Machine Learning und entdecke mit uns, was damit alles möglich ist. * Schnelles Lernen: Übernimm früh Verantwortung in einem jungen Unternehmen und entwickle dich schnell weiter. * Mitwirken: Bringe ein, was du im Laufe deines Studiums lernst und gestalte dida mit. * Teamkultur: Arbeite in einem offenen, diversen und internationalen Team. * Flexibles Arbeiten: Wir bieten hybride Arbeitsmodelle mit Remote-Optionen – abgestimmt auf dein Studium. * Gute Lage: Unser Büro liegt zentral in Berlin-Kreuzberg – mit super ÖPNV-Anbindung und weiteren Vorteilen eines City-Office.
WHO WE ARE Foundation models transformed text and images. Structured data - the largest and most consequential data format in the world - stayed untouched. Tables run every clinical trial, every financial model, every scientific experiment, every business decision, and no one had built a foundation model that truly understood them. Until now. What LLMs did for language, we're doing for tables. The next modality shift in AI is happening, and we're hiring the team that makes it. Momentum. We pioneered tabular foundation models and are now the world-leading organization in structured-data ML. Our TabPFN v2 model was published as a Nature cover story and set a new state of the art for tabular machine learning. Since release we've scaled model capabilities 20x+, passed 3.5M+ downloads and 7,500+ GitHub stars, and are seeing accelerating adoption across research and industry - from detecting lung disease with Oxford Cancer Analytics to preventing train failures with Hitachi to improving clinical-trial decisions with BostonGene. The hardest work is ahead. We're scaling tabular foundation models to millions of rows, thousands of features, real-time inference, and entirely new data modalities, while building the infrastructure to run them in production across some of the most demanding industries on earth. These are open problems no one else is working on at this level. Our team. We're a small, highly selective team of 30+ engineers, researchers, and GTM specialists, with backgrounds spanning Google, Apple, Amazon, DeepMind, Meta, Microsoft Research, G-Research, Jane Street, Goldman Sachs, and CERN. We're led by Frank Hutter, Noah Hollmann, and Sauraj Gambhir, and advised by world-leading AI researchers including Bernhard Schölkopf and Turing Award winner Yann LeCun. We ship fast, do top-tier research, and hold each other to an extremely high bar. What's next. In 2025 we raised €9m pre-seed led by Balderton Capital, backed by leaders from Hugging Face, DeepMind, and Black Forest Labs. The next phase of growth is here, which makes this an ideal time to join. ABOUT THE ROLE We spend tens of millions per year on GPU compute to train tabular foundation models. That's not a target, it's what we're running today, and it's growing. The person who owns this infrastructure makes decisions worth millions of dollars: cluster architecture, scheduling efficiency, provider strategy, hardware selection. A wrong call costs six figures. Today we run Slurm on GCP across multiple clusters. We're scaling to multi-cluster, multi-provider infrastructure and evaluating new hardware generations as they come online. You own the full stack, from cluster operations and cost optimization to distributed training performance and the tooling layer that keeps researchers moving fast. You work directly with the research team and understand what they're doing well enough to make infrastructure decisions that actually help them. And this isn't a pure support role. We operate an open environment. If you've got the next SOTA tabular architecture up your sleeve, go ahead and train it. What you'll work on: * Own and evolve multi-cluster GPU infrastructure. Slurm on GCP today, multi-provider and new hardware tomorrow. Architecture, scheduling, reliability, cost optimization * Drive GPU utilization and training throughput: profiling, memory optimization, communication bottlenecks, systems-level debugging of distributed training across large runs * Architect the next generation of our infrastructure: multi-cluster orchestration, new GPU generations, provider diversification, capacity planning against growing compute demands * Build the developer productivity layer: CI pipelines, experiment tracking, model registry, data processing, and internal tooling that keeps research iteration speed high * Own the compute budget. You understand cost per FLOP across providers and hardware, and you hate wasted compute Tech stack: Slurm, GCP, Docker, wandb, GitHub Actions, uv, PyTorch, Triton You may be a good fit if you have: * 5+ years building and operating production GPU infrastructure or distributed training systems at scale. At a major AI lab, a well-funded ML startup, or an HPC environment * Deep hands-on experience with Slurm and cluster management. You've debugged scheduling failures, optimized utilization across multi-tenant GPU workloads, and operated infrastructure where downtime has real cost * Expert-level systems thinking: memory bandwidth, GPU profiling. You reason about hardware, not configs * Strong Python and genuine fluency with PyTorch internals. Enough to profile a training run and tell whether the bottleneck is data loading, communication, or compute * Track record of making infrastructure decisions that measurably improved training throughput or cost efficiency * Strong AI tooling skills. You use Claude Code, Cursor, or similar fluently to move fast without sacrificing quality Bonus: * Experience operating at tens-of-millions-scale GPU spend * Multi-cloud or hybrid HPC/cloud infrastructure experience * Triton, CUDA, or custom kernel experience * Experience scaling from single cluster to multi-cluster orchestration * Background building experiment tracking, model registry, or ML pipeline tooling Life at Prior Labs We're a small, ambitious team solving one of the hardest problems in AI, and we're just getting started. You'll work closely with world-class researchers and builders who care deeply about the quality of their craft, the impact of their work, and the people they work with. We move fast, we think rigorously, and we take the time to do things right. If you're excited by hard problems, motivated by real-world impact, and want to be part of building something that matters, we'd love to hear from you. We're building our teams in Berlin, Freiburg, and New York and we believe that when you're working on something as hard and exciting as TabPFN, being in the same room matters. Most of our roles are based in one of our offices but great people come from everywhere, and in exceptional cases we're open to remote. This usually involves frequent travel to one of our offices and the whole company comes together regularly for offsites to think, build, and celebrate together. Our Commitments We believe the best products and teams come from a wide range of perspectives, experiences, and backgrounds. That's why we welcome applications from people of all identities and walks of life, especially anyone who's ever felt discouraged by "not checking every box." We're committed to creating a safe, inclusive environment and providing equal opportunities regardless of gender, sexual orientation, origin, disability, or any other trait that makes you who you are. We care about how your data is handled. Read our Recruiting Privacy Notice to see exactly what we collect, why, and how long we keep it.
ABOUT THE OPPORTUNITY Become part of the AI revolution at Contentful where we build real world AI solutions for thousands of active customers. At Contentful you will help imagine, drive, and build ML solutions which produce immediate customer impact by helping them get more done than they ever imagined. We are looking for an excellent Senior Software Engineer - Machine Learning (f/m/d) who understands and has experience in this field and current Generative AI trends. If you're passionate about pushing the boundaries of what's achievable with ML products on a large scale content management system, we want you on our team. WHAT TO EXPECT? * Measure and ensure the high-quality output of ML workloads for our enterprise customers. * Design, build, and measure production software in the Contentful Platform. * Conduct focused research and testing using tools like Jupyter Notebook. * Optimize generative AI products for accuracy, speed, and scalability. * Integrate ML cloud solutions and adapt them to our large-scale use cases. * Integrate prompt engineering and fine-tuning for customer use cases. * Develop software that scales to meet the demands of high-load customer workloads. * Lead design reviews with peers and stakeholders to decide amongst available technologies. * Drive the ML education in the team and establish a high standard of quality & ownership WHAT YOU NEED TO BE SUCCESSFUL? * Proven experience as a technical leader in a product software development environment * Proven Track Record in Model Development, including successfully developing and deploying LLM-based models for various NLP tasks in real-world applications with production workloads * Proven ability to work backwards from customer needs to deliver ML-based features that meet those needs * Ability to organize and prioritize competing workloads * Familiarity with distributed systems and cloud services (e.g., AWS, Azure, GCP) * Solid understanding of machine learning principles * Experience with container frameworks such as Docker or Kubernetes * Constructive Problem-Solving: As a natural problem-solver, you bring forward ideas that lead to practical solutions and contribute to product growth. * Up-To-Date on The Latest Updates on LLM Development & Research: From RAGs to Open-Source models over Agent Architectures – you have an inherent drive to stay up-to-date on what is the latest and greatest Preferred: * Background includes machine learning and AI implementation * You can translate a research paper into a PoC if the code does not exist WHAT’S IN IT FOR YOU? * Join an ambitious tech company reshaping the way people build digital experiences * Full-time employees receive Stock Options for the opportunity to share in the success of our company * Fertility and family building benefits, including a lifetime reimbursable wallet to support your growing family. * We value Work-Life balance and You Time! A generous amount of paid time off, including vacation days, sick days, education days, compassion days for loss, and volunteer days * Time off to care for and focus on your growing family * Use your personal annual education budget to improve your skills and grow in your career * Enjoy a full range of virtual and in-person events, including workshops, guest speakers, and fun team activities, supporting learning and networking exchange beyond the usual work duties * An annual wellbeing stipend to care for your physical, financial, or emotional health * A monthly communication phone/internet stipend and phone hardware upgrade reimbursement. * New hire office equipment stipend for hybrid or distributed employees. Get the gear you need to work at your best. WHO ARE WE? Contentful is a leading digital experience platform that helps modern businesses meet the growing demand for engaging, personalized content at scale. By blending composability with native AI capabilities, Contentful enables dynamic personalization, automated content delivery, and real-time experimentation, powering next-generation digital experiences across brands, regions, and channels for more than 4,200 organizations worldwide. More than 700 people from more than 70 nations contribute their energy and creativity to Contentful, working from hubs in Berlin, Denver, San Francisco, London, New York, and distributed worldwide. EVERYONE IS WELCOME HERE! “Everyone is welcome here” is a celebrated component of our culture. At Contentful, we strive to create an inclusive environment that empowers our employees. We believe that our products and services benefit from our diverse backgrounds and experiences, and we are proud to be an equal opportunity employer. All qualified applications will receive consideration for employment without regard to race, color, national origin, religion, sexual orientation, gender, gender identity, age, physical [dis]ability, or length of time spent unemployed. We invite you to apply and join us! If you need reasonable accommodations at any point during the application or interview process, please let your recruiting coordinator know. Please be aware of scammers who may fraudulently allege to be from Contentful. These types of fraud can be carried out through copycat websites, fake email addresses claiming to be from our company, or social media. We do not ask for your personal information, such as bank account numbers, identification numbers, etc, through social media or chat-based apps, nor do we request or send money for the purchase of business equipment. If you suspect fraud, please report it to your local authorities, as well as reach out to us at security-esk@contentful.com with any information you may have. By clicking “Apply for this job,” I acknowledge that I have read the “Contentful’s Candidate Privacy Notice” and hereby consent to the collection, processing, use, and storage of my personal information as described therein.