
Joko · Paris
At Joko, we help consumers shop smarter. Our mission is to revolutionize shopping, empowering people to find what they need, make informed decisions, and save m...
At Joko, we help consumers shop smarter. Our mission is to revolutionize shopping, empowering people to find what they need, make
informed decisions, and save money.
Founded in Paris, Joko is a tech company and certified B Corp with over 105 talents across Paris, Barcelona, and New York (and
beyond). More than 6 million users already save money every day at 10,000+ merchants with Joko.
From cashback and automatic coupons to price alerts and carbon tracking, we keep expanding our products to make shopping smarter.
We’re now building an AI-powered shopping assistant to help users find the best products by price, quality, and environmental
impact.
Having reached profitability in our core market, we’re now scaling globally, with a strong focus on the US.
It’s still day 1, come build the future of shopping with us!
The Data team at Joko is turning mountains of data into actionable insights for all the teams! We are part of the Operations
department led by our COO, and our mission is to empower the company to make informed decisions on solid, trustworthy foundations.
operate a modern data stack (Snowflake, dbt, Airbyte, and Metabase) and continuously raise the bar on scalability, reliability,
and performance.
for our stakeholders, seamlessly bridging data insights with the AI tools we use.
stakeholders, providing support and training on data tools, and building close relationships to ensure our solutions align
perfectly with their operational needs and evolve with them over time.
As a Data Engineer, you will own and scale our data infrastructure, working closely with Data Analysts, Product Managers, and
Engineers to build a reliable, secure, and high-performing data platform. This is a high-impact role with significant autonomy as
we invest seriously in data engineering.
Anticipate growth in data volume and complexity, and design systems with performance in mind.
across the company.
external tools) and orchestrate them efficiently.
quality and scalability of transformations across the stack.
and accuracy of data across the company.
secure, well-permissioned, and compliant.
collaboration, and long-term autonomy of the team.
within an orchestration tool.
and/or scaling data infrastructure.
streaming, etc.) in a startup or scale-up environment.
cloud environments (Snowflake, GCP, AWS).
in applications related to the data field.
(Some of the benefits listed below are available to full-time positions only)
At Joko, we believe that flexibility and trust are essential. Our work environment reflects this through:
elsewhere, we can provide access to a coworking space and a coworking budget.
months per year.
1. Intro call: Quick screening with the Hiring Manager or the Talent team.
2. Step 1 – Team interview (45 min): Conversation with two Joko team members (could include the Hiring Manager, people from the
team you’d join, or colleagues from other teams).
3. Step 2 – Role-specific assessments
think in real time. The exercise will be relevant to your role (e.g. analysis, strategy, or process design).
product thinking (with AI serving as a collaboration tool).
4. Step 3 – Leadership interview (45 min): Conversation with a SteerCo member and a Founder.
5. References: Up to 3 calls with former colleagues or managers.
☕ You may also be invited for coffee with team members to get a feel for our culture.
ABOUT US Harmattan AI is a next-generation defense prime building autonomous and scalable defense systems. Following the close of a $200M Series B, valuing the company at $1.4 billion, we are expanding our teams and capabilities to deliver mission-critical systems to allied forces. Our work is guided by clear values: building technologies with real-world impact, pursuing excellence in everything we do, setting ambitious goals, and taking on the hardest technical challenges. We operate in a demanding environment where rigor, ownership, and execution are expected. ABOUT THE ROLE As a Data Engineer on the Foundational team, you will serve as the "plumber" for deep learning, building the massive, high-performance data infrastructure required to power our foundational models. Based in Paris, you will manage terabytes—and eventually petabytes—of raw, unstructured, and noisy video data (EO and IR). Your mission is to ensure our ML engineers spend their time designing architectures, not waiting for data loaders or wrangling corrupted files. RESPONSIBILITIES * Multi-Modal Ingestion Pipeline: Build ETL/ELT pipelines to extract, decode, and store raw Electro-Optical (EO) and Infrared (IR) video from field logs into highly optimised formats like WebDataset, TFRecords, or Parquet. * Sensor Synchronisation & Alignment: Develop algorithms to programmatically synchronise EO and IR frames temporally and spatially to provide paired inputs for model training. * High-Throughput Data Loading: Architect storage-to-GPU pipelines to ensure multi-node training clusters maintain >90% GPU utilisation without I/O bottlenecks. * Distributed Processing: Write and optimise distributed data processing jobs using tools like Apache Spark, Ray, or Apache Beam to process thousands of hours of tactical video logs. * Data Quality & Versioning: Implement automated quality checks to filter corrupted or blank frames and maintain 100% reproducible training runs through robust versioning and lineage tracking. * Infrastructure Evaluation: Assess and implement advanced storage solutions (e.g., MinIO, S3 tiering) to manage growing datasets while optimising for cost and latency. CANDIDATE REQUIREMENTS * Educational Background: A BS or MS in Computer Science, Software Engineering, or Distributed Systems is highly preferred. Deep knowledge of operating systems, networking, and parallel computing is essential. * Technical Experience: 5-6+ years of experience building and maintaining terabyte-scale pipelines for unstructured data (video, images, or point clouds). * Performance Optimisation: Proven track record of maximising multi-node GPU utilisation and optimising data loaders for frameworks like PyTorch or JAX. * Tooling Expertise: Strong command of distributed computing tools (Spark, Ray, Beam) and ML data versioning tools (DVC, Apache Iceberg, or Pachyderm). * Adaptability & Ownership: A systems-thinker who thrives in a fast-paced startup environment and views messy data as an engineering problem to be solved via automation. * Commitment: 100% dedication to Harmattan AI’s mission of providing a defensive edge to allied nations through ethical, high-impact technology We look forward to hearing how you can help shape the future of autonomous defense systems at Harmattan AI.
Join Pigment: The AI Platform Redefining Business Planning Pigment is the AI-powered business planning and performance management platform built for agility and scale. We connect people, data, and processes in one intuitive, feature-rich solution, empowering every team—from Finance to HR—to build, adapt, and align strategic plans in real time. Founded in 2019, Pigment is one of the fastest-growing SaaS companies globally. Industry leaders like Unilever, Snowflake, Siemens, and DPD use Pigment daily to make more informed decisions and confidently navigate any scenario. With a team of 600+ across Paris, London, New York, Toronto, San Francisco and Austin, we've raised nearly $400M from top-tier investors and were named a Visionary in the 2024 Gartner® Magic Quadrant™ for Financial Planning Software. At Pigment, we take smart risks, celebrate bold ideas, and challenge the status quo—all while working as one team. If you're driven by innovation and ready to make an impact at scale, we’d love to hear from you. What You’ll Do: We’re looking for a talented engineer to join Pigment’s Growth team. We operate as a Product Team inside the Revenue organization, focused on building GTM systems rather than just looking for the next growth hack. Our small team designs and ships high-impact internal systems used by 100+ internal users every week. We build and maintain a centralized GTM data warehouse, internal web applications, orchestration workflows, and AI agents. These systems are treated like real products - with users (internal teams), adoption metrics, feedback loops, and continuous iteration. It’s also a unique opportunity to collaborate with amazing brains across Sales and Marketing as end users, RevOps and Data team as key stakeholders for everything systems and data governance, and even our R&D team on AI engineering. To give you an idea of projects we built, you can check this article here.
En quête d’un collectif pionnier, en mode Test & Learn permanent, qui donne réellement les moyens de vos ambitions ? Ne cherche plus ! Chez SFEIR, aucune évolution n’est imposée : tu pilotes ton propre parcours — qu'il soit vertical, horizontal ou transverse 💻 Ton Terrain de Jeu Stack : Google Cloud Platform (BigQuery, Cloud Storage, Dataflow, Cloud Composer), Python, SQL, DBT SFEIR AI : Un parcours dédié pour maîtriser l'IA générative et le coding agentique Communauté : 140 certifiés Google Cloud, 30 Google Authorized Trainers, 7 Google Developer Experts 🚀 Tes missions Concevoir et développer des pipelines de données robustes et scalables sur GCP Implémenter des solutions ETL/ELT avec DBT sur BigQuery Orchestrer des workflows data avec Cloud Composer Mettre en place du monitoring et de l'observabilité des plateformes data Participer à la modernisation des architectures data de nos clients 🎯 Ton profil Tu as au moins 3 ans d’expérience en tant que Data Engineer Tu as une expérience professionnelle confirmée sur Google Cloud Tu es passionné(e) par la Data et l’IA : tu fais de la veille et tu as des projets persos Tu souhaites relever des défis de haut vol tout en évoluant dans un environnement stimulant 💙 Pourquoi rejoindre SFEIR ? Chaque Sfeirian est accompagné(e) par un Engineering Manager Evolution vers des rôles de Lead Data ou Data Architect Formation continue avec accès à Google Cloud Skills Boost et certifications Participation aux évènements Google CloudCommunauté active avec plus de 50 événements annuels 🛠 Ton Pack Sfeirian Équipement au choix : MacBook Pro ou Dell Precision Prime Vacances : 10% de la valeur des congés pris Prime de participation Télétravail : mode hybride 5 semaines de congés + 8 à 12 RTT/an Carte ticket restaurant (11€/jour) Mutuelle et prévoyance Abonnement navigo ou remboursement de 70% pour un vélo électrique CSE : participation sports / culture / loisirs, chèques-cadeaux, activités (ski, golf, futsal, ...) Teamstarter : SFEIR verse 10€ par mois sur ta cagnotte personnelle pour financer des projets associatifs ou solidaires