
About Modash
Modash gives brands the tools to work with the right content creators and helps creators earn a living doing what they love. Behind the scenes, the Data Insights team is building the intelligence layer that turns raw social media signals into trusted, customer-facing data products — with reliable access, quality, and freshness at scale.
We’re looking for a hardened Senior Product Data Engineer to help us scale these systems end-to-end, raise our quality bar, and accelerate how quickly we turn messy public data into consistent, valuable insights customers can build on.
What Your Day-to-Day Will Look Like
We’re not a service function — Data Search & Data Insights are core product capabilities at Modash, building products for customers to use. You’ll own high-impact projects end-to-end, from idea to launch.
- Start your day with a short standup
- Heads-down focus time to plan, build, iterate, and launch
- Minimal meetings — maximum ownership
Impactful Projects
- Creating an understanding of creator location, age, and interests at scale
- Creating systems to extract collaborations between creators and brands from raw social data
- Shaping the future of AI-assisted search, exploring how LLMs and embeddings can enhance search and recommendations
You won’t be patching pipelines — you’ll be creating data products from scratch that directly impact customers.
The Data Team
At Modash, the Data Insights team is a core part of the product. You’ll join a growing group of data and backend engineers working across three closely aligned teams:
- Data Insights — Builds creator and brand-level insight products and APIs (e.g., collaborations, reports, dictionaries, contacts, audience overlap).
- Data Search — Owns our search products (including AI Search) end-to-end.
- Data Core — Responsible for raw data collection and the foundations of our data platform.
We value autonomy while working closely through pair programming, fast feedback loops, and shared wins.
Our Tech Stack
- Compute & Cloud: AWS, GCP with Pulumi (IaC), PySpark on AWS EMR
- AI & Orchestration: GCP Vertex Batch API, Airflow
- Persistence: Iceberg, Aurora (Postgres), S3, Glue, Kinesis, Lambda, ECS, Athena
- Tools: Slack, GitHub, Linear, Notion, Cursor
Skillset We’re Looking For
- Strong knowledge of Spark (Scala, Databricks, or PySpark; PySpark preferred)
- Proven track record with ETL/ELT pipelines and large-scale data processing
- Comfortable working with unstructured data
- Experience with workflow orchestration tools like Airflow or AWS Step Functions
- Familiarity with the AWS ecosystem (Glue, EMR, etc.)
- Experience shipping full features from idea to production
- Based in Europe with significant working-hours overlap with EET
- Hands-on experience building agentic / LLM-powered features in production
- Practical understanding of trade-offs between LLMs (cost, latency, capability)
Bonus Points
- Worked with AI/ML tools or LLMs
- Familiar with the GCP stack (especially Vertex AI)
- Experience with lakehouse formats like Apache Iceberg
- Used Pulumi or Terraform for IaC
- Familiar with Node.js and TypeScript
- Understand AWS cost mechanics
What We Offer
- Remote-first: Work from anywhere in Europe with regular team offsites
- Compensation: Base salary range of 90,000€ - 140,000€ plus a significant stock option package
- Time Off: Unlimited paid vacation
- Flexibility: Async-friendly culture with low bureaucracy
- Growth: Personal development support for courses, books, or conferences
Timezone overlap
UTC+0–+3
Culture
Async-friendly
Benefits
Open to
Europe
Sign in to track applications and earn points.