About Aeolus
Aeolus Data Solutions is a boutique, founder-led data engineering practice for startups and scale-ups across North America. We design, build, and audit the production data pipelines and AI-ready data foundations that analytics and AI initiatives actually run on. Our thesis is simple: AI projects don’t fail on the model — they fail on the data pipeline underneath it.
We bring Big Tech engineering discipline — CI/CD, automated testing, infrastructure-as-code, versioned pipelines — to companies too lean to hire it in-house, and a senior practitioner works directly with every client. No account managers. No junior delivery bench. No handoffs.
The role
You’ll be the senior engineer on client engagements from first audit to production handoff. That means going deep on an unfamiliar stack in week one, finding the schema drift and silent pipeline failures nobody caught, and then building the foundation that holds. One month you’re hardening a RAG ingestion pipeline; the next you’re rebuilding a scale-up’s warehouse or embedding as their fractional data lead. You own the work and the client relationship directly — this is a senior seat, not a ticket queue.
What you’ll do
- Run data & AI-readiness audits. Assess a client’s pipelines, schemas, metadata, and warehouse; write the technical report and the prioritized remediation roadmap that follows.
- Build production data platforms. ELT/ETL pipelines, dbt models and tests, orchestration (Airflow/Dagster), warehouse and lakehouse builds (Snowflake/Databricks), and Terraform-managed infrastructure with CI/CD for data.
- Make data AI-ready. Semantic layers, metadata cataloging, clean lineage, and RAG-ready ingestion with validation before context ever reaches a model endpoint.
- Provide fractional data leadership. Embed with a client’s team a few days a week — set engineering standards, standardize KPIs, guide platform decisions, and mentor their engineers.
- Own the relationship. Scope work, explain tradeoffs to a CTO, and deliver — directly, with no manager relaying messages in between.
What we’re looking for
- 5+ years building and operating production data pipelines. You’ve owned data infrastructure in production, not just prototypes.
- Depth in the modern data stack: dbt, at least one cloud warehouse/lakehouse (Snowflake, Databricks, or BigQuery), an orchestrator (Airflow, Dagster, or Prefect), a major cloud (AWS/GCP/Azure), and strong Python + SQL.
- Software-engineering discipline applied to data: version control, automated testing (dbt tests, Great Expectations), CI/CD, and infrastructure-as-code (Terraform).
- You ramp fast on messy, unfamiliar systems. Consulting means a new stack and real-world data quality problems every engagement — you find your footing quickly.
- You’re client-facing. You can explain a technical tradeoff to a non-engineer, write a clear audit report, and be the calm senior voice in the room.
- You’re self-directed. Small team, high ownership, no one assigning you tickets.
Nice to have
- RAG / LLM data pipeline experience — chunking, embeddings, vector databases, retrieval quality.
- Warehouse cost optimization / FinOps (query tuning, clustering, spend containment).
- Experience at Big Tech scale or high-growth data teams.
- Prior consulting, agency, or fractional/embedded experience.
- Relevant certifications (dbt, SnowPro, Databricks).
How we work
- Remote-first, async-friendly, working across North American time zones.
- Direct founder collaboration — flat, fast, no bureaucracy.
- Variety by design — many stacks, industries, and problems; greenfield builds and rescue jobs.
- Senior-only shop — your name is on the work, and the quality bar is the whole product.
Compensation & logistics
- Employment type: Full-time, or contract if that suits you better.
- Location / work authorization: remote within Canada or the United States; you must be authorized to work where you’re based. We’re incorporated in both British Columbia and California.
- Compensation:
- United States: US$135,000–US$175,000 / year
- Canada: CA$95,000–CA$125,000 / year
- Start: Flexible — we move at the pace of finding the right person.
How to apply
Email [email protected] with a short note on a data platform you built or rescued — what was broken, what you did, and how you knew it worked — plus your resume, GitHub/portfolio, and (if you have it) anything you’ve shipped with dbt, Snowflake, Databricks, or Airflow. No cover letter needed. We reply to every candidate.
Sound like you?