G
Principal Data Scientist & Engineer
G-P
Contract type
Ongoing
Work mode
100% remote
Experience
Senior · 7+ years
Job description
Key details
- Work across data science, data engineering, and analytics to build reliable, scalable data pipelines and models
- Use SQL and Python daily for complex queries, data modeling, scripting, and analysis
- Apply LLM techniques such as fine-tuning, prompt engineering, embeddings, RAG, and evaluation
- Build traditional ML solutions including classification, regression, clustering, NLP, and feature engineering
- Connect analysis to product decisions and present findings clearly to non-technical stakeholders
- Instrument and measure new features from scratch, including event taxonomy, dashboards, and self-serve reporting
- Company mission
- Information not specified
Primary stack
Core technologies
Python (Programming Language)SQL
Benefits
- Generous paid parental leave
- Flexible time off
- Spending accounts
- Medical insurance
- Dental insurance
- Vision insurance
- Sabbatical after 5 years
- Annual bonus for non-sales roles, dependent on individual and company performance
- Commission structure for sales roles in addition to base salary
Requirements & details
- 7+ years across data science, data engineering, and analytics
- Strong SQL and Python skills for complex queries, data modeling, scripting, and analysis
- Databricks or equivalent modern data platform experience (Snowflake, BigQuery)
- LLM experience including fine-tuning, prompt engineering, embeddings, RAG, and evaluation
- Traditional ML depth in classification, regression, clustering, NLP, and feature engineering
- Product mindset to filter signal from noise and connect analysis to product decisions
- Pipeline engineering experience building reliable, scalable data pipelines
- Clear communicator able to present findings to non-technical stakeholders
- Preferred: early-stage startup or founding data hire experience
- Preferred: built product analytics from scratch including instrumentation, event taxonomy, dashboards, and self-serve reporting
- Preferred: legal or HR domain experience
- Preferred: LLM evaluation and observability experience (tracing, scoring, drift detection)
- Preferred: familiarity with dbt, Airflow/Dagster, Spark, or similar orchestration and transformation tools
- Location: Remote, with additional information for applicants in California or Philadelphia, Pennsylvania
- Python, SQL, Databricks, Snowflake, BigQuery, LLMs, dbt, Airflow, Dagster, Spark
- Python (Programming Language)
- SQL
