Senior Data Scientist
Job Description
At Archer, the AI Products team brings together the resources and momentum of a real software organization inside a company of aerospace engineers. Products are still early, and you will have the chance to help shape what gets built and how it scales as the team grows.
As a Senior Data Scientist based in San Jose, CA (onsite), you will work in Archer’s AI Products org, bridging raw, messy aviation data and intelligence that supports their AI platform. The work spans deriving signals from telemetry and audio, building multi-modality models, and taking those models into production streaming pipelines.
What you’ll do
- Derive new signals from high-throughput, real-time data streams, inferring complex physical and operational states that are not directly observable in raw telemetry.
- Build, evaluate, and iterate on models across diverse modalities, including dense time-series, spatial data, and highly noisy, unstructured sequential data.
- Design evaluation strategies for problems with little or no labeled ground truth using weak supervision and programmatic labeling, and clearly defend those definitions to engineers and domain experts.
- Partner with backend and platform engineers to move models from offline analysis into production streaming pipelines, and own model behavior after deployment.
- Work closely with aviation subject-matter experts to convert operational knowledge into features, labels, and evaluation criteria.
What you’ll bring
- BS/MS/PhD in Computer Science, Statistics, Applied Mathematics, Operations Research, Aerospace Engineering, or a related quantitative field.
- 2+ years of applied data science or machine learning experience, including experience with models that have run in production.
- Strong Python and strong SQL on large analytical datasets.
- Depth in time-series and sequence modeling.
- Experience with sparse, incomplete, or unlabeled data, including weak supervision, programmatic labeling, and creating evaluation sets where no ground truth exists.
- Experience designing data validation and drift detection for continuously arriving data.
- Ability to leverage AI assistants to increase velocity while maintaining full ownership of every model, feature, and analysis shipped.
- Strong communication and collaboration skills, including explaining methodology and limitations to a non-specialist audience.
Tools you’ll use
- Python, SQL
- Apache Pulsar, Kafka
- Apache Flink
Bonus qualifications
- Prior experience or strong interest in aerospace, aviation, or high-throughput tracking systems, including familiarity with flight telemetry, ATC phraseology, or airspace procedures.
- Familiarity with geographic/spatial data processing, including projections, spatial indexing, and great-circle geometry.
- Awareness of distributed messaging and stream-processing systems such as Apache Pulsar, Kafka, and Apache Flink, plus columnar or lakehouse analytical storage.
Salary: USD 145,000 - 180,000 per year.