| Location | Santa Clara, CA |
| Salary | 130,000 - 220,000 USD, yearly |
| Commitment | Full-time |
| Role | Data Infrastructure |
| Remote | 🏢🌴Hybrid |
| First listed | In the last 2 weeks |
Get help on your job search
Need help in your climate job search? Dive deep into climate with Terra.do’s 12-week climate bootcamp course.
Terra.do has partnered with ClimateTechList to give ClimateTechList users a 15% discount for its flagship Climate Change: Learning for Action program.
Job Description
Finding the right data is central to improving autonomous-driving models. Among petabytes of fleet data, you will develop methods that identify and rank the most valuable moments for training and evaluation, then turn those methods into reliable tools that autonomy and ML engineers use to search, review, and curate datasets. You will work at the intersection of applied machine learning, information retrieval, large-scale data processing, and product engineering. We welcome candidates with ML or data-mining foundations who are excited to grow across scalable systems and the product stack.
We are open to candidates at either the Software Engineer or Senior Software Engineer level. Level will be determined by experience, technical depth, scope of ownership, and demonstrated impact. You do not need experience with every technology in our stack; we value strong fundamentals, ownership, and the ability to learn.
Responsibilities:
- Develop and evaluate mining, retrieval, and ranking methods using signals such as model confidence, disagreement, embeddings, anomalies, temporal behavior, and learned representations
- Build and evolve semantic image/video/scenario search, including text-to-image/video and image-to-image or video-to-video retrieval, vector search, metadata and temporal or spatial filters, task-specific ranking, and search quality, freshness, latency, and reliability
- Build and operate distributed mining, inference, and indexing pipelines over fleet-scale imagery, video, time-series, and autonomy-system data, including GPU batch inference, embedding generation, reproducible candidate datasets, and reliable index refreshes
- Design and ship mining products end to end: Python APIs and services, relational data models, asynchronous jobs, modern TypeScript/React search and review experiences, deployment, access control, testing, observability, and production reliability
- Ensure that your work is performed in accordance with the company’s Quality Management System (QMS) requirements and contribute to continuous improvement efforts
Plus number of job openings over time by month
ClimateTechList is the web's largest aggregator of climate, clean tech, renewable energy & green jobs. Contact us if you'd like to use partner or use our current or historical jobs data in any way.
Apply to Job
👉 Please mention that you found the job on ClimateTechList, this helps us get more climate tech companies listed here, thanks!
Get a referral to Plus
If possible, try to get a warm intro/referral to Plus before applying! Do a LinkedIn search to see who you may know at the company. See this LinkedIn post from Steven for more details on this tactic.
Join ClimateTechList Talent Collective
Want to be matched with companies directly? Apply to the talent collective.
Here's how it works:
You submit an application
We'll share your profile with climate tech companies potentially interested in chatting with you
We'll reach out if there's a company interested in talking to you.
No spam. Unsubscribe any time.
Join ClimateTechList Talent Collective
Want to be matched with companies directly? Apply to the talent collective.
Here's how it works:
You submit an application
We'll share your profile with climate tech companies potentially interested in chatting with you
We'll reach out if there's a company interested in talking to you.