Job Opportunities API

The Public Ledger of Openings

← Back to the ledger

Applied AI/ML Engineer

Overtone
CompanyOvertone
CategoryUncategorised
LocationNew York City
RemoteOn-site (inferred)
EmploymentNot stated
LevelNot stated
SalaryUSD 200k–300k
Posted29 Jul 2026
Last verified31 Jul 2026
SourceEmployer career page (ashby)
Applications are handled by the employer, not by us.Apply on the employer's site →
Description
YOUR ROLE As a founding AI Engineer at Overtone, you will build and own the AI systems that power our core product experience, from conversational intelligence to matchmaking insight generation. You'll design prompting, evaluation, and model infrastructure that makes our AI reliable, trustworthy, and continuously improving. This is a hands-on, early-stage role: you'll run experiments, ship infrastructure, and build the tools that give the team real visibility into model behavior and quality. What You’ll Do - LLM systems & prompt architecture: Design and maintain the prompts, structured outputs, and orchestration systems that power Overtone's AI-driven features. These include both text-based LLM output, as well as transcripts generated from our code voice AI expereince. - Evaluation frameworks: Build evaluation infrastructure to measure AI quality at scale using tools such as Braintrust, Langfuse, TensorZero, and Promptfoo. You'll work with our team to make sure LLM judges are aligned. Evals should cover both generated text and voice transcripts, including transcription fidelity and response latency. - Model experimentation: Run structured experiments across models, prompts, and configurations to optimize quality, cost, and latency. - AI observability & tooling: Build internal tools (dashboards, admin interfaces, debugging workflows) that give the team visibility into system behavior and model performance. - Matching & ML systems: Apply data science and lightweight modeling to improve match quality, identifying when traditional ML models or ranking systems outperform prompt-based approaches. - Model serving & production inference: Own the model serving layer: deploying models, managing inference infrastructure (API providers and self-hosted), model versioning, and cost/latency optimization in production. Work with the backend lead to define clean service contracts between the AI layer and the rest of the system. - Model development & fine-tuning: Identify opportunities to fine-tune open-source models using techniques such as SFT, DPO, and RLHF for cost and latency improvements. - Memory & personalization systems: Design and maintain vector databases and memory architectures that enable personalization and context-aware experiences. - Responsible AI systems: Ensure systems are fair, unbiased, and auditable, building evaluation pipelines and human-in-the-loop processes where appropriate. - Cross-functional collaboration: Work closely with product, engineering, and research to bring AI capabilities into real user-facing features with scalability and reliability in mind. Who You Are You likely have: - 5+ years of engineering experience, with meaningful time spent building LLM-powered products in production - Strong prompt engineering and evaluation instincts. You think in terms of evals, not vibes - Experience building evaluation pipelines for LLM systems - Familiarity with vector databases and retrieval-augmented architectures - Experience with fine-tuning techniques such as SFT, DPO, or RLHF - Strong understanding of when ML models outperform prompt-based systems - Experience working in early-stage environments with high ownership - Experience with Python, SQLAlchemy, FastAPI, PostgreSQL Nice to have: - Experience with recommender systems or ranking systems - Experience with MLOps and model lifecycle infrastructure - Experience building AI evaluation frameworks or model observability systems - Experience building internal developer tooling (React/TypeScript + Python) - Experience with production voice pipelines: streaming ASR (Deepgram, AssemblyAI, Whisper), TTS (ElevenLabs, Cartesia, PlayHT), or realtime voice agent frameworks - Experience with latency-sensitive or streaming AI surfaces, including WebRTC/LiveKit and evaluating voice quality (WER, naturalness, perceived responsiveness) COMPENSATION & BENEFITS Salary of $200k to $300k. This position will receive mea
HOUSE AD991,236 openings. Erioun finds yours.Scored against your own profile, every hour.Try the radar →