Remote: Yes
Willing to relocate: Yes (open to discussion)
Technologies: AI/LLM: Whisper, Deepgram, ElevenLabs, Gemini 2.5/3, Gemini Pro Vision, Vertex AI (Model Garden, Vector Search, Workbench, Pipelines), LangGraph, LlamaIndex, RAG, multimodal & multilingual embeddings, FAISS/pgvector, Redis vector serving, ONNX, Triton, tool-calling agents, real-time voice agents (WebRTC, LiveKit) Backend: Python (FastAPI, Django), Node.js, Nest.js, Go, Java/Spring Boot, .NET Core, WebSockets, microservices, event-driven systems, Redis, PostgreSQL Frontend: React, Next.js, Angular, TypeScript, Redux Cloud/DevOps: GCP (Cloud Run, GKE, BigQuery, Pub/Sub, Vertex AI), AWS (EC2, S3, Lambda, RDS), Azure, Docker, Terraform, GitHub Actions, CI/CD Other: FFmpeg, OpenCV, Databricks, Delta Lake, hybrid search, semantic ranking
Résumé/CV: https://flowcv.com/resume/4dicl4tsnk2h
Email: davidmryland@outlook.com
About: Senior AI Engineer & Full-Stack Engineer with deep experience shipping production, latency-sensitive AI systems.
Most recently at Striveworks (Austin) I designed and launched a real-time AI voice agent platform on GCP using LiveKit/WebRTC, Whisper/Deepgram STT, ElevenLabs TTS, Gemini embeddings, and LangGraph multi-agent orchestration — delivering sub-second response times with interrupt handling and barge-in.
Previously at Zillow I built multimodal, multilingual embedding retrieval systems serving 20M+ users at 10–100 ms P99 / 10k+ QPS, plus conversational search with RAG, query rewriting, and observability (LangSmith/Langfuse). Earlier roles at Banner Health and GoDaddy covered large-scale retrieval, clinical decision support, and full-stack platform work.
I focus on end-to-end ownership: audio ingestion → streaming inference → agent orchestration → low-latency serving → observability. Looking for senior AI/Full-Stack roles where I can help ship production AI products quickly.