Founding Full-Stack AI Engineer
- Built COROS's 24/7 AI coaching platform from zero — a microservice backbone serving LLM-powered coaching to professionals, founders, and leaders.
- Cut AI response latency 78% (27–32s → 6–10s) via semantic search optimization.
- Engineered a semantic memory system with topic isolation and time-based decay, plus zero-downtime LLM failover across OpenAI, DeepSeek, Together AI, and Llama.
- Shipped a production ElevenLabs voice clone and AI guardrails that block harmful coaching advice before delivery.
- Shipped the iOS + Android app end-to-end — full CI/CD and store release lifecycle — while cutting GCP spend with request-based scaling.






