AI & DataRemoteFull-timeSenior (4+ yrs)
Senior AI / Applied Machine Learning Engineer
San Francisco, CA / Remote
$175,000 - $225,000 USD
About the Role & Mission
CoralSwift is expanding its Applied AI practice. In this role, you will build production-ready LLM pipelines, implement domain-specific fine-tuning, and optimize low-latency vector search systems for global enterprises.
What You Will Do
Build enterprise-grade RAG architectures with advanced chunking, reranking, and semantic search
Fine-tune and evaluate open-source foundation models (Llama, Mistral, DeepSeek) for specific domains
Deploy high-throughput inference endpoints using vLLM, TensorRT-LLM, and Triton
Implement automated LLM evaluation benchmarks, guardrails, and toxicity filtering
Collaborate with client stakeholders to identify high-ROI AI automation opportunities
Qualifications & Requirements
4+ years experience in machine learning and software engineering with strong Python/PyTorch skills
Hands-on experience deploying LLMs, vector databases (Qdrant, Pinecone, pgvector), and LangChain/LlamaIndex in production
Solid understanding of distributed GPU computing, CUDA fundamentals, and quantization techniques
Strong foundation in data structures, API design, and cloud deployments
Passion for keeping up with the rapid evolution of generative AI research
Compensation & Benefits
Competitive salary + equity package
Access to dedicated GPU clusters for R&D
Remote work freedom with top tier benefits
Conference attendance and paper publishing support
Ready to Join CoralSwift?
Submit your CV/resume. Our technical hiring team reviews every engineering application carefully.
Hiring Process
- CV / Technical Portfolio Review
- 30-Min Initial Scope & Culture Sync
- Architectural Deep Dive with Principal Engineer
- Executive Offer & Leadership Meet
