AI Engineer
CodeRound is hiring for this role for a VC-backed startup.
- AI Engineer
- Founding Engineer
- LLMs
- RAG
- Vector Databases
- Transformers
- PyTorch
- Model Optimization
- MLOps
- Production ML
We’re looking for a Founding AI Engineer to build and scale production-grade conversational intelligence systems. You’ll work deeply at the model and infrastructure level—owning LLM pipelines, retrieval systems, and real-time inference optimization from scratch.
What you'll do
- Design, build, and optimize end-to-end LLM pipelines for conversational AI
- Develop and deploy RAG systems using vector databases in production
- Optimize model inference for latency, memory, and cost (GPU-level optimizations)
- Own production ML systems including model serving, monitoring, and reliability
- Collaborate closely with product and engineering to ship real-world AI features
Must have
3+ years of experience in AI / ML Engineering Strong hands-on experience with LLMs and conversational AI systems Experience building and deploying RAG pipelines in production Strong understanding of vector databases Experience with Transformers / modern NLP models Hands-on experience in model serving, inference, and production ML systems Ability to optimize latency, memory, and cost for inference workloads Experience building end-to-end AI systems from development to deployment
Good to have
Experience as a Founding Engineer or early startup engineer Exposure to health-tech / wearable tech domains Experience with real-time inference infrastructure Familiarity with GPU-level optimization Experience with monitoring, reliability, and ML system observability Strong product thinking and ability to ship real-world AI features Experience working in a 0→1 / seed-stage startup environment Ability to take full ownership of AI systems end-to-end