Greetings!
AI/ML Engineer
Plano Texas,5 days onsite since day 1
We are seeking an AI Engineer to design and deploy production-grade AI solutions across LLM applications, RAG pipelines, retrieval systems, and scalable ML services. This role blends model orchestration, search relevance, MLOps, and cross-functional collaboration to deliver reliable, high-impact AI products.
Responsibilities
Build and optimize LLM, RAG, and multi-agent workflows for production use cases.
Design retrieval systems using vector search, reranking, and metadata filtering.
Fine-tune, evaluate, and benchmark models to improve output quality and business impact.
Deploy AI/ML services with strong standards for latency, scalability, and observability.
Develop data, inference, and experimentation pipelines in partnership with product and engineering teams.
Required Qualifications
Strong Python and SQL skills with solid machine learning and NLP fundamentals.
Experience with LangChain, LangGraph, LlamaIndex, prompt engineering, and retrieval optimization.
Familiarity with vector databases and search tools such as Pinecone, OpenSearch, FAISS, or pgvector.
Experience with Docker, Kubernetes, MLflow, CI/CD, and cloud platforms such as Azure.
Preferred Qualifications
Experience fine-tuning open-source models using LoRA, QLoRA, SFT, or DPO.
Background in high-scale, low-latency AI systems and enterprise data platforms.
Ability to work across research, engineering, and product teams to ship AI solutions into production.