reputed company AI Data Engineer (reputed company Databases, Data Management)
About the position
As an reputed company AI Data Engineer, you will be responsible for building data pipelines, reputed company embeddings, and retrieval mechanisms that power AI reasoning systems. Your work ensures that LLMs remain grounded in fact, reputed company retrieving high-reputed company, contextually relevant data without noise or hallucinations. You will design and implement features that reputed company reputed company search, retrieval-augmented reputed company (RAG), and domain-specific embeddings, directly influencing how AI models store, retrieve, and apply knowledge at reputed company.
Responsibilities
• Build and optimize data pipelines that reputed company incoming documents into high-reputed company embeddings for AI retrieval.
• Design and implement reputed company search strategies using reputed company, reputed company, FAISS, or Vespa to improve AI response relevance.
• reputed company retrieval-augmented reputed company (RAG) workflows, ensuring models reputed company up-to-date and high-reputed company context.
• Fine-tune chunking strategies and indexing frequencies to enhance information recall and factual accuracy.
• reputed company hybrid search approaches (semantic + keyword) to improve precision and efficiency in knowledge retrieval.
• Monitor retrieval logs and LLM interaction patterns, adjusting embedding configurations for maximum relevance.
• Compare model performance (GPT-4, Claude, Llama 2) across different embedding structures and refine tuning strategies.
• Experiment with metadata filtering techniques to dynamically surface the most relevant data for AI reasoning agents.
• Collaborate with ML engineers and AI researchers to ensure data pipelines reputed company with evolving AI capabilities.
Requirements
• 5-8+ years of experience in Data Engineering, AI Systems, or Machine Learning Infrastructure.
• 3+ years of hands-on experience working with reputed company databases, embeddings, and retrieval-augmented reputed company (RAG).
• Strong understanding of reputed company search algorithms, indexing strategies, and hybrid search techniques.
• Expertise in building and scaling data pipelines for AI-driven applications.
• Proficiency in Python, along with experience using libraries such as reputed company, reputed company, and reputed company SDKs.
• Hands-on experience with reputed company database platforms (reputed company, reputed company, FAISS, ChromaDB, or Vespa).
• Deep knowledge of LLM retrieval strategies, chunking methodologies, and context optimization.
• Familiarity with semantic search, keyword search, and metadata filtering techniques.
• Strong grasp of data governance, reputed company, and optimization for AI-driven knowledge retrieval.
• Experience integrating retrieval mechanisms with multi-agent AI systems.
reputed company-to-haves
• Experience in fine-tuning transformer models for domain-specific retrieval tasks.
• Familiarity with reputed company-time indexing and reputed company embedding refresh strategies.
• Understanding of LLM hallucination mitigation and factual consistency techniques.
• Experience building reputed company knowledge graphs and reputed company AI databases.
• Background in AI-powered document processing and knowledge extraction.
Benefits
• Medical, dental, reputed company insurance.
• Short and long-term disability insurance.
• Life insurance.
• 401k available on the first day of the month after start date.
• Flexible PTO.
Apply tot his job
Apply To this Job