A

AI Engineer (Data Guardrails & LLM Ingestion Pipelines)

AWISEE · Remote

🔥32 people viewed this job

About the Role

This is a remote position. About the Role We are looking for a highly skilled AI Engineer to design and build robust data ingestion, cleaning, validation, and LLM enhancement pipelines that power our AI applications. You will transform raw, unstructured data into high-quality, AI-ready datasets while implementing guardrails that ensure accuracy, consistency, and reliability. Key Responsibilities ·         Design and develop scalable data ingestion pipelines for structured and unstructured data. ·         Build automated data cleaning, normalization, and preprocessing workflows. ·         Develop AI-powered enrichment pipelines using LLMs (OpenAI, Claude, Gemini, etc.). ·         Implement data quality validation and AI guardrails. ·         Develop prompt engineering workflows for data transformation. ·         Build document processing pipelines for PDFs, Word documents, CSVs, websites, and APIs. ·         Develop Retrieval-Augmented Generation (RAG) pipelines. ·         Create evaluation frameworks for LLM quality and accuracy. ·         Build ETL/ELT workflows for AI-ready datasets. ·         Integrate vector databases for semantic search. ·         Monitor pipeline performance, cost, latency, and data quality. ·         Collaborate with cross-functional teams to deliver production AI systems. Required Technical Skills Programming ·         Python (Expert) ·         SQL ·         Git AI & LLMs ·         OpenAI API ·         Anthropic Claude API ·         Google Gemini API ·         Prompt Engineering ·         Function Calling ·         Structured Outputs AI Frameworks ·         LangChain ·         LlamaIndex ·         DSPy (Preferred) ·         PydanticAI (Nice to Have) Data Engineering ·         Pandas ·         Polars ·         ETL/ELT Pipelines ·         Apache Airflow (Preferred) ·         Data Validation Frameworks Vector Databases ·         Pinecone ·         Weaviate ·         Qdrant ·         ChromaDB ·         FAISS Cloud & Infrastructure ·         Docker ·         Kubernetes (Preferred) ·         AWS / Azure / GCP ·         Linux Databases ·         PostgreSQL ·         MongoDB ·         Redis RequirementsPreferred Qualifications ·         Experience building production-grade AI systems. ·         Strong understanding of RAG architectures. ·         Experience implementing AI guardrails and hallucination mitigation. ·         Experience with OCR and document parsing. ·         Experience with embedding models and semantic search. ·         Knowledge of data governance and security best practices. Success Metrics ·         Build scalable ingestion pipelines. ·         Deliver automated data cleaning and LLM enhancement workflows. ·         Implement AI guardrails to improve output quality. ·         Develop evaluation pipelines for LLM performance. ·         Contribute to a production-ready AI platform.

AWISEE has 1 open position on Remote Vibe Coding Jobs.

💬 Developer Questions

Ask the team a question — answers show up here

🎯

What does the interview process look like?

🤖

What AI/vibe coding tools does the team use daily?

👥

How big is the engineering team?

⏰

Is the team fully async or are there required meetings?

🚀

What does onboarding look like for remote hires?

🔧

Can you share more about the tech stack and architecture?

📈

What does career growth look like in this role?

📅

What does a typical day look like?

💰

Is there a salary range you can share?

📊

Is equity or stock options part of the package?

🌍

Are there timezone requirements or preferences?

🛂

Do you sponsor work visas?

🏢 Is this your listing? Claim it to answer questions

Similar Jobs

Helpful resources

Hiring for a similar role? Post your job here — it's free →