R

AI Engineer - Onsite - San Francisco, CA

RS Global Services · San Francisco, CA

Full-timeSeniorPythonTypeScriptAWS

About the Role

Role OverviewOne of our start-up clients is hiring a full-time AI Engineer to own the prompts, agents, evals, and pipelines behind user-facing features that ship to users. You'll take product requirements and turn them into working prompts, agents, and pipelines. You'll evaluate them rigorously, iterate until they're production-ready, and keep improving them once they ship. This role sits at the intersection of product and platform: you decide what the AI should do, prove it works, and get it in front of users. Because we're an early-stage company moving fast, we're looking for someone who can work quickly through ambiguous AI problems, measure output quality, and ship only when the system is reliable enough for production. This is an in-person role, 5 days a week in our office. The ability to tell the difference between "looks good in the demo" and "works in production" is essential. Key Responsibilities • Build new AI features end to end, from prototype to production. • Improve AI output quality through prompt engineering, model selection, retrieval, and evaluation. • Design and run evals that measure real output quality, not just first impressions. • Iterate fast on prompts, agent designs, and orchestration patterns. • Partner with the Product Engineer to translate requirements into AI features that actually work. • Partner with the AI Platform team to land features on solid infrastructure. • Evaluate new models, tools, and techniques when they improve quality, latency, cost, or reliability. What We Are Looking For • Hands-on experience building LLM-powered features that shipped to real users • Production engineering chops in TypeScript/Node (primary, especially in AWS Lambda) and/or Python • Experience with multiple LLM providers such as Anthropic, OpenAI, Google Vertex, AWS Bedrock, or similar • Practical judgment in prompt engineering, retrieval, and agent design, backed by evaluation results • Track record of building evaluation systems that actually catch regressions • Solid software engineering fundamentals: you can write production code, not just notebooks Role Requirements (please apply only if you meet "MUST HAVE" requirements) Seniority • 3 -​ 8 years of experience in hands-​on software engineering,​ building LLM-​powered features that shipped to real users (MUST HAVE) Work experience • Has shipped LLM-​powered features to real users in production at a reputable,​ high-​growth startup with a high engineering bar and can speak to what broke (MUST HAVE) • Built agents and agentic systems -​ orchestrating LLMs,​ tool use over large data sets • Has kept up with the frontier of agentic AI methods with a finger on the pulse;​ LangChain-​only experience is a yellow flag (MUST HAVE) • Experience on a small team (<​15 engineers) or as a founding engineer / former founder.​ • Enterprise vertical SaaS experience • FDE-​ or explicitly customer-​facing type experience Education • Bachelor's degree in Computer Science (MUST HAVE).​ Hard skills • Production engineering chops in TypeScript/Node (primary,​ especially in AWS Lambda) and/or Python (MUST HAVE) • Experience with eval systems,​ structured output,​ and function calling.​ (MUST HAVE) • Experience with Anyscale Ray or similar distributed compute frameworks for batch inference,​ eval pipelines,​ or scaling agent workloads • Open source contributions in the LLM or agent tooling space • Familiarity with pgvector or other vector retrieval systems • Experience with post-​training or fine-​tuning Soft skills • Genuinely excited about early-​stage work and the company/mission (MUST HAVE) Important Note : Please DO NOT apply if you have any of the below Traits • Short tenures (less than 2 years in any company excluding internships and contract work) • Big tech only (FAANG,​ Netflix) with very narrow role scope • RAG-​only or LangChain-​heavy experience -​ indicates lack of up-​to-​date agentic AI knowledge • Candidates without engineering depth -​ background in data analytics for example is a negative signal Salary : $180K - $250K Equity : Competitive equity On-site work policy: 5 days in-office in SOMA, San Francisco Visa sponsorship available: E3 visas. H1B Transfers. TNs. No new H1Bs.

RS Global Services has 1 open position on Remote Vibe Coding Jobs.

💬 Developer Questions

Ask the team a question — answers show up here

🎯

What does the interview process look like?

🤖

What AI/vibe coding tools does the team use daily?

👥

How big is the engineering team?

Is the team fully async or are there required meetings?

🚀

What does onboarding look like for remote hires?

🔧

Can you share more about the tech stack and architecture?

📈

What does career growth look like in this role?

📅

What does a typical day look like?

💰

Is there a salary range you can share?

📊

Is equity or stock options part of the package?

🌍

Are there timezone requirements or preferences?

🛂

Do you sponsor work visas?

🏢 Is this your listing? Claim it to answer questions

Similar Jobs

Helpful resources

Hiring for a similar role? Post your job here — it's free →