Hands-on Senior AI Engineer to design and ship production Generative AI systems: build RAG pipelines, embeddings, vector search, multi-agent orchestration, backend APIs (Python/Flask/FastAPI), deploy cloud-native workloads, implement LLMOps practices, and mentor through code reviews while owning end-to-end delivery.
We're looking for a hands-on AI Engineer who combines strong backend engineering fundamentals with hands-on experience building production Generative AI systems. You'll design and ship RAG pipelines, integrate LLMs into real products, and build the backend services that support them, writing code daily, not just architecting on paper.
This is a purely technical IC role, not a managerial one. You’ll lead by example, mentor through code reviews, and own end-to-end technical delivery.
Key Responsibilities
- Design and build RAG systems, embeddings, vector search, chunking, and evaluation pipelines.
- Build and maintain multi-agent orchestration workflows (LangGraph, AutoGen, CrewAI, or similar).
- Develop backend services and APIs (Python — Flask/FastAPI) that expose AI workflows to production systems.
- Deploy and scale AI workloads in cloud-native environments, using serverless or containerized patterns.
- Implement LLMOps practices: prompt versioning, cost tracking, monitoring, and evaluation.
- Write clean, tested code, and use AI-assisted tools (Copilot, Cursor, Claude Code) to move faster without cutting corners.
- Work with data and platform engineers to ship GenAI features quickly, from prototype to production.
Skills, Knowledge and Expertise
Must-Have Skills
- 5+ years of backend experience, with strong Python coding skills.
- Proven experience shipping RAG systems (vector DBs, embeddings, chunking).
- Familiarity with orchestration frameworks (LangGraph, LangChain, AutoGen, or similar).
- Experience with APIs, microservices, and cloud-native development (AWS preferred).
- Familiarity with distributed systems concepts (async, message queues, caching).
- Experience with unstructured data (PDFs, tables, images).
- Builder mindset: thrives on writing, debugging, and improving production code.
- Collaborative, humble, and open to feedback.
- Strong communicator who explains design decisions clearly.
Influences through contribution, not hierarchy.
About
Emumba is a global engineering and consulting company with strengths in software development and an established AWS cloud practice focused on Data and GenAI. For 15 years, our teams across the US, the UAE, and Pakistan have earned trust through quality delivery and ownership of work. We look for people who value the culture they work in as much as the craft they bring to it.
Similar Jobs
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Design, build, evaluate, and deploy AI/LLM-powered solutions (RAG, agents, classifiers, summarizers) for consumer-facing products. Develop prompt strategies, guardrails, and evaluation frameworks; partner with data engineering and cross-functional teams to ensure trusted data, observability, safety, and continuous improvement of model performance in production.
Top Skills:
Agent FrameworksData PipelinesEmbeddingsEvaluation ToolingLlmsModel ApisOrchestration FrameworksProduction Ai Monitoring ToolsPrompt EngineeringPythonRagVector Search
Digital Media • Information Technology • News + Entertainment
Senior Applied AI Engineer builds and deploys AI/ML solutions for a SaaS cybersecurity platform. Responsibilities include data preparation, modeling, LLM-powered applications, RAG and vector search, MLOps pipelines, microservices, production deployments, incident mitigation, and mentoring global engineering teams while ensuring secure, scalable, and compliant AI practices.
Top Skills:
Agile ScrumAnthropicApache FlinkSparkAPIsContainerizationDevsecopsDockerEmbeddingsGeminiGitJIRALarge Language Models (Llms)MicroservicesMlopsOpen-Source LlmsOpenaiOrchestration FrameworksPrompt EngineeringPythonPython Unit Test FrameworksRetrieval-Augmented Generation (Rag)SaaSSemantic SearchVector Databases
Healthtech • Social Impact • Software • Telehealth
Design, implement, and maintain core infrastructure and tooling to enable agentic AI development. Drive AI enablement, developer relations, training, and shared platform tooling to support engineering teams and production LLM/agent integrations.
Top Skills:
Agent FrameworksAgentic Coding ToolsCi/Cd Automation PipelinesCloud EnvironmentsLlm Integrations
What you need to know about the Los Angeles Tech Scene
Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.
Key Facts About Los Angeles Tech
- Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
- Key Industries: Artificial intelligence, adtech, media, software, game development
- Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
- Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
- Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering



.jpg)