Design and implement ML pipelines that encode visual inputs (pose, face, objects) into shared embeddings. Build and fine-tune CNN and transformer models, perform training, evaluation, augmentation, and hyperparameter tuning. Apply similarity metrics and clustering for behavioral inference and action prediction. Collaborate with cross-functional engineers and produce documented, version-controlled code and model artifacts.
About Our Client
Our client is a technology company developing next-generation intelligent systems at the intersection of AI, XR, robotics, autonomy, and spatial computing. Their products support mission-critical applications across defense, public safety, and critical infrastructure. They are seeking passionate professionals who thrive in fast-paced environments and enjoy building impactful products from concept to deployment.
The Role
Our client is seeking a Machine Learning Engineer to help design and implement intelligent systems that extract meaning and predictive value from computer vision and behavioral datasets. This is a junior-level, in-person role suited for candidates with 2–3 years of experience and a solid foundation in deep learning, embeddings, and modern neural architectures.
As a member of the AI team, the ideal candidate will work on projects that leverage CNNs, transformer models, and embedding architectures to encode and reason over pose, facial, and action-based visual data. These systems support downstream tasks such as future action prediction, semantic matching, and similarity-based inference.
Key Responsibilities
- Design and implement machine learning pipelines that encode visual input (pose, face, object/classification) into shared embedding spaces for similarity and predictive tasks.
- Build and fine-tune convolutional and transformer-based neural architectures optimized for visual recognition and representation learning.
- Develop encoding and embedding techniques that allow consistent comparison across multiple data types (e.g., pose vectors, facial landmarks, class labels).
- Apply techniques such as cosine similarity, distance metrics, and latent clustering to perform behavioural inference and action prediction.
- Contribute to model training, evaluation, and deployment workflows, including data preprocessing, augmentation, hyperparameter tuning, and performance profiling.
- Collaborate closely with engineers in computer vision, embedded systems, software, and UI/UX to ensure seamless integration of AI pipelines into real-time systems.
- Produce clean, well-documented code and maintain version-controlled model artefacts and experiment logs.
- Write technical documentation for models, training procedures, evaluation criteria, and system integration.
Skills, Knowledge and Expertise
- Bachelor's or Master's degree in Artificial Intelligence, Data Science, Computer Science, Machine Learning, or a closely related discipline.
- 2–3 years of experience in machine learning roles through internships, academic labs, or early career positions.
- Strong understanding of Convolutional Neural Networks (CNNs) for image and video-based tasks.
- Strong understanding of transformer architectures and their applications in vision or multimodal learning.
- Strong understanding of embedding systems and vector space modeling for semantic and similarity-based tasks.
- Strong understanding of encoding mechanisms and dimensionality reduction techniques for latent representation.
- Proficiency in Python and deep learning frameworks such as PyTorch or TensorFlow.
- Familiarity with pose estimation, facial recognition, or classification models (e.g., OpenPose, MediaPipe, FaceNet, ResNet variants).
- Experience training models with structured and unstructured visual datasets.
- Exposure to techniques like cosine similarity, triplet loss, contrastive learning, or temporal prediction modeling.
- Strong computer science fundamentals, including data structures, algorithms, and software design patterns.
- Comfort working in Linux-based development environments and version control systems (Git).
- A collaborative mindset, with excellent communication skills and a willingness to learn across domains.
Bonus (Nice to have):
- Experience integrating vision-based AI models into embedded or robotics systems.
- Familiarity with ONNX or TensorRT for model optimization and deployment.
- Background in sequence modeling, recurrent architectures, or video-based action recognition.
- Exposure to multimodal AI systems that blend image, pose, and metadata representations.
- Familiarity with techniques like CLIP, DINO, or self-supervised representation learning.
- Experience with MLOps or training orchestration tools such as MLflow, Weights & Biases, or DVC.
Other Requirements:
- Must be a US Citizen or a valid Green Card holder. Visa sponsorship is not available for this role at this time.
- Candidates must reside within a commutable distance of Santa Monica, California.
Benefits
- Compensation: $100,000 to $120,000 per year
- Comprehensive health coverage and flexible PTO
- Opportunity to work on innovative AI, robotics, XR, and autonomous technologies
- Collaborative multidisciplinary engineering environment
- Career growth and professional development opportunities
Similar Jobs
Defense • Manufacturing
Own end-to-end deployment of ML models to resource-constrained edge hardware: optimize and convert models (quantization, pruning, distillation, operator fusion), profile inference on accelerators, write and maintain C++ inference host code, build test harnesses and tooling, and set best practices while mentoring engineers to meet real-time, memory, power, and latency constraints.
Top Skills:
Arm CortexC++CudaDspsFpgasLinuxNpusNvidia GpuNvidia JetsonOnnxOnnx RuntimePyTorchQualcomm SocsRtosTensorflow LiteTensorrt
Defense • Manufacturing
Develop and optimize real-time computer vision algorithms and ML models for drone detection, tracking, and classification; integrate CV systems with embedded hardware and sensors; test and harden systems for military-grade, safety-critical deployment.
Top Skills:
C++CamerasEmbedded SystemsLidarPythonPyTorchRadarTensorFlow
Defense • Manufacturing
Develop and optimize real-time computer vision and ML systems for detecting, tracking, and classifying drones on an autonomous gun turret. Implement resource-efficient models, integrate vision systems with hardware, and validate performance across diverse, challenging environments to harden prototypes to military-grade standards.
Top Skills:
C++CameraEmbedded SystemsLidarNvidia JetsonPythonPyTorchRadarTensorFlow
What you need to know about the Los Angeles Tech Scene
Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.
Key Facts About Los Angeles Tech
- Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
- Key Industries: Artificial intelligence, adtech, media, software, game development
- Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
- Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
- Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

