Tiger Analytics Logo

Tiger Analytics

Gen AI Data Engineer

Reposted One Month Ago
Remote
Hiring Remotely in United States
Expert/Leader
Remote
Hiring Remotely in United States
Expert/Leader
The Gen AI Data Engineer will design and build distributed data systems, develop data pipelines, manage data infrastructure, and integrate technologies for real-time and batch processing, contributing to scalable analytics solutions.
The summary above was generated by AI

Tiger Analytics is looking for experienced Machine Learning Engineers with Gen AI experience to join our fast-growing advanced analytics consulting firm. Our employees bring deep expertise in Machine Learning, Data Science, and AI. We are the trusted analytics partner for multiple Fortune 500 companies, enabling them to generate business value from data. Our business value and leadership has been recognized by various market research firms, including Forrester and Gartner.

We are looking for top-notch talent as we continue to build the best global analytics consulting team in the world. You will be responsible for:

Technical Skills Required:

Programming Languages: Proficiency in Python, SQL, and PySpark.

Data Warehousing: Experience with Snowflake, NOSQL and Neo4j.

Data Pipelines: Proficiency with Apache Airflow.

Cloud Platforms: Familiarity with AWS (S3, RDS, Lambda, AWS batch, SageMaker processing Job, CloudFormation, etc.) or GCP (Vertex AI RAG, Data pipeline, Bigquery, GKE)

Operating Systems: Experience with Linux.

Batch/Realtime Pipelines: Experience in building and deploying various pipelines.

Version Control: Experience with GitHub.

Development Tools: Proficiency with VS Code.

Engineering Practices: Skills in testing, deployment automation, DevOps/SysOps.

Communication: Strong presentation and communication skills.

Collaboration: Experience working with onshore/offshore teams.


Requirements

Desired Skills:

·        Big Data Technologies: Experience with Hadoop and Spark.

Data Visualization: Proficiency with Streamlit and dashboards.

·        APIs: Experience in building and maintaining internal APIs.

·        Machine Learning: Basic understanding of ML concepts.

·        Generative AI: Familiarity with generative AI tools and techniques.

Additional Expertise:

·        Knowledge Graphs: Experience with creation and retrieval.

·        Vector Databases: Proficiency in managing vector databases.

·        Data Persistence: Ability to develop and maintain multiple forms of data persistence and retrieval methods (RDMBS, Vector Databases, buckets, graph databases, knowledge graphs, etc.).

·        Cloud Technologies: Experience with AWS, especially SageMaker, Lambda, OpenSearch.

·        Automation Tools: Experience with Airflow DAGs, AutoSys, and CronJobs.

·        Unstructured Data Management: Experience in managing data in unstructured forms (audio, video, image, text, etc.).

·        CI/CD: Expertise in continuous integration and deployment using Jenkins and GitHub Actions.

·        Infrastructure as Code: Advanced skills in Terraform and CloudFormation.

·        Containerization: Knowledge of Docker and Kubernetes.

·        Monitoring and Optimization: Proven ability to monitor system performance, reliability, and security, and optimize them as needed.

·        Security Best Practices: In-depth understanding of security best practices in cloud environments.

·        Scalability: Experience in designing and managing scalable infrastructure.

·        Disaster Recovery: Knowledge of disaster recovery and business continuity planning.

·        Problem-Solving: Excellent analytical and problem-solving abilities.

·        Adaptability: Ability to stay up-to-date with the latest industry trends and adapt to new technologies and methodologies.

·        Team Collaboration: Proven ability to work well in a team environment and contribute to a positive, collaborative culture.

GenAI Engineer Specific Skills:

·        Industry Experience: 8+ years of experience in data engineering, platform engineering, or related fields, with deep expertise in designing and building distributed data systems and large-scale data warehouses.

·        Data Platforms: Proven track record of architecting data platforms capable of processing petabytes of data and supporting real-time and batch ingestion processes.

·        Data Pipelines: Strong experience in building robust data pipelines for document ingestion, indexing, and retrieval to support scalable RAG solutions. Proficiency in information retrieval systems and vector search technologies (e.g., FAISS, Pinecone, Elasticsearch, Milvus).

·        Graph Algorithms: Experience with graphs/graph algorithms, LLMs, optimization algorithms, relational databases, and diverse data formats.

·        Data Infrastructure: Proficient in infrastructure and architecture for optimal extraction, transformation, and loading of data from various data sources.

·        Data Curation: Hands-on experience in curating and collecting data from a variety of traditional and non-traditional sources.

·        Ontologies: Experience in building ontologies in the knowledge retrieval space, schema-level constructs (including higher-level classes, punning, property inheritance), and Open Cypher.

·        Integration: Experience in integrating external databases, APIs, and knowledge graphs into RAG systems to improve contextualization and response generation.

·        Experimentation: Conduct experiments to evaluate the effectiveness of RAG workflows, analyze results, and iterate to achieve optimal performance.


Benefits

This position offers an excellent opportunity for significant career development in a fast-growing and challenging entrepreneurial environment with a high degree of individual responsibility.

Similar Jobs

One Month Ago
Remote
United States
60K-210K Annually
Senior level
60K-210K Annually
Senior level
Agency • Information Technology
Design, evaluate, and productionize generative AI/ML solutions (RAG, agents, embeddings, retrieval). Build evaluation frameworks for hallucination detection, benchmark LLMs, optimize prompts/models, create datasets, fine-tune models, and collaborate to deploy and monitor enterprise-scale GenAI systems.
Top Skills: Ai Observability PlatformsAws BedrockAzure Ai FoundryClaudeDatabricksEmbeddingsGeminiKubernetesMlflowNeo4JNumpyOpen-Source LlmsOpenaiPandasPythonPyTorchRagScikit-LearnSemantic SearchSparkTensorFlowVector Databases
An Hour Ago
Easy Apply
Remote or Hybrid
Arizona, USA
Easy Apply
16-16 Hourly
Junior
16-16 Hourly
Junior
Automotive • Big Data • Insurance • Software • Transportation
Handle high-volume inbound roadside assistance calls, gather location and vehicle details, dispatch tow and service providers, and support distressed motorists. The role requires empathetic de-escalation, accurate multitasking across digital systems, sound judgment under pressure, reliable schedule adherence, and effective collaboration in a remote contact-center environment.
Top Skills: Crm SoftwareDispatch SoftwareEthernetGmailGoogle ChatGoogle ChromeGoogle DocsGoogle MapsGoogle SheetsGoogle WorkspaceHarverMozilla FirefoxSwoopWindows 11Zoom
An Hour Ago
Easy Apply
Remote or Hybrid
Alabama, USA
Easy Apply
16-16 Hourly
Junior
16-16 Hourly
Junior
Automotive • Big Data • Insurance • Software • Transportation
Provides real-time roadside assistance to stranded motorists by gathering location and vehicle details, dispatching tow and service providers, de-escalating stressful calls, documenting interactions, and tracking service progress across digital systems. The role requires empathy, sound judgment, multitasking, reliable schedule adherence, and strong technology skills in a remote contact-center environment. Associates must work full time, support weekend and holiday coverage, complete mandatory training, and provide an approved home-office setup and equipment.
Top Skills: Ai ToolsCrm SoftwareDispatch SoftwareEthernetGoogle ChatGoogle ChromeGoogle MapsGoogle WorkspaceHarverMozilla FirefoxWindows 11Zoom

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

  • Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
  • Key Industries: Artificial intelligence, adtech, media, software, game development
  • Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
  • Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account