Cogniify Logo

Cogniify

Data engineer - USA

Posted 2 Days Ago
Remote
Hiring Remotely in United States
134K-140K Annually
Mid level
Remote
Hiring Remotely in United States
134K-140K Annually
Mid level
Build and maintain production ETL/ELT pipelines, data models, and curated datasets for analytics and AI applications. Ingest data from databases, APIs, SaaS tools, and event streams; implement quality checks, monitoring, documentation, governance, and access controls. Collaborate with analysts, data scientists, and ML engineers on reporting, feature datasets, AI search, and RAG applications. Optimize query performance and pipeline costs while using Git, code reviews, and CI/CD for safe releases.
The summary above was generated by AI

The Role

Cogniify is hiring a Mid Level Data Engineer to build reliable data pipelines and analytics datasets for reporting, business decisions and AI initiatives. You will own data work from ingestion through transformation and delivery, working with analysts, data scientists and engineers to make large datasets accurate, accessible and ready for use. This is a hands-on role for someone who enjoys both data engineering and the analytical questions behind the data.

What You Will Do

  • Build and maintain ETL/ELT pipelines using SQL, Python, dbt and tools such as Apache Spark, PySpark or Airflow.

  • Ingest data from databases, APIs, SaaS tools and event streams using connectors or custom pipelines.

  • Develop tested data models and curated datasets in Snowflake, Databricks, BigQuery or Redshift for reporting and self-service analytics.

  • Work with data scientists and ML engineers to prepare feature datasets for model training and inference.

  • Prepare and refresh structured business data that can support AI search, retrieval-augmented generation (RAG) or other Generative AI applications.

  • Build clear dashboards and analyses in Tableau, Looker, Power BI or similar tools when the work calls for it.

  • Add data quality checks, monitoring and documentation so teams can trust the data and identify pipeline issues early.

  • Improve query speed and pipeline cost; use Git, code reviews and CI/CD to release changes safely.

  • Help manage data access, lineage and sensitive information, including personally identifiable information (PII).

What We Are Looking For

  • 3 to 6 years of professional experience in data engineering, analytics engineering or a related data role with production delivery.

  • Strong SQL skills and experience writing complex transformations and improving query performance.

  • Hands-on experience with a cloud data platform. Snowflake is preferred; Databricks, BigQuery or Redshift experience is also relevant.

  • Production experience with dbt for transformation, testing and documentation.

  • Working knowledge of Python and either Pandas or PySpark for data processing.

  • Experience scheduling pipelines with Airflow, Dagster, Prefect or a similar orchestration tool.

  • A good understanding of data modeling and how to build datasets that analysts and business teams can use.

Preferred Experience

  • Apache Spark, Databricks and large-scale data processing.

  • Data quality or observability tools such as Great Expectations, Soda or Monte Carlo.

  • Streaming data with Kafka or Kinesis, or ingestion tools such as Fivetran or Airbyte.

  • Experience preparing data for ML features, AI search, embeddings or RAG applications.

  • Cloud services across AWS, Azure or Google Cloud, and data governance tools such as Unity Catalog or DataHub.

Why Join Cogniify

You will work on data products used for analytics and emerging AI applications, with room to own your pipelines and improve how teams use data. We would like to hear from engineers who care about clean data, dependable systems and useful outcomes.

Perks And Benefits Of Working With Us

  • Unlimited PTO.

  • Please ask us about our very generous parental leave, much above industry standards!.

  • Entrepreneurial culture where pushing limits and taking risks is everyday business.

  • Open communication with management and company leadership.

  • Small, dynamic teams = massive impact.

  • Medical, Dental and Vision coverage for employees.

  • Access to Disability & Life insurance.

  • Mental health and wellbeing support

  • Annual bonus program

  • Employer Stock Purchase Program (ESPP)

  • Yearly Team building experiences

  • Mentorship and sponsorship opportunities

  • Manager resources and support

We are an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or any other protected characteristic.

Similar Jobs

8 Days Ago
In-Office or Remote
99K-168K Annually
Senior level
99K-168K Annually
Senior level
Consulting
Develops Scala and Python data solutions for Medicare and Medicaid claims cost measures. Responsibilities include ETL, data analysis, SQL querying, validation, quality checks, documentation, requirements communication, and collaboration with external partners. The role participates in Scrum ceremonies, test meetings, and Agile planning while supporting CMS’s Quality Payment Program. It is remote within the United States, follows East Coast hours, and may require occasional travel for planning events.
Top Skills: SparkConfluenceGitGitJIRAPysparkPythonScalaScaled Agile FrameworkScrumSparkrSQL
10 Days Ago
Remote
USA
135K-150K Annually
Senior level
135K-150K Annually
Senior level
Healthtech
Build and operate Danaher’s enterprise data platform automation and reliability layer. Responsibilities include Infrastructure-as-Code, CI/CD, self-service developer enablement, data platform provisioning, observability, SLOs, auto-remediation, FinOps, compliance guardrails, and event-driven operations across Snowflake, Azure, Matillion, dbt, and Airflow. The role also develops AI-driven operators and copilots using Azure AI Foundry, Anthropic Claude, and related agent frameworks.
Top Skills: Anthropic ClaudeApache AirflowAutogenAzureAzure Ai FoundryAzure MonitorAzure OpenaiDbtGithub ActionsGitlabLanggraphLog AnalyticsMatillionOpenaiPrompt FlowPulumiPythonSemantic KernelServicenowSnowflakeSnowflake CortexTerraform
17 Days Ago
In-Office or Remote
99K-168K Annually
Senior level
99K-168K Annually
Senior level
Consulting
Design, build, and maintain scalable data pipelines and processing workflows (Spark, Hive, Databricks, Airflow). Develop APIs, dashboards (QuickSight), ensure data quality, security, and performance optimizations. Write tests, perform code reviews, collaborate on CI/CD and IaC, and support data governance and integration across cloud/SaaS sources.
Top Skills: AirflowAmazon SnsAthenaAws GlueAws Glue DatabrewAws QuicksightAzkabanC++CassandraDatabricksEmrGithub ActionsHiveJavaJupyter NotebooksLuigiPostgresPythonRdsRedshiftS3ScalaSparkSpark-StreamingSQLStormTerraform

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

  • Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
  • Key Industries: Artificial intelligence, adtech, media, software, game development
  • Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
  • Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account