Skip to main content

Cognition

Data Engineer

San FranciscofulltimemidAdded 1 month ago

About this role

Join an applied AI lab building software agents like Devin and Windsurf. As a Data Engineer, you'll design and manage the complete data infrastructure—from databases and pipelines to analytics and reporting—across GTM, Product, or Finance teams.

What you'll do

  • Design and manage database architecture, data models, and warehouse systems
  • Build and maintain ETL/ELT pipelines and orchestration workflows
  • Create integrations between internal and external systems
  • Own business reporting including dashboards, metrics, and self-serve analytics
  • Ensure data quality, observability, governance, and documentation

What they're looking for

  • SQL and Python
  • Data modeling and warehouse architecture
  • ETL/ELT tools (dbt, Airflow, Dagster)
  • BI tool development (Metabase preferred)
  • Scalable backend application development
  • Statistics and experimentation design
  • Data integration and API management
Apply with Autofill

Opens the application — the Jobs AI extension fills it for you. Set up autofill

Opens the official application on the employer’s site. No login required.

Cognition

Cognition builds AI software engineers and developer tools, including Devin (an AI software engineer) and Windsurf (an AI-native IDE) that help developers automate tasks and write code more efficiently. The company is hiring Deployed Engineers to work directly with customers on adoption and integration, SREs to manage production reliability and infrastructure, federal engineers for government deployments, and IT specialists to support internal operations.

View all jobs at Cognition

Likely interview questions

  • Walk us through how you've designed a data warehouse or database schema from scratch. What were the key decisions you made around data modeling, and how did you handle evolving requirements?
  • Describe your experience building and maintaining ETL/ELT pipelines. What tools have you used, and how have you handled data quality issues or pipeline failures in production?