all rolesEngineering · Bangalore, IND · Full-time
Data Engineer
Design and operate the data infrastructure behind AI systems for sophisticated hedge funds — reliable pipelines, warehouse models, and the datasets that investment teams and agents depend on every day.
What you’ll do
- Build and maintain ETL/ELT pipelines that land market, portfolio, and research data with reliability and clear ownership
- Model analytical datasets for research, risk, IR, and ops workflows — performant, documented, and easy for agents and analysts to query
- Partner with Applied AI Engineers to expose clean, MCP-ready data sources for agentic tools
- Improve observability: freshness SLAs, lineage, anomaly detection, and recovery playbooks
- Work directly with client and internal stakeholders to scope data needs and ship in weeks, not months
What we’re looking for
- 3+ years of data engineering experience building production pipelines and analytical data models
- Strong Python and SQL; comfort with modern data tooling (dbt, Airflow/Prefect, Spark, or equivalents)
- Experience with cloud data platforms (Azure, AWS, or GCP) and warehouse/lakehouse patterns
- Familiarity with OLAP stores, partitioning, incremental loads, and data quality monitoring
- Curiosity about financial market data and how investment teams consume it
- Strongly preferred: experience with Databricks, Polars/Pandas, or real-time/streaming ingestion