SummitTX Capital→
Intern - Research Engineer at SummitTX Capital in New York, NY
InternshipOn-siteFull-timeNew York, NY$83k–$104k/yr
Skills
pythonsqldata modelingdbtawsmachine learninggitvisualization toolsclear communication
Job Description
Summary: SummitTX Capital is a multi-manager, multi-strategy hedge fund managing over $3 billion in AUM. They are seeking exceptional master’s candidates for their Research Engineer Internship, where the intern will help build and scale a systematic data platform that powers alpha research and production signals.
Responsibilities:
- Design, build, and maintain systematic data pipelines, including ingestion, medallion-style data modeling, feature engineering, and experiment tracking
- Operationalize robust ELT workflows using DBT/SQL and Python on Databricks, with strong enforcement of data quality, lineage, and documentation
- Develop research-grade datasets and features across market, alternative, and fundamental domains to support L/S Equity and systematic strategies
- Productionize models and alpha signals with CI/CD pipelines, model registries, monitoring, and cost/performance optimization on Databricks and AWS
- Partner with PMs and Analysts to translate investment hypotheses into testable research artifacts, delivering clear results, visualizations, and readouts to guide decision-making
- Contribute to the evolution of the data platform roadmap, including observability, governance, access controls, cataloging, and documentation standards
Required Qualifications:
- BS or pursuing an MS in Data Science, Data Engineering, Statistics, Business Analytics, Applied Math, or related field with strong academic performance
- Strong Python and SQL fundamentals; comfort with Git and testing frameworks
- Coursework or internship experience in data modeling, ETL/ELT, artificial intelligence/machine learning/statistics, or time-series analysis
- Clear communication skills and ability to partner with investment, risk, and operations stakeholders
Preferred Qualifications:
- Hands-on experience with Python, SQL, DBT, Spark, and modern data-quality toolkits
- Exposure to ML frameworks (pandas, scikit-learn, PyTorch, MLflow) and feature pipelines
- Familiarity with Databricks (Lakehouse, Unity Catalog) and AWS data services (S3, Glue/Athena, Lake Formation)
- Experience with visualization and BI tools (e.g., Plotly, Tableau/Power BI), and Financial Data Platform (e.g. Bloomberg Terminal)
- Experience in GenAI/LLM applications (prompt engineering, agentic workflow, RAG)
Required Skills: Python, SQL, Data Modeling
Important Skills: DBT, AWS, Machine Learning
Nice-to-Have Skills: Git, Visualization Tools, Clear Communication
Benefits: Eligible for overtime
Benefits
Eligible for overtime