SB Energy→
Data Engineering & AI Enablement Intern 2026 at SB Energy · San…
InternshipHybridSan Francisco Bay Area, CA$52k–$62k/yr
Skills
pythonpandassqldata structuresalgorithmssystem design fundamentalsazure blob storagedata migration pipelinesdata schemaspartitioning strategiesparquet formatapiscloud platformstime-series dataetl/elt conceptsdistributed systemsai/llm applications
Job Description
Summary: SB Energy is a leading company backed by SoftBank and Ares, focused on providing reliable and affordable energy solutions. They are seeking a highly motivated undergraduate Data Engineering & AI Enablement Intern to lead the migration and modernization of operational time-series data, working closely with IT and Data teams to enhance their data platform and AI chatbot capabilities.
Responsibilities:
- Analyze and understand time-series data stored in the Canary Historian system
- Design and implement components of a scalable data migration pipeline to Azure Blob Storage
- Transform and structure large datasets for efficient storage, query performance, and cost optimization
- Define data schemas, partitioning strategies, and storage formats (e.g., Parquet)
- Validate data integrity and ensure reliability throughout the migration process
- Contribute to building mechanisms for AI chatbot access (e.g., APIs, query layers, or retrieval pipelines)
- Collaborate with IT and Data Science teams to integrate data into downstream systems
- Evaluate trade-offs between different architectural approaches (performance, cost, scalability)
- Document technical designs, decisions, and best practices
- Present technical solutions and business impact to stakeholders
Required Qualifications:
- Currently pursuing a Bachelor's degree in Computer Science, Data Science, Engineering, Mathematics, or related field
- Strong programming skills in Python (or similar language)
- Experience working with data (e.g., Pandas, SQL, or similar tools)
- Solid understanding of data structures, algorithms, and system design fundamentals
Preferred Qualifications:
- Experience with cloud platforms (Azure, AWS, or GCP)
- Familiarity with data pipelines, ETL/ELT concepts, or distributed systems
- Experience working with large datasets or time-series data
- Exposure to APIs, backend systems, or data access layers
- Familiarity with AI/LLM applications or interest in AI-powered systems
- Demonstrated ability to work independently and in a team environment
Required Skills: Python, Pandas, SQL, Data structures, Algorithms, System design fundamentals, Azure Blob Storage, Data migration pipelines, Data schemas, Partitioning strategies, Parquet format, APIs, Cloud platforms, Time-series data, ETL/ELT concepts, Distributed systems, AI/LLM applications
Internship Start Date: Start in 2026
Benefits: Comprehensive, value-added project(s), Work in teams and with cross-functional colleagues in a professional environment, Develop technical skills specific to your major, Gain opportunities for professional development by building relationships and learning from industry professionals in a live business environment, Final project summary presentation to the team
Benefits
Comprehensive, value-added project(s)
Work in teams and with cross-functional colleagues in a professional environment
Develop technical skills specific to your major
Gain opportunities for professional development by building relationships and learning from industry professionals in a live business environment
Final project summary presentation to the team