LOG IN
SIGN UP
Canary Wharfian - Online Investment Banking & Finance Community.
Sign In
or continue with e-mail and password
Forgot password?
Don't have an account?
Join Canary Wharfian
or continue with e-mail and password
By signing up, you agree to our Terms & Conditions and Privacy Policy.

Data Engineering Internship (Summer 2027)

SummerNo visa sponsorship
Castleton Commodities logo

at Castleton Commodities

Commodities

Posted 13 days ago

No clicks

**Data Engineering Intern - Summer 2027, NYC** Join Castleton Commodities International (CCI), a global energy commodities merchant, to build our leading-edge data science platform. The Data Engineering Intern will collaborate with cross-functional teams to develop, optimize, and maintain robust data pipelines, crucial for our analytics and investment decisions. This hands-on role involves working with diverse data sources, including APIs and cloud providers, using Python, SQL, and Snowflake. Responsibilities: - Ingest data from various sources, parlaying it into Snowflake for analytics and forecasting tools. - Design and implement ETL processes, ensuring data quality and integrity. - Automate workflows using Python, SQL, and orchestration tools like Airflow. - Transition legacy datasets to scalable, cloud-native workflows aligned with modern data architecture. Requirements: - Pursuing a bachelor's degree in Computer Science, Engineering, MIS, or related field, with an expected graduation of Winter 2027 or Spring/Summer 2028. - Proficient in Python, SQL, and exposure to cloud platforms (AWS, Azure). - Familiarity with web scraping frameworks and handling large datasets. Application deadline: September 1st, 11:59m EST.

Compensation
Not specified

Currency: Not specified

City
London
Country
United Kingdom

Full Job Description

Application Deadline: September 1st, 11:59m EST

Program Summary - Data Science & Technology Internship

Company Overview:

Castleton Commodities International is a leading global energy commodities merchant and infrastructure asset investor. As a trader, CCI deploys capital on a proprietary basis in the physical and financial commodity markets, providing the Company with market insights and access. As a strategic investor and developer, CCI leverages its market expertise, operations capabilities, and industry knowledge to invest in, and develop, select commodity infrastructure assets. Our strategically integrated platform has generated strong risk-adjusted returns for our investors since our formation.

Position Overview:

CCI is developing a leading-edge Data Science platform, as staying at the forefront of data management and analytics is essential to our investment strategy. We are looking for motivated and detail-oriented Data Engineering Interns to join our Global Data Science & Technology team in our London office. The Data Engineering Intern will work closely with our Data Science, Data Engineering and Commercial teams to build and optimize data pipelines that power our analytics, forecasting, and investment decision-making processes. This is a hands-on technical internship ideal for someone who enjoys solving real-world data challenges, especially around ingesting, scraping, and managing large datasets across the commodity markets.

Responsibilities:

  • Develop and maintain robust data ingestion pipelines from various internal and external sources, including APIs, FTP endpoints, and cloud data providers.
  • Develop data ingestion and transformation pipelines using Python and SQL, publishing Snowflake for downstream use in analytics and forecasting tools.
  • Work on data architecture and data management projects for both new and existing data sources.
  • Design and implement ETL processes to clean, normalize, and store structured and semi-structured data in Snowflake, our core relational data warehouse.
  • Analyze data pipeline performance and implement optimizations to improve efficiency and reliability.
  • Conduct data quality checks and build validation logic to identify anomalies and ensure data integrity for use by commercial trading and analytics teams.
  • Automate data workflows using Python, SQL, and orchestration tools (e.g., Airflow or similar).
  • Assist in transitioning legacy datasets and codebases into scalable, cloud-native workflows aligned with our modern data architecture.
  • Document data sources, pipeline logic, and data models to ensure maintainability and knowledge transfer.

Qualifications:

  • Currently pursuing a Bachelors or higher degree in Computer Science, Engineering, Management Information Systems, or related technical field.
  • Expected graduation date of Winter 2027 or Spring/Summer 2028.
  • Strong programming experience in Python (preferred libraries: pandas, NumPy, SQL alchemy, etc.).
  • Strong understanding of SQL and experience querying relational databases (Snowflake a plus).
  • Exposure to or interest in cloud platforms (e.g., AWS, Azure), particularly with cloud data storage and compute.
  • Familiarity with web scraping frameworks and handling large-scale structured and unstructured data sources.

Data Engineering Internship (Summer 2027)

Compensation

Not specified

City: London

Country: United Kingdom

Castleton Commodities logo
Commodities

13 days ago

No clicks

at Castleton Commodities

SummerNo visa sponsorship

**Data Engineering Intern - Summer 2027, NYC** Join Castleton Commodities International (CCI), a global energy commodities merchant, to build our leading-edge data science platform. The Data Engineering Intern will collaborate with cross-functional teams to develop, optimize, and maintain robust data pipelines, crucial for our analytics and investment decisions. This hands-on role involves working with diverse data sources, including APIs and cloud providers, using Python, SQL, and Snowflake. Responsibilities: - Ingest data from various sources, parlaying it into Snowflake for analytics and forecasting tools. - Design and implement ETL processes, ensuring data quality and integrity. - Automate workflows using Python, SQL, and orchestration tools like Airflow. - Transition legacy datasets to scalable, cloud-native workflows aligned with modern data architecture. Requirements: - Pursuing a bachelor's degree in Computer Science, Engineering, MIS, or related field, with an expected graduation of Winter 2027 or Spring/Summer 2028. - Proficient in Python, SQL, and exposure to cloud platforms (AWS, Azure). - Familiarity with web scraping frameworks and handling large datasets. Application deadline: September 1st, 11:59m EST.

Full Job Description

Application Deadline: September 1st, 11:59m EST

Program Summary - Data Science & Technology Internship

Company Overview:

Castleton Commodities International is a leading global energy commodities merchant and infrastructure asset investor. As a trader, CCI deploys capital on a proprietary basis in the physical and financial commodity markets, providing the Company with market insights and access. As a strategic investor and developer, CCI leverages its market expertise, operations capabilities, and industry knowledge to invest in, and develop, select commodity infrastructure assets. Our strategically integrated platform has generated strong risk-adjusted returns for our investors since our formation.

Position Overview:

CCI is developing a leading-edge Data Science platform, as staying at the forefront of data management and analytics is essential to our investment strategy. We are looking for motivated and detail-oriented Data Engineering Interns to join our Global Data Science & Technology team in our London office. The Data Engineering Intern will work closely with our Data Science, Data Engineering and Commercial teams to build and optimize data pipelines that power our analytics, forecasting, and investment decision-making processes. This is a hands-on technical internship ideal for someone who enjoys solving real-world data challenges, especially around ingesting, scraping, and managing large datasets across the commodity markets.

Responsibilities:

  • Develop and maintain robust data ingestion pipelines from various internal and external sources, including APIs, FTP endpoints, and cloud data providers.
  • Develop data ingestion and transformation pipelines using Python and SQL, publishing Snowflake for downstream use in analytics and forecasting tools.
  • Work on data architecture and data management projects for both new and existing data sources.
  • Design and implement ETL processes to clean, normalize, and store structured and semi-structured data in Snowflake, our core relational data warehouse.
  • Analyze data pipeline performance and implement optimizations to improve efficiency and reliability.
  • Conduct data quality checks and build validation logic to identify anomalies and ensure data integrity for use by commercial trading and analytics teams.
  • Automate data workflows using Python, SQL, and orchestration tools (e.g., Airflow or similar).
  • Assist in transitioning legacy datasets and codebases into scalable, cloud-native workflows aligned with our modern data architecture.
  • Document data sources, pipeline logic, and data models to ensure maintainability and knowledge transfer.

Qualifications:

  • Currently pursuing a Bachelors or higher degree in Computer Science, Engineering, Management Information Systems, or related technical field.
  • Expected graduation date of Winter 2027 or Spring/Summer 2028.
  • Strong programming experience in Python (preferred libraries: pandas, NumPy, SQL alchemy, etc.).
  • Strong understanding of SQL and experience querying relational databases (Snowflake a plus).
  • Exposure to or interest in cloud platforms (e.g., AWS, Azure), particularly with cloud data storage and compute.
  • Familiarity with web scraping frameworks and handling large-scale structured and unstructured data sources.