PitchMeAI
Sundayy

Data Engineer

Sundayy · United States

  • Hybrid
  • Full-time
  • $140,000 / year
  • United States
Tailored resumekeyword-matched to this role.
Hiring managerwe find who's hiring.
Intro emaildrafted to reach them directly.

Job highlights

  • Build foundational data systems for AI platform.
  • Design and implement scalable data pipelines.
  • Transform operational data into reliable products.
  • Collaborate with cross-functional teams.
  • Shape core data strategies and architecture.

About the role

About The Company

Nscale is a pioneering GPU cloud platform specifically engineered for artificial intelligence applications. Our mission is to provide cost-effective, high-performance infrastructure tailored for AI start-ups and large enterprise customers. By simplifying the complexities inherent in AI development, Nscale empowers companies to achieve superior results through optimized resource utilization and streamlined workflows. Our platform enhances technical capabilities and directly supports strategic business outcomes such as cost management, rapid innovation, and environmental responsibility. We foster a culture rooted in relentless innovation, ownership, and accountability, where every team member takes pride in delivering excellence with a sense of urgency. At Nscale, transparency and openness are core values that inspire our team to do their best work. Joining us means contributing to the development of cutting-edge technology that shapes the future of AI and powering the next generation of intelligent solutions.

About The Role

We are seeking a highly skilled Data Engineer to join our dynamic team. In this pivotal role, you will be responsible for designing, building, and maintaining the foundational data systems that support Nscale’s platform, internal operations, and customer-facing services. This is an early-stage, high-impact position that offers the opportunity to influence core data strategies and architecture. You will collaborate closely with Operations, Infrastructure, Platform Engineering, Product, and Commercial teams to transform raw operational signals—collected from GPUs, clusters, customers, and internal systems—into reliable, scalable data products. Your work will enable the company to operate at an unprecedented pace, ensuring data is accurately collected, modeled, served, and trusted across various domains. This role is ideal for someone who enjoys building data systems from first principles, thrives in ambiguous environments, and desires a direct impact on product decisions, platform reliability, and customer success. You will also have the chance to work with advanced tools such as Palantir Foundry and contribute to establishing data standards and best practices as the company scales.

Qualifications

  • Extensive hands-on experience with Palantir Foundry, including ontology modeling, pipeline development, API integration, and large-scale data platform design.
  • Proficiency in Python, with practical experience applying data engineering libraries such as Spark, PySpark, Dask, and pandas.
  • Familiarity with API-driven data integration methods, including REST, GraphQL, and Foundry Action APIs.
  • Experience working within Git-based development workflows, including code reviews, version control, and CI/CD pipelines.
  • Comfort working in ambiguous, early-stage environments where requirements evolve rapidly.
  • Strong communication skills, capable of articulating complex data concepts clearly to both technical and non-technical stakeholders.
  • A proactive ownership mindset, pragmatic approach to problem-solving, and focus on delivering practical solutions.

Responsibilities

  • Design and develop scalable, reliable data pipelines to ingest data from infrastructure, platform services, and business systems.
  • Define data models and schemas that support operational workflows, monitoring, and analytics use cases.
  • Clean, transform, and structure data to create a comprehensive digital twin of Nscale’s environment.
  • Implement permissioning, manage access, and ensure security within the Foundry platform.
  • Create trusted datasets and metrics that power internal workflows, tools, and customer-facing insights.
  • Establish self-serve analytics capabilities through clear data contracts, documentation, and semantic layers.
  • Develop use cases such as capacity planning, cost optimization, reliability analysis, and customer reporting to drive strategic growth.
  • Collaborate with Product and Commercial teams to translate business questions into robust data solutions.
  • Implement data quality checks, monitoring, and alerting to ensure data accuracy and availability.
  • Codify data lineage, freshness, and consistency across multiple systems.
  • Establish best practices around data versioning, access control, and governance suitable for a rapidly scaling organization.
  • Continuously improve system resilience, observability, and overall data infrastructure.
  • Take end-to-end ownership of data projects from conception through deployment and iteration.
  • Help define standards, tooling, and best practices for data management at Nscale.
  • Participate in technical decision-making processes as the platform and customer base expand.
  • Act as a thought partner to engineering and operational teams, fostering a culture of data-driven decision making.

Benefits

  • Highly competitive salary package with annual reviews and potential bonuses.
  • Equity participation, offering long-term growth opportunities.
  • Collaborative, innovative, and supportive work environment that values your contributions.
  • Opportunities for professional development and career progression in a fast-growing tech startup.
  • Flexible, remote-first work policy with seamless virtual collaboration tools.
  • Comprehensive benefits including medical, dental, and vision insurance.
  • Flexible paid time off, parental leave, and retirement plan options.
  • Chance to work on cutting-edge AI infrastructure and influence the future of AI technology.

Equal Opportunity

Nscale is committed to creating an inclusive environment and is proud to be an equal opportunity employer. We encourage applications from individuals of all backgrounds, including people of color, members of the LGBTQ+ community, individuals with disabilities, neurodivergent individuals, parents, caregivers, and those from lower socio-economic backgrounds. We strive to provide accommodations to support your application process and ensure an equitable workplace where everyone can thrive. Diversity, equity, and inclusion are integral to our culture and success.

Key skills/competency

  • Data Engineering
  • Palantir Foundry
  • Python
  • Data Pipelines
  • Data Modeling
  • API Integration
  • Scalability
  • Data Governance
  • CI/CD
  • Cloud Infrastructure

Skills & topics

  • Data Engineer
  • Data Pipelines
  • Palantir Foundry
  • Python
  • Spark
  • PySpark
  • Dask
  • pandas
  • API Integration
  • Cloud Computing
  • AI Infrastructure
  • Data Modeling
  • Data Governance
  • CI/CD
  • Remote Work
  • Startup

How to get hired

  • Tailor your resume: Highlight experience with Palantir Foundry, Python, and data pipeline development.
  • Showcase ownership: Emphasize your proactive mindset and experience in ambiguous environments.
  • Quantify impact: Provide examples of how your data solutions drove business outcomes.
  • Prepare for technical questions: Be ready to discuss data modeling, API integration, and CI/CD.
  • Demonstrate cultural fit: Express your understanding of Nscale's innovative and accountable culture.

Technical preparation

Master Palantir Foundry's ontology and pipelines.,Practice Python with Spark, PySpark, Dask, pandas.,Build data pipelines for diverse data sources.,Understand API-driven data integration methods.

Behavioral questions

Describe building systems from first principles.,How do you handle ambiguous early-stage environments?,Give an example of taking ownership of a project.,How do you translate business needs into data solutions?

Frequently asked questions

What specific experience with Palantir Foundry is required for the Data Engineer role at Nscale?
For the Data Engineer position at Nscale, extensive hands-on experience with Palantir Foundry is essential. This includes proficiency in ontology modeling, developing data pipelines within the platform, integrating with its APIs, and designing large-scale data platforms using Foundry. Demonstrating practical application of these skills will be key.
How does Nscale leverage data to drive strategic business outcomes?
Nscale leverages data to drive strategic business outcomes by transforming raw operational signals into reliable, scalable data products. This enables precise capacity planning, cost optimization, reliability analysis, and customer reporting, directly supporting AI start-ups and enterprise customers in managing costs, fostering innovation, and ensuring environmental responsibility.
What is the work environment like for a Data Engineer at Nscale?
The Data Engineer role at Nscale is in an early-stage, high-impact, and remote-first environment. The culture emphasizes relentless innovation, ownership, accountability, transparency, and openness. You'll collaborate closely with various teams and have the opportunity to build data systems from first principles in an ambiguous, fast-paced setting.
What Python data engineering libraries are most relevant for this Data Engineer role at Nscale?
For this Data Engineer role at Nscale, proficiency in Python is required, with practical experience in libraries such as Spark, PySpark, Dask, and pandas. These tools are crucial for building scalable and efficient data pipelines, transforming data, and developing robust data products within the Nscale ecosystem.
How does Nscale support professional development for its Data Engineers?
Nscale offers opportunities for professional development and career progression, especially within its fast-growing startup environment. As a Data Engineer, you'll have the chance to work on cutting-edge AI infrastructure, influence core data strategies, and contribute to defining standards and best practices, providing significant learning and growth potential.