Vulcury · Remoto · United States
Vulcury LLC is a venture studio and advisory firm that supports the development of data-driven businesses across multiple industries. The firm provides shared services—including research, analytics, and technical leadership—to early-stage operating companies as they move from concept through execution.
This role will support an internal initiative that is currently focused on preparing high-quality data assets and workflows before a broader software and analytics build planned for later this year.
Vulcury is seeking aSenior Data Scientistto lead an early-stagedata preparation and data-quality initiative. The primary objective of this role is to ensure that datasets are well-structured, well-documented, and operationally usable before more extensive development work begins.
The role is structured aspart-time initially, with a clear expectation that it may transition tofull-timeas the project moves into its next phase. The Senior Data Scientist will serve as the technical lead for data-related work and will oversee a small, India-based team of junior developers and/or data scientists.
This position emphasisesdata rigour, process design, and team leadership, rather than production model deployment at this stage.
Define and oversee standards for data organization, labelling, and validation across multiple datasets.
Establish practical frameworks to improve consistency, accuracy, and usability of structured and semi-structured data.
Ensure datasets are reviewable, auditable, and suitable for downstream analytical or software development work.
Develop repeatable workflows for data ingestion, cleaning, documentation, and version control.
Identify appropriate tools and practices to support efficient data operations without premature complexity.
Create clear documentation that enables continuity as the team scales.
Manage and mentor a distributed, India-based team performing data preparation and technical support tasks.
Translate high-level requirements into clear, actionable work-streams.
Review outputs and provide feedback to maintain quality and consistency.
Work directly with Vulcury leadership to align data efforts with near- and medium-term development plans.
Support future engineering and analytics hires by delivering well-prepared datasets and clearly defined processes.
Contribute input to planning for subsequent phases of technical development.
Expected commitment of approximately10–20 hours per weekduring the initial phase.
Flexible structure with a clearly defined path to full-time engagement.
7+ years of experience in data science, analytics, or data-focused engineering roles.
Strong proficiency in Python and common data-analysis libraries.
Demonstrated experience transforming imperfect or unstructured data into reliable, usable datasets.
Prior experience setting data standards, documentation practices, and quality-control processes.
Experience managing or mentoring distributed or offshore technical teams.
Comfort operating in early-stage or ambiguous environments with limited upfront specification.
Background supporting research-heavy, analytical, or data-intensive products.
Familiarity with text-heavy or document-based data sources.
Exposure to data modelling or schema design concepts.
Experience working in startups, venture studios, or early product teams.
Foundational responsibility:You will shape the data practices that future development depends on.
Growth opportunity:The role is expected to expand as the initiative progresses into a full build phase.
Leadership visibility:Direct interaction with senior leadership and meaningful ownership of outcomes.
Global team management:Opportunity to lead and develop a distributed technical team.
Aligned incentives:Long-term upside participation anticipated upon full-time conversion.
Originally posted on Himalayas
Inicia sesión para generar una carta de presentación para esta vacante.
Iniciar sesiónInicia sesión para ver cómo encaja este empleo con tu perfil.
Iniciar sesión