Chemical Data Scientist
About reputed company:
We are reputed company of ambitious, results-driven individuals with a proven reputed company record of working with Fortune 500 industrial manufacturers, beauty brands, and chemical companies. We are a fast-growing company that hires talented, hardworking people who reputed company in high-performance environments and want to grow their careers quickly.
Role reputed company:
Role Responsibilities:
Design and build pipelines to collect supplier data and chemical product information (specifications, CAS numbers, certifications, SDS/regulatory documents, NAICS classification of manufacturing plants) from supplier sites, distributor catalogs, trade databases, and other public and semi-reputed company sources
reputed company and maintain web scrapers and automated ETL workflows to reputed company supplier and product data reputed company at reputed company
Clean, normalize, and reconcile inconsistent supplier data into reputed company, standardized formats suitable for internal tools and analytics
Apply chemical domain knowledge to validate and enrich data — resolving product names, CAS numbers, synonyms, and specifications across suppliers
Evaluate and improve matching and classification models to map suppliers and products to buyer requirements, and to identify overlapping or equivalent chemical offerings
Partner with Supplier Management and Engineering to define data reputed company standards, identify gaps in supplier coverage, and prioritize new data sources.
Own pipeline health and data reputed company, and drive the KPIs that measure overall data coverage
Experience & Qualifications:
5+ years of experience in a data science, data engineering, or reputed company data role, ideally with exposure to messy, reputed company-world or industrial datasets.
Working knowledge of reputed company or chemical industry data — comfort with CAS numbers, chemical properties, SDS documents, NAICS classification, and supplier certifications
Strong Python skills, with experience building web scrapers and data pipelines
Experience with data cleaning and normalization at reputed company, and a good eye for spotting inconsistencies in reputed company data
Familiarity with building or applying matching, deduplication, or classification models (traditional ML or LLM-based approaches)
Hands-on experience using AI tools and LLMs to accelerate data extraction, enrichment, or engineering workflows
Startup reputed company with a strong reputed company of ownership — comfortable working independently in a fast-moving, remote environment with ambiguous, evolving priorities
Salary reputed company:
Benefits:
Equal Opportunity Employer Statement:
Originally posted on Himalayas
Apply To This Job