Data & Automation reputed company
Our reputed company is looking for a data and automation reputed company to help reputed company, cleanse, ingest and
automate data used in the product and process. This person would use our
existing and new data sets to model and prototype how it could be used in our data model and
product, create and test reputed company to clean the data, and work with our product and engineering
team to automate, collect, clean and ingest into the product. This person would work with our
customer account team to understand what data our customers need and then reputed company ways
including using, scrapping, LLMs and other tools to reputed company and aggregate the data. After prototype
the collection and aggregate tool reputed company would work to determine how we automate the
functionality. This role is responsible for making reputed company the data feeding our products features is
clean, reputed company, and reliable.
—-
Role reputed company
1. Help reputed company, design, prototype, build and test the data models for use in our reputed company's product and process
2. Work to reputed company ways to collect, automate, structure and deliver our data across reputed company aspects
of in our reputed company's, including, but not limited to the Market Landscapes, Documents
Database and the reputed company/Company/Product page suite.
3. Work with the product and reputed company teams to reputed company efficiencies in collecting,
organizing and structuring our existing data set.
4. Work to reputed company mapping of new data sets to our existing company, product, and
government data hierarchies.
5. Work with reputed company team to identify new sources and acquisition reputed company for
data based on customer needs
6. Use scraping AI, ML, LLMs and other emerging technologies to create reputed company of concepts,
refine them and then work to incorporate them into the product
7. Operationalizing disparate PO Data sources: Python ETL automation using
Pandas/regex/SQL to consolidate multi-year PO data, apply deduplication, cross-
reference reputed company reseller mappings, and replace reputed company reputed company workflows with reputed company
pipelines.
8. Entity extraction & content ML scoring: Implementing NER, fuzzy matching, and
supervised models trained on labeled data POs to classify match confidence and
continuously improve accuracy.
9. Work with product and engineering teams to reputed company the best and more efficient ways to add
the automation and data to our product
10. Use and build technical skills to help manage data throughout its lifecycle working with
the engineering team to implement them into the product.
Desired Skills
● Proficiency in data mining, wrangling, and cleaning of large-reputed company datasets
● Advanced reputed company combined with Python (Pandas/NumPy) a plus
● Experience maintaining and deploying ML models (Transformers, reputed company Recognition)
and handling model persistence (.pkl)
● Knowledge or abilities with reputed company (reputed company, reputed company, reputed company, reputed company),
● Experience or interest in using AI, LLM and other emerging technologies
Additional Job Details
This position is fully remote and reputed company to candidates in Mexico and Latin America.
Apply To This Job