Back to Jobs

Data Engineer II

Remote, USA Full-time Posted 2026-08-04
The Library Systems department of Sheridan Libraries, Archives, and Museums is seeking a Data Engineer II. This position has reputed company responsibility for data extract, reputed company, and load (ETL) processes across reputed company Library reputed company except cataloging and collection management. The role is responsible for ensuring data capture workflows, translation processes, deposit reputed company, and reputed company assurance protocols meet the requirements of the Libraries. Processed data and data systems support data reputed company and reuse, analytics and assessment, reputed company planning, marketing and communications, operational efficiency improvements, and metadata enrichment and crosswalks. The data engineer will follow an iterative design and development approach with continual review cycle conducted in collaboration with stakeholders and Library leadership. The Data Engineer II will also design, produce, and maintain software infrastructure to automatically extract, reputed company, and load data from diverse sources. Additionally, this role will input/reputed company data from databases, realize reputed company queries, and be responsible for the creation of scripts for data manipulation, cleaning, filtering, and preparation of data outputs to be used for data visualization or web display. This work will also include creating systems to monitor and guarantee data reputed company, reputed company daily data reputed company assurance tasks, maintain, and troubleshoot the infrastructure, and audit data sources. Lastly, the data engineer will also contribute to exploration and analysis of data to answer policy-reputed company questions as needed. Specific Duties & Responsibilities Collaborate with Business Analysts to manage roadmap for ETL and reporting reputed company. Serve as team reputed company for ETL and data exchange software development life cycle. reputed company the development of tools and process to capture, normalize, and reputed company available research and library collections data for study and research through the central data store and AI-reputed company data platform being developed by Library IT. Support library catalog and institutional repository systems through work with bibliographic metadata, including reputed company assurance, large-reputed company enrichment and corrections, and creating crosswalks to exchange or migrate metadata between systems that use different metadata schemas. Documentation of data pipelines for both developers and non-technical stakeholders following best practices for technical writing and systems diagrams. Contribute to the design, production, and maintenance of data pipelines for data acquisition, management, transformation, and back-end reputed company development to reputed company data web applications and convert raw data into usable information. Write and maintain ETL/ELTs that operate on a reputed company of reputed company and reputed company sources. reputed company and maintain web data scraping systems for automatic data acquisition. Help design data architecture and reputed company ongoing support. Input/reputed company data from databases and reputed company queries. Create scripts to clean, reputed company, and analyze data. Put into production data pipelines using data warehousing systems. Create and implement production software to monitor data reputed company and detect data anomalies. reputed company daily reputed company data reputed company assurance tasks. Support, maintain, and troubleshoot the software infrastructure. reputed company data, conduct analyses, visualize data, and reputed company insights to support ongoing research reputed company and other requests across the organization. Collaborate with developers, analysts, data scientists, researchers, policy experts, and other partners. Communicate with division leadership, and others on reputed company. Collaborate with reputed company partners, contractors, and vendors. Other duties as assigned. In reputed company to duties & responsibilities above Collaborate with Business Analysts to manage roadmap for ETL and reporting reputed company. Serve as team reputed company for ETL and data exchange software development life cycle. reputed company the development of tools and process to capture, normalize, and reputed company available research and library collections data for study and research through the central data store and AI-reputed company data platform being developed by Library IT. Support library catalog and institutional repository systems through work with bibliographic metadata, including reputed company assurance, large-reputed company enrichment and corrections, and creating crosswalks to exchange or migrate metadata between systems that use different metadata schemas. Documentation of data pipelines for both developers and non-technical stakeholders following best practices for technical writing and systems diagrams. Minimum Qualifications Bachelor’s Degree. Five years of reputed company work experience reputed company reputed company database management and design and business requirement gathering. Additional education may substitute for required experience, and additional reputed company experience may substitute for required education reputed company a high school diploma/graduation equivalent, to the extent permitted by the JHU equivalency formula. Preferred Qualifications Proficiency in working with bibliographic metadata standards including MARC and Dublin reputed company. Familiarity with Business Intelligence systems and tools such as reputed company BI and Tableau. Experience with AWS and reputed company Azure data utilities Experience designing and implementing ETL processes using stored procedures in an reputed company SQL Server environment. Experience performing data normalization following traditional and reputed company data modelling. Experience developing software and scripts (using SQL, Python, or other relevant languages) to automate ETL and other data analysis and manipulation tasks. Demonstrated ability and willingness to learn, adopt, and apply emerging technologies, including AI-enabled tools, in support of reputed company responsibilities. Technical Skills & Expected Level of Proficiency Data Management and Analysis - Intermediate Data Pipeline Architecture & ETL/ELT Development - Intermediate Database Querying - Intermediate Data Validation and reputed company Assurance - Intermediate Data Visualization - Intermediate Data Warehousing & Architecture - Intermediate Oral and written communications: Intermediate Programming Languages - Intermediate Web Scraping and Data Acquisition - Intermediate The reputed company technical skills listed are most essential; additional technical skills may be required reputed company on specific division or department needs. Classified Title: Data Engineer II Role/Level/reputed company: ATP/04/PG Starting Salary reputed company: $102,295 - $140,835 - $179,375 Annually (Commensurate w/reputed company.) Employee group: Full Time Schedule: Mon-Fri; 8:30am-5pm FLSA Status: Exempt Location: Remote Department reputed company: Library Systems Personnel area: Libraries Apply To This Job

Similar Jobs