Position Title : IT Service Desk Specialist – Level II–MEADE
Position Type : Full Time, Remote (Telework in Metro DC area)
Position Location : Arlington, VA
Tracking Code : 01131
Daily Responsibilities
- Positionsupports the development and maintenance of the agency’sweb scraping infrastructure.The positionis responsible forextracting data from various websites and APIs, ensuring data quality and accuracy, andoptimizingthe scraping process for efficiency.Dutiesinclude:
- Develop andmaintainweb scraping scripts and tools to extract data from websites and APIs.
- Collaborate with cross-functional teams to understand data requirements and implement scraping solutions accordingly.
- Monitor and troubleshoot scraping processes to ensure data quality and accuracy.
- Optimizescraping scripts for performance and efficiency, considering factors such as speed, scalability, and resourceutilization.
- Stayup-to-datewith the latest web scraping techniques, tools, and best practices.
- Conduct data analysis and validation to ensure the integrity of scraped data.
- Collaborate with data engineering and data science teams to integrate scraped data into our data pipelines and systems.
- Document and communicate technical solutions, processes, and best practices to team members.
Required Experience
- 3+ years of professional experience in web scraping or a similar role.
- Proficiencyin Pythonand Javaand experience with web scraping libraries such asBeautifulSoup, Scrapy, or Selenium.
- Knowledge of AI/machine learning techniques for data extraction and classification.
- Experience working with APIs and handling different data formats (JSON, XML, etc.).
- Familiarity with database systems and SQL for data storage and retrieval.
- Familiarityof data cleaning and preprocessing techniques to ensure data quality.
- Strong problem-solving skills and ability to troubleshoot and debug scraping issues.
- Excellent communication and collaboration skills to work effectively in a team environment.
- Attention to detail and ability to handle large volumes of data efficiently.
Preferred Experience
- Experience with cloud platforms for scalable web scraping infrastructure.
- Familiarity with data visualization tools and techniques.
- Understanding oflegal and ethical considerations related to web scraping.
Required Degree
- BS/BA degree in Computer Science, Information Sciences, or related IT discipline OR Allowable Substitution: Additional ten (10) years of related professional experience can be substituted for a BS/BA degree.
Required Clearance
- Secret