Unknown Company

Web Scraping Engineer (Mid-Level)

arlington, va • Posted 3 days ago
Remote Full Time IT & Technology

Position Title : IT Service Desk Specialist – Level II–MEADE

Position Type : Full Time, Remote (Telework in Metro DC area)

Position Location : Arlington, VA

Tracking Code : 01131

Daily Responsibilities

  • Positionsupports the development and maintenance of the agency’sweb scraping infrastructure.The positionis responsible forextracting data from various websites and APIs, ensuring data quality and accuracy, andoptimizingthe scraping process for efficiency.Dutiesinclude:
  • Develop andmaintainweb scraping scripts and tools to extract data from websites and APIs.
  • Collaborate with cross-functional teams to understand data requirements and implement scraping solutions accordingly.
  • Monitor and troubleshoot scraping processes to ensure data quality and accuracy.
  • Optimizescraping scripts for performance and efficiency, considering factors such as speed, scalability, and resourceutilization.
  • Stayup-to-datewith the latest web scraping techniques, tools, and best practices.
  • Conduct data analysis and validation to ensure the integrity of scraped data.
  • Collaborate with data engineering and data science teams to integrate scraped data into our data pipelines and systems.
  • Document and communicate technical solutions, processes, and best practices to team members.

Required Experience

  • 3+ years of professional experience in web scraping or a similar role.
  • Proficiencyin Pythonand Javaand experience with web scraping libraries such asBeautifulSoup, Scrapy, or Selenium.
  • Knowledge of AI/machine learning techniques for data extraction and classification.
  • Experience working with APIs and handling different data formats (JSON, XML, etc.).
  • Familiarity with database systems and SQL for data storage and retrieval.
  • Familiarityof data cleaning and preprocessing techniques to ensure data quality.
  • Strong problem-solving skills and ability to troubleshoot and debug scraping issues.
  • Excellent communication and collaboration skills to work effectively in a team environment.
  • Attention to detail and ability to handle large volumes of data efficiently.

Preferred Experience

  • Experience with cloud platforms for scalable web scraping infrastructure.
  • Familiarity with data visualization tools and techniques.
  • Understanding oflegal and ethical considerations related to web scraping.

Required Degree

  • BS/BA degree in Computer Science, Information Sciences, or related IT discipline OR Allowable Substitution: Additional ten (10) years of related professional experience can be substituted for a BS/BA degree.

Required Clearance

  • Secret

#J-18808-Ljbffr
Back to Job Search