← Jobs

Python & AI Developer (m/w/d)

Budget: $2000.0 FIXED / ⭐ 5.00 (27) Germany

google-sheets

About the Role We are looking for an experienced Python & AI Developer to transform our existing AI solution into a robust, scalable, and fully automated production pipeline. We already have a working prototype (currently implemented as a Claude Skill) that successfully identifies topic-specific articles on individual publisher websites. The goal is to evolve this prototype into a production-ready architecture capable of processing thousands of domains reliably and cost-effectively. You will have full ownership of the technical design and are free to choose the architecture and technology stack. Your Responsibilities Develop a scalable Python application for automated website analysis. Build a fully unattended pipeline that can be manually triggered. Process 5,000–6,000 domains per month within a defined processing window. Crawl and analyze publisher websites. Identify topic-specific articles based on predefined rules. Extract relevant structured information from matching articles. Automatically export the results to Google Sheets or other spreadsheet formats. Optimize the solution for performance, reliability, scalability, and API costs. Project Goals The completed solution should meet the following requirements: Reliably process 5,000–6,000 domains per month. Run fully automatically after being manually triggered. Achieve a minimum 75% success rate on domains where our labeled test dataset confirms that a valid article exists. Keep total monthly tool and API costs below approximately €1,000 at the target processing volume. Include robust error handling, logging, and recovery mechanisms. Existing Logic Our current prototype already performs the following tasks for individual domains: Discover relevant article URLs. Retrieve and analyze webpage content. Validate content against predefined business rules. Extract structured information. Export the results into a spreadsheet. Your task is to preserve and improve this logic while rebuilding it as a scalable, production-ready infrastructure. Your Profile Strong proficiency in Python. Experience with web crawling, web scraping, and data extraction. Experience designing and building scalable data pipelines. Experience working with LLMs, AI agents, or AI workflows (e.g., Claude, OpenAI, or similar models). Experience integrating third-party APIs and optimizing operational costs. Familiarity with the Google Sheets API or comparable export solutions. Ability to work independently and design scalable software architectures. Nice to Have Experience with any of the following is a plus: Async Python Playwright, Selenium, or Scrapy BeautifulSoup or lxml Docker Cloud platforms (AWS, GCP, or Azure) Queue systems such as Redis or RabbitMQ LangGraph, LangChain, or similar AI orchestration frameworks What We Expect We care more about the outcome than the specific technology stack. The solution should be reliable, scalable, cost-efficient, and easy to maintain while automatically processing several thousand domains each month. You will have significant freedom in choosing the architecture and implementation, provided the final solution meets our objectives for accuracy, scalability, reliability, and cost efficiency.
Open job