Jr. Data Engineer
indonesia indicator- Posted 3 hours ago
- Be among the first 10 applicants
Job Description
Company Description
PT. Indonesia Indicator is a strategic intelligence company that focus on Big Data and AI. We provide intelligence insight and forecasting using Osint (Open Source Intelligence) Dataint (Data Intelligence) and Humint (Human Intelligence) to make a better decisions.
We are looking for Jr. Data Engineer position for working onsite at Bintaro, South Tangerang.
Role Description
As a Jr. Data Engineer, you will be responsible for building and maintaining reliable data pipelines that support analytics and AI-driven applications. You will prepare and deliver high-quality structured and unstructured data for model training, inference, forecasting, and strategic intelligence. You will work closely with product, analyst, and AI/ML teams to ensure data accessibility, scalability, governance, and quality across systems.
Key Responsibilities
- Design and implement ETL pipelines to process structured and unstructured data.
- Develop automation scripts for data scraping, cleansing, and loading into storage.
- Build and manage relational and non-relational databases, including writing SQL queries.
- Collaborate on data flow architecture and ensure data availability for analytics.
- Utilize Docker, Git, and Linux for deployment, versioning, and process automation.
- Monitor and debug data pipeline processes with proper logging and error handling.
- Integrate web data sources using HTTP protocols, asynchronous programming, and browser automation tools.
- Prepare structured and unstructured data pipelines to support analytics and practical AI use cases.
Requirements
- Bachelor's degree in Informatics, System Information, Engineering, or a related fields.
- Strong understanding of database types (relational & non-relational) and ERD modeling.
- Experience with Extract Transform Load (ETL) processes.
- Proficient in SQL (SELECT, INSERT, UPDATE, DELETE) and integration into code.
- Experience with data scraping tools (BeautifulSoup, Selenium, Playwright) and Python libraries (requests, re, asyncio).
- Familiarity with message queue systems and asynchronous processing.
- Hands-on experience with Linux commands and Docker containerization.
- Knowledge of version control (Git) and basic unit testing in data pipelines.
- Comfortable with HTTP protocols and using CLI tools for debugging.
- Basic understanding of AI/ML concepts and familiarity with preparing or integrating data for AI applications; knowledge of LLMs or AI APIs is a plus.
