Why This Role Stands Out
This fully remote Data Engineer role at ITBrainiac Inc. offers significant growth potential and the chance to work on impactful projects within a reputable technology company. You'll thrive here if you're a seasoned professional with extensive experience, seeking a dynamic and flexible work environment. Apply today to leverage your expertise and contribute to innovative solutions!
Quick Overview
Job Description
Hello,
I hope you’re doing well. I wanted to share a new opportunity that might interest you:
Data Engineer
Location: Denver, Colorado
Duration: 6–8 Months Contract – 100% remote
Note:
- 15 Years + Exp Mandatory
- The client is willing to make this position 100% remote, but the final interview will be face-to-face.
- Face-to-face interview is mandatory in Denver, Colorado.
- Flight ticket reimbursement can be requested from the client for the final face-to-face interview.
- No fake profiles. Visa candidates are fine.
Position Overview
The Data Engineer IV designs, builds, and maintains the ETL pipelines and data infrastructure that feed the IIA Data Lake, anomaly detection models, and AI agents. This role focuses on constructing robust, scalable data pipelines using Spark/Scala, ensuring data quality and availability across a growing portfolio of network data sources, and enabling downstream consumers (data scientists, agents, dashboards) to access reliable, well-structured data.
Responsibilities
Design, develop, and maintain scalable ETL pipelines using Apache Spark (Scala) to ingest, transform, and load network data into the IIA Data Lake.
Onboard new data sources (network telemetry, syslogs, SNMP traps, device configuration data, ticketing systems) by building ingestion pipelines from raw source to query-ready format.
Implement monitoring and alerting solutions to ensure data pipeline reliability and performance.
Develop and manage deployment pipelines to facilitate continuous integration and delivery of data engineering solutions.
Manage and optimize data storage solutions, including distributed file systems, relational databases, flat files, and external source access via API.
Implement data quality checks, validation rules, and automated testing to ensure pipeline reliability and data integrity.
Optimize pipeline performance for large-scale data processing (billions of events per day) across batch and mini-batch processing patterns.
Manage and evolve data schemas, partitioning strategies, and storage formats to support efficient querying and downstream consumption.
Support data backfills and recovery when upstream issues or schema changes require reprocessing.
Collaborate across teams to ensure data solutions align with existing production architectures and business requirements.
Work with data scientists and agent developers to understand data requirements and deliver datasets that support anomaly detection models and AI agent workflows.
Provide technical guidance on data engineering best practices and methodologies.
Document and communicate data engineering processes and standards to business intelligence, data, and analytics professionals with varied backgrounds.
Continuously evaluate and improve data engineering tools and approaches to enhance performance and efficiency.
Perform other duties as required.
Required Qualifications
Expertise in Scala (preferred) or Java, with proficiency in Python.
Strong experience with Apache Spark for distributed data processing.
Proficiency in building and maintaining ETL pipelines at scale.
Experience with AWS services: S3, Glue, Athena, EMR.
Strong understanding of relational databases and SQL.
Knowledge of data architecture, data warehousing, partitioning strategies, and columnar storage formats (e.g., Parquet).
Experience implementing data quality checks and validation frameworks.
Experience with workflow orchestration tools (Airflow preferred).
Proficiency with Linux-based operating systems and shell scripting.
Experience with Git-based version control and collaborative development workflows.
Demonstrated ability and desire to continually expand skill set and learn from and teach others.
Preferred Qualifications
Experience with streaming or mini-batch data processing (Spark Streaming, Structured Streaming, or similar).
Experience with Apache Kafka or similar messaging/streaming platforms.
Experience with NoSQL databases.
Experience in the telecommunications industry or other large-scale network operations environments.
Familiarity with network data sources: telemetry, syslogs, SNMP traps, device configuration.
Experience with data integration via REST APIs and cloud SDKs (e.g., boto3).
Experience writing automated tests for data pipelines.
Knowledge of text analysis or log parsing techniques.
Thanks!
Ranjeeth Sesham
Senior IT & Non-IT Recruiter
ITBrainiac Inc.
116 Village Blvd, Princeton
Forrestal Village, Suite 200
Princeton - New Jersey 08540
CPUC, NCMSDC – MBE
Notice: This email contains confidential or proprietary information which may be legally privileged. It is intended only for the named recipient (s). If an addressing or transmission error has misdirected the email, please notify the author by replying to this message.
Disclaimer: This is not meant to be an unsolicited email, So If you want to be removed, please reply with REMOVE in the Subject & I will promptly remove upon receipt
Similar jobs
- TR
Senior Data Engineer-Azure and Databricks
NewTechnology Recruiting Solutions,Inc.
Houston, TX🇺🇸Hybrid21 hours agoOraclePL/SQLSQL+14Technology - TA
Data Engineer (Ab-initio)
NewTandavix llc
Mississauga, ON🇺🇸Hybrid21 hours agoOracleSQLETL+2Technology - RI
AWS Lead Data Engineer
NewRivago infotech inc
Newark, NJ🇺🇸On-site21 hours agoSQLScalaShell+7Technology - VC
Principal Microsoft Fabric and Data Engineer (W2/NoOPT)
NewVoto Consulting LLC
Irving, TX🇺🇸Hybrid21 hours agoSQLSQL ServerETL+9Technology - KS
Lead Data Engineer
NewKeypixel Software Solutions
United States🇺🇸Hybrid21 hours agoMongoDBAPI GatewayETL+5Technology - CS
Data Engineer - Remote / Telecommute
NewCynet Systems
Charlotte, NC🇺🇸$50 - $55/hrRemote21 hours agoSQLSOC 2Agile+1Technology