Gratitude Inc banner
Gratitude Inc logo

Data Engineer

Gratitude Inc
25 Views
2 days ago

Data Engineer

8-10 Year(s)
Mumbai, Kolkata, Delhi
Mumbai, Kolkata, Delhi

Job Description

Key Skills

Understanding of AWS/Data bricks concepts Experience with Python/R, SQL, NoSQL Cloud experience (i.e. AWS, AZURE or GCP) Experience with GitLab, GitHub Experience with Jenkins, GitLab

4 candidate(s) have already applied for this Job. Apply now

Experience in data engineering, building data pipelines to manage heterogenous data ingestions or similar in data integration across multiple sources including collected data. • Experience with Python/R, SQL, NoSQL • Cloud experience (i.e. AWS, AZURE or GCP)

Educational Qualification:
• Bachelor's degree in computer science, statistics, biostatistics, mathematics, biology or other health related field or equivalent experience that provides the skills and knowledge necessary to perform the job.

Experience:
• BS with ~8+ years experience. Minimum of 5 years’ experience in data engineering, building data pipelines to manage heterogenous data ingestions or similar in data integration across multiple sources including collected data.
• Experience with Python/R, SQL, NoSQL
• Cloud experience (i.e. AWS, AZURE or GCP)
• Experience with GitLab, GitHub
• Experience with Jenkins, GitLab
• Experience deploying data pipelines in the cloud
• Experience with Apache Spark (databricks)
• Experience setting up and working with data warehouse, data lakes (eg: snowflake, Amazon RedShift etc.,)
• Experience setting up ELT and ETL
• Experience with unstructured data processing and transformation
• Experience developing and maintaining data pipelines for large amounts of data efficiently
• Must understand database concepts. Knowledge of XML, JSON, APIs.
• Demonstrated ability to lead projects and work groups. Strong project management skills. Proven ability to resolve problems independently and collaboratively.
• Must be able to work in a fast-paced environment with demonstrated ability to juggle and prioritize multiple competing tasks and demands.
• • Ability to work independently, take initiative and complete tasks to deadlines.

Special Skills/Abilities:
• Strong attention to detail, and organizational skills
• Strong Project Management skills
• Strong understating of end-to-end processes for data collection, extraction and analysis needs by end users
• Strong ability to communicate with cross functional stakeholders
• Strong ability to develop technical specifications based on communication from stakeholders
• Quick learner and comfortable asking questions, learning new technologies and systems
• Good knowledge of office software (Microsoft Office).
• Experience creating custom functions Python/R
• Cloud computing (AWS, Snowflakes, Databricks)
• Ability to visualize large datasets
• R shiny/ Python App experience a plus

Preferable:
• Experience developing R shiny and Python apps
• Experience with Hadoop
• Experience with Agile development methods

Behavioral Competencies:
• • Is comfortable with ambiguity.
• • Excellent teamwork, organizational, interpersonal, conflict resolution and problem-solving skills.

Job Complexity:
• • Medium-High complexity project.

Supervision:
• Supervision required, should be able to function collaboratively (with guidance) with all levels of employees.

License/Certifications:
• Preferred to have R or Python certification,

Physical Demands:
• Ability to sit and stand for long periods of time.
• Carrying, handling, and reaching objects.
• Manual dexterity to operate office equipment i.e. computers, phones, etc.

Key Accountabilities:
• Experience building data pipelines for various heterogenous data sources.
• Identifying, designing and implementing scalable data delivery pipelines and automating manual processes
• Building required infrastructure for optimal data extraction, transformation and loading of data using cloud technologies like AWS, Azure etc.,
• Develop end to end processes on the enterprise level for use by the clinical data configuration specialist to prepare data extraction and transformations of raw data quickly and efficiently from various sources at the study level
• Coordinate with downstream users such as statistical programmers, SDTM programming, analytics, and clinical data programmers to ensure that outputs meet requirements of end users
• Experience creating ELT and ETL to ingest data into data warehouse and data lakes
• Experience creating reusable data pipelines for heterogenous data ingestions
• Manage and maintain pipelines and troubleshoot data in data lake or warehouse
• Provide visualization and analysis of data stored in data lake
• Define and track KPIs and provide continuous improvement
• Develop and maintain, tools, libraries, and reusable templates of data pipelines and standards for study level consumption by data configuration specialist
• Collaborate with various vendors and cross functional teams to build and align on data transfer specification and ensure a streamlined process of data integration
• Provide ad-hoc analysis and visualization as needed
• Ensure accurate delivery of data format and data frequency with quality deliverables per specification
• Participate in the development, maintenance and training rendered by standards and other functions on transfer specs and best practices used by business.
• Collaborate with system architecture team in designing and developing data pipelines as per business needs
• Network with key business stakeholders on refining and enhancing the integration of structured and non-structured data.
• Provide expertise for structured and non-structured data ingestion
• Develop organizational knowledge of key data sources, systems and be a valuable resource to people in the company on how to best integrate data to pursue company objectives.
• Provides technical leadership on various aspects of clinical data flow including assisting with the definition, build, and validation of application program interfaces (APIs), data streams, data staging to various systems for data extraction and integration
• Experience in creating data integrity and data quality checks for data ingestion
• Coordinates with data base builders, clinical data configuration specialists and data management (DM) programmers ensuring accuracy of data integration per SOPs
• Provide technical support / consultancy and end-user support, work with Information Technology (IT) in troubleshooting, reporting, and resolving system issues
• Develop and deliver training programs to internal and external team, ensure timely communication of new and/or revised data transfer specs
• Continuous Improvement/Continuous Development
• Efficiently prepare and process large datasets for various end users for downstream consumption
• Understand end to end requirements for stakeholders and contribute to process and conventions for clinical data ingestion and data transfer agreements
• Adhere to SOPs for computer system validation and all GCP (Good Clinical Practice) regulations. Ensure compliance with own Learning Curricula, corporate and/or GxP requirements
• Assists with quality review of above activities performed by a vendor, as needed
• Assess and enable clinical data visualization software in the data flows
• Performs other duties as assigned within timelines
• Performs clinical data engineering tasks according to applicable SOPs (standard operating procedures) and processes.

**PAN and DOB is required to create employee profile in the database**
**Graduation is compulsory**

Role

Customer Success Specialist

Timings

Day Shift (Permanent)

Industry

BPO

Work Mode

Work from office

Process

Non-Voice

Functional Area

ITES / BPO / Customer Service

Note: Myglit doesn't charge any money from candidates. If you have been asked to pay money to get this job then report to us immediately at support@myglit.com.

Recruiter profile

Beatrice Kiprono

Recruiter - Gratitude Inc

NA, kenya

0+ Followers

500+ Posts

Interview Tips

  • Giving the VNA round?
  • What are the most important skills you acquired as a Soft Skills/VNA trainer?
  • How would you handle an irate customer?

Get the Best Jobs
on your Fingertips

Similar Jobs

1 - 10 Year(s)

Banking Banking Insurance Customer Service

₹ 70 - ₹ 75 Thousand p.m

Mumbai, India

1 - 4 Year(s)

Confident Multitasking Oracle

₹ 50 - ₹ 55 Thousand p.m

Kolkata, India

5 - 8 Year(s)

Analysis Analysis Risk Management Negotiating Skill

₹ 80 - ₹ 90 Thousand p.m

Hyderabad, India

2 - 8 Year(s)

Enterprise systems BPO Operations BPO industry

₹ 60 - ₹ 70 Thousand p.m

Hyderabad, India

6 - 11 Year(s)

Team Handling Team Management BPO Skills

₹ 1.4 - ₹ 1.5 Lacs p.m

Mumbai, India

2 - 5 Year(s)

Quality assurance Vendor Management Stakeholder Management

₹ 60 - ₹ 75 Thousand p.m

Delhi, India

1 - 5 Year(s)

Vendor Master Data Transactional Procurement

₹ 50 - ₹ 55 Thousand p.m

Hyderabad, India

5 - 10 Year(s)

SQL AWS Cloud NLE and mastering tools

₹ 100 - ₹ 160 Thousand p.m

Bangalore, India

2 - 8 Year(s)

UG/PG from a computer application background with basic knowledge of HTML/XML and JavaScript

₹ 8 - ₹ 9 Lacs p.m

Thane, India

5 - 8 Year(s)

Minimum of 5 years’ experience in data engineering, building data pipelines to manage heterogenous data ingestions or similar in data integration across multiple sources including collected d Experience with Python/R, SQL, NoSQL Cloud experience (i.e. AWS, AZURE or GCP)

₹ 1.33 - ₹ 1.42 Lacs p.m

Mumbai, India