Role Overview & Specifications

We are seeking a skilled Data Engineer with a strong background in data pipeline creation, data visualization, and analytics. As a key member of our team, you will play a crucial role in designing, developing, and optimizing data solutions to drive actionable insights.

Responsibilities

Data Pipeline Creation

oProficiently design, develop, and maintain data pipelines using the Palantir Foundry tool suite, including apps like Slate, Workshop, Contour, Taurus, Ontology Manager, Object View, and Report.

oUtilize JavaScript and Python to create robust data pipelines, ensuring efficient data flow and transformation.

Data Visualization

oDevelop interactive dashboards using JavaScript, providing stakeholders with clear insights into complex datasets.

oLeverage Python for data transformation and visualization, enhancing the user

experience.

Big Data Processing

oWork on the Azure platform with Databricks, developing Spark applications using Spark-SQL.

oExtract, transform, and aggregate data from various file formats (e.g., txt, avro, parquet, csv) to analyze customer data and derive meaningful insights.

Spark And Py

Spark: oUtilize Spark and PySpark to streamline data processing tasks, enhance scalability, and optimize performance.

oDesign and implement efficient data workflows, ensuring data quality and reliability.

Machine Learning And Deep Learning

oDesigned regression models and employed ensemble techniques to predict life expectancy, achieving a 60% decrease in RMSE compared to the baseline model using gradient boosting regression.

oProficient in using Spark APIs, including Spark SQL, Spark DataFrames, and User-Defined Functions (UDFs).

oHands-on

experience with various functions, transformations, and actions performed on Spark RDDs.

Data Modeling And Analysis

oImported data onto the Hadoop Distributed File System using MapReduce.

oAnalyzed data using Hive and reported insights using Tableau 2020.1.

oWorked with different file formats (Excel, CSV, Parquet, Avro) for Hive querying and processing.

Data Visualization Tools

oConducted Exploratory Data Analysis (EDA) using Tableau 2020.1, Matplotlib, Seaborn (Python), and GGPlot ®.

oProficiently designed both logical and physical data models, addressing OLTP and OLAP

requirements.

Version Control And Collaboration

oProficiently used Git and GitHub for version control, creating and managing branches.

Tableau Expertise: div>

oGenerated Tableau visualizations and dashboards using Tableau Desktop.

oCreated dashboards with quick filters, parameters, and sets to handle views more efficiently.

oCombined visualizations into interactive Tableau dashboards and published them to web portals such as Tableau Public.

oExtensively used advanced chart visualizations in Tableau, including dual-axis, box plots, bullet graphs, tree maps, bubble charts, waterfall charts, and funnel charts to assist business users in solving complex problems.

Bachelor’s degree in Computer Science, Data Science, or a related field.Proven

experience in data engineering, data visualization, and big data processing.Strong analytical

skills with meticulous attention to detail.

Excellent Communication And Collaboration

skills.

Partner Recommendation

About Tata Consultancy Services

Tata Consultancy Services is an actively verified employer hiring talent across technology, engineering, and operations.

  • Headquarters: United States / United Kingdom
  • Company Size: 1,000+ employees
  • Trust Rating: 95 / 100 (Official Registry Audited)

View all openings at Tata Consultancy Services →