Cloudesign

Python Data Engineer

Bangalore, India
Python FastAPI SQL Oracle SQL Server MongoDB Spark PySpark Shell Message Queue Redis Elasticsearch Kafka Spark Streaming Docker Kubernetes
Description

Python Data Engineer

Location: Bangalore, India

Department: Software Development

Experience: 5-10

Job Description:
Skills: Python, FastAPI, LLMs.
Responsibilities: Scrape public data, build pipelines (bronze/silver/gold), and create APIs.
What you will do
  • Bring in industry best-practices around creating and maintaining robust data pipelines for complex data projects with / without AI component:
  • programmatically retrieve (unstructured mostly) data from several static and real-time sources (incl. web scraping, API use)
  • Structure this data into a structured format o Harmonize the data, into a common format and store it in a dedicated database.
  • Schedule the different jobs into a dedicated pipeline
  • rendering results through dynamic interfaces incl. web / mobile / dashboard with ability to log usage and granular user feedbacks
  • performance tuning and optimal implementation of complex Python scripts, SQL,…
  • Industrialize ML / DL solutions and deploy and manage production services; proactively handle data issues arising on live apps
  • Perform ETL on large and complex datasets for AI applications - work closely with data scientists on performance optimization of large-scale ML/DL model finetuning
  • Build data tools to facilitate fast data cleaning and statistical analysis
  • Build and ensure data architecture is secure and compliant
  • Resolve issues escalated from Business and Functional areas on data quality, accuracy, and availability
  • Work closely with APAC IT Transformation and coordinate with a fully decentralized team across different locations in APAC and global HQ (Paris).
You should be
  • Expert in structured and unstructured data in traditional and Big data environments Oracle / SQLserver, MongoDB, Hive / Pig, BigQuery and Spark
  • Have excellent knowledge of Python programming both in traditional and distributed models (PySpark)
  • Expert in shell scripting and writing schedulers
  • Hands-on experience with Cloud - deploying complex data solutions in hybrid cloud / on-premise environment both for data extraction / storage and computation
  • Experience working on industry standard services like Message Queue, Redis, Elastic Search, Kafka, or Spark Streaming
  • Well versed with DevOps best practices like containerization, CICD pipeline (Jenkins and Maven)
  • Hands-on experience in deploying production apps using large volumes of data with state-of-the-art technologies like Dockers, Kubernetes and Kafka
  • Strong knowledge of data security best practices
Cloudesign
Cloudesign

0 applies

0 views

There are more than 50,000 engineering jobs:

Subscribe to membership and unlock all jobs

Engineering Jobs

60,000+ jobs from 4,500+ well-funded companies

Updated Daily

New jobs are added every day as companies post them

Refined Search

Use filters like skill, location, etc to narrow results

Become a member

🥳🥳🥳 452 happy customers and counting...

Overall, over 80% of customers chose to renew their subscriptions after the initial sign-up.

To try it out

For active job seekers

For those who are passive looking

Cancel anytime

Frequently Asked Questions

  • We prioritize job seekers as our customers, unlike bigger job sites, by charging a small fee to provide them with curated access to the best companies and up-to-date jobs. This focus allows us to deliver a more personalized and effective job search experience.
  • We've got over 200,000 jobs from 15,000+ vetted companies. No fake or sleazy jobs here!
  • We aggregate jobs from 15,000+ companies' career pages, so you can be sure that you're getting the most up-to-date and relevant jobs.
  • We're the only job board *for* software engineers, *by* software engineers… in case you needed a reminder! We add thousands of new jobs daily and offer powerful search filters just for you. 🛠️
  • Every single hour! We add 2,000-3,000 new jobs daily, so you'll always have fresh opportunities. 🚀
  • Typically, job searches take 3-6 months. EchoJobs helps you spend more time applying and less time hunting. 🎯
  • Check daily! We're always updating with new jobs. Set up job alerts for even quicker access. 📅

What Fellow Engineers Say