Skip to main content
Upload your CV and find your next job on Indeed!

Spark jobs

Sort by: -
    • Legacy Data Conversion specializes in transforming legacy and unstructured data into clean, structured, database-ready information.
    • View all Legacy Data Conversion jobs - Remote jobs
    • Salary Search: Data Scientist salaries in Remote
    • This internship provides hands-on experience in business operations, process management, project coordination, and cross-functional collaboration.
  • View similar jobs with this employer
    • Build and maintain strong relationships with clients, institutions, administrators, and users to ensure a seamless platform experience.
    • Flexible Hours: Dedicate just 3 hours a day at your convenience.
    • Engage with prospective clients to promote Skill Spark’s offerings.
    • Solid computer foundation and programming skills, familiar with common data structures and algorithms.
    • Excellent in one of the following languages: Go/Python.
    • Skills:* Ab Initio, ETL, Parquet File Handling, Data Processing, SQL.
    • Time: Between 6 PM IST to 12 AM IST.
    • (3 hours per day) |.
    • Ab Initio: 5 years (Required).
  • View similar jobs with this employer
    • Job Type: Full-time(Remote).
    • The ideal candidate will be responsible for overseeing technical projects, mentoring developers, conducting technical interviews…
    • We are seeking a skilled Data Engineer to join our team and help build scalable, reliable, and high-performance data solutions.
    • Production Operator works on the factory floor to operate machinery, assemble products, monitor production processes, and maintain quality and safety standards.
    • As an Engineering Intern at Eightfold AI, you will get a chance to work on highly scalable distributed systems and AI platforms spanning several TBs of data…
    • The MDP Ingestion team manages the ingestion of data into the Marketplace Data Platform (MDP), ensuring various types of data (raw events, facts, externally-…
    • We are seeking an entry-level Spark Engineer with at least 1 year of experience working with Apache Spark to join our Data engineering team.
    • Our company is searching for a talented and experienced CNC machine operator to oversee our computer numeric controlled (CNC) machines.
    • Build and maintain strong relationships with clients, institutions, administrators, and users to ensure a seamless platform experience.
    • The ideal candidate will be responsible for overseeing technical projects, mentoring developers, conducting technical interviews and product demonstrations,…

Job Post Details

Data Scientist - job post

Legacy Data Conversion
Remote
₹5,99,267.41 - ₹19,78,083.45 a year

Job details

Pay

  • ₹5,99,267.41 - ₹19,78,083.45 a year

Job type

  • Student job
  • Permanent
  • Fresher
  • Full-time

Benefits

Pulled from the full job description

  • Food provided
  • Health insurance
  • Paid sick time
  • Life insurance
  • Leave encashment
  • Provident Fund
  • Work from home

Full job description

About Us

Legacy Data Conversion specializes in transforming legacy and unstructured data into clean, structured, database-ready information. We work with enterprise clients across industries, converting data from PDFs, scanned documents, Excel files, text files, images, and legacy systems into modern formats with exceptional accuracy.

We are looking for a Data Engineer who enjoys solving complex data challenges and building reliable, scalable data processing solutions.

Responsibilities

  • Design, build, and maintain scalable ETL/ELT pipelines.
  • Extract, transform, and load data from multiple sources including PDFs, Excel, CSV, text files, APIs, and databases.
  • Develop automated data extraction and validation workflows.
  • Ensure high data quality, consistency, and integrity.
  • Optimize data processing for large-scale datasets.
  • Work with structured and unstructured data.
  • Collaborate with software engineers and business stakeholders to understand data requirements.
  • Create and maintain technical documentation.
  • Troubleshoot and improve existing data processing systems.
  • Implement data validation, monitoring, and error-handling processes.

Required Skills

  • 2+ years of experience as a Data Engineer or similar role.
  • Strong SQL skills.
  • Proficiency in Python.
  • Experience with ETL processes and data transformation.
  • Experience working with relational databases such as PostgreSQL, MySQL, SQL Server, or Oracle.
  • Knowledge of Pandas and data processing libraries.
  • Familiarity with Git.
  • Strong analytical and problem-solving skills.
  • Excellent communication skills.

Preferred Qualifications

  • Experience with OCR technologies.
  • Experience processing PDF and scanned document data.
  • Knowledge of cloud platforms such as AWS, Azure, or Google Cloud.
  • Experience with Docker.
  • Familiarity with Apache Spark or similar big data technologies.
  • Experience with workflow orchestration tools such as Airflow.

What We Offer

  • Competitive salary.
  • Fully remote work environment.
  • Flexible working hours.
  • Opportunity to work on enterprise-scale data transformation projects.
  • Exposure to large and complex datasets.
  • Professional growth and learning opportunities.
  • Collaborative and supportive team culture.

Nice to Have

  • Experience handling millions of records.
  • Knowledge of data quality frameworks.
  • Experience with AI-assisted document processing.
  • Understanding of data governance and security best practices.

Application

If you're passionate about data engineering and building robust data pipelines, we'd love to hear from you. Please apply with your resume and a brief summary of your relevant experience.

Pay: ₹599,267.41 - ₹1,978,083.45 per year

Benefits:

  • Flexible schedule
  • Food provided
  • Health insurance
  • Leave encashment
  • Life insurance
  • Paid sick time
  • Provident Fund
  • Work from home

Work Location: Remote

Let Employers Find YouUpload Your Resume