
Closed
Posted
We are looking for an experienced Databricks Engineer to support an ongoing data engineering project. Responsibilities Design, develop, and maintain solutions using Databricks Build and optimize data pipelines Work with large-scale datasets Develop data transformations using PySpark and SQL Work with Delta Lake and Databricks Lakehouse architecture Troubleshoot pipeline and performance issues Improve reliability, scalability, and performance of existing workloads Collaborate with our engineering team on technical requirements Required Skills Strong hands-on Databricks experience PySpark / Apache Spark Python SQL Delta Lake ETL/ELT and data pipeline development Data modeling Git and CI/CD experience Experience with cloud-based data platforms Nice to Have Databricks Workflows / Jobs Unity Catalog Structured Streaming Performance optimization AWS, Azure, or GCP experience Databricks certification What We're Looking For We need someone who has real production experience with Databricks, not only theoretical knowledge or training projects. When applying, please briefly describe: Your Databricks experience A production Databricks project you worked on Your experience with PySpark and Delta Lake The largest dataset or pipeline you have worked with Location : Philippine Please start your proposal with "DATABRICKS" so we know you read the full job description.
Project ID: 40660421
107 proposals
Remote project
Active 2 days ago
Set your budget and timeframe
Get paid for your work
Outline your proposal
It's free to sign up and bid on jobs
107 freelancers are bidding on average $13 USD/hour for this job

DATABRICKS Production Databricks work is where I would focus here, especially pipeline reliability, PySpark transformations, SQL optimisation and Delta Lake workloads rather than only notebook development. I have worked with Databricks-based data engineering workflows involving PySpark, Python, SQL, ETL/ELT and Delta Lake, including pipeline optimisation and troubleshooting. I am also comfortable with Git-based development and CI/CD practices for data workloads. For a production engagement, I would first understand the existing Lakehouse structure, pipeline dependencies and current bottlenecks, then work on reliability, performance and maintainability without disrupting existing workloads. My relevant experience includes [insert your actual production Databricks project], where I worked with [actual dataset/pipeline size] using PySpark and Delta Lake. Could you share whether the current Databricks environment is AWS, Azure or GCP and whether the immediate priority is new pipeline development or optimisation of existing production jobs? Jenifer
$12 USD in 40 days
8.0
8.0

Hello, DATABRICKS I can confidently say that I am the experienced Databricks Engineer you are looking for to support your project. With over 7 years of experience as a senior engineer, I have dedicated a substantial portion of my career to practicing and honing my skills in working with Databricks. Not only do I have a deep theoretical understanding of the platform, but I've also successfully completed numerous projects using Databricks in real production environments. I have extensively worked with PySpark and Delta Lake, building and optimizing complex data pipelines, often dealing with large-scale datasets. This experience has familiarized me with troubleshooting performance or pipeline issues efficiently while ensuring high reliability and scalability. Moreover, my expertise spans across ETL/ELT, data modeling, and Git and CI/CD - all crucial elements for handling valuable and sensitive data seamlessly along iterative process improvements. Additionally, geographical flexibility shouldn't be an issue as I'm based in the Philippines. My commitment to quality workmanship, timely deliveries, and open communication enables me to build strong professional relationships that ultimately drive project success. Choose me for this role and rest assured, your ongoing project will be in safe hands! Thanks!
$25 USD in 29 days
7.6
7.6

((DATABRICKS)) Hello, Databricks Engineer {{{ I HAVE CREATED SIMILAR BEFORE AND I CAN SHOW YOU }}} I have carefully reviewed your requirements and understand that you need a production-focused Databricks engineer who can build reliable pipelines, optimize workloads and work confidently with large-scale datasets. I have 11+ years of software development experience and can work with Databricks, PySpark, Python, SQL, Delta Lake and ETL/ELT pipelines. I can design and maintain scalable data workflows, implement transformations, optimize Spark performance and troubleshoot reliability or pipeline issues. >>> 40-45 hours weekly I am available for work<<<< >>> you will track all progress of the project thru the tracker <<< I can also work with Lakehouse architecture, Databricks Jobs/Workflows, data modeling, Git/CI/CD and cloud-based data environments where required. My focus will be on clean, maintainable pipelines with proper validation, monitoring and performance optimization. I can review your existing workloads, identify bottlenecks and improve processing efficiency without unnecessarily rebuilding components that are already working correctly. I WILL PROVIDE 2 YEARS OF FREE ONGOING SUPPORT AND COMPLETE SOURCE CODE. Thanks, Christina
$10 USD in 40 days
6.6
6.6

DATABRICKS ★★★ DATABRICKS SPECIALIST ★★★ Hi, I can design and develop solutions using Databricks to support your data engineering project, including building data pipelines, optimizing performance, and troubleshooting issues. What I'll do: - Build and optimize data pipelines using PySpark and SQL. - Develop data transformations and work with Delta Lake architecture. - Collaborate with your engineering team on technical requirements. - Troubleshoot pipeline and performance issues to improve reliability. Please let me know if you have specific datasets or pipelines you want to focus on. Looking forward to your reply. Thanks!
$10 USD in 40 days
6.5
6.5

DATABRICKS Hi, I can support your Databricks environment across PySpark, Python, SQL, Delta Lake, ETL/ELT pipelines, and Lakehouse workloads, with emphasis on scalable transformations, reliability, and performance. I can also work within Git-based CI/CD workflows and troubleshoot existing pipelines systematically. I’d first review your current architecture and workloads, then focus on measurable improvements to pipeline stability, execution time, and maintainability. A few questions: * Which cloud platform hosts your current Databricks workspace: AWS, Azure, or GCP? * Are the existing pipelines primarily batch-based, streaming, or a combination of both? * Which area currently needs the most attention: pipeline reliability, performance optimization, or new development? Best regards, Muhammad Usman
$12 USD in 40 days
6.2
6.2

DATABRICKS As an experienced data engineer, I have a solid grasp on several technologies you're looking for in your Databricks project, including Python, SQL, Delta Lake, PySpark, and more. My robust experience with Databricks is not just theoretical; it extends into important work I've done mixing data and artificial intelligence. Building AI systems that stand up to the demands of real-world scenarios is what we specialize in. In one notable project, I designed and implemented a sophisticated Databricks solution for a major retail company where millions of data points were being handled daily. It involved creating efficient data pipelines, working on large datasets with superior performance optimization using Databricks Workflows and PySpark. Similarly, my experience with Delta Lake will be particularly relevant to your requirements. Being based in the Philippines, I offer cost-effective yet high-quality solutions for your endeavor. Not to mention that my exposure to multiple cloud platforms like AWS, GCP, and Azure perfectly aligns with the preferred technology environment you mentioned. Let's collaborate to bring value to your data ecosystem and propel your ongoing project towards better performance and reliability.
$12 USD in 40 days
6.3
6.3

Hi there, DATABRICKS , I can see the hidden pressure here: you need someone who can keep pipelines stable while the lakehouse grows, not just someone who knows the buzzwords. I’ve worked on production Python, SQL, and ETL-style systems where reliability, performance, and clean transformations mattered more than flashy demos. I’m comfortable supporting Databricks-style workflows, PySpark logic, Delta-based transformations, and troubleshooting slow or failing jobs end to end. I’ve shared an initial estimate based on your description, and once we go over a few technical or functional details, I’ll confirm the exact cost and delivery schedule. If you’d like, I can review the current pain points, the size of the data, and the parts of the pipeline that are most fragile first, then map out the safest fix plan. What’s the biggest issue right now — slow jobs, broken pipelines, or scaling the current Databricks setup? Best regards, Asad
$10 USD in 77 days
5.6
5.6

DATABRICKS Your project’s success hinges on production-tested pipelines, not just demos. I’ll focus on the hidden bottlenecks: optimizing Delta Lake table layouts for faster queries, tuning Spark shuffle partitions to avoid stragglers, and implementing incremental processing to cut costs on large datasets. I’ll also enforce Unity Catalog governance early to prevent permission chaos later. My approach is to first audit your existing workloads, then apply targeted fixes to reliability and performance—not rewrite what works. For the largest dataset, I’ve handled terabytes with mixed structured and semi-structured data, ensuring exactly-once semantics and idempotent writes. This project gets a pragmatic engineer who treats Databricks as a tool, not a showcase.
$12 USD in 40 days
5.7
5.7

DATABRICKS. As an seasoned professional with 13+ years of experience, I have spent a significant amount of time working with Databricks and PySpark. My team and I have successfully deployed over 300 projects, many of which are long‑term enterprise solutions mirroring the nature of your ongoing data project. Our main focus is to enable clean, scalable, and efficient coding structures that not only deliver required functions but also provide space for future expansion – a quality pivotal in dealing with large datasets on Databricks. During my career, I have worked on numerous production Databricks projects incorporating elements akin to your project description like ETL/ELT, data modeling, pipeline optimization using Delta Lake. Having expericence AWS services such as Glue, S3 etc complements our strong hands-on Databricks profiency which will surely bring great value to your business. When it comes to teamwork, transparency and clear communication is vital. With direct access to me as your main contact, you can be assured there will be no communication gaps. Additionally, we are meticulous in breaking down the project into clear milestones ensuring frequent demos and transparent progress updates for allaying any concerns you might have regarding the ongoing work. Let's collaborate and construct something that not only meets but surpasses your expectations; a solution that's stable today && easy to extend tomorrow
$12 USD in 40 days
5.4
5.4

Nice to meet you ,The requirements of your project match my areas of work and skills, to introduce myself. My name is Anthony Muñoz and i am the lead engineer for DS Pro IT agency. I have worked for over 10 years as a Full-Stack and software development engineer and have successfully done multiple jobs. It will be a pleasure to work together to make your project. Feel free to discuss about the project with me, greetings.
$12 USD in 40 days
5.8
5.8

Greetings, DATABRICKS. It looks like you're seeking an experienced Databricks engineer to enhance your data engineering project. I’d love to help you design and optimize solutions that work efficiently with large datasets. My approach would involve building robust data pipelines and utilizing PySpark and SQL to develop data transformations, ensuring everything integrates smoothly with Delta Lake and the Lakehouse architecture. I've worked extensively with Databricks in production environments, tackling real-world challenges. One project involved creating a data pipeline for a retail client, where I managed millions of records and implemented performance optimizations that significantly improved processing times. My experience with PySpark and Delta Lake allows me to troubleshoot issues effectively and enhance the scalability of workloads. Looking forward to the opportunity to collaborate with your engineering team. Best regards, Saba Ehsan
$12 USD in 40 days
4.9
4.9

DATABRICKS Hi, I reviewed the project and I’ll support your ongoing data engineering work by designing and maintaining Databricks solutions, focusing on data pipelines and transformations. I will build optimized data pipelines for large-scale datasets using PySpark, Python, and SQL, applying solid data modeling and ETL/ELT practices. I’ll work with Delta Lake and Databricks Lakehouse architecture to keep workloads reliable, scalable, and easy to troubleshoot. I can improve pipeline performance, reliability, and maintain clean, well-managed code suitable for production. Let’s discuss here now.
$15 USD in 18 days
5.0
5.0

DATABRICKS Hi, I’m an experienced Python/data engineer with hands-on experience building scalable data pipelines, ETL/ELT workflows, cloud infrastructure, and analytics systems using Python, SQL, Spark-based architectures, and AWS/Azure environments. I’m comfortable with PySpark transformations, Delta Lake, Lakehouse patterns, pipeline optimization, CI/CD, and troubleshooting performance/reliability issues across large datasets. I can work within your existing production environment, optimize workloads, and collaborate closely with your engineering team to deliver reliable, scalable pipelines. Best regards, Shakila Naz
$12 USD in 40 days
5.2
5.2

★•══•★ Hi client ★•══•★ DATABRICKS here! I've got solid hands-on experience designing and optimizing Databricks pipelines in production, especially with PySpark and Delta Lake. One project I handled involved building scalable ETL workflows processing multi-terabyte datasets daily, ensuring smooth performance and reliability. I focus on clear data transformations, troubleshooting bottlenecks, and collaborating closely with teams to meet technical needs. I'm comfortable with Git-based CI/CD setups and cloud platforms like AWS. Your project's all about real-world skills, and that’s exactly what I bring—no theory, just results. What’s the biggest challenge you’re facing with your current pipelines? Best regards, Rico
$12 USD in 40 days
5.0
5.0

With over a decade of experience in full-stack development, there is no data challenge that my team and I haven't tackled. As a cohesive unit, we have built scalable, high-quality software solutions with a strong focus on performance, maintainability, and clean architecture – values that align perfectly with your project's requirements. Moreover, we are not only theoretically knowledgeable but have also executed real production projects with Databricks. In terms of Databricks expertise, I have worked with the platform extensively - leveraging its powerful tools and technologies including PySpark and Delta Lake for ETL/ELT processes, building robust data pipelines that can handle large-scale datasets efficiently. In fact, I've managed a dataset upwards of 50TBs using Databricks as well as designed and developed entire workflows for streaming and batch processes using Structured Streaming. Last but not least, my team is well-acquainted with cloud-based data platforms including AWS which further empowers our ability to work seamlessly with Databricks. And while technical prowess is essential in a project like yours, our commitment to clear communication, punctual updates, and superior project management are equally vital for successful collaboration. With us on board, you can be assured of a game-changing data engineering journey.
$12 USD in 40 days
4.8
4.8

Hi, Have a pleasant day! I carefully reviewed your project posting on Freelancer.com and believe I have the skills and experience needed to complete it successfully. I have worked on similar projects in the past and would be happy to share relevant examples during our discussion. Please share your detailed project requirements so I can understand your expectations and recommend the best approach for your project. My goal is to deliver high-quality work that meets your requirements. The final cost and delivery timeline will depend on the project's scope and complexity. My pricing is flexible, and I am happy to work within your budget. I do not require any upfront payment. I prefer milestone-based payments to ensure transparency and satisfaction throughout the project. I look forward to hearing from you and hope to have the opportunity to work together. Kind regards, Sneha Kanwar
$12 USD in 40 days
4.7
4.7

Hi there, I am A.R.M. MASUD with a strong background in Data Science.I am an experienced Machine Learning developer with expertise in designing, training, and deploying intelligent models that deliver real-world value. My background includes supervised and unsupervised learning, deep learning with TensorFlow and PyTorch, and data preprocessing using Pandas, NumPy, and Scikit-learn. I specialize in developing classification, regression, clustering, and predictive models, as well as computer vision and NLP solutions. I follow best practices in feature engineering, hyperparameter tuning, and model evaluation to ensure high accuracy and scalability. My focus is on building end-to-end ML pipelines that are efficient, reliable, and tailored to your project’s requirements to maximize impact. https://www.freelancer.com/u/MZITSERVICES I appreciate the opportunity to submit this proposal and am excited about the possibility of working with you to bring your project to life. Thanks A.R.M MASUD
$12 USD in 40 days
4.4
4.4

DATABRICKS Hello, I am excited to apply for the Databricks Engineer position supporting your data engineering project. I understand you need a professional who can design, develop, and maintain effective solutions using Databricks, while also optimizing data pipelines and working with large-scale datasets. With extensive hands-on experience in Databricks and a strong background in PySpark and SQL, I have successfully managed production projects that involved building and maintaining complex data pipelines and transformations. I have worked extensively with Delta Lake and have a solid understanding of the Databricks Lakehouse architecture, allowing me to ensure reliability and performance in data workloads. To achieve your project goals, I would take the following approach: - Analyze existing data workflows and identify optimization opportunities using Databricks. - Develop robust ETL/ELT pipelines that leverage PySpark for efficient data transformations. - Collaborate closely with your engineering team to align technical requirements and ensure seamless integration. - Implement best practices for performance optimization and scalability within the Databricks environment. I am eager to contribute my expertise and start work on this project. I am confident in my ability to deliver high-quality results on time and would love to discuss further details at your convenience. Thank you for considering my proposal.
$8 USD in 40 days
4.7
4.7

I’m experienced in Databricks, PySpark, Python, SQL, Delta Lake, ETL/ELT, and production data pipelines. Build and optimize Databricks pipelines using PySpark and SQL, with clean transformations and reliable Delta Lake workflows. Work with Lakehouse architecture, data modeling, Databricks Jobs/Workflows, and performance optimization for large datasets. Troubleshoot existing workloads, improve pipeline reliability and scalability, and maintain clean Git and CI/CD practices. Collaborate closely with your engineering team on requirements, implementation, testing, and production support. Q: Can you share the current Databricks environment and a brief overview of the production pipeline you need support with? Location: Pakistan. I’m not based in the Philippines, so I want to be transparent about that requirement. If the location requirement is flexible, I’d be happy to discuss the project. Best Regard, Shawana
$12 USD in 40 days
4.4
4.4

DATABRICKS Built and maintained production Databricks pipelines in AWS with PySpark, Delta Lake, and Unity Catalog, handling datasets up to 50TB in a lakehouse architecture. Integrated ETL workflows via CI/CD, optimized Spark jobs with partitioning and Z-Ordering, and troubleshot cluster performance using Ganglia logs and Spark UI. Experience with Databricks Workflows, Jobs, and structured streaming for near-real-time pipelines. Familiar with data modeling in Delta and SQL transformations in Databricks SQL warehouses. Can start immediately. Thanks, Andrii
$11.50 USD in 40 days
4.1
4.1

San Salvador de Jujuy, Argentina
Payment method verified
Member since May 12, 2026
$30-250 USD
$250-750 USD
₹1250-2500 INR / hour
₹1500-12500 INR
₹12500-37500 INR
₹750-1250 INR / hour
₹1500-12500 INR
₹12500-37500 INR
$30-250 USD
₹12500-37500 INR
$25-50 USD / hour
₹1500-12500 INR
₹1500-12500 INR
₹750-1250 INR / hour
$750-1500 AUD
$250-750 USD
$15-25 AUD / hour
₹1500-12500 INR
₹600-610 INR
$8-15 USD / hour
₹1500-12500 INR
$15-25 USD / hour