Python & Apache Spark Engineer / Data Processing Specialist @ Square One Resources v Česko
Square One Resources
Na dálkuPráce odkudkoli
Česko
Odpovědnosti
Původní popisek.
Software & Data Engineering
Design, develop, and maintain robust, scalable, and production-ready Apache Spark applications.
Develop efficient data processing, transformation, and analysis workflows using Python and Spark.
Build solutions capable of handling large-scale datasets while maintaining high performance and reliability.
Work with columnar data formats, particularly Parquet, and implement data solutions using Delta Tables.
Develop clean, maintainable, testable, and efficient code in accordance with established engineering standards.
Ukázat vše (21)
O pozici / o projektu
Původní popisek.
About the Project
You will join a dynamic technology team responsible for designing, developing, and deploying innovative enterprise technology and AI-driven solutions supporting the delivery of tax services.
The team combines expertise across software engineering, data engineering, AI, tax technology, change management, and project management . The projects cover the full technology lifecycle — from solution design and development through deployment and optimization, to training, adoption, and stakeholder engagement.
You will work in a modern, cloud-based environment and contribute to the development of scalable data processing solutions, with a strong focus on Python, Apache Spark, cloud technologies, microservices, and modern data platforms .
Technology Stack
Cloud: Microsoft Azure
Data Processing: Apache Spark, Databricks, Azure Synapse / Microsoft Fabric
... zkušenosti s Databricks, Delta Lake a Spark; znalost Kafka, Flink nebo podobných frameworků je výhodou. o Zkušenosti s budováním datových kanálů pomocí nástrojů jako Apache Airflow, Data Factory nebo Apache Beam. o Znalost Terraformu, Pythonu a nástrojů CI/CD, jako jsou GitHub Actions. o Důkladná znalost správy dat, monitorování, ...
... engineering best practices. - Provide guidance and informal mentorship to junior team members. What We're Looking For - 5+ years of hands-on experience with Apache Spark, ideally on Databricks specifically. - Strong PySpark and SQL skills, including performance tuning of Spark jobs. - Practical experience with Delta Lake ...
... with tools such as MS Fabric, Databricks, Snowflake, ADF, dbt, Kafka, Apache Spark Strong knowledge of data governance and compliance principles Proficiency in Python, PySpark and SQL is an advantage ➕ Ability to translate business requirements into scalable, maintainable data solutions Ability to lead technical discussions ...
... development tools and share best practices - Collaborate in a distributed team, taking ownership of impactful projects O pozici / o projektu Původní popisek. Join the Data Platform team to build and maintain data-driven applications, including AI-powered data access tools. Work on a mix of legacy, improved, and greenfield projects ...
... konzultanty - Code review a mentoring kolegů - Odhady pracnosti a rozdělování větších úkolů - Aktivní podíl na technických rozhodnutích Váš profil - Výborná znalost Pythonu - Komerční zkušenosti s backendovým vývojem - Znalost REST API - Práce s databázemi - Docker - Schopnost navrhovat technická řešení - Komunikativní angličtina ...
... and scale robust automated pipelines (using tools like Jenkins and GitHub Actions) to handle software builds, seamless deployments, and security scanning. - Python-Based Validation: Design and implement intelligent scripts and quality checks in Python to automatically validate environment stability, API integrations, and ...
... About the Role We are looking for a Data Engineer who actually likes data. We could be writing that we want “someone who is passionate about building robust data solutions and enabling advanced analytics in the pharmaceutical domain”. But we simply want You to care about data, learn and understand the data core logic. ...
Odpovědnosti Původní popisek. - Backend Development: - Design, develop, and maintain APIs using Python and other relevant frameworks - Work with SQL databases to create, optimize, and manage queries and data models - Develop and maintain backend services using Java and .NET frameworks - Automate infrastructure and deployment ...
... performance and efficiency, identifying bottlenecks and implementing solutions to improve response times and resource utilization; - Implement complex algorithms and data structures to handle data processing, operations research (OR), recommendation systems, and optimization problems; Ukázat vše (11) O pozici / o projektu Původní ...