Remote
(Anywhere)
Salary Range
Not informed
Experience Level
Senior
Requirements
Tasks and Responsibilities
Show originalSenior Data Platform Engineer (DataOps) | Databricks + AWS
📍 Work Model: Remote
📄 Contract Type: Contractor or Cooperative
About the Opportunity
You will be the operational owner of a financial services company's Databricks platform on AWS, with autonomy to automate, monitor, and optimize costs. The data team already operates with an SRE culture, and you will join to take the platform to the next level.
What You'll Do
- Administer the Databricks account and workspaces: users, groups, service principals, SCIM/SSO, cluster policies, SQL warehouses, and serverless
- Operate Unity Catalog day-to-day (grants, external locations, storage credentials, onboarding new domains)
- Manage the platform's AWS infrastructure via Terraform (VPC/PrivateLink, S3, IAM, KMS, Secrets Manager, CloudWatch)
- Build and maintain CI/CD for pipelines using Databricks Asset Bundles and Git, with promotion from dev → staging → prod
- Orchestrate and monitor workflows (Lakeflow Jobs, Airflow/MWAA) with SLI/SLO, alerts, and runbooks
- Implement observability and health and cost dashboards (system tables, Datadog, Grafana, Prometheus)
- Work with FinOps: tagging, budgets, right-sizing, and reducing DBU and AWS costs
- Automate tests and quality checks in deployment, together with the Quality team
- Ensure security and compliance: access auditing, secrets, private network, backups, and DR
- Plan runtime upgrades, new platform features, and capacity
What We're Looking For
- 5+ years in data engineering, DevOps, or SRE, with 2+ years administering Databricks in production
- Strong AWS: IAM, S3, VPC, KMS, EC2, CloudWatch
- Terraform (AWS and Databricks providers) and Git
- CI/CD applied to data (Asset Bundles, Databricks CLI, or API)
- Python and SQL
- Hands-on Unity Catalog: permissions, external locations, and system tables
Nice to Have
- Familiarity with Spark and Delta Lake for troubleshooting and tuning
- Glue, Lambda, or MWAA
- Datadog, Grafana, or Prometheus
- Experience with monitoring, alerting, and incident management
- Kubernetes or Ansible
- AWS certifications (Solutions Architect, DevOps Engineer) or Databricks Platform Administrator
- Experience in the financial sector or credit bureaus, and knowledge of LGPD
- Use of generative AI in day-to-day work
Only at Code Group will you find the best IT opportunities with innovative technologies. Our main focus is Recruitment & Selection and Talent Hunting, and we are proud to have clients and partners with the same objective: Engaged People. We seek innovative professionals who are driven by challenges and focused on results. Do you see yourself here?!
Share job:
Share job: