Mid-Level Data Engineer
Crisis Text Line, Inc. Atlanta, Georgia, United States · $100K–$117K/yr
Non-profit Organizations · 201-500 employees
About the role
Design, build, and maintain reliable production data pipelines while optimizing workflows using SQL, Python, and Spark. Collaborate with cross-functional teams to translate data needs into scalable solutions and ensure system performance through monitoring and automation.
What they look for
Requirements
Requires three to five years of professional experience in data or software engineering with a focus on production data pipelines. Candidates must have proficiency in SQL, Python, and distributed processing technologies like Spark, along with experience in cloud-based environments.
Benefits
Full description
We’re looking for a Data Engineer to build, improve, and support the data pipelines and infrastructure that power our products and organization. You’ll work with SQL, Python, Spark/PySpark, Databricks, AWS, and Infrastructure as Code to develop reliable, scalable data solutions and help ensure the performance and quality of our production data environment.
You’ll collaborate with engineers, data scientists, and teams across Crisis Text Line to translate data needs into thoughtful, maintainable solutions. We’re looking for someone with strong data engineering fundamentals and hands-on experience building and supporting production data pipelines. You don’t need experience with every technology in our environment—we value transferable experience, curiosity, and the ability to learn and solve problems collaboratively.
Our Stack
You’ll work across a modern, cloud-based environment, including:
- Languages & Processing: SQL, Python, Spark/PySpark
- Data Platform: Databricks
- Cloud: AWS
- Infrastructure as Code: Terraform
- Observability: Datadog
- Data Stores: MySQL, PostgreSQL, and Redis
- Delivery & Reliability: Automated CI/CD, monitoring, and shared ownership of production data pipelines
Responsibilities:
- Design, build, maintain, and improve reliable production data pipelines as part of our broader data and engineering environment.
- Develop and optimize data processing workflows using SQL, Python, Spark/PySpark, and Databricks.
- Partner with engineers, data scientists, and business teams to understand data needs and translate them into practical, maintainable solutions.
- Build and maintain integrations between data sources and our data infrastructure.
- Use Infrastructure as Code practices to support and improve our AWS data environment.
- Monitor pipeline health, data quality, and performance; troubleshoot issues and contribute to solutions when something isn’t working as expected.
- Automate repeatable processes that improve the reliability, efficiency, and maintainability of our data pipelines.
- Participate in shared production support and develop confidence diagnosing and resolving pipeline issues independently.
- Contribute to code reviews, technical documentation, and engineering standards that help the team build reliable and maintainable systems.
- Bring ideas, ask questions, and identify opportunities to improve our data tools, processes, and engineering practices over time.
Required Qualifications:
- Approximately three to five years of professional data engineering, software engineering, or related experience, including meaningful responsibility for production data pipelines.
- Professional experience using SQL and Python to build, transform, and troubleshoot data.
- Experience building, maintaining, and supporting production data pipelines or ETL/ELT workflows.
- Experience with distributed data processing technologies such as Spark or PySpark.
- Familiarity with cloud-based environments such as AWS and modern data engineering practices.
- Experience testing, debugging, monitoring, and improving the reliability and performance of production data systems.
- The ability to independently complete well-defined engineering work, investigate problems, and clearly communicate technical decisions, risks, and tradeoffs.
- A collaborative approach to working with engineers, data scientists, and cross-functional partners, with a commitment to building secure and reliable data systems.
Preferred Qualifications
The following experience would be useful, but we encourage you to apply even if you do not meet every item:
- Databricks or a comparable cloud-based data platform.
- AWS services and cloud-based data architecture.
- Terraform or another infrastructure-as-code tool.
- Lakehouse modeling such as medallion architecture and data warehousing.
- Datadog or comparable monitoring and observability tools.
- Data quality, automated testing, CI/CD, or deployment automation.
- Designing data systems for reliability, performance, scalability, and security.
- Supporting data used for analytics, reporting, or machine learning workloads.
- Claude or other AI-assistance code tools.
- Working in a mission-driven, regulated, safety-sensitive, or high-trust environment.
Reliable High-Speed Internet Required: Must have a stable high-speed internet connection to support seamless remote collaboration, virtual meetings, online job tasks, etc.
For United States-based candidates:
The target salary range for this position, across the United States, is $99,704 - $116,990. Starting salary will vary based on location, qualifications, and prior experience. Candidates will learn the range specific to their location during the interview process. We pay competitively in the tech-forward nonprofit space and offer a robust benefits package.
This is a remote position within the United States. At the time of hire and throughout employment, the employee’s primary residence and regular home work location must be in one of the following approved hiring states: California, Colorado, Connecticut, Florida, Georgia, Illinois, Indiana, Maryland, Massachusetts, Michigan, New Jersey, New Mexico, New York, North Carolina, Pennsylvania, Tennessee, Texas, Utah, Virginia, or Washington state.
No visa sponsorship available for this position.
Benefits & Well-Being
Crisis Text Line recognizes that we are all unique human beings with unique life circumstances, and our benefits package aims to be as flexible as possible to support your needs as you work to promote mental well-being for people, wherever they are. Our benefits package is thoughtfully designed using an equity lens, with input from our team and from industry best practices.
Highlights include:
- Comprehensive medical, dental, and vision options that prioritize accessibility and financial peace of mind
- Employer-funded HSA contributions
- Generous PTO, sick time, and 19 paid holidays with a winter break
- 12 weeks of fully paid parental leave after 26 consecutive weeks of service
- Monthly internet and mental health stipends
- Annual Wellness Stipend
- Home office and professional development stipends
- 403(b) retirement plan with employer contribution
- Sabbatical after 3 years of service
Benefits are for U.S.-based employees; international benefits may vary.
#LI-KR1
This is a remote-only position
Similar roles
-
TS/SCI w/Poly - AI/ML Data Engineer
Leading Path Consulting Chantilly, Virginia, United States
-
Principal Consultant, Data Engineer
Lovelytics Chicago, Illinois, United States · $130K–$170K/yr
-
Database Administrator (Data Engineer)
Navy Federal Credit Union Pensacola, Florida, United States · $78K–$123K/yr
-
Data Engineer - AWS, GCP, Snowflake (CDI - H/F)
Talan Toulouse, Occitania, France
-
Data Engineer II
InVita Healthcare Technologies Baltimore, Maryland, United States · $100K–$120K/yr
-
Data Engineer
EXL Haryana, India