Data Engineer
Company hidden until unlock
- Work model
- Remote
- Location
- United States - Remote
- Posted
- 5d ago
Company hidden until unlock
You've been carrying an urgency to make an impact. Join to give it velocity. 01 THE ROLE Chemistry is the invisible layer of all critical industries, and the old supply chain is being weaponized. We don't refine petroleum. We create the molecules that matter most — from the atom up. As a Data Engineer at Ohr, you'll coordinate data pipelines across every front of the company: connecting the platforms we work in every day, integrating the instruments our scientists rely on, and building the cohesive data platform that our analytics, AI systems, and scale-up decisions depend on. THE KIND OF PEOPLE THIS WORK DEMANDS Our mission demands more than great science. It requires a particular kind of team — visionary builders, hustlers, scientists, operators. Together, we carry both the technical depth to make the chemistry work and the commercial instinct to make the business thrive. People are our engine. Collaboration is our multiplier. 02 WHAT YOU'LL DESIGN AND BUILD - Build and maintain connectors across our current data platforms — Google Drive, Microsoft GCC High (for ITAR export-controlled information), Slack, and Claude — so information moves reliably and securely between the systems - Enforce data boundaries as a first-class design constraint: export-controlled data stays in compliant environments, with access control, isolation, and audit trails built into every pipeline - Partner with laboratory scientists in our San Diego office to integrate lab data collection tools — including Invert Bio and data outputs from analytical chemistry instruments (HPLC, GC) — into a cohesive, queryable data platform - Design ingestion, cleaning, and standardization pipelines for messy real-world data: instrument exports, spreadsheets, PDFs, and semi-structured lab records - Create the unified data foundation that powers analytics, dashboards, and machine learning across R&D, manufacturing, and operations - Own pipelines end-to-end: schema design, orchestration, monitoring, data quality checks, and documentation 03 WHAT YOU BRING WITH YOU You're the engineer who gets quietly excited about a gnarly integration problem — four platforms, multiple data ingress routes from lab instruments, one compliance regime, zero existing pipeline. You believe good data infrastructure is invisible when it works and invaluable when decisions depend on it. You likely have: - 4–8 years of professional data engineering or backend engineering experience - A four-year undergraduate degree in computer science or engineering; a science or engineering degree in another field may substitute when paired with 6–10 years of relevant experience - Obsessed over data quality - experience with defining, monitoring, and root-causing data quality problems. - Strong SQL and Python, and hands-on experience building production data pipelines (batch and/or streaming) with modern orchestration tools - Experience with building reports and reporting infrastructure, queryable lakehouse architecture, and/or medallion architecture. - Experience integrating third-party platforms via APIs — authentication, rate limits, webhooks, and the unglamorous edge cases - Comfort with messy, heterogeneous data: instrument files, CSVs, PDFs, and formats that were never meant to be integrated - A security- and compliance-aware mindset — you treat access control and data isolation as requirements, not afterthoughts STRONG PLUS - Experience with laboratory or scientific data (LIMS, ELN, analytical instrument outputs such as HPLC/GC chromatography data) - Familiarity with government or regulated cloud environments (Microsoft GCC High, ITAR/CMMC contexts) - Experience supporting ML/AI workloads — feature pipelines, retrieval corpora, or data prep for LLM systems - Strong technical writer and experience developing technical specifications and standards that software developers can use to integrate as both producers and consumers. - Experience with selecting tools and infrastructure and designing the system the data pipelines will be built on. WHAT SUCCESS LOOKS LIKE AFTER 3 MONTHS - Connectors across our core platforms are live and trusted, with export-controlled data provably isolated AFTER 6–12 MONTHS - Lab instrument data (InvertBio, HPLC, GC) lands automatically in a unified platform scientists actually query - Analytics and AI systems across the company build on your pipelines rather than ad hoc exports 04 EXPORT CONTROL COMPLIANCE This role involves access to data and systems subject to U.S. export control regulations, including ITAR. Candidates must be a U.S. citizen or national, lawful permanent resident, or a protected individual as defined under 8 U.S.C. § 1324b(a)(3). 05 COMPENSATION & BENEFITS This is a remote role, with some travel required open to candidates located in the United States. The annual base salary range for this position is $140,00 - $200,000, commensurate with experience, plus equity. Final compensation will be determined based on skills, experience, and qualifications. Ohr offers a competitive benefits package, including: MEDICAL COVERAGE A wide range of plans, with many options fully covered for employees and dependents at no cost to the employee. 401(K) WITH MATCH Employer match up to 4% of eligible compensation. PAID TIME OFF Vacation, paid company holidays, and an end-of-year shutdown. EQUITY You're joining at an early and consequential moment in Ohr's growth — equity commensurate with experience, with meaningful upside as we scale. Applicants must be authorized to work for any employer in the United States. Ohr is an equal opportunity employer. We do not discriminate on the basis of race, color, religion, sex, national origin, age, disability, or any other characteristic protected by law. ABOUT OHR We are the integrators of industrial alchemy. Ohr exists because the molecules that matter most deserve a better origin story. Our chemistry is not abstract: it