{"id":1204322,"url":"https://alion.io/job/grassrootcarbons-data-platform-engineer","title":"Data Platform Engineer","company":{"id":7827,"name":"Grassroots Carbon","domain":"grassrootcarbons.com","url":"https://alion.io/company/grassrootcarbons","size_band":null,"is_staffing_agency":false,"is_intermediary":false,"listed_via":null,"ats_vendor":"Workable","truth_index":null},"role":"Data Science","role_family":"Data Science","seniority":null,"employment_type":"full_time","work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["San Antonio, United States"],"countries":["US"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":97000,"max_usd":207000,"period":"year","method":"role_country_seniority_unknown","sample_n":1667},"experience_years_min":null,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Amazon S3","optional":false},{"name":"AWS","optional":false},{"name":"CI/CD","optional":false},{"name":"Least Privilege","optional":false},{"name":"PostGIS","optional":false},{"name":"Python","optional":false},{"name":"SOC 2","optional":false},{"name":"SQL","optional":false},{"name":"Amazon SageMaker","optional":true},{"name":"GDAL","optional":true},{"name":"MLFlow","optional":true},{"name":"PostgreSQL","optional":true}],"status":"live","first_seen_at":"2026-09-24T22:51:23Z","employer_posted_date":"2026-09-24","last_verified_at":"2026-09-24T22:51:23Z","board_verified":true,"closed_at":null,"days_open":0,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":0},"description":"Grassroots Carbon is hiring a Data Platform Engineer to build and operate the core data infrastructure powering our soil carbon modeling work and PastureMap grazing management platform. In this role, you will integrate and serve soil measurements, satellite imagery, ranch records, and partner data through a unified database to ensure reliable, high-integrity access across the company.\nAs the first dedicated owner of our data platform, you will drive data ingestion, dataset accessibility, and production pipeline architecture. Your immediate mission is to replace manual and ad-hoc data preparation with dependable, automated workflows that support science, operations, and rancher-facing products.\nYou will work closely across our engineering and science teams. While Data & Soil Science handles model development and scientific validation, you will own the production pipelines and execution runtimes that put approved models into live use. You will also partner with backend engineers on application APIs and integrations, aligning key priorities with the VP of Product & Technology.\nAbout You\n\nWe are looking for a pragmatic, hands-on builder who designs sound systems and gets things done. You take pride in making sensible architecture decisions, weighing trade-offs clearly, and methodically tackling production challenges when they arise. Working alongside scientists, product managers, and our engineering team, you will rely on clear documentation and strong cross-time-zone collaboration to keep projects moving forward. If you love shipping resilient infrastructure and want to apply your skills to agriculture and environmental technology, this is the role for you.\nRequirements\nData storage and pipelines\nDesign and build our unified data platform from our AWS stack of S3 and PSQL with PostGIS DB, selecting the right storage and processing tools to balance workload, reliability, and cost.\nDevelop tested, scheduled datasets with clear lineage. Maintain consistent, versioned ranch and pasture identifiers across all internal systems.\nMigrate priority PastureMap history, reconcile it to source records, and document gaps. Preserve original locations and flag records that cannot be mapped reliably.\nData ingestion and integrations\nBuild and maintain reliable, automated ingestion pipelines for APIs, hardware feeds (IoT), and imagery services, replacing ad-hoc manual data extraction.\nNormalize spatial and temporal incoming data from external partners and hardware, establishing clear validation and error-recovery protocols.\nModel operations & Security\nBuild secure data services with backend engineers so PastureMap, internal tools, and approved partners can use shared datasets and model outputs.\nPackage, deploy, schedule, and monitor science models. Ensure code and inputs are versioned so runs are reproducible and verifiable for carbon crediting.\nProvide documented datasets and accessible query tools for operations and commercial teams. Evaluate AI-assisted access where it solves a defined need and respects permissions.\nApply least-privilege access, ranch-level data isolation, secrets management, and encryption. Protect rancher and landowner information and support SOC 2 readiness.\nMust Haves\n5+ years in data engineering or data platform work, including ownership of a production data system from design through ongoing operation.\nStrong Python and SQL, with maintainable code, automated tests, and practical debugging skills.\nProduction orchestration and deployment experience, including scheduled jobs, backfills, failure recovery, and CI/CD.\nHands-on AWS experience, infrastructure as code, and an understanding of cloud operating costs.\nExperience taking ML or scientific models into production, ensuring reproducible runs, versioning, monitoring, and automated recovery.\nExperience handling third-party API integrations, specifically dealing with rate limits, schema changes, and unreliable source data.\nNice to Haves\nWorking knowledge of geospatial data, coordinate systems, spatial joins, and tools/formats like H3, PostGIS, GDAL, GeoPandas, or cloud-optimized GeoTIFFs.\nExperience implementing data access controls, secrets management, and encryption to support SOC 2 readiness or audited MRV systems.\nFamiliarity with model deployment tools (MLflow or SageMaker) and serverless batch processing (e.g. Modal).\nExperience ingesting IoT/telemetry data from livestock devices, environmental sensors, or satellite pipelines.\nBS or MS in Computer Science or similar technical studies\nWhat success looks like\nIn Your First 90 Days:\nAlign on Priorities: Partner with Product, Tech, and Soil Science to select and map out your first high-value production workflow based on data readiness.\nShip to Production: Deploy that first pipeline into active use for science or product teams, fully equipped with data validation, access controls, monitoring, and runbooks.\nBuild the Blueprint: Establish shared identifiers, source lineage, and repeatable ingestion patterns to form the foundation for all future pipelines.\nIn Your First Year:\nMigrate Historical Data: Successfully migrate and reconcile priority PastureMap legacy data, systematically documenting known gaps.\nProductionize Core Modeling: Power core science inputs and approved production models through versioned, maintainable pipelines feeding a unified, reliable database.\nScale Data Services: Deliver key integrations and data services built with end-to-end lineage, clear SLAs, and documentation that enables seamless team coverage.\nAbout Grassroots Carbon\nGrassroots Carbon works with ranchers to improve grazing practices, restore grasslands, and store carbon in soil. We combine field measurements, data, and scientific models to quantify soil carbon outcomes. PastureMap is our ranch management platform, and Grassroots Carbon is a portfolio company of Soilworks Natural Capital.\nWhat the teams are building\nProduct & Technology powers PastureMap and its core data services, while Data & Soil Science builds the quantification models for carbon crediting and Scope 3 insetting. Together, these teams turn complex datasets and science into actionable grazing recommendations and practical decision-support tools for ranchers on the ground.\nWhy join Grassroots Carbon\nBuild systems used in carbon measurement, rancher payments, and grazing management.\nWork with soil measurements, satellite imagery, and livestock activity data.\nHave a direct role in platform architecture and implementation.\nWork closely with the people who use the data and see the results of your work.\nBenefits\nBenefits include health, dental, and vision insurance; health plan options with a $0 deductible and $0 co-pay; a flexible spending account option; open paid time off and 9 paid holidays; a 401(k); company-paid Life and AD&D coverage; and support for continuing education.","description_format":"text","description_chars":6851,"description_truncated":false,"requirements":{"experience_years_min":null,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":["Vision insurance"],"hiring_locations":[],"hiring_excludes":[],"relocation_offered":false,"industries":["Farming & Agriculture","Science & Engineering","Environmental Science"],"lifecycle":[{"event":"open","at":"2026-09-24T22:51:23Z"}],"liveness":{"score":86,"band":"hot","label":"Hiring now","p_open":1,"p_active":0.86,"p_room":1,"age_days":0,"expected_fill_days":20,"reasons":["conf:1","win:early"],"computed_at":"2026-09-25T00:36:12Z"},"pay":null,"html_url":"https://alion.io/job/grassrootcarbons-data-platform-engineer","json_url":"https://alion.io/job/grassrootcarbons-data-platform-engineer.json","meta":{"generated_at":"2026-09-25T00:36:12Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":496,"day_limit":5000,"remaining_today":4504,"minute_limit":60,"resets_at":"2026-09-26T00:00:00Z"}}}