CareerPlanGet AI match score →

Senior Software Engineer, Platform

🌐 Remote💼 Full-time🗓 2026-06-24

Core

Design, extend, and operate the core Python services powering agent orchestration, connector management, schema resolution, streaming chat, and sandboxed execution for enterprise customers.

Role type

Senior Software Engineer, Platform

Builds

Real-time agent workflows, multi-connector data pipelines, sandboxed execution environments, and versioned artifact delivery.

Domain

Enterprise AI/LLM infrastructure, data integration, and agent orchestration.

Deliverable

production ML models | product features | infrastructure

Required skills

Python, FastAPI, async programming, Celery, Redis, RabbitMQ, PostgreSQL, SQLAlchemy, Pydantic, Snowflake, DuckDB, Databricks, BigQuery, Prometheus, Grafana, OpenTelemetry, Terraform, AWS, Docker, Nginx, GitHub Actions

Preferred skills

None stated

Technologies

FastAPI, Celery, Redis, RabbitMQ, PostgreSQL, SQLAlchemy, Alembic, Pydantic v2, Snowflake, DuckDB, Databricks, BigQuery, Prometheus, Grafana, OpenTelemetry, Terraform, AWS, Docker, Nginx, GitHub Actions

Responsibilities

Own the FastAPI platform and design core services for agent orchestration and connector management. Build and scale async workers for task routing and real-time notifications. Manage the context layer pipeline for processing enterprise documents and building knowledge layers. Build and harden runtime connectors to various data warehouses and SaaS sources. Instrument the system with observability stacks and define error budgets. Ship and operate on AWS using Docker, Terraform, and CI/CD pipelines.

Seniority

Senior, hands-on IC

Rewrite
## About the role Experience: 4+ years building and operating production-grade Python services. Location: Remote ## Why This Role Matters Every insight Terrabase delivers travels through a Python service you will own. Our platform powers real-time agent workflows, multi-connector data pipelines, sandboxed execution, and versioned artifact delivery, all streaming live to enterprise customers. Reliable async workers, low-latency APIs, and precise observability are not nice-to-haves here. They decide whether customers trust the system. Your mission: keep this engine reliable and scale it as we grow. ## What You Will Do - Own the FastAPI platform. Design, extend, and operate the core services powering agent orchestration, connector management, schema resolution, streaming chat, and sandboxed execution. Async handlers, SSE and WebSocket support, Pydantic v2 validation, SQLAlchemy with Alembic migrations against PostgreSQL. - Build and scale async workers. Operate Celery workers backed by Redis and RabbitMQ for schema fetching, task routing, stuck-task detection, and real-time notifications. Understand failure modes at the worker level, not just the API level. - Own the context layer pipeline. Build and operate the ingestion pipeline that processes enterprise documents, extracts and ranks business concepts, and builds the structured knowledge layer that agents reason over. This covers connector integrations, chunking strategies, and the data contracts between upstream sources and the agent layer. - Manage data connections at scale. Build and harden runtime connectors to Snowflake, DuckDB, Databricks, BigQuery, and other warehouse and SaaS sources. Handle encrypted credentials, OAuth flows, and live schema discovery. Make connections stay alive, fail cleanly, and recover fast. - Instrument everything. Own the observability stack: Prometheus and Grafana, structured logging with correlation IDs, OpenTelemetry tracing, health endpoints. P99 latency and error budgets are yours to define and defend. - Ship and operate on AWS. Docker-based deployments, Nginx, Terraform, GitHub Actions CI/CD. Write runbooks and post-mortems anyone can use to debug at 2am. Harden secrets management and SOC 2 logging. - Collaborate across teams. The platform serves LangGraph-based agent workflows and React frontends. Design API contracts that enable sub-second streaming responses and zero-downtime releases. ## What We Are Looking For - 4+ years building and operating production Python services - Strong bias for ownership: you identify problems, propose fixes, and drive them to closure without supervision - Deep Fa
Sourced via wellfound · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Wellfound ↗