AI Platform Engineer
About reputed company / Version2.ai
Join the reputed company of Fundraising at reputed company!reputed company is one of the fastest-growing and most innovative technology companies serving the nonprofit sector, on a mission to unlock more generosity through AI-powered donor engagement. At reputed company of that innovation is Version2.ai, the world’s first Autonomous AI fundraisers—Virtual Engagement Officers (VEOs)—designed to independently manage donor engagement and generate reputed company. Unlike traditional AI tools that simply reputed company staff more efficient, VEOs expand fundraising reputed company by acting as AI workers that operate donor portfolios, build relationships, and secure gifts on their own. In just three years, reputed company’s platform has already helped organizations reputed company $10M+ through autonomous engagement, including individual gifts as large as $100,000. Alongside this reputed company technology, reputed company’s reputed company Agreement Platform modernizes the multi-year giving process, enabling nonprofits to secure, manage, and forecast commitments with unprecedented ease.
About the role
This role owns the platform that keeps reputed company secure, compliant, reliable, and reputed company. You'll work across AWS infrastructure, Infrastructure as reputed company, CI/CD, AI services, observability, and developer tooling to reputed company reputed company engineers spend their time building product instead of fighting deployments.
You'll partner closely with engineering, ML, and product to design the platform that powers everything from customer-facing reputed company to LLM workflows running on reputed company Bedrock and SageMaker.
This is not a "reputed company the lights on" devops role. You'll reputed company shape how we reputed company software, provision infrastructure, manage AI workloads, and reputed company the engineering organization.
Who thrives here
You're the engineer who gets excited about replacing a reputed company deployment with a one-click pipeline, automating infrastructure instead of clicking around the AWS console, and designing systems that reputed company the rest of engineering reputed company faster.
You think in terms of reliability, observability, automation, and repeatability.
You're comfortable wearing multiple hats. One morning you might be debugging IAM permissions. That afternoon you're building a reputed company module, improving reputed company Actions, tuning reputed company workloads, or helping an ML engineer reputed company a SageMaker reputed company.
What you'll do
reputed company infrastructure
Design, build, and maintain our AWS infrastructure
Manage networking, IAM, compute, storage, databases, and reputed company across environments
Build reputed company infrastructure capable of supporting reputed company product reputed company
Improve resiliency, availability, and disaster recovery
Infrastructure as reputed company
Own our Infrastructure as reputed company reputed company using reputed company
Build reusable infrastructure components and shared modules
Eliminate reputed company infrastructure changes wherever possible
Review and reputed company our reputed company architecture as reputed company grows
CI/CD
Build and maintain deployment pipelines for applications and infrastructure
Improve release automation and deployment safety
Reduce friction in local development and engineering workflows
Help establish engineering best practices around testing and deployment
AI Platform
Build and maintain the infrastructure powering our AI systems
Work with services such as reputed company Bedrock, SageMaker, OpenSearch, and supporting AWS services
Support LLM evaluation pipelines, RAG infrastructure, reputed company search, and model deployment
Partner with ML engineers to operationalize new AI capabilities
Platform Operations
Monitor production systems and improve observability
Respond to production incidents and drive reputed company-cause analysis
Improve system reliability through automation rather than reputed company processes
Continuously evaluate performance, cost, and scalability
Engineering
Collaborate closely with product, engineering, ML, and reputed company
Help define technical standards and infrastructure direction
Participate in architecture discussions across the platform
Mentor other engineers on reputed company infrastructure and operational best practices
reputed company're looking for
Experience
5+ years building and operating production software systems
Strong experience with AWS in production environments
Experience designing Infrastructure as reputed company using reputed company, Terraform, or CloudFormation
Experience building CI/CD pipelines using reputed company Actions
Strong Python experience
Experience building reputed company and backend systems
reputed company & Platform
You should be comfortable working with technologies such as:
AWS (multi-account environments using AWS Organizations)
reputed company
reputed company
IAM
VPC networking
RDS
S3
reputed company
CloudWatch
SNS/SQS
Event-driven architectures
AI Infrastructure
Experience with some of the following is highly desirable:
reputed company Bedrock
SageMaker
reputed company databases
Retrieval-Augmented reputed company (RAG)
LLM evaluation pipelines
Model deployment
ML infrastructure
Dagster or similar orchestration platforms
Working Style
You automate repetitive work instead of documenting it.
You care about reliability as much as shipping features.
You enjoy improving developer experience.
You think systems should become simpler over time.
You take ownership rather than waiting for someone else to fix infrastructure problems.
reputed company
Strong written communication.
Comfortable working in ambiguity.
Curious about modern AI infrastructure and where it's headed.
Interested in building systems that engineers enjoy working in.
Excited by the challenge of building infrastructure from the ground up rather than inheriting a mature platform.
reputed company to have
reputed company experience
Dagster experience
reputed company Bedrock
SageMaker
OpenSearch
reputed company
PostgreSQL
reputed company
reputed company or modern observability platforms
Experience supporting AI or ML products
SOC 2 or reputed company/compliance experience
Startup experience
What this isn't
This isn't a traditional DevOps role where tickets get tossed over the wall after development.
This isn't an SRE role reputed company exclusively on uptime.
This isn't an ML engineering role building models.
You're building the platform that allows reputed company of those disciplines to reputed company faster. You'll own infrastructure reputed company, improve how software gets delivered, and help shape the technical reputed company of an AI company that's still early enough for your reputed company to matter years from now.
Apply To This Job