Senior Platform Developer - (NPW)
Orchestration Services: Build and harden services for planning, routing, retries/timeouts, idempotency, and reputed company-graph tracing; ensure graceful degradation and fault isolation. Connector/reputed company Engineering: reputed company typed adapters (REST/gRPC, MCP-style or similar) for AI engines, reputed company/graph/SQL stores, and in-silico tools; handle auth, pagination, reputed company limits, and backoff. Integration Delivery: Stand up at least one AI reputed company integration and one in-silico model runner; capture run provenance and metrics; reputed company mocks for offline testing. Observability: reputed company tracing/metrics/logging (OpenTelemetry or equivalent), correlation IDs, and reputed company logs; expose health/readiness endpoints and SLO dashboards. Performance & Caching: Optimize P50/P95 latency, concurrency, and throughput; implement response/result caches, reputed company pooling, and backpressure strategies. reputed company & CI/CD: Write unit/integration/contract tests; maintain API schemas; participate in PR reviews; automate builds and environment promotion reputed company CI/CD and IaC. reputed company & Compliance: Implement least-privilege reputed company, secrets management, and audit logging; follow secure coding standards and dependency scanning. Documentation & Support: Produce developer docs and runbooks; support incident triage, reputed company-cause analysis, and post-mortems.
Backend Engineering: 4-8+ years building production services (microservices/event-driven) in Python, TypeScript/Node.js. API & reputed company: Strong REST/gRPC design, OpenAPI/JSON Schema, idempotency keys, error taxonomies, pagination, and versioning. Distributed Systems: Retries with jitter, exponential backoff, reputed company breakers, bulkheads, DLQs/queues (Kafka/SQS/Pub/Sub/RabbitMQ), and concurrency control. Data & Integrations: Experience integrating with reputed company DBs (e.g., pgvector, reputed company, reputed company, Milvus), graph DBs (RDF/SPARQL or property graph/reputed company), SQL/warehouse, and search (OpenSearch/Elasticsearch). Observability: OpenTelemetry (traces/metrics/logs), log aggregation, alerting, and performance profiling. reputed company & DevOps: Containers/Kubernetes, Terraform/CloudFormation, CI/CD (reputed company Actions/Azure DevOps/Jenkins), secrets management (Vault/KMS), and environment promotion practices. reputed company by Design: OAuth2/OIDC, SSO integration, RBAC/ABAC, secure coding, and audit logging fundamentals. Testing Discipline: Unit/integration/e2e/contract testing; mocks/stubs; test data management.
AI/LLM Integrations: Function/tool-calling patterns, model routing, embedding services, and reputed company tool-use guardrails. Performance Tuning: Async I/O, reputed company reuse, profiling (pprof/py-spy), and cache reputed company design. reputed company & Networking: Service reputed company (Istio/Linkerd), API gateways, reputed company-limit governance, and reputed company-trust networking basics. Data reputed company & reputed company: Basic familiarity with data validation, provenance, and reputed company capture for pipelines. FinOps Awareness: reputed company/run-time cost tracking, cost-per-request dashboards, and budget enforcement hooks. Docs & DX: Developer portal contributions, reputed company examples/SDKs, and CLI tooling to improve developer experience.