- Страна
- США
- Зарплата
- 207 000 $ – 243 000 $
Откликайтесь
на вакансии с ИИ

Staff Software Engineer – Core AI Platform
Высокая оценка обучена конкурентной зарплатой, работой над передовыми технологиями (AI Agents, MCP) в известной компании и возможностью влиять на архитектуру продукта. Единственный минус — отсутствие спонсорства виз.
Сложность вакансии
Роль требует редкого сочетания глубокой экспертизы в распределенных системах и практического опыта работы с новейшими протоколами AI-агентов (MCP). Высокий уровень ответственности за архитектуру платформы корпоративного уровня делает эту позицию крайне сложной.
Анализ зарплаты
Предлагаемый диапазон ($207k - $243k) полностью соответствует рыночным ожиданиям для позиции Staff Engineer в Кремниевой долине, находясь на уровне или чуть выше медианы для компаний технологического сектора.
Сопроводительное письмо
I am writing to express my strong interest in the Staff Software Engineer position for the Core AI Platform at Sumo Logic. With over 8 years of experience in building large-scale distributed systems and a deep focus on AI agent infrastructure, I am excited about the opportunity to architect the MCP-based platform that powers Dojo AI. My background in designing fault-tolerant microservices and integrating complex event-driven systems aligns perfectly with your mission to enable seamless, secure agent interactions across enterprise environments.
In my previous roles, I have successfully led the development of extensible API frameworks and managed production-grade multi-tenant services on AWS. I am particularly impressed by Sumo Logic's commitment to unifying security and operational data through agentic AI. I am confident that my expertise in Model Context Protocol (MCP), tool-calling frameworks, and observability will allow me to make immediate contributions to your Core AI Platform team and help set the standard for AI infrastructure at scale.
Составьте идеальное письмо к вакансии с ИИ-агентом

Откликнитесь в sumologic уже сейчас
Присоединяйтесь к команде Sumo Logic, чтобы создавать будущее AI-агентов на острие технологий MCP!
Описание вакансии
Staff Software Engineer – Core AI Platform (MCP & Agent Infrastructure)
As a Staff Software Engineer on the Core AI Platform team, you will lead the design and development of the foundational platform that powers Dojo AI agents at Sumo Logic. You will own the architecture and implementation of a robust, scalable MCP (Model Context Protocol)–based platform that enables AI agents to securely access context, invoke tools securely, and interact with enterprise systems in real time.
In this role, you will build and operate the core MCP server infrastructure, including frameworks for hosting first‑party MCP servers, federating with external and third‑party MCP servers, and orchestrating agent interactions across distributed environments. You will design systems that allow agents to reason over real‑time and historical context, manage conversation state and memory, and reliably execute tool calls with strong guarantees around fault tolerance, retries, idempotency, and isolation.
You will also architect the agent communication layer, enabling seamless integrations with systems such as Slack, Microsoft Teams, and other event‑driven communication platforms, allowing agents to be invoked from messages, threads, and workflows. Your work will ensure secure handling of credentials, webhooks, and streaming events while supporting multi‑tenant execution at enterprise scale.
A key part of your responsibility will be building deep observability and reliability into the platform—providing visibility into MCP interactions, agent decision paths, and tool executions; enabling monitoring, alerting, and debugging of non‑deterministic agent behavior; and enforcing rate limits, quotas, and backpressure to ensure platform stability.
Responsibilities
- Architect MCP-first platforms
- Design scalable, fault-tolerant infrastructure for hosting and operating MCP servers
- Define standards for MCP server onboarding, versioning, and interoperability
- Build federated context systems
- Enable agents to retrieve and reason over context from multiple internal and external MCP servers
- Design secure, low-latency context propagation and caching strategies
- Lead agent-to-tool communication design
- Build resilient tool invocation frameworks that handle partial failures gracefully
- Ensure deterministic execution paths where possible in probabilistic AI systems
- Enable conversational agent ecosystems
- Architect integrations with Slack, Teams, and similar platforms for real-time agent interactions
- Design event-driven systems for message ingestion, agent response, and feedback loops
- Drive technical leadership
- Lead architecture and design reviews across AI, platform, and product teams
- Mentor engineers and establish best practices for building AI infrastructure
- Operate at scale
- Continuously improve platform scalability, reliability, latency, and cost efficiency
- Own production readiness, incident response patterns, and operational excellence
Required Qualifications and Skills
- B.S. in Computer Science or related discipline (M.S. preferred)
- 8+ years of experience building large-scale, distributed backend systems
- Deep distributed systems expertise
- Microservices, async/event-driven systems, and fault-tolerant architectures
- Strong backend programming skills
- Java, Scala, Go, or Python with solid object-oriented design principles
- Concurrency & async programming
- Multi-threading, non-blocking I/O, and message-driven architectures
- API & protocol design
- Experience designing extensible APIs and protocol-based integrations
- Production systems experience
- Operating 24x7 multi-tenant services with SLAs and on-call ownership
- MCP (Model Context Protocol) expertise
- Hands-on experience building or operating MCP servers or similar agent protocols
- Federated systems
- Experience integrating with external services across trust boundaries
- Agent & LLM platforms
- Experience building AI agent infrastructure (LangChain, LangGraph, CrewAI, AutoGen, etc.)
- AWS cloud-native
- EC2, ECS/EKS, Lambda, SQS, DynamoDB, CloudWatch
- Infrastructure as Code
- Terraform, OpenAPI, CI/CD pipelines
- Security
- OAuth, token exchange, secrets management, and multi-tenant isolation
Desired Qualifications and Skills
- Tool calling / plugin systems
- Designed extensible tool registries or function-calling frameworks
- Communication platforms
- Slack, Microsoft Teams, or webhook-based event systems
- Observability
- Distributed tracing, metrics, structured logging (OpenTelemetry a plus)
About Us
Sumo Logic, Inc. helps make the digital world secure, fast, and reliable by unifying critical security and operational data through its Intelligent Operations Platform. Built to address the increasing complexity of modern cybersecurity and cloud operations challenges, we empower digital teams to move from reaction to readiness—combining agentic AI-powered SIEM and log analytics into a single platform to detect, investigate, and resolve modern challenges. Customers around the world rely on Sumo Logic for trusted insights to protect against security threats, ensure reliability, and gain powerful insights into their digital environments. For more information, visitwww.sumologic.com.
Sumo Logic Privacy Policy. Employees will be responsible for complying with applicable federal privacy laws and regulations, as well as organizational policies related to data protection.
The expected annual base salary range for this position is $207,000 - $243,000. Compensation varies based on a variety of factors, which include (but aren’t limited to) role level, skills and competencies, qualifications, knowledge, location, and experience. In addition to base pay, certain roles are eligible to participate in our bonus or commission plans, as well as our benefits offerings and equity awards.
Must be authorized to work in the United States at the time of hire and for the duration of employment. At this time, we are not able to offer non-immigrant visa sponsorship for this position.
Создайте идеальное резюме с помощью ИИ-агента

Навыки
- AWS
- Python
- Terraform
- OAuth
- Kubernetes
- OpenTelemetry
- Microservices
- Docker
- Distributed Systems
- Java
- DynamoDB
- Go
- LangChain
- Scala
Возможные вопросы на собеседовании
Проверка понимания ключевого протокола, указанного в вакансии.
Расскажите о вашем опыте работы с Model Context Protocol (MCP). Какие основные сложности возникают при федерации контекста из нескольких источников?
Вакансия требует навыков проектирования отказоустойчивых систем.
Как бы вы спроектировали систему вызова инструментов (tool calling) для AI-агента, чтобы обеспечить идемпотентность и обработку частичных отказов?
Работа предполагает интеграцию с мессенджерами в реальном времени.
Какие архитектурные паттерны вы бы использовали для обработки всплесков трафика (backpressure) в событийно-ориентированной системе интеграции со Slack/Teams?
Позиция уровня Staff подразумевает лидерство.
Опишите случай, когда вам пришлось принимать сложное архитектурное решение в условиях неопределенности (например, с вероятностным поведением LLM). Как вы убеждали стейкхолдеров?
Важно для обеспечения стабильности корпоративной платформы.
Как обеспечить изоляцию данных и безопасность учетных записей в многопользовательской (multi-tenant) среде при выполнении агентами произвольного кода или вызовов API?
Похожие вакансии
AI Engineer (CV & Navigation)
Senior / Lead LLM Engineer
Middle, Middle+, Senior GenAI/LLM Разработчик
Senior Python AI Developer
GenAI/LLM Разработчик
Middle / Senior GenAI Engineer (CV)
1000+ офферов получено
Устали искать работу? Мы найдём её за вас
Quick Offer улучшит ваше резюме, подберёт лучшие вакансии и откликнется за вас. Результат — в 3 раза больше приглашений на собеседования и никакой рутины!
- Страна
- США
- Зарплата
- 207 000 $ – 243 000 $