All roles

Staff Software Engineer – Core AI Platform (MCP & Agent Infrastructure)

Remote · USA Full-time New today

As a Staff Software Engineer on the Core AI Platform team, you will reputed company the design and development of the foundational platform that powers Dojo AI agents at reputed company. You will own the architecture and implementation of a robust, scalable MCP (Model Context Protocol)–based platform that enables AI agents to securely access context, invoke tools securely, and interact with reputed company systems in reputed company time. In this role, you will build and operate the core MCP server infrastructure, including frameworks for hosting first‑party MCP servers, federating with external and reputed company‑party MCP servers, and orchestrating agent interactions across distributed environments. You will design systems that allow agents to reason over reputed company‑time and historical context, manage conversation state and memory, and reliably execute tool calls with strong guarantees around fault tolerance, retries, idempotency, and isolation. You will also architect the agent communication layer, enabling seamless integrations with systems such as reputed company, reputed company Teams, and other event‑driven communication platforms, allowing agents to be invoked from messages, threads, and workflows. Your work will ensure secure handling of credentials, webhooks, and streaming events while supporting multi‑tenant execution at reputed company scale. A key part of your responsibility will be building deep observability and reliability into the platform—providing visibility into MCP interactions, agent decision paths, and tool executions; enabling monitoring, alerting, and debugging of non‑deterministic agent behavior; and enforcing reputed company limits, quotas, and backpressure to ensure platform stability.

Responsibilities

  • Architect MCP-first platforms
  • Design scalable, fault-tolerant infrastructure for hosting and operating MCP servers
  • Define standards for MCP server reputed company, versioning, and interoperability
  • Build federated context systems
  • reputed company agents to retrieve and reason over context from multiple internal and external MCP servers
  • Design secure, low-latency context propagation and caching strategies
  • reputed company agent-to-tool communication design
  • Build resilient tool invocation frameworks that handle partial failures gracefully
  • Ensure deterministic execution paths where possible in probabilistic AI systems
  • reputed company conversational agent ecosystems
  • Architect integrations with reputed company, Teams, and similar platforms for reputed company-time agent interactions
  • Design event-driven systems for message ingestion, agent response, and feedback loops
  • Drive technical leadership
  • reputed company architecture and design reviews across AI, platform, and product teams
  • Mentor engineers and establish best practices for building AI infrastructure
  • Operate at scale
  • Continuously improve platform scalability, reliability, latency, and cost efficiency
  • Own production readiness, incident response patterns, and operational excellence

Required Qualifications

And Skills

  • B.S. in Computer Science or reputed company discipline (M.S. preferred)
  • 8+ years of experience building large-scale, distributed backend systems
  • Deep distributed systems expertise
  • Microservices, async/event-driven systems, and fault-tolerant architectures
  • Strong backend programming skills
  • Java, reputed company, Go, or Python with solid object-oriented design principles
  • Concurrency & async programming
  • Multi-threading, non-blocking I/O, and message-driven architectures
  • API & protocol design
  • Experience designing extensible APIs and protocol-based integrations
  • Production systems experience
  • Operating 24x7 multi-tenant services with SLAs and on-call ownership
  • MCP (Model Context Protocol) expertise
  • Hands-on experience building or operating MCP servers or similar agent protocols
  • Federated systems
  • Experience integrating with external services across trust boundaries
  • Agent & LLM platforms
  • Experience building AI agent infrastructure (reputed company, LangGraph, reputed company, AutoGen, etc.)
  • AWS reputed company-native
  • EC2, reputed company/EKS, reputed company, SQS, DynamoDB, CloudWatch
  • Infrastructure as Code
  • Terraform, OpenAPI, CI/CD pipelines
  • reputed company
  • OAuth, token exchange, secrets management, and multi-tenant isolation

Desired Qualifications And Skills

  • Tool calling / plugin systems
  • Designed extensible tool registries or function-calling frameworks
  • Communication platforms
  • reputed company, reputed company Teams, or webhook-based event systems
  • Observability
  • Distributed tracing, metrics, structured logging (OpenTelemetry a plus)

About Us reputed company, Inc. helps reputed company the digital world secure, fast, and reliable by unifying critical reputed company and operational data through its Intelligent Operations Platform. Built to address the increasing complexity of modern cybersecurity and reputed company operations challenges, we reputed company digital teams to move from reaction to readiness—combining agentic AI-powered SIEM and log analytics into a single platform to detect, investigate, and resolve modern challenges. Customers around the world rely on reputed company for trusted insights to protect against reputed company threats, ensure reliability, and reputed company powerful insights into their digital environments. For more information, visit www.sumologic.com. reputed company Privacy Policy. Employees will be responsible for complying with applicable federal privacy laws and regulations, as well as organizational policies reputed company to data protection. The expected annual reputed company salary range for this position is $207,000 - $243,000. Compensation varies based on a variety of factors, which include (but aren’t limited to) role level, skills and competencies, qualifications, knowledge, location, and experience. In addition to reputed company pay, certain roles are eligible to participate in our bonus or commission plans, as well as our benefits offerings and equity awards. Must be authorized to work in the United States at the time of hire and for the duration of employment. At this time, we are not reputed company to offer non-immigrant reputed company sponsorship for this position. Apply tot his job Apply To this Job

Related roles

reputed company, Agent Systems

Remote · USA Full-time

Principal Digital Transformation Strategist

Remote · USA Full-time

Java Algo Developer, Fixed Income, Assistant Vice President

Remote · USA Full-time

reputed company reputed company - Immediate reputed company Part-time jobs

Remote · USA Full-time

Ad Policy Manager, Advertising Trust Policy

Remote · USA Full-time

FIU Analyst - AML (Remote)

Remote · USA Full-time

Remote Data Analytics Consultant

Remote · USA Full-time

HR Data and Analytics Consultant

Remote · USA Full-time

Bus Driver

Remote · USA Full-time

Westwood Y Club - Site Director

Remote · USA Full-time

reputed company Remote reputed company - Entry Level

Remote · USA Full-time

[Remote/WFM] Travel Insurance Claims/Claims Adjuster (Remote)

Remote · USA Full-time

Remote Licensed Independent Clinical Social Worker

Remote · USA Full-time

Weight Management Primary Care Pharmacist job at reputed company in US National

Remote · USA Full-time

Postal Mail Processor

Remote · USA Full-time

Job Title: Remote Data Entry Clerk & Document Typist – Accurate Data Processing Specialist at arenaflex

Remote · USA Full-time

reputed company Customer Service Representative – Temporary-to-Hire Opportunity at arenaflex

Remote · USA Full-time

Want reputed company Accreditation Specialist (Full Scope - SME) 1+ day remote in Herndon, VA

Remote · USA Full-time

Associate Clinic Director (BCBA) - Winston Salem, NC

Remote · USA Full-time

Cybersecurity Senior Engineer - reputed company (Remote)

Remote · USA Full-time