> Markdown version of [/jobs/ext/617169-senior-software-engineer-backend-infra](https://www.wearedevelopers.com/jobs/ext/617169-senior-software-engineer-backend-infra). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Software Engineer, Backend/Infra - **Company:** PIKA, LLC - **Location:** Palo Alto, CA, United States - **Experience:** Expert - **Salary:** $185,000.0 - $300,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Systems Engineering, Automated Storage and Retrieval Systems, Software as a Service, Cloud Computing, Code Review, Continuous Integration, Database Design, Database Models, Software Debugging, Distributed Systems, Python (Programming Language), Node.Js, NoSQL, Open Source Technology, Performance Tuning, Queueing Systems, Software Engineering, SQL Databases, TypeScript, Speech Recognition, WebSocket, Large Language Models, Multi-Agent Systems, Prompt Engineering, Generative AI, Backend, Fastapi, Event Driven Architecture, Kubernetes, Low Latency, Api Design, GPT, Golang, Microservices - **Published:** June 18, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=ebb9d9b1e56e523d ## About the Role Do you have experience in Systems engineering?, Experience: 5+ years of software engineering experience building production services at scale, with 2+ years hands-on with LLM-based orchestration, multi-agent systems, or agentic solutions. * Backend Mastery: Deep proficiency in modern backend technologies (Node.js, Python, Go) and frameworks (Express, FastAPI, TypeScript, etc.). * Systems & Infra Knowledge: Strong understanding of distributed systems, event-driven microservices, message queues, cloud infrastructure (AWS/GCP), Kubernetes, and CI/CD workflows. * AI & Agentic Expertise: Solid grasp of LLM capabilities and limitations, prompt engineering (system prompts, chain-of-thought, structured output), tool-use execution, and embedding models. * Real-Time Patterns: Comfort designing and debugging real-time streaming pipelines, long-polling, and highly concurrent networking setups. * Product & Data Intuition: Deep understanding of database design and the product sense needed to make an AI system feel "alive" and responsive. * Mindset: Ownership mentality-identify systemic bottlenecks and ship solutions without waiting for exact specifications. Strong communication and a collaborative, team-first attitude., Experience with multi-modal AI architectures (image generation, TTS, speech-to-text, video generation) * Experience with agent frameworks (LangChain, CrewAI, AutoGPT) or building custom, high-performance execution runtimes * Experience with fine-tuning, RLHF, or DPO pipelines * Background in multi-tenant SaaS or internal tooling and operational automation * Previous startup experience-comfortable with ambiguity and rapid experimentation * Competitive coding background (IOI, ICPC, Olympiad medalists, etc.) ## Description We are seeking a Senior Backend & Infra Engineer to help shape the core infrastructure powering Pika's products. In this role, you will operate at the intersection of systems engineering and advanced AI, taking ownership of the unified architecture that supports our platform-from real-time messaging and scalable APIs to cognitive agent runtimes and orchestration frameworks. You will architect and build robust, distributed backend systems enabling autonomous AI agents to reason, use tools, recall memories, and act across multiple platforms at scale. This high-ownership position means your architectural decisions will directly impact how millions of users experience generative AI. As a senior engineer, you'll also play a key role in raising the technical bar through guidance, RFCs, code reviews, and mentorship. What You'll Do * Architect Distributed Systems: Build scalable backend, infrastructure, and agentic services for Pika's web, mobile, and multi-platform products. * Evolve the Agent Runtime: Design and optimize the core execution loop that handles agent reasoning, tool-use frameworks, function calling, memory retrieval, and multi-step orchestration. * Design Real-Time Architecture: Own and scale real-time messaging infrastructure, event-driven architectures, WebSocket connections, and pub/sub patterns for throughput, latency, and reliability. * Implement Core AI Capabilities: Optimize LLM integrations, multi-provider model routing (Claude, GPT, Gemini, open-source), context window management, cost optimization, and streaming responses. * Build Memory & Retrieval Systems: Design semantic search and vector-based embedding infrastructure to handle long-term memory, working memory, and episodic recall for autonomous agents. * Own Backend Logic End-to-End: Drive database modeling (SQL/NoSQL), API design, performance tuning, and production reliability for high-traffic pipelines. * Drive Technical Excellence: Write RFCs, evaluate complex technical trade-offs, mentor junior engineers through code reviews, and build alignment across engineering and product teams. ## Related Videos - [ Evaluating AI models for code comprehension](https://www.wearedevelopers.com/videos/1462-evaluating-ai-models-for-code-comprehension) - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Agentic employees in world's most downloaded FinTech app](https://www.wearedevelopers.com/videos/100123-agentic-employees-in-world-s-most-downloaded-fintech-app) - [Speak, Code, Deploy: Transforming Developer Experience with Voice Commands](https://www.wearedevelopers.com/videos/1159-speak-code-deploy-transforming-developer-experience-with-voice-commands) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) ## Related Articles - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [From Prototype to Production: Build AI Agents with This Free 4-Course Learning Path](https://www.wearedevelopers.com/magazine/655-from-prototype-to-production-build-ai-agents-with-this-free-4-course-learning-path) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)