> Markdown version of [/jobs/ext/135764-founding-engineer-software](https://www.wearedevelopers.com/jobs/ext/135764-founding-engineer-software). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Founding Engineer - Software - **Company:** Urun LLC - **Location:** San Francisco, CA, United States - **Experience:** Expert - **Salary:** $200,000.0 - $350,000.0 - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Artificial Intelligence, Component-Based Software Engineering, Cloud Computing, Code Review, Software Debugging, Fault Tolerance, Python (Programming Language), Node.Js, Performance Tuning, Data Streaming, TypeScript, Web Services, WebSocket, WebRTC, AI Infrastructure, Datadog, Delivery Pipeline, Concurrency, Backend, Event Driven Architecture, Kubernetes, Crud - **Published:** May 21, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=58300ff209341372 ## About the Role Do you have experience in Web services design?, * 7+ years building and shipping backend software in production * Proficiency in one or more backend languages - Python, Go, or TypeScript/Node.js * Experience designing APIs, service-oriented systems, and distributed application components * Solid understanding of cloud infrastructure, containers, and modern deployment workflows * Ability to reason about performance, concurrency, reliability, and debugging in complex systems * Experience with real-time, interactive, streaming, or latency-sensitive systems, this is central to the role, not a bonus Things that will give you an edge * WebRTC or WebSockets for real-time communication * AI infrastructure, inference-adjacent systems, media pipelines, or event-driven architectures * Kubernetes, observability tooling, and hands-on production operations * Early-stage startup experience - owning problems end-to-end and moving quickly with limited scaffolding ## Description AI inference today is slow, expensive, and stateless. Send a query, wait, get a response, reset. That's fine for batch, but AI is becoming interactive, and interactive means inference has to respond instantly, hold context across a session, and be steerable in real time. Nobody had built infrastructure that does all three at once. The bottleneck isn't the models. It's the runtime underneath them. What we're building to fix it uRun - Universal Runtime is the layer that makes real-time, stateful inference possible. Our platform lets AI respond instantly, hold context across a session, and be directed as it runs. We prove it through the hardest problem in the stack: real-time AI video generation. Not pre-rendered clips. Not queued jobs. Live, steerable, continuous video that responds as you speak. Solve that, and the rest of the inference stack follows - and that's what we've done. We're an infrastructure company; we build the layer model labs, builders, and research teams ship on top of. Where you come in You'll build the services, APIs, and core application systems that power uRun's runtime the software layer that turns our real-time inference platform into something product and applied AI teams can actually build on. This is not a conventional CRUD backend role. The work centres on low-latency, high-throughput systems: real-time interaction, evolving session state, and request handling that stays reliable under heavy compute and concurrency. You'll work closely with product, infrastructure, and applied AI teams, in an early-stage environment where the architecture is still being set. What you'll actually be doing day-to-day * Build and maintain backend services, APIs, and internal platform components that power uRun's real-time inference runtime * Design systems for real-time interaction, evolving session state, and scalable request handling across production environments * Partner with infrastructure and platform engineers to keep services observable, reliable, and efficient under heavy compute and concurrency * Translate experimental AI capabilities into robust, user-facing software, working closely with product and applied AI teams * Shape architecture decisions on data flow, service boundaries, performance optimisation, and fault tolerance for interactive systems * Raise engineering quality through testing, monitoring, code review, documentation, and sound operational practice ## Related Videos - [Rest API Antipatterns](https://www.wearedevelopers.com/videos/100208-rest-api-antipatterns) - [Coffee with Developers - Adam Wiggins](https://www.wearedevelopers.com/videos/919-coffee-with-developers-adam-wiggins) - [Developing the Backend with Stefan Lingler, CTO at Shpock](https://www.wearedevelopers.com/videos/100360-developing-the-backend-with-stefan-lingler-cto-at-shpock) - [Transforming Education: A Journey from interactive Markdown to Remote-Labs](https://www.wearedevelopers.com/videos/941-transforming-education-a-journey-from-interactive-markdown-to-remote-labs) - [Meet Your New BFF: Backend to Frontend without the Duct Tape](https://www.wearedevelopers.com/videos/682-meet-your-new-bff-backend-to-frontend-without-the-duct-tape) - [Nest.js - TypeScript in the backend can also be clean](https://www.wearedevelopers.com/videos/1033-nest-js-typescript-in-the-backend-can-also-be-clean) ## Related Articles - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 137 - AI'm not sure about this](https://www.wearedevelopers.com/magazine/485-dev-digest-137-ai-m-not-sure-about-this) - [Dev Digest 132 - Binging WADFlix?](https://www.wearedevelopers.com/magazine/473-dev-digest-132-binging-wadflix) - [The Best X (Twitter) Accounts for Developers](https://www.wearedevelopers.com/magazine/294-the-best-x-twitter-accounts-for-developers) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift)