Full time, Remote
Software Engineer, Infrastructure
As a Software Engineer on the Infrastructure team, you'll own the platform and infrastructure that powers Hermes Agent, Nous Portal, and our model serving. You'll design and build the systems, deployment pipelines, and platform services that keep our products reliable, scalable, and easy to operate.
This role spans infrastructure and platform engineering, with room to ship product where it counts. You'll work closely with Product, Research, and Engineering to keep everything from inference to agent execution running smoothly as usage grows.
Responsibilities
- Design, build, and operate the infrastructure and platform behind Hermes Agent, Nous Portal, and our managed AI services.
- Own systems end-to-end across backend services, platform infrastructure, deployment pipelines, and the tooling that supports them.
- Architect scalable, reliable infrastructure that supports inference, agent execution, model serving, and enterprise deployments.
- Build and improve CI/CD pipelines, deployment workflows, and infrastructure-as-code to accelerate engineering velocity.
- Implement observability, monitoring, alerting, incident response, and reliability improvements across production systems.
- Optimize infrastructure performance, scalability, and cost as product usage grows.
- Evaluate and integrate third-party infrastructure and managed services where appropriate, balancing speed, reliability, and long-term maintainability.
- Collaborate closely with Research, Product, and Engineering teams to support new AI capabilities and production launches.
Qualifications
- 3+ years of software engineering experience with significant ownership of infrastructure or platform systems.
- Strong programming skills in one or more of TypeScript/Node.js, Python, Go, or Rust.
- Hands-on experience with modern cloud platforms such as AWS, Azure, or GCP.
- Experience with containerization, orchestration, and deployment technologies including Docker and Kubernetes.
- Familiarity with CI/CD systems, infrastructure-as-code, observability, monitoring, and production incident response.
- Strong software engineering fundamentals with the ability to build across backend, infrastructure, and platform systems.
- Excellent problem-solving skills and the ability to operate effectively in a fast-moving environment.
- Curiosity about AI infrastructure and enthusiasm for building reliable systems that support cutting-edge AI applications.
Preferred
- Experience operating high-throughput, latency-sensitive distributed systems.
- Familiarity with LLM infrastructure, inference systems, or OpenAI-compatible APIs.
- Experience with serverless platforms, managed cloud infrastructure, or multi-cloud deployments.
- Experience building or operating AI infrastructure, model serving platforms, or agent systems.
- Contributions to open-source software or the AI ecosystem.
- Experience optimizing production systems for reliability, scalability, and cost efficiency.
Email recruiting@nousresearch.com with the following information:
- The role you are applying for in the subject
- Resume or CV
- Cover letter
- A portfolio showcasing your work