AI Hiring Index

Ollama · Engineering · Posted 2026-07-09

Software Engineer, Runtime

Ollama · Palo Alto

Apply on Ollama's site Watch Ollama for new roles

Job description from Ollama's careers page.

Ollama is the most popular way for developers to access open models. What started as an open-source, local-first runtime is now the largest developer network in the open-model ecosystem: 8.9 million monthly active developers and over 67,000+ community-built integrations. We're backed by Y Combinator, Benchmark, 8VC, and Theory Ventures.

Our team is small and talent dense. We're flat, low-ego, and fast-moving. We like people who are truth-seeking, passionate, design-driven, and who enjoy shipping code.

About the role

You'll work on the heart of Ollama — the local runtime that runs open models on developers' own machines. It loads models, manages memory, drives GPU acceleration across NVIDIA, AMD, Intel, Qualcomm, and Apple Silicon (including our MLX integration), and makes all of it feel instant. You'll work in Go and C/C++ and touch the model formats and inference engines underneath, shipping to macOS, Linux, and Windows across an enormous range of hardware.

What you'll do

Example projects

You may be a fit if

Apply on Ollama's site

Report this listing or ask us to remove it

A removal request from the employer takes effect at the next refresh, within 6 hours. Other reports go to the site owner.

New engineering roles at AI companies, every Monday. The week's openings in this function across 277 companies, plus the weekly index. Free.

More engineering roles at Ollama

See also: Software Engineer jobs · AI jobs in San Francisco Bay Area · C++ jobs · CUDA jobs.

This listing is reproduced from Ollama's public careers feed and links to the original. AI Hiring Index is not the employer and does not accept applications. All Ollama roles · AI salaries.