# Resume Source: Roger Fleig

## Profile

- Name: Roger Fleig
- Location: San Jose, CA
- Email: rogerfleig@gmail.com
- Phone: 425-516-9049
- LinkedIn: https://linkedin.com/in/rogerfleig
- Websites:
  - https://fleig.us

## Headline

Engineering Leader - Developer Productivity, Infrastructure, and AI Platform Engineering

## Summary

- Engineering leader who builds and scales engineering organizations across developer productivity, large-scale infrastructure, and ML platform engineering.
- Track record of hiring and developing senior technical talent, founding new functions from zero, and setting the systems, standards, and tooling that let large engineering orgs move fast.
- Experienced leading engineering teams of up to 100 and driving productivity and code health for organizations of 1,200+ behind mission-critical infrastructure and foundational ML systems.
- Writes widely on agentic engineering, developer productivity, and the narrow pipe problem in AI-accelerated software delivery.

## Skills

- Leadership: Developer Productivity, Developer Experience (DevEx), Engineering Excellence, Code Health, Technical-Debt Reduction, Org Building, Hiring and Talent Development, Vendor Management, Platform Governance
- AI and Agent Tooling: Claude, Codex, Windsurf, multi-agent orchestration, enterprise AI tool adoption and enablement
- AI Infrastructure: MCP / Model Context Protocol, RAG, vector databases
- CI/CD and Platform: CI/CD, GitHub, GitHub Actions, GitLab, Argo CD, GKE, Kubernetes, Docker, Tailscale, NATS JetStream, Cloudflare
- Infrastructure as Code: Terraform, OpenTofu, Ansible, secrets management, Doppler, Ansible Vault

## Projects

### codex-fleet

- URL: https://github.com/grubbyhacker/codex-fleet
- Description: Multi-agent orchestration layer that lets any orchestrator delegate coding tasks to ephemeral worker agents over MCP, making delegation, task state, model choice, git worktrees, and logs explicit and observable through an OpenTUI dashboard. Removes the human-as-relay bottleneck in multi-agent coding workflows.

### YouKnowMe

- URL: https://github.com/grubbyhacker/youknowme
- Description: Personal-memory MCP server exposing a curated, Git-versioned Markdown corpus to AI agents via authenticated RAG over a LanceDB vector store, secured via Cloudflare Tunnel and Cloudflare Access JWT assertion validation, with a Curator agent that turns uploads into reviewable pull requests.

## Experience

### Senior Director, Engineering Productivity, Crusoe AI | San Francisco, CA | 2024-12 - 2026-03

- Company: Crusoe AI
- Location: San Francisco, CA
- Start: 2024-12
- End: 2026-03
- Description: Founded and built the engineering productivity function from zero, personally hiring and growing the team to 10 engineers and establishing developer experience as a discipline within a rapidly scaling AI cloud company.
- Bullets:
  - Set the enterprise AI coding platform strategy, directing the rollout of Windsurf and Claude Code to 500+ users, roughly 70% of engineering, owning enablement, adoption, and vendor management across Anthropic, Windsurf, and GitLab.
  - Built the operating model for enterprise AI tool adoption, covering provisioning, usage reporting, support channels, and security review, establishing the governance framework the org now runs on.
  - Directed the modernization of CI/CD infrastructure, migrating off aging local hardware to an autoscaling, multi-architecture GKE runner platform with build and container caching, roughly halving build times and ending the hardware failures that stalled pipelines, while engineering grew from about 70 to over 200.
  - Absorbed peak build load and increased code churn from AI tooling, and enabled rapid adoption of ARM64 builds for new NVIDIA AI fleet platforms.
  - Established repository and platform governance standards across about 400 repositories, driving technical-debt reduction and compliance campaigns through Terraform-managed infrastructure and health tooling.

### Engineering Director, Google | Mountain View, CA | 2018-08 - 2024-11

- Company: Google
- Location: Mountain View, CA
- Start: 2018-08
- End: 2024-11
- Description: Led developer productivity and code-health organizations, typically 40 to 100 engineers, that set engineering-excellence standards and drove down technical debt across much larger engineering populations. Ran several organizations concurrently, including a Google-wide Code Health and Reliability mandate alongside domain-specific productivity leadership for Core ML / Cloud AI and Core Data Infrastructure.
- Bullets:
  - Core ML Infrastructure and Cloud AI: Led concurrent productivity organizations for Core ML and Vertex AI, managing a 60-engineer team supporting about 1,200 Core ML engineers and directing technical-debt reduction across TensorFlow, JAX, XLA, and evaluation systems. Directed CI/CD strategy for the generative AI development lifecycle in Google Cloud, supporting Gemini and other first-party models.
  - Core Data Infrastructure: Led the developer productivity and engineering-excellence organization serving about 1,200 engineers building the data foundations behind Search, Ads, YouTube, and Geo. Drove code-health initiatives across one of Google's most critical infrastructure organizations.
  - Google-wide Code Health and Reliability: Built and led the team responsible for developer tooling used broadly across Google engineering. The team built Sensenmann, which automatically deleted dead code at scale, submitting over 1,000 changelists a week and removing more than 5% of all C++ at Google, nearly half a billion lines in total, and AutoCommenter, which applied machine learning to code review for every Google engineer.
  - Code Coverage and Engineering Standards: Built Google's code coverage infrastructure, making code review 5% faster and cutting review cost 11%, and introduced Productive Coverage, a new coverage measure that reduced review cost a further 2%. Elevated configuration management as a first-class engineering concern.
  - Research and Patents: Sponsored applied research and patent work by teams under my leadership in partnership with academic collaborators, including publications at ICSE, ICSE-SEIP, ESEC/FSE, IEEE TSE, and AIware, and issued U.S. patents in code coverage, mutation testing, and ML-assisted developer tooling.

### Senior Engineering Manager, Google | Mountain View, CA | 2011-04 - 2018-08

- Company: Google
- Location: Mountain View, CA
- Start: 2011-04
- End: 2018-08
- Description: Progressed through developer-productivity leadership roles across Shopping, Ads, and Analytics. Led teams of up to 100 engineers across Mountain View, Pittsburgh, Irvine, and Zurich, resolving long-standing developer-velocity bottlenecks in Ads and Analytics.

### Principal Test Manager, Microsoft | Redmond, WA | 1998-01 - 2011-03

- Company: Microsoft
- Location: Redmond, WA
- Start: 1998-01
- End: 2011-03
- Description: Advanced through roles of increasing scope in software quality and test engineering, culminating as Principal Test Manager. Built and led teams across Exchange Server, SQL Server, and Microsoft Advertising. Worked deeply in query processing, XML systems, and ML-based ad relevance and selection across large-scale server, data, and advertising systems.

## Education

### Missouri University of Science and Technology

- Location: Rolla, MO
- Degree: Master of Science
- Field: Computer Science
- Start: 1995
- End: 1997

### Missouri University of Science and Technology

- Location: Rolla, MO
- Degree: Bachelor of Science
- Field: Geology and Geophysics
- Start: 1990
- End: 1995
