Z
Generative AI Engineer
Zof AI, Inc.
1 hour ago
Full-time
On-site
San Francisco, California, United States
Compensation
Competitive salary
Plus meaningful equity
San Francisco, CA
Position Generative AI Engineer
Full-time, On-site
Zof AI is hiring for this role in San Francisco, CA. This is a full-time opportunity for candidates who want to contribute directly to the development of ambitious AI products in a high-performance environment.
Deep, hands-on experience building production features on top of LLM APIs is required.
About this Role Zof AI is seeking a Generative AI Engineer to build product features on top of frontier model APIs. This role owns the model-facing layer of our products: retrieval and RAG pipelines, agent planning and orchestration, multi-agent frameworks, and the context engineering that makes AI systems reliable in production. If you have worked as a GenAI Engineer, LLM Engineer, RAG Engineer, or Algorithm Engineer, this is that discipline at Zof AI. The ideal candidate has shipped LLM-powered features to real users and treats quality, cost, and latency as engineering constraints, not afterthoughts.
Responsibilities
Design and build product features on top of frontier model APIs.
Build and tune retrieval and RAG pipelines end to end.
Design agent planning, tool use, and multi-agent orchestration.
Own prompt and context engineering as a disciplined, tested practice.
Wire evals into the development loop so quality is measured, not assumed.
Optimize the cost, quality, and latency of AI features in production.
Collaborate with product and engineering to ship reliably and fast.
Own model-layer systems from design through production.
Requirements
Experience shipping LLM-powered features to production.
Strong software engineering foundation.
Working knowledge of retrieval, RAG, and agent patterns.
Fluency with the modern model API and tooling ecosystem.
Judgment about quality, cost, and latency trade-offs.
Clear written and verbal communication.
Comfort operating in a fast-moving environment.
Evidence of building high-quality technical work.
Nice to have
Experience with multi-agent frameworks or agent orchestration systems.
Experience building or using eval harnesses.
Experience with TypeScript, Python, Node.js, Postgres, or similar technologies.
Experience with fine-tuning or model adaptation.
What we provide in San Francisco
MacBook Pro
Premium AI development tools
Cursor Ultra
Claude Code Ultra
OpenAI Codex Max or equivalent advanced AI tooling
Access to a high-performance AI product environment
Close collaboration with leadership, engineering, and customers
Opportunity to work in the San Francisco AI ecosystem
Wellness and productivity support where applicable
Competitive startup environment
High ownership
Direct product impact
Benefits may depend on role and final offer terms.
#J-18808-Ljbffr
Competitive salary
Plus meaningful equity
San Francisco, CA
Position Generative AI Engineer
Full-time, On-site
Zof AI is hiring for this role in San Francisco, CA. This is a full-time opportunity for candidates who want to contribute directly to the development of ambitious AI products in a high-performance environment.
Deep, hands-on experience building production features on top of LLM APIs is required.
About this Role Zof AI is seeking a Generative AI Engineer to build product features on top of frontier model APIs. This role owns the model-facing layer of our products: retrieval and RAG pipelines, agent planning and orchestration, multi-agent frameworks, and the context engineering that makes AI systems reliable in production. If you have worked as a GenAI Engineer, LLM Engineer, RAG Engineer, or Algorithm Engineer, this is that discipline at Zof AI. The ideal candidate has shipped LLM-powered features to real users and treats quality, cost, and latency as engineering constraints, not afterthoughts.
Responsibilities
Design and build product features on top of frontier model APIs.
Build and tune retrieval and RAG pipelines end to end.
Design agent planning, tool use, and multi-agent orchestration.
Own prompt and context engineering as a disciplined, tested practice.
Wire evals into the development loop so quality is measured, not assumed.
Optimize the cost, quality, and latency of AI features in production.
Collaborate with product and engineering to ship reliably and fast.
Own model-layer systems from design through production.
Requirements
Experience shipping LLM-powered features to production.
Strong software engineering foundation.
Working knowledge of retrieval, RAG, and agent patterns.
Fluency with the modern model API and tooling ecosystem.
Judgment about quality, cost, and latency trade-offs.
Clear written and verbal communication.
Comfort operating in a fast-moving environment.
Evidence of building high-quality technical work.
Nice to have
Experience with multi-agent frameworks or agent orchestration systems.
Experience building or using eval harnesses.
Experience with TypeScript, Python, Node.js, Postgres, or similar technologies.
Experience with fine-tuning or model adaptation.
What we provide in San Francisco
MacBook Pro
Premium AI development tools
Cursor Ultra
Claude Code Ultra
OpenAI Codex Max or equivalent advanced AI tooling
Access to a high-performance AI product environment
Close collaboration with leadership, engineering, and customers
Opportunity to work in the San Francisco AI ecosystem
Wellness and productivity support where applicable
Competitive startup environment
High ownership
Direct product impact
Benefits may depend on role and final offer terms.
#J-18808-Ljbffr