Skip to main content
I

Applied AI Engineer & Researcher San Francisco, CA View role ->

Interfaze JigsawStack, Inc.
3 hours ago
Full-time
On-site
San Francisco, California, United States
At Interfaze , we're building a new model for deterministic developer tasks that expects high accuracy like OCR, web scraping, classification, STT, and more. We currently power thousands of systems in industries like healthcare, finance, government, SaaS and more which has processed billions of tokens every month.

About this role We're hiring an AI Engineer & Researcher who is looking to push the boundaries of our models by implementing and experimenting with new research and techniques. As part of the Interfaze lab, you will help improve the performance and capabilities of our models by managing our AI lifecycle, including fine-tuning, automating data collection and cleaning, benchmarking, and deployment.

As a founding member at Interfaze, your role will be dynamic, offering opportunities to spearhead and launch innovative research, papers, and products into production while being the go-to expert for all things AI!

Area

What we use

Languages Python, Cython, Typescript, C++ (nice to have), Rust (nice to have)

Frameworks Docker, PyTorch, Transformers, CUDA

Inference engines SGLang, vLLM, TensorRT-LLM, Ollama, Llama.cpp

Frontend NextJS/React

Backend NodeJS with Typescript

Functions & sandboxes Modal, Vercel

Database Postgres (Supabase)

Query SQL/GraphQL

Payment/Billing Stripe

What you will do

Research and understand different machine learning techniques and papers, then experiment with real-world implementation

Experience with deep learning frameworks (e.g., TensorRT-LLM, PyTorch), and AI development tools is a must

Working closely with the team to train, deploy and serve models at scale on popular cloud providers like AWS, GCP, and more

Write papers on your research and experiments contributing to the growing open-source AI world

Writing detailed benchmarks on comparisons and analysis (e.g. comparing performance for two embedding models or benchmarking whisper 3 with another model)

Quantization and fine-tuning of LLMs and other models for our developer use cases

Optimizing models and GPU infra for cost-to-scale ratio based on load

Writing Python code and Jupyter Notebooks

Who you are

A strong AI Researcher with at least 5 years of experience, and paper publications at venues like ICML, NeurIPS, ACL, NAACL, and IEEE

Worked with most of our GPU infra and have a good understanding of our tech stack

Comfortable with working across the stack - model deployment, benchmarking, fine-tuning, monitoring, and scaling the service

Up-to-date with the latest papers, techniques, and experiments

Using AI to 100x your productivity

Natural curiosity to experiment with new concepts and tools while working closely with our community of developers

You care about building great products and love to geek out about the latest tech in AI

Open to feedback and have a growth mindset

Enjoy working in a startup environment where you can take initiative and contribute to the zero-to-one phase of development

Previously ran a startup or worked as a founding team member in a startup

Great compensation package and equity/shares

Flexible and remote-friendly

Annually run off-sites

Learn and Grow - we provide mentorship and send you to events that help you build your network and skills

The entire process is fully remote and all communication will happen over email or via video chat.

#J-18808-Ljbffr