ML/AI Research Engineer — Agentic AI Lab (Founding Team)

Fabrion
San Francisco, CA

ML/AI Research Engineer — Agentic AI Lab (Founding Team)

Location: San Francisco Bay Area
Type: Full-Time
Compensation: Competitive salary + meaningful equity (founding tier)

Backed by 8VC, we're building a world‑class team to tackle one of the industry’s most critical infrastructure problems.

About the Role

We’re designing the future of enterprise AI infrastructure — grounded in agents, retrieval‑augmented generation (RAG), knowledge graphs, and multi‑tenant governance.

We’re looking for an ML/AI Research Engineer to join our AI Lab and lead the design, training, evaluation, and optimization of agent‑native AI models. You'll work at the intersection of LLMs, vector search, graph reasoning, and reinforcement learning — building the intelligence layer that sits on top of our enterprise data fabric.

This isn’t a prompt engineer role. It’s full‑cycle ML: from data curation and fine‑tuning to evaluation, interpretability, and deployment — with cost‑awareness, alignment, and agent coordination all in scope.

Core Responsibilities

  • Fine‑tune and evaluate open‑source LLMs (e.g. LLaMA 3, Mistral, Falcon, Mixtral) for enterprise use cases with both structured and unstructured data

  • Build and optimize RAG pipelines using LangChain, LangGraph, LlamaIndex, or Dust — integrated with our vector DBs and internal knowledge graph

  • Train agent architectures (ReAct, AutoGPT, BabyAGI, OpenAgents) using enterprise task data

  • Develop embedding‑based memory and retrieval chains with token‑efficient chunking strategies

  • Create reinforcement learning pipelines to optimize agent behaviors (e.g. RLHF, DPO, PPO)

  • Establish scalable evaluation harnesses for LLM and agent performance, including synthetic evals, trace capture, and explainability tools

  • Contribute to model observability, drift detection, error classification, and alignment

  • Optimize inference latency and GPU resource utilization across cloud and on‑prem environments

Desired Experience

Model Training

  • Deep experience fine‑tuning open‑source LLMs using HuggingFace Transformers, DeepSpeed, vLLM, FSDP, LoRA/QLoRA

  • Worked with both base and instruction‑tuned models; familiar with SFT, RLHF, DPO pipelines

  • Comfortable building and maintaining custom training datasets, filters, and eval splits

  • Understand trade‑offs in batch size, token window, optimizer, precision (FP16, bfloat16), and quantization

RAG + Knowledge Graphs

  • Experience building enterprise‑grade RAG pipelines integrated with real‑time or contextual data

  • Familiar with LangChain, LangGraph, LlamaIndex, and open‑source vector DBs (Weaviate, Qdrant, FAISS)

  • Experience grounding models with structured data (SQL, graph, metadata) + unstructured sources

  • Bonus: Worked with Neo4j, Puppygraph, RDF, OWL, or other semantic modeling systems

Agent Intelligence

  • Experience training or customizing agent frameworks with multi‑step reasoning and memory

  • Understand common agent loop patterns (e.g. Plan→Act→Reflect), memory recall, and tools

  • Familiar with self‑correction, multi‑agent communication, and agent ops logging

Optimization

  • Strong background in token cost optimization, chunking strategies, reranking (e.g. Cohere, Jina), compression, and retrieval latency tuning

  • Experience running models under quantized (int4/int8) or multi‑GPU settings with inference tuning (vLLM, TGI)

Preferred Tech Stack

  • LLM Training & Inference : HuggingFace Transformers, DeepSpeed, vLLM, FlashAttention, FSDP, LoRA

  • Agent Orchestration : LangChain, LangGraph, ReAct, OpenAgents, LlamaIndex

  • Vector DBs : Weaviate, Qdrant, FAISS, Pinecone, Chroma

  • Graph Knowledge Systems : Neo4j, Puppygraph, RDF, Gremlin, JSON-LD

  • Storage & Access : Iceberg, DuckDB, Postgres, Parquet, Delta Lake

  • Evaluation : OpenLLM Evals, Trulens, Ragas, LangSmith, Weight & Biases

  • Compute : Ray, Kubernetes, TGI, Sagemaker, LambdaLabs, Modal

  • Languages : Python (core), optionally Rust (for inference layers) or JS (for UX experimentation)

Soft Skills & Mindset

  • Startup DNA: resourceful, fast‑moving, and capable of working in ambiguity

  • Deep curiosity about agent‑based architectures and real‑world enterprise complexity

  • Comfortable owning model performance end‑to‑end: from dataset to deployment

  • Strong instincts around explainability, safety, and continuous improvement

  • Enjoy pair‑designing with product and UX to shape capabilities, not just APIs

Why This Role Matters

This role is foundational to our thesis: that agents + enterprise data + knowledge modeling can create intelligent infrastructure for real‑world, multi‑billion‑dollar workflows. Your work won’t be buried in research reports — it will be productionized and activated by hundreds of users and hundreds of thousands of decisions. If this is your dream role - we would love to hear from you.

#J-18808-Ljbffr
Posted 2026-08-15

Recommended Jobs

GEOSPATIAL ENGINEER

Redlands, CA

Job Detail Organization: GeoAcuity Title: GEOSPATIAL ENGINEER Location: On-site - St. Louis, MO | DC Metro | Springfield, VA | Redlands, CAPosted: 2026-07-10 Application Deadline: Position Desc…

View Details
Posted 2026-07-11

Dishwasher

Marriott
Truckee, CA

POSITION SUMMARY Operate and maintain cleaning equipment and tools, including the dish washing machine, hand wash stations pot-scrubbing station, and trash compactor. Wash and disinfect kitche…

View Details
Posted 2026-08-15

Track Operator

K-1 Speed Inc
Chula Vista, CA

Job Description Job Description Do you have the need for speed? Do you thrive in a fast paced, energetic work environment that focuses on serving our customers? If so, K1 Speed is the place for y…

View Details
Posted 2026-07-29

Sr. Credit Representative

Callaway Golf
Carlsbad, CA

ABOUT THE BRAND: Callaway Golf Company is a premium golf equipment, gear and apparel company with a portfolio of global brands, including Callaway Golf, Odyssey, TravisMathew, and OGIO. Through an u…

View Details
Posted 2026-08-15

Urology APP

CommonSpirit Health
Woodland, CA

Job Summary and Responsibilities As a Urology APP with Woodland Clinic Medical Group (Woodland Clinic), you'll join a physician-owned, physician-led group that treats a full spectrum of conditio…

View Details
Posted 2026-07-29

Director of Engineering & Operations- The Infinity

Action Property Management
San Francisco, CA

Job Description Job Description Who We Are With a legacy spanning four decades, Action Property Management has become the premier choice for homeowner’s association management. Founded in 1984…

View Details
Posted 2026-08-08

Physician, Post?Acute & Geriatric Medicine (Sacramento, CA)

Sutter Health
Sacramento, CA

Opportunity Information Sutter Medical Group is seeking a Post?Acute & Geriatric Medicine Physician to join our physician?led post?acute care team in Sacramento, CA. This role is ideal for physi…

View Details
Posted 2026-07-30

ABA Associate Clinical Supervisor (Masters level RBT)

Lasting Behavioral Solutions
Santa Ana, CA

At Lasting Behavioral Solutions (LBS), we're dedicated to creating a vibrant community of RBTs, BTs, and Clinical Supervisors who are passionate about making a real difference in the lives of indi…

View Details
Posted 2026-07-30

Sales Account Manager - Existing Installation

Schindler
San Francisco, CA

Location: San Francisco, CA, United States  Job ID: 89862  We Elevate... Quality of urban life Our elevators, escalators, and moving walks safely transport more than two billion of us up and d…

View Details
Posted 2026-08-09

WATER RESOURCE CONTROL ENGINEER

State Water Resources Control Board
Sacramento County, CA

Job Description and Duties Please note, the Water Boards do not participate in E-Verify. Currently this position requires a minimum of 4 days in the office located at 1001 I Street, Sacrament…

View Details
Posted 2026-08-03