Member of Technical Staff, Applied Research at LlamaIndex — AI Jobs | Neura Market
    Neura Market
    Neura Market
    /Jobs
    Marketplace
    Directories
    Resources
    AI JobsMember of Technical Staff, Applied Research
    LlamaIndex

    Member of Technical Staff, Applied Research

    LlamaIndex

    San Francisco

    Marketplace

    • Prompts
    • Workflows
    • Agent Hub
    • Workflow Packs
    • Categories
    • Marketplace

    Directories

    • AI Tools Directory
    • ChatGPT
    • Claude
    • Gemini
    • Cursor
    • Grok
    • DeepSeek
    • Perplexity
    • CoPilot
    • Midjourney
    • Stable Diffusion
    • MCP Servers
    • .md Directory
    • All Directories

    Free Tools

    • AI Text Humanizer
    • AI Content Detector
    • Workflow Generator
    • Model Comparison
    • AI Pricing Calculator
    • AI Benchmarks
    • ROI Calculator
    • All Free Tools

    Resources

    • AI News
    • Blog
    • AI Answers
    • Error Solutions
    • AI Tutorials
    • AI Agent Guides
    • AI Models
    • AI Research Papers
    • Integrations
    • Alternatives
    • n8n vs Zapier
    • Make vs Zapier
    • n8n vs Make
    • Resource Library
    • Documentation
    • API Access to Our Data

    Community

    • AI Newsletter
    • AI Jobs
    • AI Events
    • AI Companies
    • Start Selling
    • Sell n8n Workflows
    • Sell AI Agents
    • Sell Prompts
    • Creator Guide
    • Advertise
    • Affiliates

    Company

    • About
    • Contact
    • Help
    • Careers
    • Pricing
    • Terms
    • Privacy
    • License
    • DMCA

    The #1 Newsletter in AI

    Weekly updates, news, and content that matter.

    Neura Market Logoneuramarket

    © 2026 Neura Market. All rights reserved.

    Senior-level / Expert
    Full-time
    Remote
    7/8/2026
    Apply

    About This Role

    Join us and help shape the future of AI by defining the narrative around document understanding.

    About the Role

    We are looking for an AI Research Engineer to join our document understanding team.

    This role is ideal for someone who sits between applied research and strong engineering. You will work on vision-language models, document processing, data curation, synthetic data generation, benchmarking, training, fine-tuning, and post-training. The goal is simple: make our document AI systems more accurate, faster, and more cost-effective in production.

    You should be excited by frontier AI work, but equally motivated by practical product impact. This is not a pure research role where ideas stay in papers. You will be expected to prototype quickly, evaluate rigorously, and help turn promising approaches into production systems used by customers.

    What You’ll Do

    • Develop and train vision-language models for document processing and document understanding.

    • Build data pipelines for data curation, synthetic data generation, labeling, and benchmark creation.

    • Evaluate base models and perform post-training or fine-tuning to hit specific performance targets.

    • Improve model accuracy, latency, and cost-effectiveness across real-world document workflows.

    • Design and maintain benchmarks to measure extraction quality, layout understanding, OCR performance, reasoning accuracy, and end-to-end system reliability.

    • Work with messy real-world documents, including PDFs, scanned documents, tables, charts, forms, and multi-page enterprise documents.

    • Collaborate with engineering to move successful research prototypes into production.

    • Work directly with customers when needed to translate product requirements into benchmarks, experiments, and model improvements.

    • Stay close to the latest research in vision-language models, document AI, post-training, synthetic data, and agentic systems.

    • Use modern AI coding workflows and tools to move quickly.

    What We’re Looking For

    • 3–7 years of experience in machine learning engineering, applied research, or research engineering.

    • Strong ML foundation, including hands-on experience benchmarking and training models.

    • Strong Python skills and comfort with modern ML tooling, especially PyTorch.

    • Experience with computer vision, vision-language models, NLP, document AI, OCR, extraction, or agentic AI systems.

    • Ability to build experiments, evaluate results, and iterate quickly toward measurable performance improvements.

    • Strong engineering judgment and ability to write clean, production-quality code.

    • Comfort working in a fast-paced startup environment with high ownership and limited structure.

    • Adaptable, scrappy, and self-directed — someone who can figure things out without waiting to be told.

    • Strong technical writing and communication skills.

    Nice to Have

    • Prior startup experience, especially at an early-stage or high-growth AI company.

    • Experience as a founder or early startup engineer.

    • Experience building or improving document processing systems.

    • Experience with synthetic data generation, post-training, fine-tuning, or benchmark design.

    • Familiarity with tools such as vLLM, Pydantic, uv, ruff, mypy, Claude Code, Cursor, or similar modern AI engineering workflows.

    • Experience with open-source AI infrastructure or developer tools.

    Who You’ll Work With

    You will work closely with the CTO and the document understanding team, partnering across research, engineering, product, and customer-facing teams.

    Why Join LlamaIndex

    • Work on a core AI infrastructure problem: making complex documents understandable and actionable for AI systems.

    • Build production systems at the frontier of vision-language models and document AI.

    • Join a fast-growing startup with strong open-source adoption and commercial traction.

    • Work directly with technical founders and a highly ambitious engineering team.

    • Have real ownership over model quality, product capability, and technical direction.

    Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.

    LlamaIndex does not accept unsolicited agency resumes. Please do not forward resumes to our jobs alias, employees, or any other organization location. LlamaIndex is not responsible for any fees related to unsolicited resumes.

    Tasks

    • •Develop and train vision-language models for document processing and document understanding.
    • •Build data pipelines for data curation, synthetic data generation, labeling, and benchmark creation.
    • •Evaluate base models and perform post-training or fine-tuning to hit specific performance targets.
    • •Improve model accuracy, latency, and cost-effectiveness across real-world document workflows.
    • •Design and maintain benchmarks to measure extraction quality, layout understanding, OCR performance, reasoning accuracy, and end-to-end system reliability.
    • •Work with messy real-world documents, including PDFs, scanned documents, tables, charts, forms, and multi-page enterprise documents.
    • •Collaborate with engineering to move successful research prototypes into production.
    • •Work directly with customers when needed to translate product requirements into benchmarks, experiments, and model improvements.
    • •Stay close to the latest research in vision-language models, document AI, post-training, synthetic data, and agentic systems.
    • •Use modern AI coding workflows and tools to move quickly.
    • •3–7 years of experience in machine learning engineering, applied research, or research engineering.
    • •Strong ML foundation, including hands-on experience benchmarking and training models.
    • •Strong Python skills and comfort with modern ML tooling, especially PyTorch.
    • •Experience with computer vision, vision-language models, NLP, document AI, OCR, extraction, or agentic AI systems.
    • •Ability to build experiments, evaluate results, and iterate quickly toward measurable performance improvements.

    Perks & Benefits

    Work on a core AI infrastructure problem: making complex documents understandable and actionable for AI systems.Build production systems at the frontier of vision-language models and document AI.Join a fast-growing startup with strong open-source adoption and commercial traction.Work directly with technical founders and a highly ambitious engineering team.Have real ownership over model quality, product capability, and technical direction.

    Skills & Tech Stack

    PythonPyTorchComputer VisionNLP

    Location

    Region

    North America

    Country

    United States

    State / Province

    California

    City

    San Francisco

    Topics

    Engineering

    Related AI Jobs

    Databricks

    Staff Backend Software Engineer- (AI Platform)

    Databricks·Full-time·San Francisco, California
    Engineering
    Snowflake

    Staff Security Engineer - Threat Detection

    Snowflake·Full-time·US, Remote
    Engineering
    Snowflake

    Software Engineer - Openflow

    Snowflake·Full-time·US-CA-Menlo Park
    Engineering
    Decagon

    IT Engineer

    Decagon·Full-time·New York City

    $96,000 - $132,000/yr

    Engineering
    Replit

    Product Engineer, New Products

    Replit·Full-time·Foster City, CA
    Engineering
    Sierra AI

    Security Technical Program Manager

    Sierra AI·Full-time·San Francisco, CA
    Engineering
    ← Back to all jobs