AI System Research and Development Engineer - Optimization at Snowflake — AI Jobs | Neura Market
    Neura Market
    Neura Market
    /Jobs
    Marketplace
    Directories
    Resources
    AI JobsAI System Research and Development Engineer - Optimization
    Snowflake

    AI System Research and Development Engineer - Optimization

    Snowflake

    US-WA-Bellevue

    Marketplace

    • Prompts
    • Workflows
    • Agent Hub
    • Workflow Packs
    • Categories
    • Marketplace

    Directories

    • AI Tools Directory
    • ChatGPT
    • Claude
    • Gemini
    • Cursor
    • Grok
    • DeepSeek
    • Perplexity
    • CoPilot
    • Midjourney
    • Stable Diffusion
    • MCP Servers
    • .md Directory
    • All Directories

    Free Tools

    • AI Text Humanizer
    • AI Content Detector
    • Workflow Generator
    • Model Comparison
    • AI Pricing Calculator
    • AI Benchmarks
    • ROI Calculator
    • All Free Tools

    Resources

    • AI News
    • Blog
    • AI Answers
    • Error Solutions
    • AI Tutorials
    • AI Agent Guides
    • AI Models
    • AI Research Papers
    • Integrations
    • Alternatives
    • n8n vs Zapier
    • Make vs Zapier
    • n8n vs Make
    • Resource Library
    • Documentation
    • API Access to Our Data

    Community

    • AI Newsletter
    • AI Jobs
    • AI Events
    • AI Companies
    • Start Selling
    • Sell n8n Workflows
    • Sell AI Agents
    • Sell Prompts
    • Creator Guide
    • Advertise
    • Affiliates

    Company

    • About
    • Contact
    • Help
    • Careers
    • Pricing
    • Terms
    • Privacy
    • License
    • DMCA

    The #1 Newsletter in AI

    Weekly updates, news, and content that matter.

    Neura Market Logoneuramarket

    © 2026 Neura Market. All rights reserved.

    Full-time
    On-site
    6/16/2026
    Apply

    About This Role

    At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done.

    We are looking for talented System Developers and Researchers to join the Snowflake AI Research team and contribute to LLM inference and training system development, optimizations, and agentic systems. Our mission is to build the most efficient and scalable generative AI systems.

    Recent releases from our team include SwiftKV, an advanced inference optimization, and Arctic LLM, one of the largest open-source MoE foundation models. This is an exciting opportunity to collaborate with a world-class team, including founding members of DeepSpeed, vLLM, and TensorFlow. Together, we will push the boundaries of deep learning systems and drive cutting-edge innovations in AI.

    Responsibilities:

    • Analyze and optimize GPU kernel performance for training and inference of LLMs.

    • Develop and implement strategies to enhance the efficiency and scalability of deep learning systems.

    • Profile and benchmark deep learning systems using tools and techniques to identify bottlenecks.

    • Design and implement optimizations to reduce latency and improve resource utilization for training and inference.

    • Stay updated with the latest advancements in GPU kernel optimization, deep learning, and LLM system development.

    • Contribute to the development of agentic frameworks and applications for LLM-driven workflows, enhancing automation, reasoning, and decision-making capabilities.

    • Open-source and publish innovations, optimizations, and engineering practices in technical blogs, top-tier conferences and journals.

    Requirements:

    • Bachelor’s degree in Computer Science, Electrical Engineering, or a related field. A Master’s degree or PhD is preferred.

    • 5 years of experience in GPU kernel optimization, deep learning system optimization, or high-performance computing (HPC).

    • Proficiency in deep learning frameworks such as PyTorch, TensorFlow, JAX.

    • Strong understanding of GPU architectures and experience with CUDA or similar frameworks.

    • Experience with frameworks like CUTLASS, Triton, cuDNN, etc.

    • Experience with profiling tools (e.g., nvprof, Nsight) and performance analysis methodologies.

    • Solid problem-solving skills and ability to debug complex performance issues.

    • Excellent communication skills and ability to work effectively in a cross-functional team environment.

    Join us in optimizing deep learning systems and pushing the boundaries of AI efficiency. Apply now to be part of our dynamic and pioneering team!

    Snowflake is growing fast, and we’re scaling our team to help enable and accelerate our growth. We are looking for people who share our values, challenge ordinary thinking, and push the pace of innovation while building a future for themselves and Snowflake.

    How do you want to make your impact?

    For jobs located in the United States, please visit the job posting on the Snowflake Careers Site for salary and benefits information: careers.snowflake.com

    Tasks

    • •Analyze and optimize GPU kernel performance for training and inference of LLMs.
    • •Develop and implement strategies to enhance the efficiency and scalability of deep learning systems.
    • •Profile and benchmark deep learning systems using tools and techniques to identify bottlenecks.
    • •Design and implement optimizations to reduce latency and improve resource utilization for training and inference.
    • •Stay updated with the latest advancements in GPU kernel optimization, deep learning, and LLM system development.
    • •Contribute to the development of agentic frameworks and applications for LLM-driven workflows, enhancing automation, reasoning, and decision-making capabilities.
    • •Open-source and publish innovations, optimizations, and engineering practices in technical blogs, top-tier conferences and journals.
    • •Bachelor’s degree in Computer Science, Electrical Engineering, or a related field. A Master’s degree or PhD is preferred.
    • •5 years of experience in GPU kernel optimization, deep learning system optimization, or high-performance computing (HPC).
    • •Proficiency in deep learning frameworks such as PyTorch, TensorFlow, JAX.
    • •Strong understanding of GPU architectures and experience with CUDA or similar frameworks.
    • •Experience with frameworks like CUTLASS, Triton, cuDNN, etc.
    • •Experience with profiling tools (e.g., nvprof, Nsight) and performance analysis methodologies.
    • •Solid problem-solving skills and ability to debug complex performance issues.
    • •Excellent communication skills and ability to work effectively in a cross-functional team environment.

    Skills & Tech Stack

    PyTorchTensorFlowJAXLLMsCUDATriton

    Education

    PhDComputer Science

    Roles

    Development EngineerEngineer

    Location

    Region

    North America

    Country

    United States

    City

    WA

    Topics

    Engineering

    Related AI Jobs

    Replit

    Software Engineering Intern (Summer 2027)

    Replit·Internship·Foster City, CA
    Engineering
    Snowflake

    Frontend AI Principal Engineer - Cortex Code

    Snowflake·Full-time·US-CA-Menlo Park
    Engineering
    Snowflake

    Software Engineer - AIM Virtualization

    Snowflake·Full-time·US-CA-Menlo Park
    Engineering
    Snowflake

    Senior Software Engineer, Cortex Quality

    Snowflake·Full-time·US-CA-Menlo Park
    Engineering
    Snowflake

    Senior Security Engineer, AI Incident Response

    Snowflake·Full-time·US-CA-Menlo Park
    Engineering
    Snowflake

    Senior Software Engineer - Cortex AI - FDE

    Snowflake·Full-time·US-CA-Menlo Park
    Engineering
    ← Back to all jobs