Get Hired Faster With COMPANY_NAME!
Don't you ever think you landed here by any accident, You are here because you are searching for something bigger. You know what?
- A better Job
- A better Future
- A better Knowledge
- A better Paycheck
- A greater Path to walk on.
And COMPANY_NAME is here to give you exactly what you've been missing for so long. The reality is that most job seekers chase job postings, but successful job seekers attract job offers by chasing the accurate information. Therefore, that's the shift COMPANY_NAME is going to help you make. Here are the top 10 ideas to up-skill yourself, so lean in to begin:
1: COMPANY_NAME Smart Tools and Direct Employer Connections Help Speed Up Your Hiring Process
COMPANY_NAME is a career-changing advantage that most seekers never get access to. Imagine...
- Instead of applying for job after job and still not getting any callbacks, you suddenly bump into a tool that can do the heavy lifting for you.
- Instead of wondering, "What do employers actually want?", you are getting insights straight from the employer's desk.
- Instead of hoping your resume gets noticed, it’s kept on the table of decision-makers who are hiring right now.
That's the difference COMPANY_NAME makes. Our tools will let you reach employers directly, which automatically speeds up your hiring process.
2: With Better Matches, Real-time Job Alerts, and Direct Employer Responses, COMPANY_NAME Helps Many Candidates Secure Interviews and Job Offers Within 15 to 30 Days!
How does COMPANY_NAME make this possible?
On COMPANY_NAME, you get notified for roles aligned with your profile right from the start. When an employer posts a role that matches your qualifications and skills, you’ll know first. When you apply early, your chances of getting noticed and shortlisted increase by 20%.
COMPANY_NAME also offers direct employer responses—no more waiting for weeks. Here you engage with hiring managers who are actively looking for candidates.
When all these features combine in one place, you move from your first match to your first interview within days. And ultimately, from application to offer—all within 15 to 30 days!
3: The Type of Resume You Need to Get Priority Placement
With COMPANY_NAME, you don’t just need a resume—you need a strategy. A system that pushes your name to the right tables. We’ll show you exactly how the most successful candidates take initiative and get noticed.
4: Browse Full-Time, Part-Time, and Freelancing Roles With COMPANY_NAME
The job market isn’t one-size-fits-all—and your career shouldn’t be either. COMPANY_NAME gives you access to a wide range of opportunities including full-time, part-time, and freelancing roles all in one place.
5: COMPANY_NAME Helps You Grow Your Career
COMPANY_NAME provides insights, tools, and role-matching that help you find the right direction, the right skills, and the opportunities aligned with your ambition.
6: The Easiest Way To Find A Job
COMPANY_NAME cuts the noise, the endless scrolling, and the confusion. With accurate matches, direct employer connection, and real-time updates, you get a clear and simple path from application to interview.
7: Find Roles That Offer Growth, Culture & Benefits
COMPANY_NAME helps you find roles where you grow, feel supported, and thrive—not just survive. With us, you discover opportunities that elevate your professional life.
8: Get Support With Resume, Interviews & Career Planning
COMPANY_NAME provides expert guidance on resumes, interviews, and planning so employers instantly recognize your strengths and value.
9: Your Future Starts Today
COMPANY_NAME gives you everything you need—tools, guidance, and opportunities—to step forward confidently and begin a new chapter where your potential is seen and supported.
10: Get Hired Within 15 to 30 Days With COMPANY_NAME
COMPANY_NAME follows a smart, strategic, and proven approach that gets your profile noticed faster and moves you toward interviews and offers within 15 to 30 days.
AI Systems Research and Development Engineer - LLM Inference Systems & Optimization
We are looking for talented systems developers and researchers to join the Snowflake AI Research team and advance the state of the art in **LLM inference systems and optimization**. Our mission is to build the next generation of **high-performance and intelligent inference systems**. We optimize not only how fast and efficiently models run, but also how quickly inference systems can adapt to new models, architectures, hardware, and workloads. Our work spans the full inference stack-from distributed serving and runtime systems to GPU kernels and model-system co-design. We explore techniques such as **adaptive parallelism, speculative and parallel decoding, disaggregated inference, scheduling and batching, KV-cache optimization, model swapping, quantization, and GPU kernel optimization** to push the frontier of latency, throughput, scalability, and cost. Beyond optimizing individual models, we are building **intelligent and adaptive inference systems** that can automate performance optimization-rapidly profiling new models and workloads, identifying bottlenecks, selecting effective execution strategies, and adapting system configurations with minimal manual tuning. We embrace **AI-native engineering**, using AI not only as the workload we optimize, but also as a tool to accelerate system development, experimentation, debugging, optimization, and adaptation to new models. Our goal is to accelerate both the **speed of inference and the agility of inference development**. Recent innovations from Snowflake AI Research include **Arctic Inference**, our open-source inference system, and technologies such as **Shift Parallelism**, which dynamically adapts parallelism to workload characteristics; **SwiftKV**, which reduces redundant prefill computation; **Arctic Speculator and SuffixDecoding** for fast speculative decoding; **Jacobi Forcing** for causal parallel decoding; and **Semi-Persistence** for fast model swapping and dynamic multi-model serving. This is an exciting opportunity to collaborate with a world-class team, including founding members of DeepSpeed, vLLM, and TensorFlow. Together, we will push the boundaries of AI systems and bring cutting-edge research into production-scale AI. **Responsibilities** - Design and develop **high-performance LLM inference systems**, spanning distributed serving, runtime systems, GPU execution, and performance-critical kernels. - Develop novel techniques to improve **inference latency, generation speed, throughput, memory efficiency, scalability, and cost**. - Explore advanced inference techniques including **speculative and parallel decoding, prefill/decode disaggregation, adaptive parallelism, continuous batching and scheduling, KV-cache management, quantization, and communication optimization**. - Develop **adaptive and intelligent inference systems** that automatically optimize execution for new model architectures, hardware platforms, workload characteristics, and deployment environments. - Apply **AI-driven and AI-native approaches to systems engineering**, including automated profiling, bottleneck identification, configuration search, code generation, experimentation, runtime strategy selection, debugging, and performance tuning. - Independently identify high-impact performance and systems problems, formulate hypotheses, prototype solutions, and drive promising ideas from research through production. - Design distributed inference strategies across GPUs and nodes, including tensor, sequence, pipeline, data, and expert parallelism. - Develop efficient approaches for **multi-model serving, dynamic resource management, model loading and swapping, and workload-aware scheduling**. - Analyze and optimize GPU kernels and operators for attention, MoE, communication, and other performance-critical model components. - Explore **model-system co-design**, including model or post-training techniques that unlock substantially more efficient inference. - Profile and benchmark end-to-end workloads to identify bottlenecks across compute, memory, communication, networking, scheduling, and model execution. - Collaborate closely with model researchers, infrastructure teams, and product teams to deploy research innovations in production. - Open-source and publish innovations through technical blogs and top-tier systems and machine learning conferences. **Requirements** - Bachelor's degree in Computer Science, Electrical Engineering, or a related field. A Master's degree or PhD is preferred. - 5+ years of experience in one or more of the following areas: **LLM inference systems, distributed AI systems, GPU systems, or high-performance computing**. - Strong understanding of modern LLM inference architectures and the performance tradeoffs involved in serving large-scale models. - Hands-on experience with modern **LLM inference and serving frameworks**, such as **vLLM, SGLang, TensorRT-LLM**, or similar systems. - Experience designing, extending, or optimizing inference runtimes, including areas such as **scheduling, batching, KV-cache management, distributed execution, parallelism, speculative decoding, or disaggregated serving**. - Strong understanding of GPU architectures and experience with **CUDA, Triton**, or similar GPU programming environments. - Experience with performance-oriented libraries and frameworks such as **CUTLASS, cuBLAS, cuDNN**, or related technologies. - Experience profiling and diagnosing end-to-end system performance using **Nsight Systems, Nsight Compute**, or equivalent tools. - Demonstrated ability to operate as an **independent problem identifier and solver**-recognizing important problems with limited direction, defining the right technical questions, and driving solutions through ambiguity. - Strong ability to work across **model, runtime, distributed system, and hardware layers** and reason about end-to-end performance tradeoffs. - Experience using **AI-native engineering approaches** to accelerate software development, experimentation, debugging, optimization, or system adaptation is a strong plus. - Excellent communication skills and the ability to collaborate effectively across research, engineering, and product teams. Snowflake is growing fast, and we're scaling our team to help enable and accelerate our growth. We are looking for people who share our values, challenge ordinary thinking, and push the pace of innovation while building a future for themselves and Snowflake. How do you want to make your impact? For jobs located in the United States, please visit the job posting on the Snowflake Careers Site for salary and benefits information: careers.snowflake.com