Refolk
Shared search

Find engineers who have merged PRs to vllm-project/vllm or ggml-org/llama.cpp in the last 12 months and currently work at a cloud provider or inference startup.

111 results across 111 people

1
Amine Remache
Amine Remache11y exp
SDE @ AWS | llama.cpp, vLLM, Quantization, Inference, MCP, Agentic AI | Co-host @ Wled Horma Podcast
Dublin, County Dublin, Ireland· 1y at co.Possibly open
2
Aydin Abiar
Aydin Abiar7y exp
AI Systems Engineer | Anyscale, Ray, vLLM | LLM inference & post-training at scale | Agent infrastructure, harnesses, evals
San Francisco Bay AreaRecently started
3
Chi Wang
Chi Wang19y exp
AI Infrastructure & Inference Leader | LLM Serving, vLLM & AI Accelerators | AWS Neuron | O’Reilly & Manning Author
Redmond, Washington, United StatesRecently started
4
Benjamin Chislett
Benjamin Chislett
@vllm-project Maintainer and Senior Systems Software Engineer @NVIDIA
Toronto, Canada
JavaScriptPythonC++TypeScript
5
YZ
Yuan Zhang14y exp
Technical Lead | Senior Software Engineer (Ex-Google) | LLM Inference Systems | ML Infrastructure | Distributed Systems
San Francisco Bay AreaRecently started
6
Sameer Wadkar
Sameer Wadkar25y exp
LLM Infrastructure · Training & Inference Systems · Distributed Compute · Cloud-Native AI Platforms
Germantown, Maryland, United StatesRecently started
7
Jie Gong
Jie Gong11y exp
Applied AI Inference Engineer | GenAI Model Evaluation, Accuracy Debugging & Performance Optimization
Denver Metropolitan AreaRecently started
8
Zachary Keener
Zachary Keener12y exp
Forward Deployed Engineer @ Baseten | Scaling Inference
San Francisco Bay AreaRecently started
9
Yi Su
Yi Su11y exp
AI/LLM Research @ Fireworks AI | Advisor
San Francisco Bay Area· 1y at co.
10
Tran Le
Tran Le10y exp
MTS @ Fireworks AI | Ex-Meta ML | LLM Post-training/Fine-tune Infrastructure
San Francisco, California, United StatesRecently started
11
Junming Chen
Junming Chen11y exp
LLM Inference@FireworksAI | Xoogler
San Francisco Bay Area· 1y at co.
12
Ishan G.
Ishan G.6y exp
Inference Support Engineer @ Together AI · Diagnosing & resolving production LLM inference issues at scale · Kubernetes · Python · Observability · RAG · MLOps
Bellevue, Washington, United StatesRecently started
13
Antoni Baum
Antoni Baum
Member of Technical Staff @anthropics | ex-@openai, @anyscale | Retired Commiter @ray-project @vllm-project @pycaret
Python
14
SangBin Cho
SangBin Cho
MTS @ xAI | Previously @anyscale | Senior committer @ray-project | Committer @vllm-project & @sglang
PythonJavaScript
15
Ricky Xu
Ricky Xu
hacking on @ray-project @vllm-project with @anyscale | ex @cmu-db
GoCRubyC++
16
Huamin Chen
Huamin Chen
Inventor | Inspirer | Influencer. Creator of vLLM Semantic Router, Kepler Project, MicroShift, etc | Kubernetes SIG-Storage Founding Member
United States
GoShellC
17
Raman Shinde
Raman Shinde
AI Inference & Performance Optimization | TVM | llama.cpp | vLLM
Bangalore
Jupyter NotebookHTMLPython
18
Fei Wu
Fei Wu10y exp
SDE @ AWS Annapurna Labs | Ph.D. Computer Science | ML Performance, LLM Systems & Inference Acceleration
San Francisco Bay Area· 1y at co.
19
Syed Mujtaba
Syed Mujtaba6y exp
Software Engineer @ AWS | MS CS @ UC San Diego | AI/ML HPC, LLM Infrastructure, Distributed systems, Kubernetes | C/C++, Go, Python, Rust | NCCL, NIXL, Collective communication
San Francisco Bay AreaRecently started
20
Swapnil Tiwari
Swapnil Tiwari8y exp
Applied AI & Inference Engineer | GenAI Solutions Architect | LLM Serving, Post-Training, GPU Systems | Building Production AI at Scale
Greater Delhi Area· 2y at co.Likely open
21
Indra Kumar V
Indra Kumar V4y exp
SDE-II LLM Post-training @ AWS Sagemaker AI
Santa Clara, California, United StatesRecently started
22
Andrei Kulik
Andrei Kulik19y exp
LLM / On-Device ML Inference / Principal SWE @ Google | Angel Investor
Zurich, Zurich, Switzerland· 9y at co.Likely open
23
Vadim Gimpelson
Vadim Gimpelson24y exp
Speedup models inference @ NVidia
Abu Dhabi, Abu Dhabi Emirate, United Arab Emirates· 1y at co.
24
Mengdi H.
Mengdi H.14y exp
Deep Learning Engineer | Multimodal AI, LLM Inference, MLOps, Federated Learning, Open Source
San Francisco Bay Area· 7y at co.Likely open
25
Zhipeng Wang
Zhipeng Wang17y exp
Senior Staff Research Scientist - TLM, GenAI | Distributed Training & Inference Optimization, Long-horizon Reasoning/RL/LLM post-training, Efficient AI and AI-powered Search | TSC @ DeepSpeed and Liger Kernel
San Francisco Bay AreaRecently started
Woosuk Kwon
Woosuk Kwon
@Inferact | @vllm-project
Berkeley, CA
PythonShell
Xunzhuo
Xunzhuo
Intelligent Routing @vllm-project | Open Source & AI @AMD
JavaCSS
Robert Shaw
Robert Shaw
vllm and llm-d
Boston
Jupyter NotebookPython
Simon Mo
Simon Mo
cofounder of @Inferact, lead maintainer of @vllm-project
Berkeley, CA
JustJupyter NotebookRustPython
Yikun Jiang
Yikun Jiang
vllm-ascend @vllm-project / Apache Spark Committer / PMC Member @apache / Volcano Reviewer @volcano-sh / openEuler / ex OpenStack Core
Xi'an, China
PythonJavaJavaScript
WeiQing Chen
WeiQing Chen
vLLM-Omni Committer, vLLM, verl, Ray, Kubernetes
China HangZhou
HTMLPython
Kunshang Ji
Kunshang Ji
building vLLM @intel, vLLM committer. Opinions are my own.
Shanghai
ShellC++ScalaC
Jared Wen
Jared Wen
Open Source @vllm-project & @Inferact
Beijing
PythonHTML
Misha Goin
Misha Goin
Open source inference optimization @vllm-project | @redhatofficial | @neuralmagic
Boston
PythonBrainfuckHTML
Jee Jee Li
Jee Jee Li
@Inferact | @vllm-project
Chengdu, China
Python
Yihua Cheng
Yihua Cheng
CTO @ TensorMesh; Core developer at @LMCache and @vllm-project
United States
ShellJupyter NotebookC++Python
youkaichao
youkaichao
Ph.D. from Tsinghua University. Core maintainer of @vllm-project . Co-Founder & Chief Scientist @Inferact .
Beijing, China
JavaJupyter NotebookDockerfile
Jiangyun Zhu
Jiangyun Zhu
[ML]SYS. MTS@Inferact , building vLLM and vLLM-Omni. Previously at @ISCAS-OSLab & NJU.
ShellCTeX
Chao-Ju Chen
Chao-Ju Chen
Founder @Infinirc Co-maintainer @vllm-project/vllm-metal · @llmmanorg/llmman
Taiwan
SwiftShellC
Chauncey
Chauncey
Karmada/OpenELB/HAMi/vLLM approver cilium/istio member
C++Go
Chukwuma Nwaugha
Chukwuma Nwaugha
Bridging abstract concepts. Contributor vLLM
San Francisco
TypeScriptJavaScript
Zhewen Li
Zhewen Li
@Inferact | @vllm-project
SF Bay Areea
Huazhang Hu
Huazhang Hu
AIGC & VLLM
Shanghai
Python
Rajath Bharadwaj
Rajath Bharadwaj
AI Engineer | Langgraph | SGLang | vLLM | IaC
Toronto, Canada
Python
Canlin Guo
Canlin Guo
vLLM-Omni Maintainer
Shenzhen, China
HTMLC++PythonC
Tyler Michael Smith
Tyler Michael Smith
vLLM Core Maintainer | MTS at Red Hat
TeXCPythonJust+1
Isotr0py
Isotr0py
@Inferact | @vllm-project | M.S. @ Sun Yat-Sen University
Guangdong, China
RustJupyter NotebookPython
Haco
Haco
vLLM & vLLM-Omni Contributor
JavaScriptTypeScriptPythonC++
Flora Feng
Flora Feng
Maintainer @vllm-project
Ekagra Ranjan
Ekagra Ranjan
Cohere.ai - Contributor to ML frameworks like vLLM, vLLM-omni, Pytorch Geometric, TorchVision, Huggingface, TorchSharp, Pytorch Lightning.
PythonJupyter Notebook
Alexander Matveev
Alexander Matveev
@vllm-project
Boston, MA, USA
ShellPython
Muhammad Faizan Khan
Muhammad Faizan Khan
AI Engineer | LLM Infra & Agentic Systems | RAG • MCP • LoRA • Distributed Inference | PyTorch • vLLM • Ray • CUDA
Karachi
Jupyter NotebookPython
Yueqian Lin
Yueqian Lin
PhD candidate at Duke | Core Maintainer at vLLM-Omni
Durham, NC
PythonTeX
Ti Zhou
Ti Zhou
Linux AI & Data Foundation TAC Member, KunLun XPU, vLLM
Shanghai, China/Sunnyvale, CA
GoCSS
Sy03
Sy03
People don't understand. |Love @eunomia-bpf & @vllm-project
Shanghai
C++PythonJavaScript
Swapnil Parekh
Swapnil Parekh
NLP : AI Safety Research : vLLM
San Francisco
Jupyter NotebookPython
Gao Han
Gao Han
@vLLM-Omni Maintainer
Hong Kong, China
PythonJupyter Notebook
Hyunmok Choi
Hyunmok Choi
CS M.Sc. Student in System Software Research @OSSS-KU | Contributor @vllm-project (vllm-metal)
Seoul, South Korea
KotlinPythonShell
Roy Wang
Roy Wang
@Inferact Software Engineer | vLLM Committer | LLM Inference
PythonJava
Thomas Parnell
Thomas Parnell
Principal Research Scientist @IBM | Committer @vllm-project
Zürich, Switzerland
Dockerfile
HDCharles
HDCharles
ML Engineer @vllm-project @neuralmagic | ex-Meta
PythonJavaScriptShell
Lik Xun Yuan (Lx)
Lik Xun Yuan (Lx)
AI Engineering @ BCG X · vllm-metal
Singapore
Jupyter NotebookVim ScriptPythonSwift+1
Moderator
Moderator
Senior Software Engineer Red Hat | GSoC Mentor | Maintainer vLLM-sr Project and Board Member @asyncapi
Remote
JavaHTML
vllmellm
vllmellm
Nils Matteson
Nils Matteson
Inferact vLLM Fellow MS CS @ Northeastern
San Jose, CA
PythonGo
Rishabh Saini
Rishabh Saini
MLE @RedHatOfficial working on llm-d and vLLM
Toronto, ON, Canada
Jupyter NotebookDartJavaScript
Ranran
Ranran
Interested in Desktop LLM inference (vllm-metal & DGX-Spark) and speculative decoding.
US
Python
Yuhan Liu
Yuhan Liu
@LMCache Team, @vllm-project/production-stack, CS PhD student at UChicago
Chicago
PythonC++
Kevin H. Luu
Kevin H. Luu
Working on CI, release, cloud infrastructure for @vllm-project @Inferact
JavaScriptPython
Abhijit Ramesh
Abhijit Ramesh
Building the webGPU backend for llama.cpp
Santa Cruz, California
Jupyter NotebookJava
Nikhil Jain
Nikhil Jain
AI Performance Engineer @ Modular | llama.cpp contributor
Santa Cruz, CA
JavaGoPython
Niklas Wenzel
Niklas Wenzel
Founder | Maintainer of Electron :electron: and llama.cpp :llama:
Berlin, Germany
JavaRubyCQML
tc-mb
tc-mb
Author@llama.cpp-omni Member@MiniCPM-V & MiniCPM-o
C++
Ma Mingfei
Ma Mingfei
PyTorch module maintainer on CPU; Software Architect from Intel DCAI; Performance optimization for PyTorch, SGlang, Llama.cpp on intel platform.
Python
Siddhesh Sonar
Siddhesh Sonar
On-Device AI Engineer | Edge Inference | llama.cpp, GGML, C++/JNI, Android | Building at RunAnywhere (YC W26) | Creator of ToolNeuron
Mumbai
KotlinC++Rust
AGmind
AGmind
LLMOps / AI Platform Engineer · self-hosted LLM/RAG · local inference · Docker · vLLM · llama.cpp
ShellPython
Nexesenex
Nexesenex
Maintainer of Croco.CPP / Kobold.CPP FrankenFork. Interested in Llama.cpp, IK_Llama, and the quantization of LLMs. Advanced copy-paster. Enthusiast, not dev!
France
Lauri P. Laux Jr
Lauri P. Laux Jr
CTO by day, GM since 1994 🎲 | Python, Azure & Docker | Homelab: self-hosted AI, OpenWrt, Tailscale and now running its own LLMs via llama.cpp | Curitiba, BR
Curitiba, Paraná, Brasil
LuaShellObjective-C
Abhishek Kumar Gupta
Abhishek Kumar Gupta17y exp
Inference & AI Factory @ NVIDIA
Morgan Hill, California, United States· 2y at co.Likely open
Ziqi Fan
Ziqi Fan12y exp
LLM Inference (Dynamo) at NVIDIA
Sunnyvale, California, United States· 1y at co.
Dmitry T.
Dmitry T.19y exp
AI Inference == NVIDIA Dynamo
United States· 1y at co.Possibly open
Joel Fernandes
Joel Fernandes19y exp
Principal Engineer with deep expertise across AI/ML and systems engineering | GenAI, inference, memory, agents, OS, drivers, GPU/CPU, real-time, embedded & firmware
Washington DC-Baltimore AreaRecently started
MG
Michael Gschwind30y exp
Sr DE (Sr. Tech Director) at NVIDIA, AI@Scale, AI Acceleration, GPU Inference, Server and Mobile LLMs, Accelerators, GPUs | IEEE Fellow | Ex VP/Dir/DE/Principal Meta AI, Huawei, IBM | Ex-Faculty Princeton, TU Wien
Santa Clara, California, United States· 1y at co.
Ajay Mallya
Ajay Mallya7y exp
Data Processing, Inference @ NVIDIA | Applied Math, Violin @ Columbia-Juilliard
Cupertino, California, United StatesRecently started
Priyam Vora
Priyam Vora10y exp
SDE III (L6) at AWS | Uber Data Conference - 2022 speaker | Ex - Uber, Hike, Innovaccer | DA-IICT
Toronto, Ontario, CanadaRecently started
Sravan Bodapati
Sravan Bodapati16y exp
Sr. Staff Agentic AI SWE @ Google || AI Startups Mentor || ex- Principal Scientist at AWS Nova Foundation Models : LLM Post Training & Inference
San Francisco Bay Area· 1y at co.Recently started
Sajeel Khan
Sajeel Khan4y exp
Software Engineer at Google | IIIT Delhi | previously - MarkovML, Nference
Bengaluru, Karnataka, India· 2y at co.Likely open
Jitendra Jalwaniya
Jitendra Jalwaniya4y exp
TPU inference @Google | IIT Ropar
Bengaluru, Karnataka, India· 2y at co.Likely open
Wei Wang
Wei Wang23y exp
Machine Learning (LLM training, inference, infra).
San Mateo, California, United States· 5y at co.Likely open
Shreyas Chandrakaladharan
Shreyas Chandrakaladharan10y exp
Staff SWE | Gemini Post Training, Inference and Compute | DeepMind
Mountain View, California, United StatesRecently started
Shanquan Tian
Shanquan Tian9y exp
Software Engineer @ Google, efficient LLM training and inference
United States· 3y at co.Likely open
Lihao Ran
Lihao Ran7y exp
LLM Inference | SWE@Google | MSIN@CMU
United States· 4y at co.Likely open
Nadiv Gold
Nadiv Gold9y exp
I'm looking to write code that makes a difference.
United States· 4y at co.Likely open
Luying Liu
Luying Liu7y exp
Senior SWE @ Google | Founding Modeling Engineer, Gmail Contextual Smart Reply (Google I/O ’24 & ’25) | LoRA Fine-tuning, RAG & Inference at Scale
San Francisco Bay Area· 5y at co.Likely open
Shawn Zhang
Shawn Zhang12y exp
Sr. Specialist SA, AI & Engineering @ AWS | Engineering GenAI for production (millions-scale) | OSS contributor | Kubernetes × LLM
Hong Kong, Hong Kong SAR· 4y at co.Likely open
Hussain Raza Sayyed
Hussain Raza Sayyed7y exp
Founder @Tirmisol | Fixycome | Data Analyst @Stararc | Community Leader @AWS | CA @Devsinc | Aspire Alumnus@ Harvard | Data Engineer | Data Science | ML | LLM | RAG | Community Builder | Artist | Social Worker
Lahore, Punjab, PakistanRecently started
LaShondra Steger
LaShondra Steger21y exp
Serial Entrepreneur at SolMir Life Solutions, LLC
Spartanburg, South Carolina, United States· 4y at co.Likely open
SO
Sinan Ozdemir17y exp
Head of AI Developer Education @ Fireworks AI · Author of 10+ AI/LLM Books
Recently started
Bashir Mohammed
Bashir Mohammed16y exp
Senior Gen AI Architect @AWS Frontier AI Team | LLM | VLM | Startups | Agentic AI | High-Speed Network| Scientist|Quantum Networks |Previously @BerkeleyLab @Intel | All opinions shared are solely #mine
San Francisco Bay Area· 1y at co.Recently started
AS
Aviral Soni17y exp
Sr. PMT-ES @ AWS | MBA @ UNC | Product & Strategy | E-Commerce | Risk & Compliance | Healthcare | Data Engineering | Analytics | Platform-as-a-Service | LLM | Generative AI
New York, New York, United States· 1y at co.Possibly open
Robert Bradley
Robert Bradley22y exp
Principal Solutions Architect @ AWS | Enterprise AI, Agentic Systems & MCP | LLM Evaluation & AI Governance in Regulated Industries | Life Sciences, Banking, Telco | re:Invent Speaker
Congleton, England, United KingdomRecently started
Barry L.
Barry L.19y exp
Seasoned Finance Leader @ AWS | Global Net-0 (Order-to-Cash) Owner | GenAI + LLM + GAAP + FP&A + Risk Controls
Fairfax, Virginia, United States· 2y at co.
Sai Nikhil Garlapati
Sai Nikhil Garlapati8y exp
Software Development Engineer @Amazon Web Services | Java | Spring boot | TypeScript | React | JavaScript | AWS | System design | Data structures | Algorithms | LLM | Microservices | web services | GenAI
Austin, Texas, United States· 4y at co.Likely open
Aruna A.
Aruna A.10y exp
AI LLM Agent Developer @ Microsoft | AI Solutions Architect LLM | GenAI | Multimodal | Real-Time Monitoring | Cloud + Edge AI RAG | Domain-Specific LLM Fine-Tuning
United States· 5y at co.Likely open
Brian Luke
Brian Luke36y exp
Welcome Global Universal Peacekeeper Ministry, LLC Members
United StatesRecently started
Farris Cunningham
Farris Cunningham12y exp
Practice Manager at AWS, Security Assurance Services, LLC (AWS)
Houston, Texas, United StatesRecently started
Arham Ikram
Arham Ikram9y exp
Amazon Dropshipping | Wholesale | Private Label | Walmart dropshipping | Shopify Dropshipping | ebay dropshipping | LLC creation | All Amazon services | virtual assistant|
Multan, Punjab, PakistanRecently started
Tushar Bhandarkar
Tushar Bhandarkar27y exp
AI Tool Builder & Automation Engineer | Building Intelligent Trading Systems & LLM-Powered Workflows | Cloud Architecture | Prompt Engineering | SwiftUI · Python · TypeScript
Aldie, Virginia, United States· 5y at co.Likely open
Madhu Samhitha V.
Madhu Samhitha V.9y exp
Agentic AI Specialist SA @ AWS | GTM SE @ Juniper(HPE) | LLM Research @ VMware, IGCAR | Data @ Barclays | Public Speaker
Jersey City, New Jersey, United States· 1y at co.
Sanjay C.
Sanjay C.24y exp
Senior Eng Leader | AI/ML Infrastructure, LLM Platforms | Conversational AI | Cloud Scale Distributed Systems | AWS, OCI
Greater Seattle Area· 5y at co.Likely open
Sharron White
Sharron White35y exp
Mechanical Technician 2 @ Catalyst Brands LLC | Ensuring Optimal Machine Performance
Atlanta Metropolitan AreaRecently started

Same prompt, your brief. Change a word and I will run it again live, against the open web, and rank what comes back.

Remix this search

500 free credits on sign-up. No card.

Try it on the search you came here for

Stop building boolean strings. Just describe the person.

Type one sentence. I plan the search, read GitHub, public LinkedIn and Crunchbase records, and the open web as it is right now, and hand back a ranked list with the reason next to every name.

  1. 01Describe them

    One plain sentence. Role, city, stack, stage, whatever matters to you.

  2. 02I read the web live

    GitHub, public LinkedIn and Crunchbase records, the open web. Not a database that went stale last quarter.

  3. 03You read the shortlist

    Ranked, with the reasoning under every name. Open a profile, ask a follow-up, narrow it down.

  • No boolean, no filters, no seat to buy. One box.
  • Read at search time, so a profile updated yesterday counts today.
  • Every step visible as it runs, every name with its reason.

500 free credits on sign-up. No card, no demo call. See real searches.